Inference Providers
Active filters: trl
linkred/stock_prediction_v8
Updated • 18
• 12
linkred/stock_prediction_v3_mini
Updated • 19
• 12
linkred/stock_prediction_v5
Updated • 7
• 11
linkred/stock_prediction_v6
Updated • 15
• 11
linkred/stock_prediction_v2_mini
Updated • 12
• 11
trl-internal-testing/tiny-Qwen2ForCausalLM-2.5
Text Generation
• 2.44M • Updated • 8.54M
• 56
SpeculativeDecoding/doctorboom-qwen2.5-coder-7b-lora
santiviquez/reward_modeling_anthropic_hh
Text Classification
• 0.3B • Updated • 68
• 3
linkred/stock_prediction_v7
Updated • 11
• 11
Rajagopal/LlamaAIRecruit-Llama-Multimodal-Reasoning-f2
prithivMLmods/Qwen3-VL-8B-Abliterated-Caption-it
Image-Text-to-Text
• 9B • Updated • 176
• 37
neigezhu/qwen3.5-27b-jailbreak-v5-last16
Text Generation
• Updated • 18
• 15
0xA50C1A1/Ministral-3-8B-Nymphaea-RP
Image-Text-to-Text
• 9B • Updated • 1.94k
• 5
mradermacher/Ministral-3-8B-Nymphaea-RP-GGUF
8B • Updated • 826
• 3
ram-lexsi/agenttune-testrun-tree-of-thoughts
ML-Intern-lab/Qwen-Image-2.1-PE-T2I-Pocket-2B
Text Generation
• 2B • Updated • 1.16k
• 9
ML-Intern-lab/Qwen-Image-2.1-PE-T2I-Pocket-0.8B
Text Generation
• 0.8B • Updated • 3.62k
• 12
Reponx/Network-Cloud-Ops-Engineer-9B
Text Generation
• Updated • 46
• 4
pmrccs/qwen3-1.7b-tool-calling-v4
Text Generation
• 2B • Updated • 333
• 2
FineEnvs/LFM2.5-2.6B-multiharness-RL
Text Generation
• 3B • Updated • 275
• 2
ConicCat/Gemma4-Writer-26BA4B
Image-Text-to-Text
• 26B • Updated • 24
• 2
773bw-h/my_first_reward_modeling
Text Classification
• 0.3B • Updated • 20
• 2
juuxn/llama-3_8b_fine_tuning_alpaca
wxzhang/dpo-selective-redteaming
Text Generation
• 7B • Updated • 72
• 2
DavideZanutto/llama3-finetuning
Updated • 8
• 1
Starxx/LLaMa3-Fine-Tuning-ChineseLaw
SiMajid/value_reward_modeling
Text Classification
• 0.3B • Updated • 11
• 1
shailja/lora_codellm_34b_verilog_model
wacc2/qwen2.5-0.5B_lora_adapters_model