Inference Providers
Active filters: reward
li-jay-cs/test2-rlhf-rm-checkpoint
li-jay-cs/gpt2-medium-rlhf-rm-checkpoint
Updated • 14
li-jay-cs/test3-rlhf-rm-checkpoint
li-jay-cs/gpt2-rlhf-rm-checkpoint
Updated • 10
li-jay-cs/gpt2-training-full-rlhf-rm-checkpoint
Updated • 10
li-jay-cs/gpt2-last_token_reward_and_full_training-rlhf-rm-checkpoint
Updated • 14
li-jay-cs/1gpu-gpt2-myepoch1-gcp-reward-model
Updated • 20
Text Classification
• 0.3B • Updated • 13
ZhangNy/2024-11-18_10-58-28
0.2B • Updated • 6
8B • Updated • 204
• 6
33B • Updated • 29
• 8
eth-nlped/Qwen2.5-1.5B-pedagogical-rewardmodel
Text Classification
• 2B • Updated • 1.62k
• 4
NiuTrans/GRAM-Qwen3-1.7B-RewardModel
2B • Updated • 24
• 6
NiuTrans/GRAM-Qwen3-14B-RewardModel
15B • Updated • 11
• 4
NiuTrans/GRAM-LLaMA3.2-3B-RewardModel
3B • Updated • 30
• 3
NiuTrans/GRAM-Qwen3-4B-RewardModel
4B • Updated • 34
• 2
NiuTrans/GRAM-Qwen3-8B-RewardModel
8B • Updated • 20
• 4
prithivMLmods/GRAM-LLaMA3.2-3B-RewardModel-GGUF
Text Ranking
• 3B • Updated • 266
prithivMLmods/GRAM-Qwen3-4B-RewardModel-GGUF
Text Ranking
• 4B • Updated • 278
mradermacher/GRAM-LLaMA3.2-3B-RewardModel-GGUF
3B • Updated • 353
mradermacher/GRAM-LLaMA3.2-3B-RewardModel-i1-GGUF
3B • Updated • 816
TIGER-Lab/EditReward-MiMo-VL-7B-SFT-2508
Image-to-Text
• 8B • Updated • 116
• 1
TIGER-Lab/EditReward-Qwen2.5-VL-7B
Image-Text-to-Text
• 8B • Updated • 287
• 5
Text Generation
• Updated • 25
• 4
Text Classification
• 8B • Updated • 4.27k
Text Classification
• 2B • Updated • 21
Text Classification
• 4B • Updated • 102
TIGER-Lab/RationalRewards-8B-T2I
Image-to-Text
• 9B • Updated • 70
• 6
TIGER-Lab/RationalRewards-8B-Edit
Image-to-Text
• 9B • Updated • 155
• 5
OpenRAL/rskill-robometer_4b-any-general-nf4
Robotics
• 3B • Updated • 53