Reward Models
updated
nvidia/Llama-3.3-Nemotron-70B-Reward-Multilingual
Text Generation
• 71B • Updated • 226
• • 10
nvidia/Llama-3.3-Nemotron-70B-Reward-Principle
Text Generation
• 71B • Updated • 246
• • 7
nvidia/Qwen-3-Nemotron-32B-Reward
Text Classification
• 32B • Updated • 237
• 20
Skywork/Skywork-Reward-V2-Llama-3.1-8B
Text Classification
• 8B • Updated • 9.21k
• 48
Text Classification
• 8B • Updated • 180
• 9
allenai/Llama-3.1-70B-Instruct-RM-RB2
Text Classification
• Updated • 161
• 1
allenai/Llama-3.1-8B-Instruct-RM-RB2
Text Classification
• Updated • 168
• 1
RLHFlow/ArmoRM-Llama3-8B-v0.1
Text Classification
• 8B • Updated • 28.7k
• 187
nvidia/Llama-3.3-Nemotron-70B-Select
Text Generation
• 71B • Updated • 135
• • 12
nvidia/Llama-3.3-Nemotron-70B-Edit
Text Generation
• 71B • Updated • 89
• 4
nvidia/Llama-3.3-Nemotron-70B-Feedback
Text Generation
• 71B • Updated • 93
• • 9
allenai/Llama-3.1-Tulu-3-8B-RM
Text Classification
• 8B • Updated • 614
• 19
Text Classification
• 73B • Updated • 52.8k
• 83
NCSOFT/Llama-3-OffsetBias-RM-8B
Text Classification
• 8B • Updated • 164
• 25
NCSOFT/Llama-3-OffsetBias-8B
Text Generation
• 8B • Updated • 79
• • 15
nvidia/Qwen2.5-CascadeRL-RM-72B
Text Generation
• 71B • Updated • 259
• 13
general-preference/GPM-Llama-3.1-8B
8B • Updated • 9
• 1
infly/INF-ORM-Llama3.1-70B
Text Classification
• 70B • Updated • 68
• 27