Reward Models
updated
nvidia/Llama-3.3-Nemotron-70B-Reward-Multilingual
Text Generation
• 71B • Updated
• 17
• 10
nvidia/Llama-3.3-Nemotron-70B-Reward-Principle
Text Generation
• 71B • Updated
• 256
• 6
nvidia/Qwen-3-Nemotron-32B-Reward
Text Classification
• 32B • Updated
• 864
• 19
Skywork/Skywork-Reward-V2-Llama-3.1-8B
Text Classification
• 8B • Updated
• 26k
• 39
Text Classification
• 8B • Updated
• 85
• 9
allenai/Llama-3.1-70B-Instruct-RM-RB2
Text Classification
• Updated
• 15
• 1
allenai/Llama-3.1-8B-Instruct-RM-RB2
Text Classification
• Updated
• 222
• 1
RLHFlow/ArmoRM-Llama3-8B-v0.1
Text Classification
• Updated
• 17.7k
• 183
nvidia/Llama-3.3-Nemotron-70B-Select
Text Generation
• 71B • Updated
• 23
• 11
nvidia/Llama-3.3-Nemotron-70B-Edit
Text Generation
• 71B • Updated
• 14
• 3
nvidia/Llama-3.3-Nemotron-70B-Feedback
Text Generation
• 71B • Updated
• 14
• 8
allenai/Llama-3.1-Tulu-3-8B-RM
Text Classification
• 8B • Updated
• 79
• 19
Text Classification
• Updated
• 41.2k
• 82
NCSOFT/Llama-3-OffsetBias-RM-8B
Text Classification
• 8B • Updated
• 164
• 24
NCSOFT/Llama-3-OffsetBias-8B
Text Generation
• 8B • Updated
• 31
• 14
nvidia/Qwen2.5-CascadeRL-RM-72B
Text Generation
• 71B • Updated
• 417
• 11
general-preference/GPM-Llama-3.1-8B
8B • Updated
• 243
• 1