Inference Providers
Active filters: Reward
Text Classification
• Updated • 48
• 10
Text Classification
• Updated • 42
• 2
Text Classification
• Updated • 74
• 26
Text Classification
• Updated • 27
• 6
Text Classification
• 2B • Updated • 24
• 2
mradermacher/SmolTulu-1.7b-RM-GGUF
2B • Updated • 237
mradermacher/SmolTulu-1.7b-RM-i1-GGUF
2B • Updated • 684
Teen-Different/squiral_maze
Reinforcement Learning
• Updated wangclnlp/GRAM-RR-LLaMA-3.1-8B-RewardModel
Text Generation
• 8B • Updated • 19
• 2
wangclnlp/GRAM-RR-LLaMA-3.2-3B-RewardModel
Text Generation
• 3B • Updated • 18
mradermacher/GRAM-RR-LLaMA-3.2-3B-RewardModel-GGUF
3B • Updated • 276
mradermacher/GRAM-RR-LLaMA-3.2-3B-RewardModel-i1-GGUF
3B • Updated • 1.03k
mradermacher/GRAM-RR-LLaMA-3.1-8B-RewardModel-GGUF
8B • Updated • 547
• 1
mradermacher/GRAM-RR-LLaMA-3.1-8B-RewardModel-i1-GGUF
8B • Updated • 879
• 1