Inference Providers
Active filters: RLHF
NousResearch/Nous-Hermes-2-Mixtral-8x7B-DPO
Text Generation
• 47B • Updated • 39.2k
• 458
NousResearch/Nous-Hermes-2-Mixtral-8x7B-DPO-GGUF
47B • Updated • 1.4k
• 76
NousResearch/Hermes-2-Pro-Mistral-7B-GGUF
7B • Updated • 42.8k
• 258
NousResearch/Hermes-2-Pro-Mistral-7B
Text Generation
• 7B • Updated • 2.74k
• • 506
prithivMLmods/VideoGuard-Qwen3.5-9B-Safety-RL-Uncensored
Video-Text-to-Text
• 9B • Updated • 23
• 1
prithivMLmods/VideoGuard-Qwen3.5-4B-Safety-RL-Uncensored
Video-Text-to-Text
• 5B • Updated • 59
• 1
prithivMLmods/VideoGuard-Qwen3.5-9B-Safety-RL-Uncensored-GGUF
Video-Text-to-Text
• 9B • Updated • 308
• 1
prithivMLmods/VideoGuard-Qwen3.5-4B-Safety-RL-Uncensored-GGUF
Video-Text-to-Text
• 4B • Updated • 74
• 1
OpenAssistant/reward-model-deberta-v3-base
Text Classification
• Updated • 2.24k
• • 13
OpenAssistant/reward-model-electra-large-discriminator
Text Classification
• Updated • 33
• 5
OpenAssistant/reward-model-deberta-v3-large
Text Classification
• Updated • 1.32k
• 25
OpenAssistant/reward-model-deberta-v3-large-v2
Text Classification
• Updated • 12.9k
• • 247
Text Ranking
• 0.4B • Updated • 17
• 4
nicholasKluge/RewardModelPT
Text Classification
• 0.1B • Updated • 26
nicholasKluge/RewardModel
Text Classification
• 0.1B • Updated • 81
• 2
fb700/chatglm-fitness-RLHF
Updated • 268
fb700/Bofan-chatglm-Best-lora
Updated • 9
• 11
kubernetes-bad/Ligma-L2-13b
Updated • 3
• 3
Text Generation
• 0.4B • Updated • 604
• 209
berkeley-nest/Starling-LM-7B-alpha
Text Generation
• 7B • Updated • 2.8k
• 560
berkeley-nest/Starling-RM-7B-alpha
Updated • 167
• 105
LoneStriker/Starling-LM-7B-alpha-3.0bpw-h6-exl2
Text Generation
• Updated • 9
LoneStriker/Starling-LM-7B-alpha-4.0bpw-h6-exl2
Text Generation
• Updated • 10
• 1
LoneStriker/Starling-LM-7B-alpha-5.0bpw-h6-exl2
Text Generation
• Updated • 10
• 2
LoneStriker/Starling-LM-7B-alpha-6.0bpw-h6-exl2
Text Generation
• Updated • 10
• 1
LoneStriker/Starling-LM-7B-alpha-8.0bpw-h8-exl2
Text Generation
• Updated • 8
• 2
TheBloke/Starling-LM-7B-alpha-GGUF
7B • Updated • 1.82k
• 94
TheBloke/Starling-LM-7B-alpha-AWQ
Text Generation
• 7B • Updated • 51
• 9
second-state/Starling-LM-7B-alpha-GGUF
Text Generation
• 7B • Updated • 119
• 3
TheBloke/Starling-LM-7B-alpha-GPTQ
Text Generation
• 7B • Updated • 53
• 10