Inference Providers
Active filters: trl
mradermacher/DeepSeek-R1-Codegeneration-COT-GGUF
8B • Updated • 923
• 10
santiviquez/reward_modeling_anthropic_hh
Text Classification
• 0.3B • Updated • 56
• 6
blai88/reward_modeling_anthropic_hh
0.3B • Updated • 34
• 3
ram-lexsi/agenttune-testrun-tree-of-thoughts
SpeculativeDecoding/doctorboom-qwen2.5-coder-7b-lora
Updated • 10
ConicCat/Gemma4-Writer-26BA4B
Image-Text-to-Text
• 26B • Updated • 64
• 3
wxzhang/dpo-selective-redteaming
Text Generation
• 7B • Updated • 104
• 5
shailja/lora_codellm_34b_verilog_model
raghu1155/DeepSeek-R1-Codegeneration-COT
Text Generation
• Updated • 1.34k
• 8
mradermacher/gpt-4o-distil-Llama-3.1-8B-Instruct-PaperWitch-heresy-GGUF
8B • Updated • 838
• 9
neigezhu/qwen3.5-27b-jailbreak-v5-last16
Text Generation
• Updated • 17
• 15
Text Generation
• 8B • Updated • 24
• 2
sthanika-ai/gemma3-12b-kcc-advisory
Text Generation
• Updated • 18
• 3
ML-Intern-lab/Qwen-Image-2.1-PE-T2I-Pocket-0.8B
Text Generation
• 0.8B • Updated • 4.22k
• 14
mradermacher/Gemma4-Writer-26BA4B-i1-GGUF
25B • Updated • 5.79k
• 2
APaul1/Llama-3-8B-sft-lora-ultrachat
Updated • 12
• 1
SiMajid/value_reward_modeling
Text Classification
• 0.3B • Updated • 14
• 2
HuggingFaceTB/SmolLM-135M-Instruct
Text Generation
• 0.1B • Updated • 27k
• 146
HuggingFaceTB/smollm-135M-instruct-v0.2-Q8_0-GGUF
0.1B • Updated • 3.17k
• 8
mlabonne/TwinLlama-3.1-8B-DPO
Text Generation
• 8B • Updated • 129
• 25
omersaidd/Prompt-Enhace-T5-base
0.2B • Updated • 21
• 3
dhirajlochib/llama-3.2-unsensored-3b
Updated • 12
abhinavasr/Hindu-Veda-Llama-3.2-3B-Instruct
3B • Updated • 3
linkred/stock_prediction_v5
Updated • 13
• 13
linkred/stock_prediction_v6
Updated • 15
• 12
linkred/stock_prediction_v7
Updated • 14
• 12
linkred/stock_prediction_v8
Updated • 36
• 12
linkred/stock_prediction_v3_mini
Updated • 25
• 13
Masa1028/gpt2-instruction-tuning-alpaca
Text Generation
• 0.1B • Updated • 21
• 1
Locutusque/Thespis-Llama-3.1-8B
Text Generation
• 8B • Updated • 47
• • 16