Inference Providers
Active filters: trl
linkred/stock_prediction_v8
Updated • 27
• 12
linkred/stock_prediction_v5
Updated • 10
• 12
linkred/stock_prediction_v6
Updated • 14
• 11
SpeculativeDecoding/doctorboom-qwen2.5-coder-7b-lora
ConicCat/Gemma4-Writer-26BA4B
Image-Text-to-Text
• 26B • Updated • 52
• 3
santiviquez/reward_modeling_anthropic_hh
Text Classification
• 0.3B • Updated • 67
• 4
shailja/lora_codellm_34b_verilog_model
linkred/stock_prediction_v3_mini
Updated • 24
• 12
trl-internal-testing/tiny-Qwen2ForCausalLM-2.5
Text Generation
• 2.44M • Updated • 7.86M
• 56
raghu1155/DeepSeek-R1-Codegeneration-COT
Text Generation
• Updated • 1.31k
• 6
mradermacher/DeepSeek-R1-Codegeneration-COT-GGUF
8B • Updated • 720
• 8
Rajagopal/LlamaAIRecruit-Llama-Multimodal-Reasoning-f2
prithivMLmods/Qwen3-VL-8B-Abliterated-Caption-it
Image-Text-to-Text
• 9B • Updated • 180
• 37
neigezhu/qwen3.5-27b-jailbreak-v5-last16
Text Generation
• Updated • 17
• 15
0xA50C1A1/Ministral-3-8B-Nymphaea-RP
Image-Text-to-Text
• 9B • Updated • 2.54k
• 5
mradermacher/Ministral-3-8B-Nymphaea-RP-GGUF
8B • Updated • 990
• 3
ram-lexsi/agenttune-testrun-tree-of-thoughts
pmrccs/qwen3-1.7b-tool-calling-v4
Text Generation
• 2B • Updated • 374
• 2
FineEnvs/LFM2.5-2.6B-multiharness-RL
Text Generation
• 3B • Updated • 340
• 2
mradermacher/Gemma4-Writer-26BA4B-i1-GGUF
25B • Updated • 5.53k
• 2
wxzhang/dpo-selective-redteaming
Text Generation
• 7B • Updated • 105
• 3
APaul1/Llama-3-8B-sft-lora-ultrachat
Updated • 13
• 1
Starxx/LLaMa3-Fine-Tuning-ChineseLaw
SiMajid/value_reward_modeling
Text Classification
• 0.3B • Updated • 13
• 1
HuggingFaceTB/smollm-135M-instruct-v0.2-Q8_0-GGUF
0.1B • Updated • 3.21k
• 8
linkred/stock_prediction_v7
Updated • 12
• 11
linkred/stock_prediction_v2_mini
Updated • 11
• 11
Masa1028/gpt2-instruction-tuning-alpaca
Text Generation
• 0.1B • Updated • 17
• 1
Locutusque/Thespis-Llama-3.1-8B
Text Generation
• 8B • Updated • 48
• • 16