Running 118 Unlocking On-Policy Distillation for Any Model Family 📝 118 Explore on-policy distillation visualization for any model
Running Featured 92 Distilling 100B+ Models 40x Faster with TRL 📝 92 TRL distillation for 100B+ teachers, 40x faster
unsloth/Mistral-Small-3.2-24B-Instruct-2506-unsloth-bnb-4bit Image-Text-to-Text • 25B • Updated Jun 23, 2025 • 2.73k • 13
Running on CPU Upgrade 267 The Synthetic Data Playbook: Generating Trillions of the Finest Tokens 📝 267 Visualize synthetic‑data experiments as an interactive bookshelf
unsloth/Qwen3-VL-8B-Instruct-unsloth-bnb-4bit Image-Text-to-Text • 9B • Updated Oct 31, 2025 • 23.1k • 22
Running on CPU Upgrade Featured 3.25k The Smol Training Playbook 📚 3.25k The secrets to building world-class LLMs
unsloth/Qwen3-4B-Instruct-2507-unsloth-bnb-4bit Text Generation • 4B • Updated Aug 6, 2025 • 76.7k • 17
intfloat/multilingual-e5-large-instruct Feature Extraction • 0.6B • Updated Jul 10, 2025 • 2.11M • • 631
unsloth/Llama-3.2-3B-Instruct-unsloth-bnb-4bit Text Generation • 3B • Updated Jun 2, 2025 • 67.2k • 10