Running 121 Unlocking On-Policy Distillation for Any Model Family 📝 121 Explore on-policy distillation visualization for any model
Running Featured 93 Distilling 100B+ Models 40x Faster with TRL 📝 93 TRL distillation for 100B+ teachers, 40x faster
unsloth/Mistral-Small-3.2-24B-Instruct-2506-unsloth-bnb-4bit Image-Text-to-Text • 25B • Updated Jun 23, 2025 • 2.01k • 13
Running on CPU Upgrade 277 The Synthetic Data Playbook: Generating Trillions of the Finest Tokens 📝 277 Visualize synthetic‑data experiments as an interactive bookshelf
unsloth/Qwen3-VL-8B-Instruct-unsloth-bnb-4bit Image-Text-to-Text • 9B • Updated Oct 31, 2025 • 27.8k • 22
Running on CPU Upgrade Featured 3.3k The Smol Training Playbook 📚 3.3k The secrets to building world-class LLMs
unsloth/Qwen3-4B-Instruct-2507-unsloth-bnb-4bit Text Generation • 4B • Updated Aug 6, 2025 • 101k • 17
intfloat/multilingual-e5-large-instruct Feature Extraction • 0.6B • Updated Jul 10, 2025 • 1.98M • • 635