view article Article Quantization-Aware Healing: a compressed, 4-bit model that outperforms its full-precision original MultiverseComputingCAI • 3 days ago • 23
Efficient Knowledge Distillation for LLMs: Offline Top-K Logits and a Fused Chunked KL Loss Paper • 2608.03796 • Published 24 days ago • 16 • 4
view article Article Making Knowledge Distillation Cheap Enough to Run at Scale MultiverseComputingCAI • 18 days ago • 39
Efficient Knowledge Distillation for LLMs: Offline Top-K Logits and a Fused Chunked KL Loss Paper • 2608.03796 • Published 24 days ago • 16
view article Article Making Knowledge Distillation Cheap Enough to Run at Scale MultiverseComputingCAI • 18 days ago • 39
Efficient Knowledge Distillation for LLMs: Offline Top-K Logits and a Fused Chunked KL Loss Paper • 2608.03796 • Published 24 days ago • 16
Refusal Steering: Fine-grained Control over LLM Refusal Behaviour for Sensitive Topics Paper • 2512.16602 • Published Dec 18, 2025
Instructing Large Language Models for Low-Resource Languages: A Systematic Study for Basque Paper • 2506.07597 • Published Jun 9, 2025
GuideX: Guided Synthetic Data Generation for Zero-Shot Information Extraction Paper • 2506.00649 • Published May 31, 2025 • 4
Cross-Lingual Transfer for Low-Resource Natural Language Processing Paper • 2502.02722 • Published Feb 4, 2025
Efficient Knowledge Distillation for LLMs: Offline Top-K Logits and a Fused Chunked KL Loss Paper • 2608.03796 • Published 24 days ago • 16