Scaling Properties of Same-Family On-Policy Distillation Paper • 2609.32722 • Published 8 days ago • 317
Raven: The Harness of Harnesses for Composable Agentic Intelligence Paper • 2609.33439 • Published 7 days ago • 552
Heoni/llama-3-KoEn-8b_sft_ep4_merged_red_teaming_20240614 Text Generation • Updated Jun 16, 2024 • 30 • 3
Beyond Teacher Assignment: Domain-Normalized Multi-Teacher On-Policy Distillation Paper • 2609.35347 • Published 6 days ago • 178
Heoni/llama-3-KoEn-8b_sft_ep5_merged_red_teaming_20240614 Text Generation • Updated Jun 16, 2024 • 16 • 3
Disaggregated Quantization: Specializing LLM Prefill and Decode Paper • 2609.26333 • Published 12 days ago • 91
Parts-of-Speech as Emergent Categories in SAE Latent Space Paper • 2609.29362 • Published 10 days ago • 19
Your Transformer Can Hold Two Thoughts at Once: Evidence of Linear Superposition in LLMs Paper • 2609.29845 • Published 10 days ago • 102
Qwen-Planner-Agent: A Closed-Loop AI-for-AI Framework for Real-World Mobile Planner Agents Paper • 2609.29892 • Published 10 days ago • 32
Just Ask Jev: Reinforcement Learning for Calibrated Decisions as a Zero-Shot Detector of AI Alignment Failures Paper • 2609.29429 • Published 10 days ago • 28
Heoni/llama-3-KoEn-8b_sft_ep3_merged_red_teaming_20240614 Text Generation • Updated Jun 16, 2024 • 35 • 4