MotionVLA: Vision-Language-Action Model for Humanoid Motion Paper • 2606.15142 • Published Jun 13 • 6
MassAlloc Attention: Let Attention Allocate Its Own Compute Paper • 2609.32712 • Published 5 days ago • 65
CoWindow Attention: Full Causal Coverage Is a Collective Property Paper • 2609.32704 • Published 5 days ago • 60
RRSI: Regularized Recursive Self-Improvement of Agent Harnesses Paper • 2609.24972 • Published 10 days ago • 219
Dream-RSI: Recursive Self-Improvement through Evolving Worlds Paper • 2609.14858 • Published 17 days ago • 250
INT8 LLMs for vLLM Collection Accurate INT8 quantized models by Neural Magic, ready for use with vLLM! • 47 items • Updated Mar 2 • 21
Nemotron Math & Reasoning Collection Datasets for building models that excel at math reasoning, proofs, and quantitative problem-solving. Covers SFT, RL, and pretraining data. • 23 items • Updated Aug 11 • 16
Nemotron Chat & Instruction Following Collection Datasets for building helpful, multi-turn, instruction-following conversational models across single and multi-turn settings. • 19 items • Updated Aug 11 • 10
Open-SWE-Collections Collection Open-SWE-Traces: Advancing Dual-Mode Multilingual Distillation for Software Engineering Agents • 4 items • Updated 19 days ago • 6
Nemotron Supervised Fine-Tuning Collection SFT datasets covering math, code, chat, safety, agentic, VLM, multilingual, and specialized domains. • 44 items • Updated Aug 11 • 22
Nemotron Code & SWE Collection Datasets for building models that write, debug, and reason about code. Covers competitive programming, software engineering, and code pretraining. • 14 items • Updated Aug 11 • 9
Mixture-of-Recursions: Learning Dynamic Recursive Depths for Adaptive Token-Level Computation Paper • 2507.10524 • Published Jul 14, 2025 • 76
VL-JEPA: Joint Embedding Predictive Architecture for Vision-language Paper • 2512.10942 • Published Dec 11, 2025 • 65
SMELT: Scaling Laws for Compute-Matched MoE Looped Transformers Paper • 2609.01343 • Published 30 days ago • 91
OpenThinker-Agent2 Collection OpenThinker-Agent2: agentic SFT/RL datasets and 8B/32B models (cold-start SFT, RL, and the OpenThinkerAgent-32B release). • 11 items • Updated Jun 11 • 11
VibeThinker-3B: Exploring the Frontier of Verifiable Reasoning in Small Language Models Paper • 2606.16140 • Published Jun 15 • 129
LoopCoder-v2: Only Loop Once for Efficient Test-Time Computation Scaling Paper • 2606.18023 • Published Jun 16 • 161