RealtimeWAM: One-Step Asynchronous World Action Models Paper • 2610.06617 • Published 6 days ago • 24
ProAR: Learning Prospective Reasoning with Autoregressive Video Models Paper • 2610.03664 • Published 9 days ago • 34
Mid-Harness: Scaling Actions Between Model and Harness for Terminal Agents Paper • 2609.39982 • Published 11 days ago • 120
Tactile-JEPA: Topology-Aware Self-Supervised Representation Learning for Distributed Tactile Sensors Paper • 2609.24385 • Published 20 days ago • 24
In-Context Learning for Robots: Methods and Applications Paper • 2609.36012 • Published 13 days ago • 322
VoxMem: Benchmarking Multimodal Memory in Large Audio Language Models Paper • 2609.32607 • Published 15 days ago • 134
WorldAttention: An Efficient Attention Architecture for Interactive Video World Models Paper • 2609.34606 • Published 13 days ago • 50
Anisotropic Representations Improve Planning in JEPA World Models Paper • 2609.37441 • Published 12 days ago • 35
WorldLine: Action-Driven Visual Simulation for Robotic Manipulation Paper • 2609.38059 • Published 12 days ago • 33
EVO-WAM: Evolving World Action Models through Video-Action Verification Paper • 2609.38057 • Published 12 days ago • 49
SAKI: Maximal-Coupling-Routed Teacher Supervision for On-Policy Distillation Paper • 2609.36601 • Published 12 days ago • 95
Uranus: Building the Next-Generation Simulation Infrastructure for Embodied AI Paper • 2609.24815 • Published 18 days ago • 12
Six Layers Less: Encoder Pruning for Whisper with Label-Free Recovery Paper • 2609.27980 • Published 18 days ago • 6