ModaLens: Measuring Image Sensitivity in Report-Conditioned Medical VLMs Paper • 2609.15635 • Published 6 days ago • 5
Continual Learning Mechanisms Compose for Long-Horizon Memorization Paper • 2609.06986 • Published 13 days ago • 336
Dream-RSI: Recursive Self-Improvement through Evolving Worlds Paper • 2609.14858 • Published 6 days ago • 230
Benchmark Radar: A Living Database and Search Engine for AI Benchmarks and Evaluation Paper • 2609.11115 • Published 10 days ago • 167
ActReview: Rebuttal-Guided Training Data and Rubric Rewards for Actionable Peer Review Generation Paper • 2609.09076 • Published 12 days ago • 24
PlannerForge: LLM Agents for Scenario-Based Testing of Motion Planners in Autonomous Driving Paper • 2609.08965 • Published 12 days ago • 9
OpenWAM: An Open, Modular Exploration Towards Systematic World-Action Model Pretraining Paper • 2609.07398 • Published 13 days ago • 76