WTF GENIUS PAPERS Collection Papers that made me appreciate my major and my life a little more. obs=Observation, innov=Innovation. Most papers are abt improving tiny models. • 369 items • Updated about 18 hours ago • 82
Beyond Solver Verdicts: Generative Reward Models for Autoformalization Paper • 2609.11085 • Published 7 days ago • 34
Beyond Solver Verdicts: Generative Reward Models for Autoformalization Paper • 2609.11085 • Published 7 days ago • 34
StochBench: A Domain-Specific Benchmark for Stochastic Processes in Lean Paper • 2609.09264 • Published 9 days ago • 9
Test-Time Scaling in Reasoning LLMs: Inference Regimes, Evaluation, and Reproducibility Paper • 2608.04001 • Published Aug 4 • 1
Overcoming Dynamics-Blindness: Training-Free Pace-and-Path Correction for VLA Models Paper • 2605.11459 • Published May 14
Mid-Think: Training-Free Intermediate-Budget Reasoning via Token-Level Triggers Paper • 2601.07036 • Published Jan 11
Beyond Solver Verdicts: Generative Reward Models for Autoformalization Paper • 2609.11085 • Published 7 days ago • 34
StochBench: A Domain-Specific Benchmark for Stochastic Processes in Lean Paper • 2609.09264 • Published 9 days ago • 9
VERGE: Formal Refinement and Guidance Engine for Verifiable LLM Reasoning Paper • 2601.20055 • Published Jan 27 • 7 • 5
VERGE: Formal Refinement and Guidance Engine for Verifiable LLM Reasoning Paper • 2601.20055 • Published Jan 27 • 7
Forte : Finding Outliers with Representation Typicality Estimation Paper • 2410.01322 • Published Oct 2, 2024 • 2
VERGE: Formal Refinement and Guidance Engine for Verifiable LLM Reasoning Paper • 2601.20055 • Published Jan 27 • 7
VERGE: Formal Refinement and Guidance Engine for Verifiable LLM Reasoning Paper • 2601.20055 • Published Jan 27 • 7
Grammars of Formal Uncertainty: When to Trust LLMs in Automated Reasoning Tasks Paper • 2505.20047 • Published May 26, 2025 • 3
Grammars of Formal Uncertainty: When to Trust LLMs in Automated Reasoning Tasks Paper • 2505.20047 • Published May 26, 2025 • 3