SanSi: A Looped Typed Decision Model for System 1.5 Thinking Paper • 2610.07730 • Published 4 days ago • 13
Foundations of Proactive Agents: Principles, Technical Layers, and Proactivity-Gym Paper • 2609.37267 • Published 11 days ago • 40
Science or Slop?: Benchmarking and Mitigating Scientific Slop in AI-Generated Papers Paper • 2610.00531 • Published 10 days ago • 60
EvoDuet: Bilevel Co-Evolution of Web Searching and Task Solving for Scientific Discovery Paper • 2609.40340 • Published 10 days ago • 111
Evolution Fine-Tuning Collection Internalizing Discovery Capability into LLM • 10 items • Updated Jul 1 • 4
Evolution Fine-Tuning: Learning to Discover Across 371 Optimization Tasks Paper • 2606.29082 • Published Jun 27 • 44
When Thoughts Meet Facts: Reusable Reasoning for Long-Context LMs Paper • 2510.07499 • Published Oct 8, 2025 • 49
Learning Explainable Dense Reward Shapes via Bayesian Optimization Paper • 2504.16272 • Published Apr 22, 2025 • 5
LawFlow : Collecting and Simulating Lawyers' Thought Processes Paper • 2504.18942 • Published Apr 26, 2025 • 4
Toward Evaluative Thinking: Meta Policy Optimization with Evolving Reward Models Paper • 2504.20157 • Published Apr 28, 2025 • 36
Toward Evaluative Thinking: Meta Policy Optimization with Evolving Reward Models Paper • 2504.20157 • Published Apr 28, 2025 • 36
ScholaWrite: A Dataset of End-to-End Scholarly Writing Process Paper • 2502.02904 • Published Feb 5, 2025 • 2
A Dataset of Peer Reviews (PeerRead): Collection, Insights and NLP Applications Paper • 1804.09635 • Published Apr 25, 2018
Read, Revise, Repeat: A System Demonstration for Human-in-the-loop Iterative Text Revision Paper • 2204.03685 • Published Apr 7, 2022
CoEdIT: Text Editing by Task-Specific Instruction Tuning Paper • 2305.09857 • Published May 17, 2023 • 10