Mechanist: AI as a Scientific Instrument for Discovering the Mechanisms of Intelligence Paper • 2608.12036 • Published 4 days ago • 82
view article Article Beyond Surface Alignment: Belief as the Gateway to Deep Alignment in LLMs Alibaba-VELLDEPTH • 13 days ago • 4
How Controllable Are Large Language Models? A Unified Evaluation across Behavioral Granularities Paper • 2603.02578 • Published Mar 3 • 25