Decoy Direction Optimization: A Post-Hoc Defense Against LLM Abliteration Paper • 2609.16204 • Published 7 days ago • 4
open-llm-leaderboard-old/details_alexredna__Tukan-1.1B-Chat-reasoning-sft-COLA Updated Jan 22, 2024 • 80 • 3
latkes/self-consistency-correction-exp23-correlation-discrimination Viewer • Updated Apr 13 • 236 • 4 • 2
WearableQA: A Benchmark for Health Reasoning over Real-World Wearable Data Paper • 2609.05405 • Published 17 days ago • 44
Mind2Dialogue: Training Human-Aware Language Models by Simulating User Mental States Paper • 2609.15972 • Published 7 days ago • 8
Drift-Constrained Optimization: Only Direction Matters in Fine-Tuning Instruct Models Paper • 2609.13680 • Published 9 days ago • 11