arxiv:2607.15434
Jasmine Brazilek
sparrow8i8
AI & ML interests
None yet
Recent Activity
new activity about 4 hours ago
evaleval/EEE_datastore:[Submission] Add ANIMA and TAC (CaML animal-welfare benchmarks) authored a paper 13 days ago
Coercion and Deception in AI-to-AI Management: An Agentic Benchmark of Unprompted Escalation authored a paper 13 days ago
Do LLMs Hold Their Values? MANTA: A Multi-Turn Adversarial Benchmark for Animal Welfare Reasoning