Lena Prasad
lprasad21
ยท
AI & ML interests
AI alignment, jailbreak detection, red teaming, model robustness, safety evaluation
Recent Activity
upvoted a paper about 11 hours ago
World Action Agent: Harnessing VLMs for Robot Manipulation via World Action Rehearsal upvoted a paper about 11 hours ago
Just Ask Jev: Reinforcement Learning for Calibrated Decisions as a Zero-Shot Detector of AI Alignment Failures upvoted a paper about 11 hours ago
StudentSim: Training LLM-based Student SimulatorsOrganizations
None yet