I’m interested in how multi-agent LLM systems exhibit failure modes that are invisible to single-agent evaluation — emergent collusion, cross-agent goal drift, deceptive coordination — and in building auditing infrastructure that can catch and steer them before deployment. The work runs along three threads: studying failure itself, especially the misalignment that emerges under conditions where standard evaluation looks clean (Betley et al., 2025); auditing mechanistically, on what these systems are actually doing inside when they coordinate, deceive, or drift; and agents and society, on how multi-agent systems get embedded in human institutions and the technical analysis that makes that coexistence work (Reuel et al., 2024).
Currently a Master’s student in CS at Dartmouth College, co-advised by Nikhil Singh (Science and Art of Human-AI Systems Lab) and Soroush Vosoughi (Minds, Machines, and Society Lab). Before this, I did my undergrad at New York University, and undergrad thesis with Multimodal Agentic Personalization Systems (MAPS) Group.
News
- May 16, 2026 Paper accepted: PolicyLLM at the ICML Technical AI Governance Research 2026.
- May 15, 2026 Awarded the Thomas D. Sayles Research Grant ($2,000) from the Dartmouth Ethics Institute
- Apr 20, 2026 Reviewer ICML 2026 workshops: AIWILD, AI4GOOD, PhilML, and TAIGR
- Mar 15, 2026 Received Recognition of Excellence Citation from Dartmouth College (awarded for two courses).
- Nov 10, 2025 Best Poster Award at Technigala, Dartmouth College.
- May 20, 2025 Selected as AI/ML Spark Fellow at Gobi Partners (12 of 2,000+ applicants).
- Nov 15, 2024 Won Best Research Poster at the NYU CS Undergraduate Research Symposium.