Syed Ali Haider — Research
about research projects cv
Themes
  • Agent alignment
  • Adversarial robustness
  • AI governance
  • NLP
Contact
  • email
  • github
  • x
PolicyLLM: Policy Extraction and Enforcement for Runtime AI Governance
S. A. Haider*, B. Huh*, H. H. Kim*, G. Nalagatla, C. G. Alvarez, Y. Raj, S. Vosoughi
ICML Technical AI Governance Research 2026 (poster) · technical-ai-governance
Compiles natural-language safety policies into symbolically validated decision graphs that act as a runtime governance layer over LLM outputs.
pdf code
Misalignment Contagion: How a Minority Shifts Aligned LLM Agents in Debate
S. A. Haider, S. Vosoughi
COLM 2026 (under review) · agent-alignment
A misaligned minority can systematically shift aligned LLM agents' private safety beliefs through multi-agent deliberation.
pdf code
Position: Belief Attributions Require Behavioral Evidence
S. A. Haider
ICML Philosophy meets ML 2026 (under review) · agent-alignment
What makes something a belief is not its form but its functional role in guiding action across perturbation.
pdf
Position: What is a Morally Aligned AI Agent? Philosophy as a Bridge to Operationalization
S. A. Haider
ICML Philosophy meets ML 2026 (under review) · agent-alignment
Alignment optimizes with precision for targets it has not clearly defined; philosophy is the discipline that can fix that.
pdf
Log-Probability Guided Adversarial Attacks on Multi-Agent LLM Debate
F. Niyigaba*, S. A. Haider*, N. Singh
Manuscript in preparation · adversarial-robustness
A novel log-probability attack surface succeeds on frontier models where prior methods fail; +26% attack success on GPT-3.5 over prior baselines.
pdf code
Harnessing Multilingual Geometry for Knowledge Transfer
F. Niyigaba*, S. A. Haider*, J. Lee, S. Vosoughi
Manuscript in preparation for ACL ARR August 2026 · nlp
Improves low-resource language representation in mBERT via shared linear mappings without target-language supervision.
pdf
MultiPetri: Auditing Tool for Multi-Agent LLM Systems
S. A. Haider
In progress · agent-alignment
An auditing framework for production multi-agent LLM systems — the case Petri leaves open.
ARES: Adaptive RL Agents for Evolving Cyber Threats
S. A. Haider, F. Leung
Position paper · adversarial-robustness
Hybrid GNN + MARL framework for real-time cyber defense, with an LLM-in-the-loop for autonomous patch synthesis under MARL guidance.
pdf
© 2026 Syed Ali Haider
email github x