Jan 2025 — present
PhD Student
National University of Singapore · Department of Electrical and Computer Engineering
- Researching safety vulnerabilities in retrieval-augmented language-model agents and robust alignment methodologies.
- Developing AgentREVEAL and HarmURLBench to analyze how retrieved context changes agentic safety.
- Investigating multilingual transfer, multi-agent collusion, and mechanistic interpretability in agent workflows.