Trustworthy AI Agents
How do we make LLM-based agents safe and reliable enough to act autonomously?
Agents are being handed real responsibilities faster than anyone can characterize how they fail. A chatbot's mistake is a bad answer; an agent's mistake is an action, and it lands somewhere. My PhD work at Polytechnique Montréal builds evaluations and guardrails for agentic systems, starting from their observed failure modes.