Applied AI safety · adversarial ML · robust systems
Sihui (Sophie) Dai
Applied Researcher working on AI safety, automated adversarial evaluation, and robust machine learning.
I am an applied researcher working on AI safety, automated adversarial evaluation, and robust machine learning.
My current work focuses on developing production safeguards for LLM applications and methods for systematically
stress-testing models and defenses. During my Ph.D. at Princeton, I studied robustness to unforeseen, adaptive,
and evolving adversaries. I am broadly interested in building AI systems that remain reliable as attacks, models,
and deployment environments change.