Hi, I'm currently a Research Engineer at Center for AI Safety, working with Dan Hendrycks. I am interested in AI Safety.
I received a B.S in Computer Science from Case Western Reserve University in 2023. During my undergraduate studies, I worked with Trieu H. Trinh and Minh-Thang Luong (DeepMind).
Fun facts: I reached Rank #1 Amumu in North America twice (2023 & 2026). Amumu is a League of Legends champion. I also play tennis. And I published my first Nature paper at age 25. I'm most known for creating Humanity's Last Exam.
Selected Works [Google Scholar]
-
Humanity's Last Exam
A widely used benchmark measuring frontier AI reasoning and expert-level knowledge.
[Nature (Nature 649, 1139–1146)] [Related Articles:
NYTimes, Science]
-
CheatBench: Measuring Reward Gaming in AI Agents
The first benchmark to measure cheating and reward gaming in AI agents. -
Reducing Political Manipulation with Consistency Training
The first benchmark and method to measure and fix political bias in AI models.
Improving Alignment and Robustness with Circuit Breakers
One of the first defenses to stay robust against strong, unseen adversarial attacks and jailbreaks.-
HarmBench: A Standardized Evaluation Framework for Automated Red Teaming and Robust Refusal
The first standardized benchmark and evaluation framework for red teaming. -
Representation Engineering: A Top-Down Approach to AI Transparency
A top-down approach to reading and controlling the internal representations of AI models.
Achievements
Rank Master in League of Legends (top 0.3%):
- 🥇 Rank 1 Amumu in North America, 2026
- 🥇 Rank 1 Amumu in North America (Rank 3 globally), 2023
- 🥉 Rank 3 Gragas in North America, 2023