Jonathn Chang
jc3683 at cornell dot edu
Hey! I’m a 3rd year Ph.D. student in the Center for Applied Mathematics at Cornell University, advised by Prof. Lionel Levine and Prof. Ziv Goldfeld. Previously, I earned my B.S. in Applied Mathematics from UCLA.
My research is broadly focused on AI alignment: ensuring that AI systems represent human values as they become more capable and autonomous. More specifically, I develop behavioral evaluations to measure the values expressed by large language models and use mechanistic interpretability to understand and steer the internal causal mechanisms that drive their behavior. Please feel free to reach out if you’re interested in chatting or collaborating!
news
| Aug 31, 2026 | I presented PLOT at the IPAM Foundations of Interpretability Workshop! |
|---|---|
| Jul 06, 2026 | I attended ICML 2026 in Seoul, South Korea and presented at the workshops for Pluralistic Alignment and Mechanistic Interpretability! |
| Apr 23, 2026 | I attended ICLR 2026 in Rio de Janeiro, Brazil and gave an oral presentation! |
| Apr 01, 2026 | With my SPAR team, we released ValueArena, a leaderboard for value alignment. |
| Jan 26, 2026 | “EigenBench: A Comparative Behavioral Measure of Value Alignment” was accepted to ICLR 2026 for an Oral presentation! |
| May 21, 2025 | I was selected as a Siegel PiTech PhD Impact Fellow to work with NYC DEP for Summer 2025. |