Jonathn Chang

jc3683 at cornell dot edu

prof_pic.jpg

Hey! I’m a 3rd year Ph.D. student in the Center for Applied Mathematics at Cornell University, advised by Prof. Lionel Levine and Prof. Ziv Goldfeld. Previously, I earned my B.S. in Applied Mathematics from UCLA.

My research is broadly focused on AI alignment: ensuring that AI systems represent human values as they become more capable and autonomous. More specifically, I develop behavioral evaluations to measure the values expressed by large language models and use mechanistic interpretability to understand and steer the internal causal mechanisms that drive their behavior. Please feel free to reach out if you’re interested in chatting or collaborating!

news

Aug 31, 2026 I presented PLOT at the IPAM Foundations of Interpretability Workshop!
Jul 06, 2026 I attended ICML 2026 in Seoul, South Korea and presented at the workshops for Pluralistic Alignment and Mechanistic Interpretability!
Apr 23, 2026 I attended ICLR 2026 in Rio de Janeiro, Brazil and gave an oral presentation!
Apr 01, 2026 With my SPAR team, we released ValueArena, a leaderboard for value alignment.
Jan 26, 2026 EigenBench: A Comparative Behavioral Measure of Value Alignment” was accepted to ICLR 2026 for an Oral presentation!
May 21, 2025 I was selected as a Siegel PiTech PhD Impact Fellow to work with NYC DEP for Summer 2025.

selected publications

  1. ICML
    PLOT: Progressive Localization via Optimal Transport in Neural Causal Abstraction
    Jonathn Chang, Arya Datla, and Ziv Goldfeld
    ICML Workshop for Mechanistic Interpretability, 2026.
  2. ICLR
    EigenBench: A Comparative Behavioral Measure of Value Alignment
    Jonathn Chang, Leonhard Piff, Suvadip Sana, Jasmine X. Li, and Lionel Levine
    ICLR, 2026. Oral (top 1.13%).