Karthik Viswanathan

PhD, University of Amsterdam

MATSPROGRAM_FINALS-25.jpg

I recently completed my PhD at the University of Amsterdam in the Institute of Physics on geometric and statistical physics-based methods for interpretability of large language models. Prior to this, I was working on analyzing the large-scale structure of the universe using topological data analysis.

My interdisciplinary journey began with competitive programming, progressed through developing machine learning models as a surveillance analyst at Goldman Sachs, and continued with my master’s in theoretical physics. This cross-domain experience has enabled me to transition smoothly across disciplines and contribute to both practical implementations and theoretical research.

news

Nov 15, 2026 I will be joining the Data Science & AI Lab (dlab) at EPFL in Lausanne as a postdoctoral researcher in November, where I will be working with Robert West on AI alignment and safe AI.
Sep 28, 2026 I am starting a three-month pilot of Verified Mechanisms, a project on formally verified autoresearch for theoretical mechanistic interpretability.
Jun 18, 2026 I have been selected for the MATS extension phase in London, where I am continuing my research from the program. My research investigates the representation level picture of alignment pretraining: comparing how models acquire a preference through post-training versus incorporating it directly during pretraining, and what mechanistic differences this leaves behind in their internal representations.
Jan 06, 2026 I will be participating in MATS 9.0 (ML Alignment & Theory Scholars Program) in Berkeley. I will be working with Adam Shai and Paul Riechers (Simplex AI Safety) on interpreting Chain of Thought in Toy Models using Computational Mechanics.
Jul 17, 2025 I’m attending ICML 2025, where I’ll be presenting some of my recent work on interpretability in large language models: