Karthik Viswanathan
I recently completed my PhD at the University of Amsterdam in the Institute of Physics on geometric and statistical physics-based methods for interpretability of large language models. Prior to this, I was working on analyzing the large-scale structure of the universe using topological data analysis.
My interdisciplinary journey began with competitive programming, progressed through developing machine learning models as a surveillance analyst at Goldman Sachs, and continued with my master’s in theoretical physics. This cross-domain experience has enabled me to transition smoothly across disciplines and contribute to both practical implementations and theoretical research.
news
| Nov 15, 2026 | I will be joining the Data Science & AI Lab (dlab) at EPFL in Lausanne as a postdoctoral researcher in November, where I will be working with Robert West on AI alignment and safe AI. |
|---|---|
| Sep 28, 2026 | I am starting a three-month pilot of Verified Mechanisms, a project on formally verified autoresearch for theoretical mechanistic interpretability. |
| Jun 18, 2026 | I have been selected for the MATS extension phase in London, where I am continuing my research from the program. My research investigates the representation level picture of alignment pretraining: comparing how models acquire a preference through post-training versus incorporating it directly during pretraining, and what mechanistic differences this leaves behind in their internal representations. |
| Jan 06, 2026 | I will be participating in MATS 9.0 (ML Alignment & Theory Scholars Program) in Berkeley. I will be working with Adam Shai and Paul Riechers (Simplex AI Safety) on interpreting Chain of Thought in Toy Models using Computational Mechanics. |
| Jul 17, 2025 | I’m attending ICML 2025, where I’ll be presenting some of my recent work on interpretability in large language models: |