AI Interpretability in the Era of LLMs: Architecture, Behavior and Beyond, Section 2 (Generative Interpretability)

Date: Spring 2026

Tutorial at University of Illinois Urbana-Champaign, Urbana, IL, USA

In this section, we turn to generative interpretability, exploring how large language models can be understood and steered through the lens of their generative behavior rather than solely their internal structure.

Slides

Recording