AI Interpretability in the Era of LLMs: Architecture, Behavior and Beyond, Section 2 (Generative Interpretability)
Date: Spring 2026
Tutorial at University of Illinois Urbana-Champaign, Urbana, IL, USA
In this section, we turn to generative interpretability, exploring how large language models can be understood and steered through the lens of their generative behavior rather than solely their internal structure.
