Introduction to Interpretability Now What
Let's dive into the details surrounding Interpretability Now What. Been Kim (Google Brain) https://simons.berkeley.edu/talks/tbd-72 Frontiers of Deep Learning.
Interpretability Now What Comprehensive Overview
Lex Fridman Podcast full episode: https://www.youtube.com/watch?v=ugvHCXCOmm4 Thank you for listening ❤ Check out our ... This is a talk I gave to my MATS 9.0 training scholars about the big picture of mech interp - as of Oct 2025, what had changed? Check out Gradient
EuroPython 2025 — South Hall 2B on 2025-07-17] *Hacking LLMs: An Introduction to Mechanistic
Summary & Highlights for Interpretability Now What
- A surprising fact about modern large language models is that nobody really knows how they work internally. At Anthropic, the ...
- Quantitative Testing with Concept Activation Vectors (TCAV) Been Kim, Senior Research Scientist, Google Brain Presented at ...
- What's happening inside an AI model as it thinks? Why are AI models sycophantic, and why do they hallucinate? Are AI models ...
- Interpretable
- Much of AI
That wraps up our extensive overview of Interpretability Now What.