Drinking with Einstein

From Visualizations to Circuits: The Origins of the Glass Box (The Glass Box Pt. 2)

23 min · 9. Feb. 2026
Episode From Visualizations to Circuits: The Origins of the Glass Box (The Glass Box Pt. 2) Cover

Beschreibung

Is AI just a "giant math soup" that happens to work, or is it a machine with parts we can understand? In Part 2 of The Glass Box series, we dig into the origin story of Mechanistic Interpretability. We trace the 7-year arc from the first "camera" that let us see inside a neural network (Zeiler & Fergus, 2013) to the "microscope" of Feature Visualization (Olah et al., 2017) and the eventual "blueprint" of Circuits (2020). We break down the four foundational papers that gave us the instruction manual for auditing AI and the plot twist (Superposition) that reminded us that the brain of these neural networks is a compression machine. Papers covered: * Visualizing and Understanding Convolutional Networks (2013) * Feature Visualization (2017) * The Building Blocks of Interpretability (2018) * Zoom In: An Introduction to Circuits (2020) Grab a drink. It’s time to see how the tools were built.

Kommentare

0

Sei die erste Person, die kommentiert

Melde dich jetzt an und werde Teil der Drinking with Einstein-Community!

Loslegen

2 Monate für 1 €

Dann 4,99 € / Monat · Jederzeit kündbar.

  • Podcasts nur bei Podimo
  • 20 Stunden Hörbücher / Monat
  • Alle kostenlosen Podcasts

Alle Folgen

2 Folgen

Episode From Visualizations to Circuits: The Origins of the Glass Box (The Glass Box Pt. 2) Cover

From Visualizations to Circuits: The Origins of the Glass Box (The Glass Box Pt. 2)

Is AI just a "giant math soup" that happens to work, or is it a machine with parts we can understand? In Part 2 of The Glass Box series, we dig into the origin story of Mechanistic Interpretability. We trace the 7-year arc from the first "camera" that let us see inside a neural network (Zeiler & Fergus, 2013) to the "microscope" of Feature Visualization (Olah et al., 2017) and the eventual "blueprint" of Circuits (2020). We break down the four foundational papers that gave us the instruction manual for auditing AI and the plot twist (Superposition) that reminded us that the brain of these neural networks is a compression machine. Papers covered: * Visualizing and Understanding Convolutional Networks (2013) * Feature Visualization (2017) * The Building Blocks of Interpretability (2018) * Zoom In: An Introduction to Circuits (2020) Grab a drink. It’s time to see how the tools were built.

9. Feb. 202623 min