mechanistic-interpretability 3 What Does a Neural Network Want to See? Sep 5, 2026 What Probes Can Tell Us About Truth, Deception, and Risk Aug 29, 2026 Visualizing Attention in Language Models Aug 16, 2026