sciport

Mechanistic Interpretability · EP 5 · Jul 30, 2026

What's Actually Hiding Inside AI's Brain?

This week in Mechanistic Interpretability

Papers covered

  • Why jailbreaks actually workarXiv:2607.23496
  • Bad fine-tuning recruits old personasarXiv:2607.21356
  • Animacy has a circuitarXiv:2607.20995
  • Code models agree on WHAT, not HOWarXiv:2607.21491
  • Catch failing reasoning earlyarXiv:2607.21433

Your field, every week

Get this for your own field.

Sciport turns the newest papers in your corner of the literature into a short weekly video like this one — auto-curated, delivered to your inbox. First episode’s on us.

Generate your first video →