Workshop on Actionable Interpretability@COLM 2026
  • General Information
  • 2025
  • Accepted Posters
  • Call for Papers

Actionable Interpretability

October 9 - COLM 2026 - San Francisco

The workshop on Actionable Interpretability@COLM2026 aims to foster discussions on leveraging interpretability insights to drive tangible advancements in AI across diverse domains. We welcome contributions that move beyond theoretical analysis, demonstrating concrete improvements in model alignment, robustness, and real-world applications. Additionally, we seek to explore the challenges inherent in translating interpretability research into actionable impact.

News

  • October 2: You can find the poster assignment here.
  • August 24: The poster size will be 34" x 34".
  • August 20: There will be no camera-ready version, see CfP page
  • July 24: Added a separate submission deadline for fast track submissions (August 9)
  • June 15: Submission Deadline extended to June 24 + clarified double submission policy in the CFP
  • May 22 2026: Call for Papers published
  • May 13 2026: Our workshop was accepted to COLM!

Important Dates

June 24 - Main Submission Deadline

August 9 - COLM Fast Track Submission Deadline (Submission Link)

October 9 - Workshop day

Dates are AOE.

Schedule

09:00Opening Remarks
09:10Keynote Tom McGrath, Goodfire - Ambitious Actionable Interpretability
09:50Keynote Dhanya Sridhar, Université de Montréal, Mila - Robust Interpretability with Causal Representation Learning
10:30Contributed Talks
10:50Poster Session 1
11:55Lunch Break
13:30Keynote Yonatan Belinkov, Technion - (Actionable) Interpretability beyond Language
14:10Contributed Talks
14:25Poster Session 2
15:30Coffee Break
16:00Keynote Christopher Potts, Stanford
16:40Panel
17:10Closing Remarks

Invited Speakers

Yonatan Belinkov

Associate Professor, Taub Faculty of Computer Science, Technion

Tom McGrath

Chief Scientist, Goodfire

Christopher Potts

Professor of Linguistics and Computer Science, Stanford University

Dhanya Sridhar

Assistant Professor, Université de Montréal, and Core Academic Member of Mila

Organizers

Tal Haklay

Member of technical staff, Goodfire

Hadas Orgad

Postdoc, Kempner Institute, Harvard University

Anja Reusch

Postdoc, Technion

Marius Mosbach

Postdoc, McGill University and Mila – Quebec AI Institute

Sarah Wiegreffe

Assistant Professor, University of Maryland

Ian Tenney

Staff Research Scientist, Google DeepMind

Mor Geva

Assistant Professor, Tel Aviv University

Asaf Avrahamy

Research Engineer, Meta FAIR and M.Sc. student, Tel Aviv University

Sponsors

© Workshop on Actionable Interpretability@COLM 2026 2026