itsqyh/Awesome-LMMs-Mechanistic-Interpretability

A curated collection of resources focused on the Mechanistic Interpretability (MI) of Large Multimodal Models (LMMs). This repository aggregates surveys, blog posts, and research papers that explore how LMMs represent, transform, and align multimodal information internally.

220

134 commits

updated Mar 4, 2026

See the code

README

🌟 Awesome LMMs Mechanistic Interpretability

Logo

Peering into the Black Box of LMMs.

Awesome License: MIT

A curated collection of resources focused on the Mechanistic Interpretability (MI) of Large Multimodal Models (LMMs).

📬 Have a new paper or collaboration idea? Reach out to me at: qyhhere@gmail.com.
🤝 Seeking Opportunities: I'm eager to discuss and collaborate, especially for industry internships in Large Multimodal Models!

🔔 News

  • [2025-10] Added new 2025 papers on mechanistic interpretability of multimodal and foundation models!!!
  • [2025-06] I created this repository to maintain a paper list on LMMs-Mechanistic-Interpretability. Contributions are welcome!!!

Table of Contents

📚 Surveys:

(Back to Table of Contents)

📚 Blog:

(Back to Table of Contents)

📚 Papers:

📜 Input-level Attribution

(Back to Table of Contents)


📜 Sparse Autoencoder

(Back to Table of Contents)


📜 Probing

(Back to Table of Contents)


📜 Beyond or Logit Lens

(Back to Table of Contents)


📜 Causal Tracing

(Back to Table of Contents)


📜 Steering

(Back to Table of Contents)


📜 Representation

(Back to Table of Contents)


📜 Neuron Analysis

(Back to Table of Contents)


📜 Attention

(Back to Table of Contents)

📚 Tool

(Back to Table of Contents)

📚 Contribution

The main maintainer is Yihao Quan (@Yihao Quan).

Yihao Quan

Future contributors are welcome, and feel free to send pull requests in hopes that Awesome-LMMs-Mechanistic-Interpretability can become a more mature repo in the Large Multimodal Models' community.

generative
generative-model
large-language-models
large-multimodal-models
large-vision-language-models
mechanistic-interpretability
paperlist
vision-foundation-model
vision-models

Significant stargazers

Patrick Barker

70 followers · starred Nov 2025

Aojie Zhou

31 followers · starred Jun 2025

itsqyh/Awesome-LMMs-Mechanistic-Interpretability

A curated collection of resources focused on the Mechanistic Interpretability (MI) of Large Multimodal Models (LMMs). This repository aggregates surveys, blog posts, and research papers that explore how LMMs represent, transform, and align multimodal information internally.

220

134 commits

updated Mar 4, 2026

See the code

README

🌟 Awesome LMMs Mechanistic Interpretability

Logo

Peering into the Black Box of LMMs.

Awesome License: MIT

A curated collection of resources focused on the Mechanistic Interpretability (MI) of Large Multimodal Models (LMMs).

📬 Have a new paper or collaboration idea? Reach out to me at: qyhhere@gmail.com.
🤝 Seeking Opportunities: I'm eager to discuss and collaborate, especially for industry internships in Large Multimodal Models!

🔔 News

  • [2025-10] Added new 2025 papers on mechanistic interpretability of multimodal and foundation models!!!
  • [2025-06] I created this repository to maintain a paper list on LMMs-Mechanistic-Interpretability. Contributions are welcome!!!

Table of Contents

📚 Surveys:

(Back to Table of Contents)

📚 Blog:

(Back to Table of Contents)

📚 Papers:

📜 Input-level Attribution

(Back to Table of Contents)


📜 Sparse Autoencoder

(Back to Table of Contents)


📜 Probing

(Back to Table of Contents)


📜 Beyond or Logit Lens

(Back to Table of Contents)


📜 Causal Tracing

(Back to Table of Contents)


📜 Steering

(Back to Table of Contents)


📜 Representation

(Back to Table of Contents)


📜 Neuron Analysis

(Back to Table of Contents)


📜 Attention

(Back to Table of Contents)

📚 Tool

(Back to Table of Contents)

📚 Contribution

The main maintainer is Yihao Quan (@Yihao Quan).

Yihao Quan

Future contributors are welcome, and feel free to send pull requests in hopes that Awesome-LMMs-Mechanistic-Interpretability can become a more mature repo in the Large Multimodal Models' community.

generative
generative-model
large-language-models
large-multimodal-models
large-vision-language-models
mechanistic-interpretability
paperlist
vision-foundation-model
vision-models

Significant stargazers

Patrick Barker

70 followers · starred Nov 2025

Aojie Zhou

31 followers · starred Jun 2025