arXiv:2407.14435 · 2 repos reference this paper in their README
saprmarks/dictionary_learning
428
·
No description
LLM-Interp/CLT-Forge
108
A Mechanistic Interpretability Toolkit for Cross-Layer Transcoder Training and Attribution-Graph…