Discovering Cross-Language Reasoning Invariance in LLMs with Geometry-Invariant Sparse Autoencoders
Igor Bogdanov, C. Huang
ICML 2026 Workshop on Mechanistic Interpretability
Geometry-invariant sparse autoencoders and causal intervention probe whether reasoning representations are shared, and functionally interchangeable, across languages.
- Interpretability