Loading signal
Preparing the next layer of the atlas.
Interpretability tools for understanding neural network internals
Mechanistic interpretability tools for neural multi-modal transformers.
Understanding how neural networks process information is critical for building trustworthy AI systems. This package provides hands-on tools for inspecting model activations, discovering circuits, and performing causal interventions.
neural-multi-modal-transformer-mechanistic-interpretability
View on GitHub