A JAX Research Toolkit for Visualizing, Manipulating, and Understanding Gemma Models with Multi-modal Support based on Penzai.
-
Updated
Jan 13, 2026 - Jupyter Notebook
A JAX Research Toolkit for Visualizing, Manipulating, and Understanding Gemma Models with Multi-modal Support based on Penzai.
A mechanistic interpretability workstation that maps neuron activations and polysemantic features across LLaMA and GPT-2 weights using Sparse Autoencoders (SAEs) and TransformerLens. Patch-isolates the Indirect Object Identification circuit to verify linear Taylor approximations through an interactive dashboard. (326 characters)
A research tool for studying how deception emerges in multi-agent LLM systems and detecting it through activation analysis.
Chemistry interpretability pilots on Gemma-2-2B via Neuronpedia — SAE feature discovery, causal steering, and circuit tracing. Preliminary evidence for a mechanistic-interpretability predoctoral proposal (no local GPU).
To associate your repository with the gemma-scope topic, visit your repo's landing page and select "manage topics."