Code for "Evidence of Learned Look-Ahead in a Chess-Playing Neural Network"
β31Jun 4, 2024Updated 2 years ago
Alternatives and similar repositories for leela-interp
Users that are interested in leela-interp are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- π¬ Interpretability for Leela Chess Zero networks.β22Updated this week
- Official Code for What Makes and Breaks Safety Fine-tuning? A Mechanistic Study (NeurIPS 2024)β11Oct 31, 2024Updated last year
- β10Jun 27, 2024Updated 2 years ago
- β13Dec 4, 2024Updated last year
- A collection of different ways to implement accessing and modifying internal model activations for LLMsβ24Oct 18, 2024Updated last year
- Wordpress hosting with auto-scaling - Free Trial Offer β’ AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- A Mechanistic Interpretability Analysis of Grokkingβ29Sep 26, 2022Updated 3 years ago
- Command Line Interface for Competitive Programmingβ12May 11, 2026Updated 3 months ago
- β26Feb 20, 2026Updated 6 months ago
- Mitigating Partial Observability in Sequential Decision Processes via the Lambda Discrepancyβ24Oct 28, 2024Updated last year
- A tiny easily hackable implementation of a feature dashboard.β17Oct 21, 2025Updated 10 months ago
- Code and Data Repo for the CoNLL Paper -- Future Lens: Anticipating Subsequent Tokens from a Single Hidden Stateβ21Oct 24, 2025Updated 10 months ago
- Code for "What really matters in matrix-whitening optimizers?"β25Oct 31, 2025Updated 10 months ago
- Repository for ACM India Summer School on Generative AI for Textβ13Jul 11, 2024Updated 2 years ago
- PyTorch and NNsight implementation of AtP* (Kramar et al 2024, DeepMind)β21Jan 19, 2025Updated last year
- 1-Click AI Models by DigitalOcean Gradient β’ AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- β18Aug 27, 2026Updated last week
- Minimal but scalable implementation of large language models in JAXβ34Nov 28, 2025Updated 9 months ago
- graphpatch is a library for activation patching on PyTorch neural network models.β22Feb 11, 2025Updated last year
- Flax (JAX) implementation of Progressive Growing of GANs for Improved Quality, Stability, and Variationβ12May 24, 2021Updated 5 years ago
- β12May 8, 2024Updated 2 years ago
- β15Sep 29, 2022Updated 3 years ago
- MishformerLens intends to be a drop-in replacement for TransformerLens that AST patches HuggingFace Transformers rather than implementingβ¦β10Oct 7, 2024Updated last year
- Source codes of Learning Causal Representations for Robust Domain Adaptation (IEEE TKDE)β12Feb 14, 2022Updated 4 years ago
- A framework to meta-train transformers for causal ICLβ11Jul 15, 2026Updated last month
- AI Agents on DigitalOcean Gradient AI Platform β’ AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- A new format for plaintext to-do listsβ15Jan 20, 2026Updated 7 months ago
- β28Oct 6, 2024Updated last year
- Official implementation of "Reasoning by Superposition: A Theoretical Perspective on Chain of Continuous Thought" (NeurIPS 2025)β44Oct 8, 2025Updated 10 months ago
- Network representation learning on drug-target-side effects-indication graphs for side effect predictionβ13Feb 4, 2020Updated 6 years ago
- β295Oct 1, 2024Updated last year
- Improving Steering Vectors by Targeting Sparse Autoencoder Featuresβ30Nov 20, 2024Updated last year
- The official code of TACL 2022, "Break, Perturb, Build: Automatic Perturbation of Reasoning Paths Through Question Decomposition".β12Oct 18, 2021Updated 4 years ago
- β10Feb 9, 2026Updated 6 months ago
- Multi-agent simulator in Jax for research and teaching in AI & ALifeβ32Apr 11, 2026Updated 4 months ago
- Deploy to Railway using AI coding agents - Free Credits Offer β’ AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- This is the official repository for the "Towards Vision-Language Mechanistic Interpretability: A Causal Tracing Tool for BLIP" paper acceβ¦β25Feb 16, 2026Updated 6 months ago
- We revisit the Platonic Representation Hypothesis using calibrated representational similarity metrics with statistical guarantees.β38Jun 24, 2026Updated 2 months ago
- [NeurIPS 2024 Spotlight] Code and data for the paper "Finding Transformer Circuits with Edge Pruning".β70Aug 15, 2025Updated last year
- Official Repository of Paper "Watch the Weights: Unsupervised monitoring and control of fine-tuned LLMs"β15Sep 25, 2025Updated 11 months ago
- β15Apr 15, 2026Updated 4 months ago
- Draw MNIST digits and classify in real time!β11Aug 27, 2024Updated 2 years ago
- β18Mar 23, 2026Updated 5 months ago