☆61Sep 17, 2025Updated 11 months ago
Alternatives and similar repositories for openCLT
Users that are interested in openCLT are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- ☆211Nov 17, 2024Updated last year
- Sparsify transformers with cross-layer transcoders☆27Nov 14, 2025Updated 9 months ago
- ☆18Jul 9, 2025Updated last year
- A toolkit that provides a range of model diffing techniques including a UI to visualize them interactively.☆82Updated this week
- ☆15Sep 29, 2022Updated 3 years ago
- Bare Metal GPUs on DigitalOcean Gradient AI • AdPurpose-built for serious AI teams training foundational models, running large-scale inference, and pushing the boundaries of what's possible.
- This repository contains the code used for the experiments in the paper "Language Models use Lookbacks to Track Beliefs".☆17Mar 14, 2026Updated 5 months ago
- Sparsify transformers with SAEs and transcoders☆739Updated this week
- ☆24Jun 16, 2024Updated 2 years ago
- Code repository for "Eliciting Secret Knowledge from Language Models"☆24Mar 30, 2026Updated 5 months ago
- The code for creating the iGSM datasets in papers "Physics of Language Models Part 2.1, Grade-School Math and the Hidden Reasoning Proces…☆90Jan 12, 2025Updated last year
- Code for simulations in "Computational mechanisms of curiosity and goal-directed exploration"☆11May 22, 2020Updated 6 years ago
- The Full Spectrum of Deepnet Hessians at Scale: Dynamics with SGD Training and Sample Size☆19May 19, 2019Updated 7 years ago
- ☆2,899Updated this week
- ☆26Feb 20, 2026Updated 6 months ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- This repo is built to facilitate the training and analysis of autoregressive transformers on maze-solving tasks.☆35Oct 28, 2025Updated 10 months ago
- ☆20Feb 8, 2024Updated 2 years ago
- 6,080-param transformer achieving 100% accuracy on 10-digit addition. Trained from scratch in 10 minutes.☆22Feb 19, 2026Updated 6 months ago
- Code for the paper "Verifying Chain-of-Thought Reasoning via its Computational Graph".☆34Dec 6, 2025Updated 8 months ago
- ☆15Dec 16, 2025Updated 8 months ago
- ☆15Dec 16, 2025Updated 8 months ago
- Code and data for paper "(How) do Language Models Track State?"☆28Mar 31, 2025Updated last year
- Modified to support crosscoder training.☆29Jul 2, 2026Updated 2 months ago
- Code for "Echo Chamber: RL Post-training Amplifies Behaviors Learned in Pretraining"☆30Oct 14, 2025Updated 10 months ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- Sparse Autoencoder Training Library☆58May 1, 2025Updated last year
- Official PyTorch Implementation for Learning a Generative Meta-Model of LLM Activations, ICML 2026☆94Apr 30, 2026Updated 4 months ago
- Official implementation of "What does CLIP know about a red circle? Visual Prompt Engineering for VLMs", ICCV 2023☆12Sep 21, 2023Updated 2 years ago
- Training Sparse Autoencoders on Language Models☆1,519Updated this week
- An implementation of Etcetera Abduction in Python☆11Aug 25, 2026Updated last week
- ☆35Jul 5, 2023Updated 3 years ago
- Building on Anthropic's Circuit Tracer, Neuronpedia, Ameisen et al. (2025) and Lindsey et al. (2025), we attempt to extend the paradigm w…☆75Aug 1, 2025Updated last year
- A Mechanistic Interpretability Toolkit for Cross-Layer Transcoder Training and Attribution-Graph Visualization☆108Jul 30, 2026Updated last month
- ☆602Jul 19, 2024Updated 2 years ago
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- Official Code for What Makes and Breaks Safety Fine-tuning? A Mechanistic Study (NeurIPS 2024)☆11Oct 31, 2024Updated last year
- ☆11Jul 11, 2023Updated 3 years ago
- Flax (JAX) implementation of Progressive Growing of GANs for Improved Quality, Stability, and Variation☆12May 24, 2021Updated 5 years ago
- Code for "Evidence of Learned Look-Ahead in a Chess-Playing Neural Network"☆31Jun 4, 2024Updated 2 years ago
- ☆62Nov 19, 2024Updated last year
- Open source interpretability artefacts for R1.☆183Apr 21, 2025Updated last year
- Monitoring the health of ARR☆33Aug 19, 2026Updated 2 weeks ago