This was designed for interp researchers who want to do research on or with interp agents to give quality of life improvements and fix some of the annoying things you get from only using Claude code out of the box
☆146Feb 8, 2026Updated 6 months ago
Alternatives and similar repositories for seer
Users that are interested in seer are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Unified access to Large Language Model modules using NNsight☆119Updated this week
- Code repo for the model organisms and convergent directions of EM papers.☆82Sep 22, 2025Updated 11 months ago
- ⚓️ Repository for the "Thought Anchors: Which LLM Reasoning Steps Matter?" paper.☆142Oct 27, 2025Updated 10 months ago
- A toolkit that provides a range of model diffing techniques including a UI to visualize them interactively.☆82Updated this week
- ☆20Mar 16, 2026Updated 5 months ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- The nnsight package enables interpreting and manipulating the internals of deep learned models.☆1,087Updated this week
- Extract residual-stream activations and apply steering vectors (including activation oracles) to any vLLM model during inference.☆121Aug 11, 2026Updated 3 weeks ago
- Code for the "Overcoming Sparsity Artifacts in Crosscoders to Interpret Chat-Tuning" paper.☆18Jul 6, 2026Updated last month
- ☆1,258Updated this week
- Agent observability and replay tooling for AI safety & interpretability research.☆115Jun 19, 2026Updated 2 months ago
- Code repository for "Eliciting Secret Knowledge from Language Models"☆24Mar 30, 2026Updated 5 months ago
- ☆22Nov 15, 2024Updated last year
- Parameter Decomposition☆142Updated this week
- ☆87Feb 18, 2026Updated 6 months ago
- End-to-end encrypted cloud storage - Proton Drive • AdSpecial offer: 40% Off Yearly / 80% Off First Month. Protect your most important files, photos, and documents from prying eyes.
- Course Materials for Interpretability of Large Language Models (0368.4264) at Tel Aviv University☆337Feb 8, 2026Updated 6 months ago
- ☆25Mar 30, 2026Updated 5 months ago
- James' cookbook of evaluations and finetuning experiments☆35Feb 19, 2026Updated 6 months ago
- ☆44Feb 18, 2026Updated 6 months ago
- ☆61Jul 4, 2025Updated last year
- Inference API for many LLMs and other useful tools for empirical research☆136May 29, 2026Updated 3 months ago
- ☆51Feb 11, 2025Updated last year
- Repository for "Training Language Models To Explain Their Own Computations"☆37Jul 7, 2026Updated last month
- A suite of interpretability tasks to evaluate agents using Scribe for notebook access☆18Oct 2, 2025Updated 11 months ago
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- Delphi was the home of a temple to Phoebus Apollo, which famously had the inscription, 'Know Thyself.' This library lets language models …☆275Updated this week
- Code for Negation Neglect☆17May 22, 2026Updated 3 months ago
- harvard-cs-2881-classroom-hw0-c2881-hw0 created by GitHub Classroom☆18Jul 26, 2025Updated last year
- Mechanistic Interpretability Visualizations using React☆367Apr 30, 2026Updated 4 months ago
- ☆16Oct 13, 2025Updated 10 months ago
- A library for mechanistic interpretability of GPT-style language models☆3,850Updated this week
- Prompts used in the Automated Auditing Blog Post☆171Jul 24, 2025Updated last year
- A curated reading list of research in Sparse Autoencoders, Feature Extraction and related topics in Mechanistic Interpretability☆33Jan 30, 2025Updated last year
- Easily deploy my zsh and tmux configuration on new machines. Includes local and remote aliases to improve workflow.☆16Apr 23, 2026Updated 4 months ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- ☆21Dec 10, 2025Updated 8 months ago
- ☆31Jul 1, 2026Updated 2 months ago
- Training Sparse Autoencoders on Language Models☆1,519Updated this week
- Attribution-based Parameter Decomposition☆35Jun 11, 2025Updated last year
- open source interpretability platform 🧠☆1,124Updated this week
- Designing a Dashboard for Transparency and Control of Conversational AI, https://arxiv.org/abs/2406.07882☆40Oct 7, 2025Updated 10 months ago
- ☆428Aug 21, 2025Updated last year