Companion code for the global workspace interpretability paper
β1,893Sep 2, 2026Updated last week
Alternatives and similar repositories for jacobian-lens
Users that are interested in jacobian-lens are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- open source interpretability platform π§β1,134Updated this week
- β941Aug 2, 2026Updated last month
- Jacobian-Brainwash : A manual alignment tool for large language models built on Anthropic's Jacobian Lens. Results are exportable.β229Updated this week
- β2,906Updated this week
- The Assistant Axis is a direction in activation space that captures how "Assistant-like" a model's behavior is. Models can drift away froβ¦β171Jan 20, 2026Updated 7 months ago
- Proton VPN Special Offer - Get 70% off β’ AdSpecial partner offer. Trusted by over 100 million users worldwide. Tested, Approved and Recommended by Experts.
- A library for mechanistic interpretability of GPT-style language modelsβ3,870Updated this week
- Repository for "Training Language Models To Explain Their Own Computations"β37Jul 7, 2026Updated 2 months ago
- ADAG: Transluce's MLP neuron-level circuit tracing libraryβ37Apr 10, 2026Updated 5 months ago
- Head Vis Public Releaseβ41May 4, 2026Updated 4 months ago
- Agentic RL Training at Scaleβ2,034Updated this week
- AI agents running research on single-GPU nanochat training automaticallyβ95,643Mar 26, 2026Updated 5 months ago
- Train Embedding Models on MLX.β17Jun 2, 2026Updated 3 months ago
- The best ChatGPT that $100 can buy.β57,964Updated this week
- The nnsight package enables interpreting and manipulating the internals of deep learned models.β1,096Updated this week
- 1-Click AI Models by DigitalOcean Gradient β’ AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- Renderer for the harmony response format to be used with gpt-ossβ4,499Apr 8, 2026Updated 5 months ago
- β99Apr 18, 2026Updated 4 months ago
- Interactive Jacobian-Lens visualizer and live steerer for GGUF models on llama.cppβ84Jul 12, 2026Updated 2 months ago
- β6,562Apr 1, 2026Updated 5 months ago
- NanoGPT (124M) in 90 secondsβ5,788Aug 9, 2026Updated last month
- Training LLMs to Report Their Learned Behaviorsβ29Apr 28, 2026Updated 4 months ago
- Decomposing and measuring evaluation awareness in existing benchmarks and our proposed EvalAwareBench.β19Jun 1, 2026Updated 3 months ago
- Local UI to run and train LLMs and diffusion models. Supports GGUF, MLX, Qwen3.8, DeepSeek-V4, MiniMax-H3, Gemma 4, FLUX and more.β76,062Updated this week
- dLLM: Simple Diffusion Language Modelingβ2,684Jul 17, 2026Updated last month
- 1-Click AI Models by DigitalOcean Gradient β’ AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- AI agent toolkit: unified LLM API, agent loop, TUI, coding agent CLIβ104,359Updated this week
- Open-source cross-modal and multimodal prompt injection test suite. 250,000+ attack payloads across text, image, document, and audio modaβ¦β74Jul 22, 2026Updated last month
- ICA Lens: compact ICA-based interpretability tools for exploring LLM activations. Code release for the paper.β43Aug 16, 2026Updated 3 weeks ago
- β17Aug 19, 2026Updated 3 weeks ago
- Code for COLM '26 paper 'Loop, Think, & Generalize: Implicit Reasoning in Recurrent-Depth Transformers'β98Sep 3, 2026Updated last week
- Democratizing Reinforcement Learning for LLMsβ5,823Updated this week
- slime is an LLM post-training framework for RL Scaling.β8,447Sep 3, 2026Updated last week
- DeepSpec: a full-stack codebase for training and evaluating speculative decoding algorithmsβ7,106Jul 9, 2026Updated 2 months ago
- Make abliterated models with transformers, easy and fastβ188Mar 24, 2026Updated 5 months ago
- Managed Database hosting by DigitalOcean β’ AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- verl/HybridFlow: A Flexible and Efficient RL Post-Training Frameworkβ23,396Updated this week
- A high-throughput and memory-efficient inference and serving engine for LLMsβ91,571Updated this week
- Tools for merging pretrained large language models.β7,347Updated this week
- Train the smallest LM you can that fits in 16MB. Best model wins!β5,182May 4, 2026Updated 4 months ago
- Extract residual-stream activations and apply steering vectors (including activation oracles) to any vLLM model during inference.β122Aug 11, 2026Updated last month
- Persona Vectors: Monitoring and Controlling Character Traits in Language Modelsβ462Apr 22, 2026Updated 4 months ago
- Genertaes control vectors for use with llama.cpp in GGUF format.β49Mar 19, 2025Updated last year