☆21Feb 17, 2023Updated 3 years ago
Alternatives and similar repositories for remix_public
Users that are interested in remix_public are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Machine Learning for Alignment Bootcamp☆28Mar 7, 2024Updated 2 years ago
- Interpreting how transformers simulate agents performing RL tasks☆90Oct 23, 2023Updated 2 years ago
- Resources for skilling up in AI alignment research engineering. Covers basics of deep learning, mechanistic interpretability, and RL.☆248Aug 11, 2025Updated last year
- ☆68Feb 16, 2023Updated 3 years ago
- ☆15Jul 12, 2024Updated 2 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- A python package for protein inference in Mass Spectrometric data analysis.☆10Jun 6, 2022Updated 4 years ago
- ☆297Oct 1, 2024Updated last year
- Sparse probing paper full code.☆68Dec 17, 2023Updated 2 years ago
- ☆32Apr 4, 2024Updated 2 years ago
- Code for the paper "A is for Absorption: Studying Feature Splitting and Absorption in Sparse Autoencoders"☆17Dec 28, 2025Updated 8 months ago
- Official Implementation of "Style Generator Inversion for Image Enhancement and Animation".☆13Dec 2, 2021Updated 4 years ago
- Mechanistic Interpretability for Transformer Models☆55Jun 1, 2022Updated 4 years ago
- Machine Learning for Alignment Bootcamp (MLAB).☆36Jan 24, 2022Updated 4 years ago
- ☆38Apr 30, 2024Updated 2 years ago
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- A plugin for Figma that draws Sigils to a document☆21Jan 5, 2023Updated 3 years ago
- Accompanying codebase for neuroscope.io, a website for displaying max activating dataset examples for language model neurons☆15Feb 13, 2023Updated 3 years ago
- Keeping language models honest by directly eliciting knowledge encoded in their activations.☆225Updated this week
- ☆11Nov 22, 2019Updated 6 years ago
- ☆12Oct 24, 2022Updated 3 years ago
- A text-based game where language models learn to lie and to detect lies.☆12Oct 4, 2023Updated 2 years ago
- Vivaria is METR's tool for running evaluations and conducting agent elicitation research.☆142May 18, 2026Updated 4 months ago
- Backwards compatible callback APIs☆19Jun 9, 2020Updated 6 years ago
- A library for mechanistic anomaly detection☆22Jan 9, 2025Updated last year
- Proton VPN Special Offer - Get 70% off • AdSpecial partner offer. Trusted by over 100 million users worldwide. Tested, Approved and Recommended by Experts.
- ☆12Jun 27, 2024Updated 2 years ago
- Tools for understanding how transformer predictions are built layer-by-layer☆612Aug 7, 2025Updated last year
- This repository contains code developed by the SRI team for the IARPA/TrojAI program.☆21Jul 1, 2021Updated 5 years ago
- ☆13Nov 15, 2023Updated 2 years ago
- Algebraic value editing in pretrained language models☆71Nov 1, 2023Updated 2 years ago
- A browser extension that generates links for all headers on the page (when it can) and makes it easier to share specific sections.☆16Mar 3, 2022Updated 4 years ago
- Mechanistic Interpretability Visualizations using React☆366Apr 30, 2026Updated 4 months ago
- Program for showing Sway window manager's key bindings associated with a mode☆14Jul 23, 2022Updated 4 years ago
- Scala/Play + Vue.js web application providing online Risk, produced for CS 2340 with Professor Simpkins☆15Dec 2, 2024Updated last year
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- Self-supervised MPFNet for realistic bokeh effect rendering(JVCIR2022)☆14Jul 5, 2022Updated 4 years ago
- Kelley Blue Book Extension for Craigslist☆12Apr 10, 2022Updated 4 years ago
- Implementation of the methods described in our paper "Explicit Planning Helps Language Models in Logical Reasoning"☆22Apr 12, 2023Updated 3 years ago
- Arrakis is a library to conduct, track and visualize mechanistic interpretability experiments.☆31Jul 8, 2026Updated 2 months ago
- Models for data stocks and training dataset sizes☆20Jul 10, 2024Updated 2 years ago
- Code and project page for ICCV 2021 paper "DisUnknown: Distilling Unknown Factors for Disentanglement Learning"☆26Oct 13, 2021Updated 4 years ago
- my dotfiles☆10Jul 24, 2021Updated 5 years ago