☆21Feb 17, 2023Updated 3 years ago
Alternatives and similar repositories for remix_public
Users that are interested in remix_public are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Interpreting how transformers simulate agents performing RL tasks☆90Oct 23, 2023Updated 2 years ago
- Resources for skilling up in AI alignment research engineering. Covers basics of deep learning, mechanistic interpretability, and RL.☆248Aug 11, 2025Updated last year
- ☆68Feb 16, 2023Updated 3 years ago
- A python package for protein inference in Mass Spectrometric data analysis.☆10Jun 6, 2022Updated 4 years ago
- ☆295Oct 1, 2024Updated last year
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- Sparse probing paper full code.☆68Dec 17, 2023Updated 2 years ago
- ☆32Apr 4, 2024Updated 2 years ago
- Code for the paper "A is for Absorption: Studying Feature Splitting and Absorption in Sparse Autoencoders"☆16Dec 28, 2025Updated 8 months ago
- An application that displays a map and graphs showing solar irradiance forecasts in solar farms in Georgia using data from the National S…☆10Oct 15, 2021Updated 4 years ago
- ☆12Aug 29, 2021Updated 5 years ago
- Official Implementation of "Style Generator Inversion for Image Enhancement and Animation".☆13Dec 2, 2021Updated 4 years ago
- TransformerLens + HuggingFace☆11Nov 4, 2023Updated 2 years ago
- Mechanistic Interpretability for Transformer Models☆54Jun 1, 2022Updated 4 years ago
- Machine Learning for Alignment Bootcamp (MLAB).☆36Jan 24, 2022Updated 4 years ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- 🧠 Starter templates for doing interpretability research☆78Jul 16, 2023Updated 3 years ago
- ☆79May 31, 2023Updated 3 years ago
- ☆38Apr 30, 2024Updated 2 years ago
- Measuring the situational awareness of language models☆42Feb 12, 2024Updated 2 years ago
- ☆13May 7, 2023Updated 3 years ago
- Keeping language models honest by directly eliciting knowledge encoded in their activations.☆225Updated this week
- ☆11Nov 22, 2019Updated 6 years ago
- Official code for "Algorithmic Capabilities of Random Transformers" (NeurIPS 2024)☆15Sep 28, 2024Updated last year
- A text-based game where language models learn to lie and to detect lies.☆12Oct 4, 2023Updated 2 years ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- Vivaria is METR's tool for running evaluations and conducting agent elicitation research.☆141May 18, 2026Updated 3 months ago
- A formalisation of Cartesian Frames, a perspective on embedded agency, in the HOL theorem prover.☆22Dec 20, 2021Updated 4 years ago
- Starter kit and data loading code for the Trojan Detection Challenge NeurIPS 2022 competition☆33Jul 26, 2023Updated 3 years ago
- A library for mechanistic anomaly detection☆22Jan 9, 2025Updated last year
- This repository contains code developed by the SRI team for the IARPA/TrojAI program.☆21Jul 1, 2021Updated 5 years ago
- Algebraic value editing in pretrained language models☆71Nov 1, 2023Updated 2 years ago
- A browser extension that generates links for all headers on the page (when it can) and makes it easier to share specific sections.☆16Mar 3, 2022Updated 4 years ago
- Mechanistic Interpretability Visualizations using React☆367Apr 30, 2026Updated 4 months ago
- Feature Store for Machine Learning, published by Packt☆13Apr 22, 2026Updated 4 months ago
- Open source password manager - Proton Pass • AdSecurely store, share, and autofill your credentials with Proton Pass, the end-to-end encrypted password manager trusted by millions.
- ☆15Jan 19, 2017Updated 9 years ago
- the supercharged twitter feed☆20Aug 7, 2026Updated 3 weeks ago
- Machine Learning Inference Graph Spec☆21Jul 27, 2019Updated 7 years ago
- Tools for running experiments on RL agents in procgen environments☆19Apr 5, 2024Updated 2 years ago
- An implementation of "Subspace Representations for Soft Set Operations and Sentence Similarities" (NAACL 2024)☆10May 31, 2024Updated 2 years ago
- Arrakis is a library to conduct, track and visualize mechanistic interpretability experiments.☆31Jul 8, 2026Updated last month
- Models for data stocks and training dataset sizes☆20Jul 10, 2024Updated 2 years ago