Code for Negation Neglect
☆17May 22, 2026Updated 3 months ago
Alternatives and similar repositories for negation_neglect
Users that are interested in negation_neglect are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- ☆62Jul 4, 2025Updated last year
- ☆33Nov 11, 2025Updated 10 months ago
- Code and materials for "Weird Generalization and Inductive Backdoors"☆43Jan 11, 2026Updated 8 months ago
- ☆26Sep 5, 2024Updated 2 years ago
- Easily deploy my zsh and tmux configuration on new machines. Includes local and remote aliases to improve workflow.☆16Apr 23, 2026Updated 4 months ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- Calibrate your predictions. Tighten your debugging feedback loop. See how underconfident or overconfident you are.☆12Jul 3, 2016Updated 10 years ago
- Improving Steering Vectors by Targeting Sparse Autoencoder Features☆30Nov 20, 2024Updated last year
- A library for training crosscoders☆18May 28, 2025Updated last year
- Code for the reproduction of counting manifolds☆16Feb 26, 2026Updated 6 months ago
- Training LLMs to Report Their Learned Behaviors☆29Apr 28, 2026Updated 4 months ago
- Repository for the "Chain-of-Thought Reasoning In The Wild Is Not Always Faithful" paper☆35Mar 31, 2026Updated 5 months ago
- Code repo for the model organisms and convergent directions of EM papers.☆82Sep 22, 2025Updated 11 months ago
- This was designed for interp researchers who want to do research on or with interp agents to give quality of life improvements and fix …☆146Feb 8, 2026Updated 7 months ago
- Code and data for editing model beliefs with SDF and other methods, and for evaluating the depth of the implanted beliefs.☆20Oct 23, 2025Updated 10 months ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- Code for paper "Leakage-Adjusted Simulatability: Can Models Generate Non-Trivial Explanations of Their Behavior in Natural Language?"☆21Oct 13, 2020Updated 5 years ago
- ADAG: Transluce's MLP neuron-level circuit tracing library☆37Apr 10, 2026Updated 5 months ago
- ⚓️ Repository for the "Thought Anchors: Which LLM Reasoning Steps Matter?" paper.☆142Oct 27, 2025Updated 10 months ago
- [ICML 2026] Official implementation for paper: Learning Self-Correction in Vision–Language Models via Rollout Augmentation☆16Jun 4, 2026Updated 3 months ago
- A collection of lightweight interpretability scripts to understand how LLMs think☆92Mar 18, 2026Updated 5 months ago
- Implementation of the Decrypto benchmark for multi-agent reasoning and theory of mind.☆23Jan 19, 2026Updated 7 months ago
- A toolkit for embedding text datasets with sparse autoencoders☆31Mar 24, 2026Updated 5 months ago
- Sparse Autoencoder Training Library☆58May 1, 2025Updated last year
- A toolkit that provides a range of model diffing techniques including a UI to visualize them interactively.☆83Sep 1, 2026Updated last week
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- Inference API for many LLMs and other useful tools for empirical research☆136May 29, 2026Updated 3 months ago
- ThoughtTrace: Understanding User Thoughts in Real-World LLM Interactions☆14Jun 28, 2026Updated 2 months ago
- A critical analysis of the Cambrian-S model and VSI-Super benchmarks☆16Nov 20, 2025Updated 9 months ago
- Code for the multi-agent computer use project.☆25Jul 3, 2026Updated 2 months ago
- P1-VL: Bridging Visual Perception and Scientific Reasoning in Physics Olympiads☆15Feb 11, 2026Updated 7 months ago
- AI Control for Claude Code in the real world☆34Jul 7, 2026Updated 2 months ago
- ☆24Jul 20, 2026Updated last month
- [preprint] Think Longer to Explore Deeper: Learn to Explore In-Context via Length-Incentivized Reinforcement Learning☆19Feb 18, 2026Updated 6 months ago
- ☆30Aug 2, 2024Updated 2 years ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- [ICML 2026] Code for V1: Unifying Generation and Self-Verification for Parallel Reasoners.☆39Mar 5, 2026Updated 6 months ago
- ☆53Updated this week
- ☆15Jun 11, 2025Updated last year
- ☆19Jul 1, 2026Updated 2 months ago
- Audit any agent decision across its past, present, and future, on one typed graph.☆23Updated this week
- ☆18Apr 6, 2026Updated 5 months ago
- Paper Implementation of Self-Rewarding Language Models☆13Feb 1, 2024Updated 2 years ago