Towards a Mechanistic Interpretation of Multi-Step Reasoning Capabilities of Language Models
☆16Nov 4, 2023Updated 2 years ago
Alternatives and similar repositories for MechanisticProbe
Users that are interested in MechanisticProbe are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Bird’s Eye: Probing for Linguistic Graph Structureswith a Simple Information-Theoretic Approach☆11Aug 1, 2021Updated 5 years ago
- Official PyTorch code for "Sample Efficient Offline-to-Online Reinforcement Learning" in TKDE'23.☆16Aug 14, 2023Updated 3 years ago
- ☆17Oct 11, 2022Updated 3 years ago
- 📚 List of Top-tier Conference Papers on Reinforcement Learning (RL),including: NeurIPS, AAAI, IJCAI, ICML, AAMAS, ICLR, ICRA, etc. | (AI…☆11Aug 20, 2023Updated 3 years ago
- [ACL 2024] "Understanding and Patching Compositional Reasoning in LLMs"☆14Aug 8, 2026Updated 3 weeks ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- Implementation of the paper "Multi-Agent Exploration via Self-Learning and Social Learning"☆20Dec 7, 2024Updated last year
- Code for the paper "A Mechanistic Interpretation of Arithmetic Reasoning in Language Models using Causal Mediation Analysis"☆20Jun 12, 2025Updated last year
- This is official project in our paper: Is Bigger and Deeper Always Better? Probing LLaMA Across Scales and Layers☆30Jan 13, 2024Updated 2 years ago
- Less is More: Mitigating Multimodal Hallucination from an EOS Decision Perspective (ACL 2024)☆59Oct 28, 2024Updated last year
- Redwood Research's transformer interpretability tools☆15Apr 15, 2022Updated 4 years ago
- ☆14Jan 6, 2025Updated last year
- ☆17Mar 22, 2025Updated last year
- What Has Been Enhanced in my Knowledge-Enhanced Language Model?☆13Oct 26, 2022Updated 3 years ago
- Baseline models for the paper: "Modeling Naive Psychology of Characters in Simple Commonsense Stories" by Hannah Rashkin, Antoine Bosselu…☆16Feb 23, 2021Updated 5 years ago
- Deploy open-source AI quickly and easily - Special Bonus Offer • AdRunpod Hub is built for open source. One-click deployment and autoscaling endpoints without provisioning your own infrastructure.
- Benchmarking Commonsense Reasoning in Real-World Tasks☆11Dec 14, 2023Updated 2 years ago
- A benchmark for assessing the strength of causal relationships between real-world events (EMNLP 2023).☆15Nov 23, 2023Updated 2 years ago
- Code for our paper: ACM-MILP: Adaptive Constraint Modification via Grouping and Selection for Hardness-Preserving MILP Instance Generatio…☆14Jan 3, 2025Updated last year
- ☆10Aug 24, 2023Updated 3 years ago
- Code for the paper: CodeTree: Agent-guided Tree Search for Code Generation with Large Language Models☆38Jun 2, 2026Updated 2 months ago
- [ICML 2024] Unveiling and Harnessing Hidden Attention Sinks: Enhancing Large Language Models without Training through Attention Calibrati…☆45Jun 30, 2024Updated 2 years ago
- Multi-camera calibration (intrinsics, extrinsics, and bundle adjustment)☆14Nov 2, 2025Updated 9 months ago
- New version of COMET ("COMET: Commonsense Transformers for Automatic Knowledge Graph Construction")☆17Sep 19, 2021Updated 4 years ago
- AdaICL: Which Examples to Annotate of In-Context Learning? Towards Effective and Efficient Selection☆19Oct 30, 2023Updated 2 years ago
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- Preprint: Asymmetry in Low-Rank Adapters of Foundation Models☆40Feb 27, 2024Updated 2 years ago
- Curation of resources for LLM mathematical reasoning, most of which are screened by @tongyx361 to ensure high quality and accompanied wit…☆160Jul 12, 2024Updated 2 years ago
- ☆68Jan 23, 2026Updated 7 months ago
- Codebase for Global Neural CCG Parsing with Optimality Guarantees☆25Apr 27, 2017Updated 9 years ago
- python file for lilab☆16Aug 11, 2026Updated 2 weeks ago
- ☆12Feb 6, 2021Updated 5 years ago
- [NeurIPS 2025@FoRLM] R1-Compress: Long Chain-of-Thought Compression via Chunk Compression and Search☆17Jan 24, 2026Updated 7 months ago
- Train your own GPT2!☆14Apr 11, 2023Updated 3 years ago
- CVPR2021: Detecting Human-Object Interaction via Fabricated Compositional Learning☆16Jul 7, 2021Updated 5 years ago
- Bare Metal GPUs on DigitalOcean Gradient AI • AdPurpose-built for serious AI teams training foundational models, running large-scale inference, and pushing the boundaries of what's possible.
- ☆19Mar 25, 2025Updated last year
- ☆17Feb 26, 2024Updated 2 years ago
- A Gymnasium-based Environment of the Abstraction and Reasoning Corpus (ARC)☆73Aug 30, 2024Updated 2 years ago
- Code for L4DC 2022 paper: Joint Synthesis of Safety Certificate and Safe Control Policy Using Constrained Reinforcement Learning.☆14Jul 31, 2023Updated 3 years ago
- ☆12Jan 10, 2025Updated last year
- ☆17Jan 27, 2026Updated 7 months ago
- This is the official implementation of Multi-Agent PPO.☆150Jan 17, 2023Updated 3 years ago