Towards a Mechanistic Understanding of Large Reasoning Models: A Survey of Training, Inference, and Failures
☆34Jan 29, 2026Updated 6 months ago
Alternatives and similar repositories for Awesome-LRM-Mechanisms
Users that are interested in Awesome-LRM-Mechanisms are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Official repository for the paper Number Cookbook: Number Understanding of Language Models and How to Improve It.☆22Mar 31, 2025Updated last year
- ☆15Jan 20, 2026Updated 6 months ago
- [ACL 2025] RuleArena: A Benchmark for Rule-Guided Reasoning with LLMs in Real-World Scenarios☆29Jul 31, 2026Updated last week
- About Data and Codes for EMNLP 2023 System Demo Paper "QACHECK: A Demonstration System for Question-Guided Multi-Hop Fact-Checking"☆19Dec 19, 2023Updated 2 years ago
- [arxiv: 2604.14142] From P(y|x) to P(y): Investigating Reinforcement Learning in Pre-train Space☆17Apr 16, 2026Updated 3 months ago
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- ☆11Oct 25, 2024Updated last year
- ☆124Dec 24, 2025Updated 7 months ago
- Codes for ACL 2023 Paper "Fact-Checking Complex Claims with Program-Guided Reasoning"☆32Jun 2, 2023Updated 3 years ago
- ☆17Jun 10, 2025Updated last year
- This repo is the official implementation of “Are Your Agents Upward Deceivers?”. The paper is accepted by ICML 2026.☆24Dec 15, 2025Updated 7 months ago
- The official repository of NeurIPS'25 paper "Ada-R1: From Long-Cot to Hybrid-CoT via Bi-Level Adaptive Reasoning Optimization"☆24May 6, 2026Updated 3 months ago
- A hierarchical multi-agent framework for exhaustive cross-document question answering.☆22Mar 14, 2026Updated 4 months ago
- LLM 时代的 Hot 100 - 大模型面试手撕代码☆36May 4, 2026Updated 3 months ago
- New testbed of interactive SWE tasks for coding agents, set in a realistic multi-turn developer driven environment☆24Jun 30, 2026Updated last month
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- [ICML'26 Spotlight] What is the right loss function for LLM supervised finetuning?☆67May 28, 2026Updated 2 months ago
- ☆75Apr 13, 2025Updated last year
- ☆22May 23, 2025Updated last year
- Code for InfoCTM: A Mutual Information Maximization Perspective of Cross-lingual Topic Modeling (AAAI2023)☆26Mar 6, 2024Updated 2 years ago
- ☆28Jul 18, 2025Updated last year
- ☆24Jun 13, 2023Updated 3 years ago
- The baseline method for CCIR 22 https://www.datafountain.cn/competitions/573☆13Aug 2, 2022Updated 4 years ago
- The repo for SHINE: A Scalable In-Context Hypernetwork for Mapping Context to LoRA in a Single Pass☆98May 23, 2026Updated 2 months ago
- A summary of must-read papers for Neural Question Generation (NQG)☆14Nov 14, 2020Updated 5 years ago
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- [EMNLP 2024 Tutorial] Language Agents: Foundations, Prospects, and Risks☆10Nov 27, 2024Updated last year
- Repo for my blogs explaining swish activation function☆13Dec 17, 2017Updated 8 years ago
- Code and models for ``Answering Open-Domain Multi-Answer Questions via a Recall-then-Verify Framework (ACL 2022)''☆12Jun 29, 2022Updated 4 years ago
- A curated list of resources on Reinforcement Learning with Verifiable Rewards (RLVR) and the reasoning capability boundary of Large Langu…☆91Dec 12, 2025Updated 8 months ago
- ☆29Aug 8, 2025Updated last year
- ☆11Mar 26, 2020Updated 6 years ago
- [ACL 2024] "Understanding and Patching Compositional Reasoning in LLMs"☆14Updated this week
- ☆18May 25, 2026Updated 2 months ago
- Repository for Teaching Broad Reasoning Skills for Multi-Step QA by Generating Hard Contexts, EMNLP22☆19Jun 23, 2023Updated 3 years ago
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- SimKO: Simple Pass@K Policy Optimization☆31Oct 24, 2025Updated 9 months ago
- Code for AAAI 2023 research track paper "Question Decomposition Tree for Answering Complex Questions over Knowledge Bases"☆17Jan 3, 2024Updated 2 years ago
- Official repository for the paper "ModelTables: A Corpus of Tables about Models"☆16Jul 14, 2026Updated 3 weeks ago
- Code repository for "The Clock and the Pizza: Two Stories in Mechanistic Explanation of Neural Networks"☆20Nov 24, 2023Updated 2 years ago
- ☆17Jul 2, 2026Updated last month
- Simple dungeon generation in Python☆12May 16, 2024Updated 2 years ago
- Repo of Paper: delta-Mem: Efficient Online Memory for Large Language Models☆52May 27, 2026Updated 2 months ago