Official implementation of "The Trace Is the State: Exact Credit Assignment for LLM Agent Teams" (arXiv:2603.06859).
☆43Oct 3, 2026Updated last week
Alternatives and similar repositories for C3
Users that are interested in C3 are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- [EMNLP 2024 Main] Official implementation of the paper "The Accuracy Paradox in RLHF: When Better Reward Models Don't Yield Better Langua…☆13Nov 11, 2024Updated last year
- AT2PO: Agentic Turn-based Policy Optimization via Tree Search☆22May 21, 2026Updated 4 months ago
- [EMNLP 2024 Main] Official implementation of the paper "Unveiling In-Context Learning: A Coordinate System to Understand Its Working Mech…☆15Oct 8, 2024Updated 2 years ago
- ☆27Jun 5, 2025Updated last year
- AI pull-request reviewer: hybrid RAG + a code graph + Claude Code/Codex. Whole-repo context, inline GitHub comments grounded on exact cod…☆20Updated this week
- Serverless GPU API endpoints on Runpod - Get Bonus Credits • AdSkip the infrastructure headaches. Auto-scaling, pay-as-you-go, no-ops approach lets you focus on innovating your application.
- Official implementation of "Diffusion Language Models Know the Answer Before Decoding"☆61Apr 28, 2026Updated 5 months ago
- This repository contains the implementation of reinforcement learning algorithms like PPO and A2C, to solve the problem: Dynamic Obstacle…☆19Jan 17, 2022Updated 4 years ago
- ⚖️ A multi-agent consensus ranking system to derive optimal weights through pairwise comparisons☆17Oct 30, 2025Updated 11 months ago
- ☆20Apr 9, 2026Updated 6 months ago
- pytorch implementation for "Mutual Information Neural Estimation"☆11Dec 13, 2019Updated 6 years ago
- Official Code For EMNLP2025 Findings: {DLPO : Towards a Robust, Efficient, and Generalizable Prompt Optimization Framework from a Deep-Le…☆10Dec 25, 2025Updated 9 months ago
- Official inference implementation of the paper "DON'T SETTLE TOO EARLY: SELF-REFLECTIVE REMASKING FOR DIFFUSION LANGUAGE MODELS". [ICLR 2…☆16Jan 28, 2026Updated 8 months ago
- ☆51Jun 27, 2025Updated last year
- ☆12Aug 15, 2020Updated 6 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- [ICLR 2023] The official code for paper "Guarded Policy Optimization with Imperfect Online Demonstrations"☆14Apr 30, 2023Updated 3 years ago
- This repository provides a summarization of recent empirical studies/human studies that measure human understanding with machine explanat…☆14Jul 24, 2024Updated 2 years ago
- ☆15Nov 22, 2025Updated 10 months ago
- ☆20Aug 14, 2025Updated last year
- grpo to train long form QA and instructions with long-form reward model☆17Jul 17, 2025Updated last year
- [NeurIPS 2026] "Rethinking How to Remember: Beyond Atomic Facts in Lifelong LLM Agent Memory"☆18Sep 26, 2026Updated 2 weeks ago
- OpenCockpit — the open Claude Code GUI for any LLM (Claude, Codex, DeepSeek, GLM, Kimi, Ollama). Cockpit IDE: parallel projects, terminal…☆40Oct 1, 2026Updated last week
- ☆117Oct 21, 2025Updated 11 months ago
- RETROAGENT: From Solving to Evolving via Retrospective Dual Intrinsic Feedback☆31Mar 30, 2026Updated 6 months ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- Cooperative Graph-based Networked Agent Challenges for Multi-Agent Reinforcement Learning☆16Jan 26, 2026Updated 8 months ago
- Repository for Skill Set Optimization☆14Jul 26, 2024Updated 2 years ago
- 🚲 Code and benchmark for our COLM 2025 paper - "Thought Tracing: Hypothesis-Driven Theory-of-Mind Reasoning for Large Language Models"☆15Aug 8, 2025Updated last year
- Secure Personal AI Assistant with TEE Support☆23Apr 11, 2026Updated 5 months ago
- ☆16Dec 13, 2022Updated 3 years ago
- [ICLR 2026] HiFo-Prompt: Prompting with Hindsight and Foresight for LLM-based Automatic Heuristic Design☆16Feb 7, 2026Updated 8 months ago
- URB - Urban Routing Benchmark - Benchmarking MARL algorithms on the fleet routing tasks.☆18Updated this week
- Part 1 project for ME5406 in NUS☆10Jun 25, 2021Updated 5 years ago
- Autonomous multi-agent orchestration engine for software repos — L0/L1/L2 agents, leases, gates, audits, git-worktree isolation. Detached…☆39Oct 2, 2026Updated last week
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- [ICLR 2026] A Framework for LLM-based Multi-Agent Reinforced Training and Inference☆565Aug 20, 2026Updated last month
- Implementation of the Decrypto benchmark for multi-agent reasoning and theory of mind.☆24Jan 19, 2026Updated 8 months ago
- The implementation of ICML'25 paper "LLM-Assisted Semantically Diverse Teammate Generation for Efficient Multi-agent Coordination".☆17May 13, 2025Updated last year
- In-Progress implementation of the 2021 ICML Paper " Differentiable Spatial Planning using Transformers "☆16Apr 21, 2022Updated 4 years ago
- MuJoCo benchmark for Deep Reinforcement Learning as provided by Tianshou framework.☆14Jan 12, 2025Updated last year
- This repository contains reference implementation for multi-LLM ToM paper (accepted to EMNLP 2023), Theory of Mind for Multi-Agent Collab…☆20Jun 11, 2024Updated 2 years ago
- Single-binary LLM agent runtime built on Elixir/OTP: chat UI, 4-tier memory, Anthropic-compatible Skills, scheduled tasks, multi-provider…☆29Jul 3, 2026Updated 3 months ago