Official implementation of the paper "Contextual Counterfactual Credit Assignment for Multi-Agent Reinforcement Learning in LLM Collaboration". (by Yanjun Chen)
☆36Mar 10, 2026Updated 5 months ago
Alternatives and similar repositories for C3
Users that are interested in C3 are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- [EMNLP 2024 Main] Official implementation of the paper "To Preserve or To Compress: An In-Depth Study of Connector Selection in Multimoda…☆16Dec 13, 2024Updated last year
- ROS (Python) package for controlling Yale Grablab Openhands☆13Feb 10, 2022Updated 4 years ago
- 宁波大学机器人创新协会的机器人驱动控制器程序☆15Apr 17, 2024Updated 2 years ago
- In cases such as remote servers, use clash☆14Feb 23, 2024Updated 2 years ago
- Offline Policy Evaluation via Adaptive Weighting with Data from Contextual Bandits☆11Oct 21, 2024Updated last year
- GPUs on demand by Runpod - Special Offer Available • AdRun AI, ML, and HPC workloads on powerful cloud GPUs—without limits or wasted spend. Deploy GPUs in under a minute and pay by the second.
- (ACL2025 Findings) Official code for the paper "STeCa: Step-level Trajectory Calibration for LLM Agent Learning"☆29Mar 2, 2026Updated 5 months ago
- Official implementation of "Diffusion Language Models Know the Answer Before Decoding"☆60Apr 28, 2026Updated 3 months ago
- This repository contains the implementation of reinforcement learning algorithms like PPO and A2C, to solve the problem: Dynamic Obstacle…☆19Jan 17, 2022Updated 4 years ago
- ⚖️ A multi-agent consensus ranking system to derive optimal weights through pairwise comparisons☆17Oct 30, 2025Updated 9 months ago
- ☆17Jan 24, 2024Updated 2 years ago
- Lightweight, safe, self-improving multi-role AI agent framework. Hybrid local + cloud LLMs with self-calibrating routing, decaying skills…☆16Aug 3, 2026Updated last week
- pytorch implementation for "Mutual Information Neural Estimation"☆11Dec 13, 2019Updated 6 years ago
- Official inference implementation of the paper "DON'T SETTLE TOO EARLY: SELF-REFLECTIVE REMASKING FOR DIFFUSION LANGUAGE MODELS". [ICLR 2…☆15Jan 28, 2026Updated 6 months ago
- ☆46Jun 27, 2025Updated last year
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- ☆12Aug 15, 2020Updated 5 years ago
- [ICLR 2023] The official code for paper "Guarded Policy Optimization with Imperfect Online Demonstrations"☆14Apr 30, 2023Updated 3 years ago
- ☆25Feb 24, 2023Updated 3 years ago
- AgentForesight: Online Auditing for Early Failure Prediction in Multi-Agent Systems☆16May 12, 2026Updated 2 months ago
- ☆16Feb 2, 2022Updated 4 years ago
- ☆14Nov 22, 2025Updated 8 months ago
- Local-first platform for managing research projects, vibe coding and vibe research.☆26Mar 13, 2026Updated 4 months ago
- Agent-Omit: Training Efficient LLM Agents for Adaptive Thought and Observation Omission via Reinforcement Learning☆33May 11, 2026Updated 2 months ago
- MrSteve: Instruction-Following Agents in Minecraft with What-Where-When Memory☆16May 1, 2025Updated last year
- End-to-end encrypted email - Proton Mail • AdSpecial offer: 40% Off Yearly / 80% Off First Month. All Proton services are open source and independently audited for security.
- ☆23Apr 5, 2026Updated 4 months ago
- [arXiv:2605.19952] "Rethinking How to Remember: Beyond Atomic Facts in Lifelong LLM Agent Memory"☆16May 20, 2026Updated 2 months ago
- The implementation for ACL 2026: MatchTIR: Fine-Grained Supervision for Tool-Integrated Reasoning via Bipartite Matching.☆20Jul 27, 2026Updated 2 weeks ago
- ☆27May 12, 2026Updated 2 months ago
- RETROAGENT: From Solving to Evolving via Retrospective Dual Intrinsic Feedback☆28Mar 30, 2026Updated 4 months ago
- [ICLR 2025] Unintentional Unalignment: Likelihood Displacement in Direct Preference Optimization☆32Jan 7, 2026Updated 7 months ago
- Official Implementation of ReALFRED (ECCV'24)☆47Oct 11, 2024Updated last year
- Cooperative Graph-based Networked Agent Challenges for Multi-Agent Reinforcement Learning☆15Jan 26, 2026Updated 6 months ago
- 🚲 Code and benchmark for our COLM 2025 paper - "Thought Tracing: Hypothesis-Driven Theory-of-Mind Reasoning for Large Language Models"☆15Aug 8, 2025Updated last year
- Managed Database hosting by DigitalOcean • AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- Official code for the paper Towards Fully Exploiting LLM Internal States to Enhance Knowledge Boundary Perception. The code is based on t…☆22Aug 5, 2025Updated last year
- (RA-L 2025) VILP: Imitation Learning with Latent Video Planning☆27Jun 21, 2025Updated last year
- Anchored Diffusion Language Model (NeurIPS 2025)☆30Oct 13, 2025Updated 9 months ago
- A Docker-first, non-preemptive multi-agent coordination runtime☆17Updated this week
- URB - Urban Routing Benchmark - Benchmarking MARL algorithms on the fleet routing tasks.☆18Updated this week
- Part 1 project for ME5406 in NUS☆10Jun 25, 2021Updated 5 years ago
- Autonomous multi-agent orchestration engine for software repos — L0/L1/L2 agents, leases, gates, audits, git-worktree isolation. Detached…☆26Jul 26, 2026Updated 2 weeks ago