Official implementation of the paper "Contextual Counterfactual Credit Assignment for Multi-Agent Reinforcement Learning in LLM Collaboration". (by Yanjun Chen)
☆36Mar 10, 2026Updated 4 months ago
Alternatives and similar repositories for C3
Users that are interested in C3 are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- 一个命令行弹幕姬 / display bullet from bilibili live stream comments in command line.☆11Sep 27, 2024Updated last year
- [EMNLP 2024 Main] Official implementation of the paper "To Preserve or To Compress: An In-Depth Study of Connector Selection in Multimoda…☆17Dec 13, 2024Updated last year
- [EMNLP 2024 Main] Official implementation of the paper "Unveiling In-Context Learning: A Coordinate System to Understand Its Working Mech…☆16Oct 8, 2024Updated last year
- The code for paper FLDCF, with various forgery detection and localization methods.☆19Mar 16, 2026Updated 4 months ago
- [ACL'24, Outstanding Paper] Emulated Disalignment: Safety Alignment for Large Language Models May Backfire!☆39Aug 2, 2024Updated last year
- Deploy open-source AI quickly and easily - Special Bonus Offer • AdRunpod Hub is built for open source. One-click deployment and autoscaling endpoints without provisioning your own infrastructure.
- ☆14Apr 22, 2024Updated 2 years ago
- In cases such as remote servers, use clash☆14Feb 23, 2024Updated 2 years ago
- ☆27Jun 5, 2025Updated last year
- AI pull-request reviewer: hybrid RAG + a code graph + Claude Code. Whole-repo context, inline GitHub comments grounded on exact code. MCP…☆18Updated this week
- (ACL2025 Findings) Official code for the paper "STeCa: Step-level Trajectory Calibration for LLM Agent Learning"☆29Mar 2, 2026Updated 4 months ago
- Official implementation of "Diffusion Language Models Know the Answer Before Decoding"☆60Apr 28, 2026Updated 2 months ago
- This repository contains the implementation of reinforcement learning algorithms like PPO and A2C, to solve the problem: Dynamic Obstacle…☆19Jan 17, 2022Updated 4 years ago
- ⚖️ A multi-agent consensus ranking system to derive optimal weights through pairwise comparisons☆16Oct 30, 2025Updated 8 months ago
- The code of paper: GCRDN: Global Context-Driven Residual Dense Network for Remote Sensing Image Super-Resolution☆25Dec 16, 2024Updated last year
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- ☆19Apr 9, 2026Updated 3 months ago
- pytorch implementation for "Mutual Information Neural Estimation"☆11Dec 13, 2019Updated 6 years ago
- Official inference implementation of the paper "DON'T SETTLE TOO EARLY: SELF-REFLECTIVE REMASKING FOR DIFFUSION LANGUAGE MODELS". [ICLR 2…☆15Jan 28, 2026Updated 5 months ago
- ☆46Jun 27, 2025Updated last year
- ☆12Aug 15, 2020Updated 5 years ago
- [ICLR 2023] The official code for paper "Guarded Policy Optimization with Imperfect Online Demonstrations"☆14Apr 30, 2023Updated 3 years ago
- This repository provides a summarization of recent empirical studies/human studies that measure human understanding with machine explanat…☆14Jul 24, 2024Updated last year
- MuJoCo benchmark for Deep Reinforcement Learning as provided by Tianshou framework.☆15Jan 12, 2025Updated last year
- AgentForesight: Online Auditing for Early Failure Prediction in Multi-Agent Systems☆16May 12, 2026Updated 2 months ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- ☆15Nov 22, 2025Updated 7 months ago
- ☆20Aug 14, 2025Updated 11 months ago
- Official repository for Adaptive Parallel Decoding (APD).☆20Oct 27, 2025Updated 8 months ago
- grpo to train long form QA and instructions with long-form reward model☆17Jul 17, 2025Updated last year
- ☆14Nov 19, 2024Updated last year
- MrSteve: Instruction-Following Agents in Minecraft with What-Where-When Memory☆15May 1, 2025Updated last year
- The implementation for ACL 2026: MatchTIR: Fine-Grained Supervision for Tool-Integrated Reasoning via Bipartite Matching.☆20Apr 18, 2026Updated 3 months ago
- ☆27May 12, 2026Updated 2 months ago
- RETROAGENT: From Solving to Evolving via Retrospective Dual Intrinsic Feedback☆26Mar 30, 2026Updated 3 months ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- [ICLR 2025] Unintentional Unalignment: Likelihood Displacement in Direct Preference Optimization☆32Jan 7, 2026Updated 6 months ago
- ☆116Oct 21, 2025Updated 9 months ago
- Official code repository for CurricuLLM: Automatic Task Curricula Design for Learning Complex Robot Skills using Large Language Models☆28Sep 26, 2025Updated 9 months ago
- Official Implementation of ReALFRED (ECCV'24)☆47Oct 11, 2024Updated last year
- Cooperative Graph-based Networked Agent Challenges for Multi-Agent Reinforcement Learning☆15Jan 26, 2026Updated 5 months ago
- Repository for Skill Set Optimization☆14Jul 26, 2024Updated last year
- 🚲 Code and benchmark for our COLM 2025 paper - "Thought Tracing: Hypothesis-Driven Theory-of-Mind Reasoning for Large Language Models"☆15Aug 8, 2025Updated 11 months ago