Official implementation of the paper "Contextual Counterfactual Credit Assignment for Multi-Agent Reinforcement Learning in LLM Collaboration". (by Yanjun Chen)
☆42Sep 19, 2026Updated this week
Alternatives and similar repositories for C3
Users that are interested in C3 are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- [EMNLP 2024 Main] Official implementation of the paper "The Accuracy Paradox in RLHF: When Better Reward Models Don't Yield Better Langua…☆13Nov 11, 2024Updated last year
- AT2PO: Agentic Turn-based Policy Optimization via Tree Search☆22May 21, 2026Updated 3 months ago
- Real-time AI governance platform implementing a Cognitive Digital Twin framework. Monitors AI decisions, traces reasoning steps, detect…☆29Jul 29, 2026Updated last month
- This is the official code implementation of Bongard-OpenWorld (ICLR 2024).☆14Jan 6, 2025Updated last year
- [ACL'24, Outstanding Paper] Emulated Disalignment: Safety Alignment for Large Language Models May Backfire!☆39Aug 2, 2024Updated 2 years ago
- End-to-end encrypted email - Proton Mail • AdSpecial offer: 40% Off Yearly / 80% Off First Month. All Proton services are open source and independently audited for security.
- ☆14Apr 22, 2024Updated 2 years ago
- ☆27Jun 5, 2025Updated last year
- (ACL2025 Findings) Official code for the paper "STeCa: Step-level Trajectory Calibration for LLM Agent Learning"☆29Mar 2, 2026Updated 6 months ago
- Official implementation of "Diffusion Language Models Know the Answer Before Decoding"☆61Apr 28, 2026Updated 4 months ago
- This repository contains the implementation of reinforcement learning algorithms like PPO and A2C, to solve the problem: Dynamic Obstacle…☆19Jan 17, 2022Updated 4 years ago
- ⚖️ A multi-agent consensus ranking system to derive optimal weights through pairwise comparisons☆17Oct 30, 2025Updated 10 months ago
- ☆20Apr 9, 2026Updated 5 months ago
- pytorch implementation for "Mutual Information Neural Estimation"☆11Dec 13, 2019Updated 6 years ago
- Official Code For EMNLP2025 Findings: {DLPO : Towards a Robust, Efficient, and Generalizable Prompt Optimization Framework from a Deep-Le…☆10Dec 25, 2025Updated 8 months ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- Official inference implementation of the paper "DON'T SETTLE TOO EARLY: SELF-REFLECTIVE REMASKING FOR DIFFUSION LANGUAGE MODELS". [ICLR 2…☆15Jan 28, 2026Updated 7 months ago
- ☆51Jun 27, 2025Updated last year
- ☆12Aug 15, 2020Updated 6 years ago
- [ICLR 2023] The official code for paper "Guarded Policy Optimization with Imperfect Online Demonstrations"☆14Apr 30, 2023Updated 3 years ago
- ☆25Feb 24, 2023Updated 3 years ago
- This repository provides a summarization of recent empirical studies/human studies that measure human understanding with machine explanat…☆14Jul 24, 2024Updated 2 years ago
- ☆16Feb 2, 2022Updated 4 years ago
- AgentForesight: Online Auditing for Early Failure Prediction in Multi-Agent Systems☆19May 12, 2026Updated 4 months ago
- ☆20Aug 14, 2025Updated last year
- Virtual machines for every use case on DigitalOcean • AdGet dependable uptime with 99.99% SLA, simple security tools, and predictable monthly pricing with DigitalOcean's virtual machines, called Droplets.
- A dual-agent framework leveraging left-brain logic (structured coding, syntax validation, debugging) and right-brain intuition (macro arc…☆20Jun 3, 2026Updated 3 months ago
- A message bus that lets AI assistants talk to each other. Works with Claude, ChatGPT, Gemini, Perplexity, and any AI that supports MCP or…☆17Sep 3, 2026Updated 2 weeks ago
- [arXiv:2605.19952] "Rethinking How to Remember: Beyond Atomic Facts in Lifelong LLM Agent Memory"☆18May 20, 2026Updated 4 months ago
- AegisSovereignAI: The Cross-Ecosystem Trust Layer for the Distributed Enterprise. Verifiable Identity, Hardware-Rooted Integrity, and Sov…☆20Aug 31, 2026Updated 2 weeks ago
- ☆171Dec 15, 2025Updated 9 months ago
- ☆117Oct 21, 2025Updated 10 months ago
- RETROAGENT: From Solving to Evolving via Retrospective Dual Intrinsic Feedback☆31Mar 30, 2026Updated 5 months ago
- [ICLR 2025] Unintentional Unalignment: Likelihood Displacement in Direct Preference Optimization☆32Jan 7, 2026Updated 8 months ago
- Official Implementation of ReALFRED (ECCV'24)☆48Oct 11, 2024Updated last year
- Open source password manager - Proton Pass • AdSecurely store, share, and autofill your credentials with Proton Pass, the end-to-end encrypted password manager trusted by millions.
- Cooperative Graph-based Networked Agent Challenges for Multi-Agent Reinforcement Learning☆16Jan 26, 2026Updated 7 months ago
- Repository for Skill Set Optimization☆14Jul 26, 2024Updated 2 years ago
- Low Level RL Controller for G1☆18Jul 27, 2026Updated last month
- Turn one-off AI collaboration prompts into long-running repository governance. · “把一次性的 AI 协作提示,疏导成项目里长期可运行的治理体系。”——Dayu Harness Skill(大禹…☆18Jun 1, 2026Updated 3 months ago
- 🚲 Code and benchmark for our COLM 2025 paper - "Thought Tracing: Hypothesis-Driven Theory-of-Mind Reasoning for Large Language Models"☆15Aug 8, 2025Updated last year
- Anchored Diffusion Language Model (NeurIPS 2025)☆29Oct 13, 2025Updated 11 months ago
- Secure Personal AI Assistant with TEE Support☆23Apr 11, 2026Updated 5 months ago