(ACL 2025 Main) Code for MultiAgentBench : Evaluating the Collaboration and Competition of LLM agents https://www.arxiv.org/pdf/2503.01935
☆304Oct 27, 2025Updated 11 months ago
Alternatives and similar repositories for MARBLE
Users that are interested in MARBLE are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- (ACL 2025 Main) Code for MultiAgentBench : Evaluating the Collaboration and Competition of LLM agents https://www.arxiv.org/pdf/2503.019…☆58Jun 21, 2025Updated last year
- REALM-Bench: A Real-World Planning Benchmark for LLMs and Multi-Agent Systems☆48Jul 20, 2026Updated 2 months ago
- Sotopia-RL: Reward Design for Social Intelligence☆53Apr 1, 2026Updated 6 months ago
- [ICML 2025] ResearchTown: Simulator of Human Research Community☆214Oct 1, 2026Updated last week
- Opensource code for ICML 2026 poster☆16Nov 26, 2025Updated 10 months ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- ☆21Apr 21, 2026Updated 5 months ago
- ☆246Jul 25, 2025Updated last year
- ☆42Oct 15, 2024Updated last year
- [NeurIPS'25] KVCOMM: Online Cross-context KV-cache Communication for Efficient LLM-based Multi-agent Systems☆18Nov 1, 2025Updated 11 months ago
- Official code repository for the paper "ToMAP: Training Opponent-Aware LLM Persuaders with Theory of Mind"☆27Sep 25, 2025Updated last year
- ☆145Mar 23, 2025Updated last year
- 🐝 SwarmBench: Benchmarking LLMs' Swarm Intelligence☆39May 21, 2025Updated last year
- Official implementation of the winning system at SemEval-2021 Task 11 - NLP Contribution Graph (Best System Paper Award 🏆)☆11Aug 24, 2025Updated last year
- ☆421Jul 23, 2025Updated last year
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- [EMNLP 2025 Demo] TinyScientist: A Lightweight Framework for Building Research Agents☆139Mar 4, 2026Updated 7 months ago
- ☆52Jun 12, 2026Updated 3 months ago
- ☆17Sep 2, 2025Updated last year
- ☆146Oct 29, 2025Updated 11 months ago
- [ICLR 2026] A Framework for LLM-based Multi-Agent Reinforced Training and Inference☆565Aug 20, 2026Updated last month
- Preview Code for Continuum Paper☆108Aug 13, 2026Updated last month
- [ICML 2026 Spotlight] Latent Collaboration in Multi-Agent Systems☆1,144Jun 18, 2026Updated 3 months ago
- ☆48May 12, 2026Updated 4 months ago
- ☆23Apr 29, 2025Updated last year
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- ☆38Jul 16, 2025Updated last year
- ☆107Mar 30, 2026Updated 6 months ago
- Multi-Agent LLM Evaluation Docs: https://maseval.readthedocs.io/☆39Updated this week
- Benchmark Test-Time Scaling of General LLM Agents☆25Apr 14, 2026Updated 5 months ago
- SkillsBench evaluates how well skills work and how effective agents are at using them.☆1,836Jul 23, 2026Updated 2 months ago
- This is Official implementation for T2S-Bench & Structure-of-Thought: Benchmarking and Prompting Comprehensive Text-to-Structure Reasonin…☆24Mar 5, 2026Updated 7 months ago
- ☆156Jan 21, 2026Updated 8 months ago
- A compilation of the best multi-agent papers☆1,697Sep 23, 2026Updated 2 weeks ago
- [NSDI'26] PolyRL is a reinforcement learning framework for LLM that harvest spot instances on the cloud to reduce cost.☆19Mar 30, 2026Updated 6 months ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- [ICML'25 Oral] Multi-agent Architecture Search via Agentic Supernet☆287Nov 13, 2025Updated 10 months ago
- [ICML 2024] Code for Pairwise Alignment Improves Graph Domain Adaptation (Pair-Align)☆15Jun 15, 2024Updated 2 years ago
- [NeurIPS 2026] GenEnv: Difficulty-Aligned Co-Evolution Between LLM Agents and Environment Simulators☆67Sep 25, 2026Updated 2 weeks ago
- Code and implementations for the ACL 2025 paper "AgentGym: Evolving Large Language Model-based Agents across Diverse Environments" by Zhi…☆850May 30, 2026Updated 4 months ago
- A Comprehensive Benchmark to Evaluate LLMs as Agents (ICLR'24)☆3,765Feb 8, 2026Updated 8 months ago
- [ICLR'26] Stronger-MAS: A RL Framework for multi LLM agent system; [arxiv] MetaAgent-X: End-to-End Reinforcement Learning Automatic Mult…☆223May 15, 2026Updated 4 months ago
- [NeurIPS'25] Router-R1: Teaching LLMs Multi-Round Routing and Aggregation via Reinforcement Learning☆155Dec 30, 2025Updated 9 months ago