COS-PLAY: Co-Evolving LLM Decision and Skill Bank Agents for Long-Horizon Game Play
☆30Jul 11, 2026Updated last month
Alternatives and similar repositories for COS-PLAY
Users that are interested in COS-PLAY are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Synthetic Video hallucination and Mitigation☆25Sep 21, 2025Updated 11 months ago
- Self-evolving vision language models from zero data☆81Mar 14, 2026Updated 5 months ago
- [ICML'26] VideoGPA is a self-supervised framework that enhances 3D consistency in Video Diffusion Models.☆71Jun 6, 2026Updated 2 months ago
- ☆31Apr 11, 2026Updated 4 months ago
- ☆23Dec 17, 2024Updated last year
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- ☆73Apr 28, 2026Updated 3 months ago
- Embodied-Planner-R1: Unleashing Embodied Task Planning Ability in LLMs via Reinforcement Learning☆28Mar 30, 2026Updated 4 months ago
- ☆27May 1, 2026Updated 3 months ago
- Reinforcement Learning of Vision Language Models with Self Visual Perception Reward☆181Mar 14, 2026Updated 5 months ago
- SkillX: Automatically Constructing Skill Knowledge Bases for Agents☆277Updated this week
- Heterogeneous Multi-agent Version of Highway-env☆17Jun 28, 2023Updated 3 years ago
- ☆14Oct 17, 2024Updated last year
- SR²AM: Efficient Agentic Reasoning Through Self-Regulated Simulative Planning☆21May 22, 2026Updated 3 months ago
- Memory Transfer Learning: How Memories are Transferred Across Domains in Coding Agents☆31Apr 16, 2026Updated 4 months ago
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- ☆48Apr 8, 2026Updated 4 months ago
- A Bi-Level Multi-Agent LLM System for Internet-Scale Information Search and Extraction☆41Aug 8, 2026Updated 2 weeks ago
- Video Content Customization Using First Frame☆193Mar 17, 2026Updated 5 months ago
- [ICCV 2025] MMReason, MLLMs, step by step, reasoning benchmark, AGI☆15Apr 25, 2026Updated 3 months ago
- Offcial Repo of Paper "Eliminating Position Bias of Language Models: A Mechanistic Approach""☆23Jun 13, 2025Updated last year
- Source Code for Paper "Large Language Models are Few-Shot Summarizers: Multi-Intent Comment Generation via In-Context Learning"☆19Jun 9, 2023Updated 3 years ago
- [ICML 2025] Closed-Loop Long-Horizon Robotic Planning via Equilibrium Sequence Modeling☆13May 5, 2025Updated last year
- To Trust Or Not To Trust Your Vision-Language Model's Prediction☆15May 30, 2025Updated last year
- ☆11May 17, 2024Updated 2 years ago
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- This respository is used for time reasoning task for mult-session dialogue system.☆18Feb 7, 2026Updated 6 months ago
- ☆65Jul 3, 2026Updated last month
- A Workbench for Autograding Retrieve/Generate Systems☆15Jun 30, 2025Updated last year
- Some microbenchmarks and design docs before commencement☆11Feb 1, 2021Updated 5 years ago
- [ACL 2025 Findings] Text2World: Benchmarking Large Language Models for Symbolic World Model Generation☆29Feb 25, 2025Updated last year
- Official implementation of VLAA-GUI series☆35Jun 20, 2026Updated 2 months ago
- [EMNLP 2025] The official implementation for paper "Agentic-R1: Distilled Dual-Strategy Reasoning"☆105Apr 21, 2026Updated 4 months ago
- ☆20May 25, 2026Updated 2 months ago
- ☆15Feb 25, 2026Updated 5 months ago
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- [ICLR '26 W] Diagnosing Retrieval vs. Utilization Bottlenecks in LLM Agent Memory https://arxiv.org/abs/2603.02473☆17Mar 15, 2026Updated 5 months ago
- [ICML 2026] Official Implementation of Prism: Efficient Test-Time Scaling via Hierarchical Search and Self-Verification for Discrete Diff…☆22Mar 4, 2026Updated 5 months ago
- ☆18Sep 22, 2024Updated last year
- Code for the paper Xiangqi-R1: Enhancing Spatial Strategic Reasoning in LLMs for Chinese Chess via Reinforcement Learning☆15Jul 23, 2025Updated last year
- [ICLR-2026] Official Implementation of our paper "THOR: Tool-Integrated Hierarchical Optimization via RL for Mathematical Reasoning".☆32Aug 4, 2026Updated 2 weeks ago
- Evaluating the Factuality of Large Language Models using Large-Scale Knowledge Graphs☆35Sep 3, 2024Updated last year
- 欢迎参加中文讽刺计算评测任务!☆14Nov 4, 2024Updated last year