[ICLR2026] The official repository for the CodeGym project: "Generalizable End-to-End Tool-Use RL with Synthetic CodeGym"
☆41Oct 14, 2025Updated 11 months ago
Alternatives and similar repositories for CodeGym
Users that are interested in CodeGym are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- [ICML2025] Official Repo for Paper "Optimizing Temperature for Language Models with Multi-Sample Inference"☆23Feb 16, 2025Updated last year
- MUA-RL: MULTI-TURN USER-INTERACTING AGENT REINFORCEMENT LEARNING FOR AGENTIC TOOL USE☆68Nov 5, 2025Updated 10 months ago
- Efficient Scaling laws and collaborative pretraining.☆24Jul 19, 2026Updated 2 months ago
- [COLM-LLA 2026] The official implementation for paper "AdaExplore: Failure-Driven Adaptation and Diversity-Preserving Search for Efficien…☆24Aug 23, 2026Updated last month
- [EMNLP 2025] The official implementation for paper "Agentic-R1: Distilled Dual-Strategy Reasoning"☆106Apr 21, 2026Updated 5 months ago
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- 🚀 LLM-I: Transform LLMs into natural interleaved multimodal creators! ✨ Tool-use framework supporting image search, generation, code ex…☆41Oct 20, 2025Updated 11 months ago
- This repository includes code and materials for the paper "Efficient PRM Training Data Synthesis via Formal Verification" (ACL 2026 Findi…☆20Apr 7, 2026Updated 5 months ago
- VitaBench: Benchmarking LLM Agents with Versatile Interactive Tasks in Real-world Applications☆23Oct 17, 2025Updated 11 months ago
- ☆11Apr 29, 2023Updated 3 years ago
- ☆11Oct 7, 2024Updated last year
- [EMNLP 2025] CompassVerifier: A Unified and Robust Verifier for LLMs Evaluation and Outcome Reward☆70Aug 10, 2025Updated last year
- [ICML'26] Learning Query-Aware Budget-Tier Routing for Runtime Agent Memory☆23Jun 10, 2026Updated 3 months ago
- [ICLR 2023] The official code for paper "Guarded Policy Optimization with Imperfect Online Demonstrations"☆14Apr 30, 2023Updated 3 years ago
- ☆17Aug 1, 2025Updated last year
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- [ASE 2025] CoSIL: Issue Localization via Iteritive Code Graph Searching☆26May 31, 2026Updated 3 months ago
- [EMNLP 2024 Findings] ProSA: Assessing and Understanding the Prompt Sensitivity of LLMs☆30May 22, 2025Updated last year
- ☆19Feb 25, 2024Updated 2 years ago
- Yet another LLM☆10Apr 6, 2023Updated 3 years ago
- Scalable Meta-Evaluation of LLMs as Evaluators☆43Feb 15, 2024Updated 2 years ago
- Official code implementation for the ACL 2025 paper: 'Dynamic Scaling of Unit Tests for Code Reward Modeling'☆26May 16, 2025Updated last year
- The official repo of "WebExplorer: Explore and Evolve for Training Long-Horizon Web Agents"☆123Sep 29, 2025Updated 11 months ago
- [ICLR'2026] R-HORIZON: How Far Can Your Large Reasoning Model Really Go in Breadth and Depth?☆18Oct 21, 2025Updated 11 months ago
- ☆11Oct 25, 2024Updated last year
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- GNOME Shell Extensions - Backup Tools / 备份工具☆16Jul 19, 2020Updated 6 years ago
- Code for paper "Concrete Subspace Learning based Interference Elimination for Multi-task Model Fusion"☆14Mar 28, 2024Updated 2 years ago
- finding new ramsey bounds through scaling autoresearch☆49May 13, 2026Updated 4 months ago
- ICML 2022, Finding Global Homophily in Graph Neural Networks When Meeting Heterophily☆45Jul 25, 2022Updated 4 years ago
- [ICLR 2026] End-to-End Reinforcement Learning for Multi-Turn Tool-Integrated Reasoning☆405Mar 30, 2026Updated 5 months ago
- Open-source repository for the OOPSLA'24 paper "CYCLE: Learning to Self-Refine Code Generation"☆10Mar 8, 2024Updated 2 years ago
- Code for paper "Towards Efficient Pareto Set Approximation via Weight-Ensembling Mixture of Experts"☆11Jul 30, 2026Updated last month
- ☆17Feb 4, 2025Updated last year
- Official code and dataset for our paper: RefineBench: Evaluating Refinement Capability of Language Models via Checklists☆17Dec 1, 2025Updated 9 months ago
- End-to-end encrypted cloud storage - Proton Drive • AdSpecial offer: 40% Off Yearly / 80% Off First Month. Protect your most important files, photos, and documents from prying eyes.
- ☆27Jun 10, 2025Updated last year
- ☆91Aug 16, 2025Updated last year
- [EMNLP'25 Industry] Repo for "Z1: Efficient Test-time Scaling with Code"☆69Apr 11, 2025Updated last year
- [ICLR 2026] BARREL: Boundary-Aware Reasoning for Factual and Reliable LRMs☆18May 21, 2025Updated last year
- Video-R2: Reinforcing Consistent and Grounded Reasoning in Multimodal Language Models☆19Sep 5, 2026Updated 2 weeks ago
- code for COLING paper "A Hybrid Model of Classification and Generation for Spatial Relation Extraction"☆10Oct 20, 2022Updated 3 years ago
- official repo for the paper "Learning From Mistakes Makes LLM Better Reasoner"☆62Dec 20, 2023Updated 2 years ago