This is AI implementation (not official) of the DreamGym framework from the paper "Scaling Agent Learning via Experience Synthesis" (arXiv:2511.03773).
☆47Nov 9, 2025Updated 11 months ago
Alternatives and similar repositories for DreamGym
Users that are interested in DreamGym are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Language Models for Code Completion: a Practical Evaluation☆13Jan 19, 2024Updated 2 years ago
- Official Implementation of "Simulating Environments with Reasoning Models for Agent Training"☆68Feb 18, 2026Updated 7 months ago
- Strefer: Empowering Video LLMs with Space-Time Referring and Reasoning via Synthetic Instruction Data☆19Jun 2, 2026Updated 4 months ago
- [ACL 2023 Findings] What In-Context Learning “Learns” In-Context: Disentangling Task Recognition and Task Learning☆21Jul 9, 2023Updated 3 years ago
- [ICML 2025] Official repository for paper "OR-Bench: An Over-Refusal Benchmark for Large Language Models"☆30Mar 4, 2025Updated last year
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- Code Repository for "A Causal Framework to Quantify the Robustness of Mathematical Reasoning with Language Models".☆15Oct 14, 2022Updated 3 years ago
- Implementation of "Investigating the Factual Knowledge Boundary of Large Language Models with Retrieval Augmentation"☆21Jul 31, 2023Updated 3 years ago
- Solving Token Gradient Conflict in Mixture-of-Experts for Large Vision-Language Model☆13Feb 11, 2025Updated last year
- Official code repository for "Web Agents with World Models [ICLR 2025]".☆32Mar 2, 2025Updated last year
- [ACL2026 MainConference]Code Repo for paper "Scaling Behaviors of LLM Reinforcement Learning Post-Training"☆26Jul 1, 2026Updated 3 months ago
- Chain-of-Thought Matters: Improving Long-Context Language Models with Reasoning Path Supervision☆18Apr 1, 2025Updated last year
- A "gym" style toolkit for building lightweight NAS systems.☆13Jun 13, 2022Updated 4 years ago
- Analysis of evidential models☆15Jun 22, 2023Updated 3 years ago
- g2-MLP: State-of-the-Art Model for Node Classification on Graphs (PPI Dataset)☆10Nov 12, 2022Updated 3 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- ☆15Aug 12, 2022Updated 4 years ago
- [EMNLP26 Findings] Official repository for DataChef: Cooking Up Optimal Data Recipes for LLM Adaptation via Reinforcement Learning☆27Feb 12, 2026Updated 7 months ago
- ☆147Mar 31, 2026Updated 6 months ago
- Source code for NeurIPS 2020 paper "Node Classification on Graphs with Few-Shot Novel Labels via Meta Transformed Network Embedding"☆10Nov 17, 2020Updated 5 years ago
- a novel self-evolving paradigm, without task, reward, or complex workflow☆36May 12, 2026Updated 4 months ago
- Search movies using RAG and LLMs☆19Sep 4, 2024Updated 2 years ago
- ☆34Sep 19, 2025Updated last year
- Dynamic dual-granularity skill bank for agentic RL, jointly evolving policy and skills to improve long-horizon decision making in agentic…☆71Apr 1, 2026Updated 6 months ago
- ☆44Jun 30, 2026Updated 3 months ago
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- ☆13Sep 26, 2024Updated 2 years ago
- A Python toolkit for the OmniLabel benchmark providing code for evaluation and visualization☆23Feb 1, 2025Updated last year
- ☆14Aug 10, 2023Updated 3 years ago
- ☆14May 20, 2022Updated 4 years ago
- Code for Paper: Benchmarking Multi-step Scientific Tool-use in LLM Agents☆47Jul 5, 2026Updated 3 months ago
- ☆13Jul 12, 2022Updated 4 years ago
- ☆20Dec 14, 2024Updated last year
- Network Together: Node Classification via Cross-Network Deep Network Embedding☆11May 5, 2021Updated 5 years ago
- ☆38Apr 3, 2026Updated 6 months ago
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- Courses on Deep Reinforcement Learning (DRL) and DRL papers for recommender systems☆13Jul 7, 2022Updated 4 years ago
- Code & experiments for MINDWALC: Mining Interpretable, Discriminative Walks for Classification of Nodes in a Graph☆13Jul 4, 2024Updated 2 years ago
- ☆32Feb 11, 2026Updated 7 months ago
- Official repo of Toucan: Synthesizing 1.5M Tool-Agentic Data from Real-World MCP Environments☆268Dec 16, 2025Updated 9 months ago
- The official implementation of NOSA (EMNLP 2026 main)☆19Sep 28, 2026Updated last week
- Official implement of RAHG: A Role-Aware Hypergraph Neural Network for Node Classification in Graphs.☆11Jul 5, 2024Updated 2 years ago
- JudgeLRM: Large Reasoning Models as a Judge☆42May 6, 2026Updated 5 months ago