π» SETA: Scaling Environments for Terminal Agents
β142Jul 28, 2026Updated last month
Alternatives and similar repositories for seta
Users that are interested in seta are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- π» SETA: Scaling Environments for Terminal Agents - Environmentsβ145Feb 16, 2026Updated 6 months ago
- β140Mar 31, 2026Updated 5 months ago
- GRPO training code which scales to 32xH100s for long horizon terminal/coding tasks. Base agent is now the top Qwen3 agent on Stanford's Tβ¦β409Aug 24, 2025Updated last year
- Multi-agent synthetic data generation pipeline capable of generating and validating long horizon terminal/coding tasks for RL trainingβ74Jul 28, 2025Updated last year
- Convert GitHub PRs into Harbor tasksβ81Jul 13, 2026Updated last month
- Bare Metal GPUs on DigitalOcean Gradient AI β’ AdPurpose-built for serious AI teams training foundational models, running large-scale inference, and pushing the boundaries of what's possible.
- SWE-Bench-plus-plusβ25Feb 5, 2026Updated 6 months ago
- Implementation for the paper "Fictitious Synthetic Data Can Improve LLM Factuality via Prerequisite Learning"β11Jan 10, 2025Updated last year
- Framework for evaluating and improving agentsβ4,845Updated this week
- Harness for running and evaluating AI agents against RL environmentsβ250Updated this week
- A Python SDK for Open Reward Standard servers and clientsβ17Mar 24, 2026Updated 5 months ago
- [ICLR 2026] The Tool Decathlon: Benchmarking Language Agents for Diverse, Realistic, and Long-Horizon Task Executionβ475Aug 18, 2026Updated 2 weeks ago
- Official Implementation of "Simulating Environments with Reasoning Models for Agent Training"β68Feb 18, 2026Updated 6 months ago
- Toolathlon-Gym for testing AI agents real-world tool-use capabilities across diverse MCP servers.β146Jul 22, 2026Updated last month
- β122Apr 1, 2026Updated 5 months ago
- Managed Database hosting by DigitalOcean β’ AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- Download Web-10K data by querying Bing Image Searchβ10Feb 1, 2022Updated 4 years ago
- β61May 26, 2026Updated 3 months ago
- π Loong: Synthesize Long CoTs at Scale through Verifiers.β508Updated this week
- SkyRL: A Modular Full-stack RL Library for LLMsβ2,213Updated this week
- Data recipes and robust infrastructure for training AI agentsβ286Updated this week
- [EMNLP 2024 Main] Official implementation of the paper "The Accuracy Paradox in RLHF: When Better Reward Models Don't Yield Better Languaβ¦β13Nov 11, 2024Updated last year
- [COLM 2025] Official repository for R2E-Gym: Procedural Environment Generation and Hybrid Verifiers for Scaling Open-Weights SWE Agentsβ326Jul 13, 2025Updated last year
- Code for Paper: Training Software Engineering Agents and Verifiers with SWE-Gym [ICML 2025]β729Jul 29, 2025Updated last year
- [NeurIPS 2025 D&B Spotlight] Scaling Data for SWE-agentsβ756Updated this week
- AI Agents on DigitalOcean Gradient AI Platform β’ AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- Nemotron-CORTEXA is an open-source software engineering agent that fixes GitHub issues.β26Aug 7, 2025Updated last year
- COLM2026β38Jul 9, 2026Updated last month
- β19Aug 9, 2026Updated 3 weeks ago
- RLAnything (ICML 2026) & AutoTool (ICML 2026), DemyAgent: Open-Source RL for LLMs and Agentic Scenariosβ629Jun 12, 2026Updated 2 months ago
- The RL Bridge for LLM-based Agent Applications. Made Simple & Flexible.β5,708Updated this week
- Agentic RL on Any Harness at Scaleβ823Aug 13, 2026Updated 2 weeks ago
- [NeurIPS'25] Official codebase for "SWE-RL: Advancing LLM Reasoning via Reinforcement Learning on Open Software Evolution"β719Mar 16, 2025Updated last year
- Official implementation of Selective Entropy Regularization (SIREN), proposed by paper 'Rethinking Entropy Regularization in Large Reasonβ¦β32Dec 10, 2025Updated 8 months ago
- open source SWE-Atlasβ70Aug 20, 2026Updated last week
- Deploy to Railway using AI coding agents - Free Credits Offer β’ AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- Official Implementation of ConceptLM.β26Mar 18, 2026Updated 5 months ago
- [FSE'2026] SWE-Factory: Your Automated Factory for Issue Resolution Training Data and Evaluation Benchmarksβ190May 12, 2026Updated 3 months ago
- β41May 16, 2026Updated 3 months ago
- FrontierSmith, a new system that uses AI to synthesize open-ended coding problems at scaleβ52May 30, 2026Updated 3 months ago
- Democratizing Reinforcement Learning for LLMsβ5,811Aug 24, 2026Updated last week
- slime is an LLM post-training framework for RL Scaling.β8,340Updated this week
- A visual representation of Dijkstra's Algorithm using Libgdx.β16Dec 24, 2021Updated 4 years ago