π» SETA: Scaling Environments for Terminal Agents
β134Jul 28, 2026Updated last week
Alternatives and similar repositories for seta
Users that are interested in seta are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- π» SETA: Scaling Environments for Terminal Agents - Environmentsβ143Feb 16, 2026Updated 5 months ago
- β137Mar 31, 2026Updated 4 months ago
- GRPO training code which scales to 32xH100s for long horizon terminal/coding tasks. Base agent is now the top Qwen3 agent on Stanford's Tβ¦β403Aug 24, 2025Updated 11 months ago
- Multi-agent synthetic data generation pipeline capable of generating and validating long horizon terminal/coding tasks for RL trainingβ72Jul 28, 2025Updated last year
- Convert GitHub PRs into Harbor tasksβ74Jul 13, 2026Updated 3 weeks ago
- Deploy on Railway without the complexity - Free Credits Offer β’ AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- SWE-Bench-plus-plusβ25Feb 5, 2026Updated 6 months ago
- Implementation for the paper "Fictitious Synthetic Data Can Improve LLM Factuality via Prerequisite Learning"β11Jan 10, 2025Updated last year
- Framework for evaluating and improving agentsβ4,076Updated this week
- Harness for running and evaluating AI agents against RL environmentsβ232Updated this week
- [ICLR 2026] The Tool Decathlon: Benchmarking Language Agents for Diverse, Realistic, and Long-Horizon Task Executionβ449Updated this week
- Official Implementation of "Simulating Environments with Reasoning Models for Agent Training"β66Feb 18, 2026Updated 5 months ago
- Toolathlon-Gym for testing AI agents real-world tool-use capabilities across diverse MCP servers.β142Jul 22, 2026Updated 2 weeks ago
- β121Apr 1, 2026Updated 4 months ago
- Download Web-10K data by querying Bing Image Searchβ10Feb 1, 2022Updated 4 years ago
- Wordpress hosting with auto-scaling - Free Trial Offer β’ AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- β58May 26, 2026Updated 2 months ago
- π Loong: Synthesize Long CoTs at Scale through Verifiers.β506Updated this week
- SkyRL: A Modular Full-stack RL Library for LLMsβ2,138Updated this week
- Data recipes and robust infrastructure for training AI agentsβ278Updated this week
- [EMNLP 2024 Main] Official implementation of the paper "The Accuracy Paradox in RLHF: When Better Reward Models Don't Yield Better Languaβ¦β13Nov 11, 2024Updated last year
- [COLM 2025] Official repository for R2E-Gym: Procedural Environment Generation and Hybrid Verifiers for Scaling Open-Weights SWE Agentsβ317Jul 13, 2025Updated last year
- Code for Paper: Training Software Engineering Agents and Verifiers with SWE-Gym [ICML 2025]β720Jul 29, 2025Updated last year
- [NeurIPS 2025 D&B Spotlight] Scaling Data for SWE-agentsβ732Aug 3, 2026Updated last week
- Nemotron-CORTEXA is an open-source software engineering agent that fixes GitHub issues.β25Aug 7, 2025Updated last year
- End-to-end encrypted cloud storage - Proton Drive β’ AdSpecial offer: 40% Off Yearly / 80% Off First Month. Protect your most important files, photos, and documents from prying eyes.
- COLM2026β36Jul 9, 2026Updated last month
- A compact high-signal benchmark for evaluating frontier agentsβ23Aug 3, 2026Updated last week
- β18Updated this week
- The RL Bridge for LLM-based Agent Applications. Made Simple & Flexible.β5,653Updated this week
- Agentic RL on Any Harness at Scaleβ760Updated this week
- [NeurIPS'25] Official codebase for "SWE-RL: Advancing LLM Reasoning via Reinforcement Learning on Open Software Evolution"β715Mar 16, 2025Updated last year
- Official implementation of Selective Entropy Regularization (SIREN), proposed by paper 'Rethinking Entropy Regularization in Large Reasonβ¦β32Dec 10, 2025Updated 8 months ago
- open source SWE-Atlasβ66Jul 20, 2026Updated 3 weeks ago
- Official Implementation of ConceptLM.β23Mar 18, 2026Updated 4 months ago
- Simple, predictable pricing with DigitalOcean hosting β’ AdAlways know what you'll pay with monthly caps and flat pricing. Enterprise-grade infrastructure trusted by 600k+ customers.
- [FSE'2026] SWE-Factory: Your Automated Factory for Issue Resolution Training Data and Evaluation Benchmarksβ184May 12, 2026Updated 2 months ago
- β38May 16, 2026Updated 2 months ago
- [COLM 2025] Official repo for Self-Steering Language Modelsβ27Aug 8, 2025Updated last year
- FrontierSmith, a new system that uses AI to synthesize open-ended coding problems at scaleβ50May 30, 2026Updated 2 months ago
- Democratizing Reinforcement Learning for LLMsβ5,774Updated this week
- slime is an LLM post-training framework for RL Scaling.β7,832Updated this week
- A visual representation of Dijkstra's Algorithm using Libgdx.β16Dec 24, 2021Updated 4 years ago