π» SETA: Scaling Environments for Terminal Agents
β156Jul 28, 2026Updated last month
Alternatives and similar repositories for seta
Users that are interested in seta are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- π» SETA: Scaling Environments for Terminal Agents - Environmentsβ147Feb 16, 2026Updated 7 months ago
- β148Mar 31, 2026Updated 5 months ago
- GRPO training code which scales to 32xH100s for long horizon terminal/coding tasks. Base agent is now the top Qwen3 agent on Stanford's Tβ¦β411Aug 24, 2025Updated last year
- Multi-agent synthetic data generation pipeline capable of generating and validating long horizon terminal/coding tasks for RL trainingβ74Jul 28, 2025Updated last year
- Convert GitHub PRs into Harbor tasksβ83Jul 13, 2026Updated 2 months ago
- 1-Click AI Models by DigitalOcean Gradient β’ AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- SWE-Bench-plus-plusβ25Feb 5, 2026Updated 7 months ago
- Implementation for the paper "Fictitious Synthetic Data Can Improve LLM Factuality via Prerequisite Learning"β11Jan 10, 2025Updated last year
- Framework for evaluating and improving agentsβ5,496Updated this week
- Harness for running and evaluating AI agents against RL environmentsβ279Updated this week
- A Python SDK for Open Reward Standard servers and clientsβ17Mar 24, 2026Updated 5 months ago
- [ICLR 2026] The Tool Decathlon: Benchmarking Language Agents for Diverse, Realistic, and Long-Horizon Task Executionβ489Aug 18, 2026Updated last month
- Official Implementation of "Simulating Environments with Reasoning Models for Agent Training"β68Feb 18, 2026Updated 7 months ago
- Toolathlon-Gym for testing AI agents real-world tool-use capabilities across diverse MCP servers.β147Jul 22, 2026Updated 2 months ago
- β123Apr 1, 2026Updated 5 months ago
- Open source password manager - Proton Pass β’ AdSecurely store, share, and autofill your credentials with Proton Pass, the end-to-end encrypted password manager trusted by millions.
- Download Web-10K data by querying Bing Image Searchβ10Feb 1, 2022Updated 4 years ago
- β62May 26, 2026Updated 3 months ago
- π Loong: Synthesize Long CoTs at Scale through Verifiers.β510Updated this week
- SkyRL: A Modular Full-stack RL Library for LLMsβ2,339Updated this week
- Data recipes and robust infrastructure for training AI agentsβ295Updated this week
- [EMNLP 2024 Main] Official implementation of the paper "The Accuracy Paradox in RLHF: When Better Reward Models Don't Yield Better Languaβ¦β13Nov 11, 2024Updated last year
- [COLM 2025] Official repository for R2E-Gym: Procedural Environment Generation and Hybrid Verifiers for Scaling Open-Weights SWE Agentsβ335Jul 13, 2025Updated last year
- Code for Paper: Training Software Engineering Agents and Verifiers with SWE-Gym [ICML 2025]β742Jul 29, 2025Updated last year
- [NeurIPS 2025 D&B Spotlight] Scaling Data for SWE-agentsβ779Updated this week
- Deploy on Railway without the complexity - Free Credits Offer β’ AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- Nemotron-CORTEXA is an open-source software engineering agent that fixes GitHub issues.β27Aug 7, 2025Updated last year
- A compact high-signal benchmark for evaluating frontier agentsβ39Aug 3, 2026Updated last month
- β20Aug 9, 2026Updated last month
- RLAnything (ICML 2026) & AutoTool (ICML 2026), DemyAgent: Open-Source RL for LLMs and Agentic Scenariosβ641Jun 12, 2026Updated 3 months ago
- The RL Bridge for LLM-based Agent Applications. Made Simple & Flexible.β5,786Updated this week
- Agentic RL on Any Harness at Scaleβ843Aug 13, 2026Updated last month
- [NeurIPS'25] Official codebase for "SWE-RL: Advancing LLM Reasoning via Reinforcement Learning on Open Software Evolution"β720Mar 16, 2025Updated last year
- Official implementation of Selective Entropy Regularization (SIREN), proposed by paper 'Rethinking Entropy Regularization in Large Reasonβ¦β32Dec 10, 2025Updated 9 months ago
- open source SWE-Atlasβ72Aug 20, 2026Updated last month
- Wordpress hosting with auto-scaling - Free Trial Offer β’ AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- Official Implementation of ConceptLM.β28Mar 18, 2026Updated 6 months ago
- [FSE'2026] SWE-Factory: Your Automated Factory for Issue Resolution Training Data and Evaluation Benchmarksβ195May 12, 2026Updated 4 months ago
- β44May 16, 2026Updated 4 months ago
- [COLM 2025] Official repo for Self-Steering Language Modelsβ29Aug 8, 2025Updated last year
- FrontierSmith, a new system that uses AI to synthesize open-ended coding problems at scaleβ52May 30, 2026Updated 3 months ago
- Democratizing Reinforcement Learning for LLMsβ5,832Sep 12, 2026Updated last week
- slime is an LLM post-training framework for RL Scaling.β8,521Updated this week