Qwen3-14B Orchestrator Agent Reinforcement Learning. **Achieved 160% improvement** on Stanford's TerminalBench
☆102Nov 3, 2025Updated 9 months ago
Alternatives and similar repositories for Orca-Agent-RL
Users that are interested in Orca-Agent-RL are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- An unofficial implementation of SOLAR-10.7B model and the newly proposed interlocked-DUS(iDUS) implementation and experiment details.☆14Mar 20, 2024Updated 2 years ago
- A repo for my CCN Coding Club talk, 'Python, Rust, and You: Modern Py-Rust Interoperation'☆45Aug 12, 2025Updated last year
- Next paradigm for LLM Agent. Unify plan and action through recursive code generation for adaptive, human-like decision-making.☆562Apr 21, 2026Updated 3 months ago
- Datastar Common Lisp SDK☆64Updated this week
- A pipeline engine library for PHP which allows to implement repetitive workflows in your applications, by allowing even non-developers to…☆70Nov 10, 2025Updated 9 months ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- Verifiers for LLM Reinforcement Learning☆82Jul 30, 2026Updated 2 weeks ago
- Benchmarking Goal-Oriented Software Engineering☆201Jul 16, 2026Updated last month
- Intelligent Model Context Protocol (MCP) server for AI-assisted API development. Generate mock servers from OpenAPI specs with advanced l…☆16Updated this week
- Official Implementation of Papar CM2☆26Apr 21, 2026Updated 3 months ago
- llama.cpp fork with additional SOTA quants and improved performance☆22Updated this week
- Data recipes and robust infrastructure for training AI agents☆281Updated this week
- ☆12Mar 3, 2023Updated 3 years ago
- This repository collects and organises state‑of‑the‑art papers on spatial reasoning for Multimodal Vision–Language Models (MVLMs).☆320Feb 17, 2026Updated 6 months ago
- A benchmarking tool for evaluating AI coding assistants on real-world software engineering tasks from the SWE-Bench dataset.☆61Jan 22, 2026Updated 6 months ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- Submodule of evalverse forked from [google-research/instruction_following_eval](https://github.com/google-research/google-research/tree/m…☆15May 4, 2024Updated 2 years ago
- Run Claude Code/Codex within AgentFS, orchestrated by LlamaIndex Workflows☆324Dec 19, 2025Updated 7 months ago
- GRPO training code which scales to 32xH100s for long horizon terminal/coding tasks. Base agent is now the top Qwen3 agent on Stanford's T…☆406Aug 24, 2025Updated 11 months ago
- A web-based client for connecting to MCP servers with OAuth support☆17Jul 29, 2026Updated 3 weeks ago
- Repo for collaboration on OSS agentic code search☆83Apr 29, 2026Updated 3 months ago
- a conversational finance assistant that provides users with real-time stock quotes, market news, and insights on market movers through na…☆17Apr 26, 2025Updated last year
- ☆16May 21, 2026Updated 2 months ago
- Semi-Structured Agentic Framework. Workflows build themselves as agents discover what needs to be done, not what you predicted upfront.☆1,183Dec 1, 2025Updated 8 months ago
- SLIM Models by LLMWare. A streamlit app showing the capabilities for AI Agents and Function Calls.☆21Feb 11, 2024Updated 2 years ago
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- 🍨 Gelato — From Data Curation to Reinforcement Learning: Building a Strong Grounding Model for Computer-Use Agents☆46Dec 22, 2025Updated 7 months ago
- Training an LLM to use a calculator with multi-turn reinforcement learning, achieving a **62% absolute increase in evaluation accuracy**.☆76May 5, 2025Updated last year
- Running LLMs against a sandbox airport to see if they can make the correct decisions in real time☆28Jul 22, 2025Updated last year
- A fast, helpful, and open-source document parser☆18Apr 23, 2026Updated 3 months ago
- Official implementation for the paper, StackEval: Benchmarking LLMs in Coding Assistance, https://arxiv.org/abs/2412.05288☆21Oct 30, 2024Updated last year
- Educational package-manager project for AI agent skills. Not actively maintained.☆450Aug 12, 2026Updated last week
- architect-agent☆19Dec 31, 2025Updated 7 months ago
- Synthetic Data Generation for Evaluation☆16Feb 21, 2025Updated last year
- GPU-optimized framework for training diffusion language models at any scale. The backend of Quokka, Super Data Learners, and OpenMoE 2 tr…☆343Nov 11, 2025Updated 9 months ago
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- The open-source adapter for working with RDF databases and SPARQL queries in Jupyter notebooks leveraging the yFiles Graphs for Jupyter p…☆24Apr 4, 2025Updated last year
- ☆16Jun 30, 2026Updated last month
- A tutorial showcasing the process of scaling up a SciML algorithm for multi-node training☆16Sep 27, 2025Updated 10 months ago
- Provides integration of any AI agent (Gemini CLI, Claude Code, Codex CLI, etc.) with the Unity editor using Agent Client Protocol☆274Nov 5, 2025Updated 9 months ago
- Run Claude Agent (Claude Code) in a sandbox, control it via websocket☆582Dec 28, 2025Updated 7 months ago
- An extention to the GaLore paper, to perform Natural Gradient Descent in low rank subspace☆19Oct 21, 2024Updated last year
- Enhanced Supertonic TTS with Docker, FastAPI, Web UI, and comprehensive API documentation☆21Dec 7, 2025Updated 8 months ago