A multi-agent system trained with GRPO for reliable long-horizon task planning and execution.
☆64Feb 9, 2026Updated 7 months ago
Alternatives and similar repositories for multi-agent-training-grpo
Users that are interested in multi-agent-training-grpo are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Contextual Engineering Pipeline☆25Mar 5, 2026Updated 7 months ago
- Outlier Detection with AI + ML☆16Sep 12, 2025Updated last year
- Encountering 14 different Naive RAG fails and using KG to solve it☆33Dec 4, 2025Updated 10 months ago
- Deep research agentic system using Time Test Diffusion☆56Dec 11, 2025Updated 9 months ago
- Forty hands-on recipes for production-grade retrieval-augmented generation, Nebius-first and provider-agnostic.☆29May 29, 2026Updated 4 months ago
- GPUs on demand by Runpod - Special Offer Available • AdRun AI, ML, and HPC workloads on powerful cloud GPUs—without limits or wasted spend. Deploy GPUs in under a minute and pay by the second.
- Training architecture for self-improving AI agents.☆75Nov 4, 2025Updated 11 months ago
- A Curated Collection of LLM resources (work in progress).☆25Jan 9, 2026Updated 8 months ago
- ☆56Jan 8, 2026Updated 8 months ago
- A Step-by-Step Implementation of RAPTOR based RAG implementation☆42Sep 1, 2025Updated last year
- A step by step implementation of a complex AI Evaluation System Designing☆22Sep 9, 2025Updated last year
- Implementation of 12 AI agents evaluation techniques☆48Jul 31, 2025Updated last year
- Building an advanced Agentic eCommerce WhatsApp bot☆24Sep 12, 2025Updated last year
- Core concepts - where to apply parallelism in agentic solution☆100Nov 20, 2025Updated 10 months ago
- Rotating cube in terminal by mouse.☆17Feb 17, 2026Updated 7 months ago
- Virtual machines for every use case on DigitalOcean • AdGet dependable uptime with 99.99% SLA, simple security tools, and predictable monthly pricing with DigitalOcean's virtual machines, called Droplets.
- ☆13May 17, 2025Updated last year
- ☆33Jun 3, 2026Updated 4 months ago
- This is a multi-page streamlit app to showcase Sharone's streamlit app examples☆14Feb 17, 2022Updated 4 years ago
- A variation on a standard Decision Tree such as that in sklearn, where nodes may be based on an aggregation of multiple splits.☆10May 24, 2024Updated 2 years ago
- Python stream processing with RisingWave☆20Aug 7, 2026Updated last month
- This repository provides the code for applying Contrastive Learning Penalty Loss (CLPL) and Mixture of Experts (MoE) to the BGE-M3 text e…☆11Dec 27, 2024Updated last year
- Creating 'deep agents' to encourage LLM's to complete long horizon tasks.☆44Sep 19, 2025Updated last year
- Layered guardrails to make agentic AI safer and more reliable.☆49Oct 5, 2025Updated last year
- 100+ Fine-tuning Tutorial Notebooks on Google Colab, Kaggle and more.☆15Jan 17, 2026Updated 8 months ago
- Proton VPN Special Offer - Get 70% off • AdSpecial partner offer. Trusted by over 100 million users worldwide. Tested, Approved and Recommended by Experts.
- Building a Claude-like agentic system.☆64May 1, 2026Updated 5 months ago
- A Deep Thinking RAG Pipeline to Solve Complex Queries☆128Oct 19, 2025Updated 11 months ago
- ☆16Jul 4, 2026Updated 3 months ago
- UArizona DataLab Workshops☆10Aug 8, 2025Updated last year
- 23 Components of the Claude Code Architecture☆298Apr 5, 2026Updated 6 months ago
- ☆16Nov 10, 2023Updated 2 years ago
- Building a Senior Staff Engineer with Sub-Agent Teams in Claude Code☆93Apr 13, 2026Updated 5 months ago
- ☆16Jun 12, 2024Updated 2 years ago
- Ocaml code from Writing an Interpreter in Go☆11Aug 16, 2019Updated 7 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- Tutorial and examples for using Apache Spark☆18Jul 21, 2017Updated 9 years ago
- Context & Guide For Reinforcement Learning with Verifiable Rewards with Large Language Models☆22Nov 3, 2025Updated 11 months ago
- Outlining and demonstrating how language models are able to understand image, video, and text content.☆18Mar 19, 2025Updated last year
- Streamlit OpenAI app to chat with custom text documents of all kinds☆13Apr 11, 2026Updated 5 months ago
- A detail Implementation of handling long-term memory in Agentic AI☆63Oct 9, 2025Updated 11 months ago
- Applying domain specific evaluations to RAG chunking and embedding functions☆20Dec 25, 2024Updated last year
- Self improving agentic rag pipeline☆228Nov 13, 2025Updated 10 months ago