☆107Oct 22, 2025Updated 10 months ago
Alternatives and similar repositories for Awesome-Agentic-RL-Papers
Users that are interested in Awesome-Agentic-RL-Papers are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- RISC-V Sv39 Page Table Entry Visualization Tool☆17May 20, 2025Updated last year
- Implement and train a Tiny LLM from scratch!☆22Jul 4, 2025Updated last year
- ☆1,886Jun 18, 2026Updated 2 months ago
- [ICLR'26] CodeBrain: Bridging Decoupled Tokenizer and Multi-Scale Architecture for EEG Foundation Models☆25Aug 5, 2026Updated last month
- [ACM MM2026] This is the official implementation of MedCCO☆17Jul 12, 2026Updated last month
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- Agent Skill Evaluation and Evolution: Frameworks and Benchmarks☆29Jul 15, 2026Updated last month
- Code for ICLR 2025 Paper "GenARM: Reward Guided Generation with Autoregressive Reward Model for Test-time Alignment"☆25Feb 10, 2025Updated last year
- 东南大学计软智部分课程作业。你的时间值得更有价值的事。☆184Jun 7, 2025Updated last year
- A Multi-Agent Approach Integrating Socratic Guidance for Automated Prompt Optimization☆18Dec 15, 2025Updated 8 months ago
- Local-first interview recording review reports with a Codex skill and CLI.☆81May 16, 2026Updated 3 months ago
- [ACL-2025-Findings] The official GitHub repo for the paper "Exploring Multimodal Challenges in Toxic Chinese Detection: Taxonomy, Benchma…☆23Jun 8, 2025Updated last year
- Weighted Reverse Convolution for Feature Upsampling☆24May 24, 2026Updated 3 months ago
- verl-agent is an extension of veRL, designed for training LLM/VLM agents via RL. verl-agent is also the official code for paper "Group-in…☆2,282Jun 9, 2026Updated 2 months ago
- This repository is aim to reproduce the R1-Zero on medical domain.☆32Jun 11, 2025Updated last year
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- ☆17Jul 10, 2026Updated last month
- An extensible RL framework for training LLM agents with advanced search capabilities, built on VERL and supporting state-of-the-art searc…☆37Dec 1, 2025Updated 9 months ago
- ☆63Sep 3, 2025Updated last year
- ☆37Oct 4, 2025Updated 11 months ago
- Under construction☆14Jan 15, 2025Updated last year
- Opensource code for ICML 2026 poster☆16Nov 26, 2025Updated 9 months ago
- ZJUT的保研分享库☆38Mar 12, 2025Updated last year
- Artificial Intelligence (AI) based Portfolio Selection Papers☆30Nov 7, 2025Updated 10 months ago
- ☆41May 26, 2026Updated 3 months ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- OpenClaw-RL: Personalize openclaw simply by talking to it☆16Feb 26, 2026Updated 6 months ago
- ☆27Nov 20, 2025Updated 9 months ago
- [ICLR 2026] Agentic Reinforced Policy Optimization (ARPO)☆1,114Aug 20, 2026Updated 2 weeks ago
- Applying Evolutionary Computing to Embeddings of Diffusion Models☆16Jun 6, 2026Updated 3 months ago
- Socratic-Zero is a fully autonomous framework that generates high-quality training data for mathematical reasoning☆37Oct 26, 2025Updated 10 months ago
- Official PyTorch implementation for the ICML 2023 paper "Out-of-Distribution Generalization of Federated Learning via Implicit Invariant …☆14Oct 31, 2023Updated 2 years ago
- ☆16Mar 8, 2026Updated 6 months ago
- ☆67Jul 12, 2025Updated last year
- ☆12Jan 9, 2025Updated last year
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- In-Context Reinforcement Learning for Tool Use in Large Language Models☆48Mar 26, 2026Updated 5 months ago
- Personal project about cpp multiplatform ui framework.☆15Mar 8, 2026Updated 6 months ago
- ☆14Jun 3, 2025Updated last year
- Search-R1: An Efficient, Scalable RL Training Framework for Reasoning & Search Engine Calling interleaved LLM based on veRL☆5,385Nov 13, 2025Updated 9 months ago
- ☆13Jul 14, 2024Updated 2 years ago
- Official github repo for SafeDialBench, a comprehensive multi-turn dialogue benchmark to evaluate LLMs' safety.☆57May 12, 2025Updated last year
- Survey and paper list on efficiency-guided LLM agents (memory, tool use, planning).☆305Aug 28, 2026Updated last week