☆108Oct 22, 2025Updated 11 months ago
Alternatives and similar repositories for Awesome-Agentic-RL-Papers
Users that are interested in Awesome-Agentic-RL-Papers are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- ☆30Oct 1, 2025Updated 11 months ago
- ☆1,900Jun 18, 2026Updated 3 months ago
- [ICLR'26] CodeBrain: Bridging Decoupled Tokenizer and Multi-Scale Architecture for EEG Foundation Models☆25Aug 5, 2026Updated last month
- Github repository for "Internalizing World Models via Self-Play Finetuning for Agentic RL" (EMNLP Findings 2026)☆38Nov 1, 2025Updated 10 months ago
- Code for ICLR 2025 Paper "GenARM: Reward Guided Generation with Autoregressive Reward Model for Test-time Alignment"☆25Feb 10, 2025Updated last year
- Virtual machines for every use case on DigitalOcean • AdGet dependable uptime with 99.99% SLA, simple security tools, and predictable monthly pricing with DigitalOcean's virtual machines, called Droplets.
- A Multi-Agent Approach Integrating Socratic Guidance for Automated Prompt Optimization☆18Dec 15, 2025Updated 9 months ago
- The official repository of "A Comprehensive Survey on Reinforcement Learning-based Agentic Search: Foundations, Roles, Optimizations, Eva…☆294Jul 30, 2026Updated last month
- Multiscale Score Matching Analysis☆12Jan 19, 2023Updated 3 years ago
- R1V, trained with AI feedback, answers open-ended visual questions.☆14Apr 12, 2025Updated last year
- ☆18Jan 6, 2025Updated last year
- ☆63Sep 3, 2025Updated last year
- Under construction☆14Jan 15, 2025Updated last year
- Opensource code for ICML 2026 poster☆16Nov 26, 2025Updated 10 months ago
- ZJUT的保研分享库☆39Mar 12, 2025Updated last year
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- Towards Long Form Audio-visual Video Understanding☆15Jan 16, 2026Updated 8 months ago
- ☆42May 26, 2026Updated 4 months ago
- [ICML 2024] Official implementation for the paper "Hierarchical Neural Operator Transformer with Learnable Frequency-aware Loss Prior for…☆16Nov 8, 2024Updated last year
- OpenClaw-RL: Personalize openclaw simply by talking to it☆16Feb 26, 2026Updated 7 months ago
- ☆27Nov 20, 2025Updated 10 months ago
- 2018研究生推免计算机类高校夏令营时间安排☆12May 14, 2018Updated 8 years ago
- [ICLR 2026] Agentic Reinforced Policy Optimization (ARPO)☆1,126Sep 12, 2026Updated 2 weeks ago
- Socratic-Zero is a fully autonomous framework that generates high-quality training data for mathematical reasoning☆37Oct 26, 2025Updated 11 months ago
- Official PyTorch implementation for the ICML 2023 paper "Out-of-Distribution Generalization of Federated Learning via Implicit Invariant …☆14Oct 31, 2023Updated 2 years ago
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- ICDE'24 "Time-aware Graph Structure Learning for Spatiao-temporal Forecasting"☆16Jan 26, 2026Updated 8 months ago
- ☆16Mar 8, 2026Updated 6 months ago
- ☆12Jan 9, 2025Updated last year
- [SIGIR 2022] The implementation of Logiformer☆28Jan 11, 2024Updated 2 years ago
- A novel template-free retrosynthesizer that can generate diverse sets of reactants for a desired product via discrete conditional variati…☆15Aug 7, 2022Updated 4 years ago
- ☆14Jun 3, 2025Updated last year
- Search-R1: An Efficient, Scalable RL Training Framework for Reasoning & Search Engine Calling interleaved LLM based on veRL☆5,454Nov 13, 2025Updated 10 months ago
- CS302 OS Notes☆10Jun 17, 2021Updated 5 years ago
- Survey and paper list on efficiency-guided LLM agents (memory, tool use, planning).☆308Aug 28, 2026Updated last month
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- The official repository for CVPR'26 Paper "APPO: Attention-guided Perception Policy Optimization for Video Reasoning"☆16Mar 19, 2026Updated 6 months ago
- Video Anomaly Detection with Motion and Appearance Guided Patch Diffusion Model☆17Apr 14, 2025Updated last year
- A Collection of Papers about Memory for Language Agents☆661Updated this week
- The code of paper *Learning Robust Policy against Disturbance in Transition Dynamics via State-Conservative Policy Optimization*.☆18Mar 26, 2022Updated 4 years ago
- UAV Reinforcement Learning Air Combat☆21Jan 20, 2025Updated last year
- This is the official GitHub repository for our survey paper "Beyond Single-Turn: A Survey on Multi-Turn Interactions with Large Language …☆212Jul 11, 2026Updated 2 months ago
- Awesome List for Agentic RL☆1,850Sep 15, 2026Updated last week