RETROAGENT: From Solving to Evolving via Retrospective Dual Intrinsic Feedback
☆29Mar 30, 2026Updated 4 months ago
Alternatives and similar repositories for RetroAgent
Users that are interested in RetroAgent are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Co-evolving policy actors and experience extractors for efficient experience-driven agent RL☆51May 12, 2026Updated 3 months ago
- ☆23Updated this week
- [arxiv: 2604.14142] From P(y|x) to P(y): Investigating Reinforcement Learning in Pre-train Space☆17Apr 16, 2026Updated 4 months ago
- Official Repo for DAC-RL: Training LLMs for Divide-and-Conquer Reasoning Elevates Test-Time Scalability☆16Feb 26, 2026Updated 6 months ago
- ☆62Apr 9, 2026Updated 4 months ago
- Virtual machines for every use case on DigitalOcean • AdGet dependable uptime with 99.99% SLA, simple security tools, and predictable monthly pricing with DigitalOcean's virtual machines, called Droplets.
- A Massive Multi-Discipline Lecture Understanding Benchmark☆34Apr 20, 2026Updated 4 months ago
- Build coherent and visually polished multimodal webpages with hierarchical planning, AIGC tools, and iterative reflection.☆17May 17, 2026Updated 3 months ago
- MarsRL: Advancing Multi-Agent Reasoning System via Reinforcement Learning with Agentic Pipeline Parallelism☆18Nov 18, 2025Updated 9 months ago
- ☆15Apr 17, 2026Updated 4 months ago
- MemOCR: an OCR-driven visual memory agent.☆33May 17, 2026Updated 3 months ago
- [NeurIPS 2025] The implementation of paper "On Reasoning Strength Planning in Large Reasoning Models"☆31Jul 6, 2025Updated last year
- EVA: Efficient Reinforcement Learning for End-to-End Video Agent☆25May 6, 2026Updated 3 months ago
- [ACL'26 Oral] AgentOCR is a token-efficient framework that compresses multi-turn agent history by rendering it into images and adopting R…☆45Mar 1, 2026Updated 5 months ago
- On Path to Multimodal Generalist: General-Level and General-Bench☆21Jul 11, 2025Updated last year
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- Will Pre-Training Ever End? A First Step Toward Next-Generation Foundation MLLMs via Self-Improving Systematic Cognition☆31May 14, 2025Updated last year
- An approach to utomatically generating browser environment with verifiable tasks☆68Mar 24, 2026Updated 5 months ago
- ☆17Apr 11, 2025Updated last year
- ☆27May 12, 2026Updated 3 months ago
- Official implementation of the ΔBelief-RL method.☆31Feb 28, 2026Updated 5 months ago
- ConfAgents: A Conformal-Guided Multi-Agent Framework for Cost-Efficient Medical Diagnosis☆15Jul 22, 2026Updated last month
- DLLM-Searcher has been accepted by SIGIR 2026! 🥳☆34Jan 23, 2026Updated 7 months ago
- ☆50Feb 4, 2026Updated 6 months ago
- [EMNLP 2026] MDM-Prime-v2: Binary Encoding and Index Shuffling Enable Scaling of Diffusion Language Models☆29May 23, 2026Updated 3 months ago
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- 上朝式科研:AI-powered research workflow showcase☆131Apr 7, 2026Updated 4 months ago
- SearchEyes: Towards Frontier Multimodal Deep Search Intelligence via Search World Simulation. A typed knowledge graph unifies data synthe…☆23Jul 8, 2026Updated last month
- The official codebase for "Experiential Reinforcement Learning" - https://arxiv.org/pdf/2602.13949v1☆77Jul 2, 2026Updated last month
- Implementation for Bridging Semantic and Kinematic Conditions with Diffusion-based Discrete Motion Tokenizer.☆36Mar 20, 2026Updated 5 months ago
- An ns-3 module for simulations of power line communication networks☆26Nov 16, 2022Updated 3 years ago
- ☆20Aug 28, 2025Updated 11 months ago
- ☆21May 1, 2026Updated 3 months ago
- Reproduction code for paper "MineExplorer: Evaluating Open-World Exploration of MLLM Agents in Minecraft"☆20Jun 12, 2026Updated 2 months ago
- An implementation of parameter server framework in PyTorch RPC.☆12Nov 12, 2021Updated 4 years ago
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- ☆45Jan 16, 2026Updated 7 months ago
- Cross-Attention Guided Loss-Based Deep Dual-Branch Fusion Network for Liver Tumor Classification☆16Sep 26, 2024Updated last year
- This repository includes code for our paper: ResRL: Boosting LLM Reasoning via Negative Sample Projection Residual Reinforcement Learning…☆15May 2, 2026Updated 3 months ago
- 🌟 SwarmAgent: A framework for simulating social group dynamics using multi-agent collaboration, aiding insights into collective behavior…☆13Dec 5, 2023Updated 2 years ago
- UniDoc-RL: Unified Document Understanding with Reinforcement Learning☆17May 21, 2026Updated 3 months ago
- Dynamic dual-granularity skill bank for agentic RL, jointly evolving policy and skills to improve long-horizon decision making in agentic…☆69Apr 1, 2026Updated 4 months ago
- ☆41Aug 11, 2026Updated 2 weeks ago