RETROAGENT: From Solving to Evolving via Retrospective Dual Intrinsic Feedback
☆31Mar 30, 2026Updated 6 months ago
Alternatives and similar repositories for RetroAgent
Users that are interested in RetroAgent are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Co-evolving policy actors and experience extractors for efficient experience-driven agent RL☆52May 12, 2026Updated 4 months ago
- ☆23Aug 25, 2026Updated last month
- [NeurIPS 2026 Pre-to-Post ] From P(y|x) to P(y): Investigating Reinforcement Learning in Pre-train Space☆18Apr 16, 2026Updated 5 months ago
- ☆14Aug 21, 2025Updated last year
- A Massive Multi-Discipline Lecture Understanding Benchmark☆34Apr 20, 2026Updated 5 months ago
- Deploy open-source AI quickly and easily - Special Bonus Offer • AdRunpod Hub is built for open source. One-click deployment and autoscaling endpoints without provisioning your own infrastructure.
- Official Repo for DAC-RL: Training LLMs for Divide-and-Conquer Reasoning Elevates Test-Time Scalability☆40Sep 20, 2026Updated 2 weeks ago
- Build coherent and visually polished multimodal webpages with hierarchical planning, AIGC tools, and iterative reflection.☆18May 17, 2026Updated 4 months ago
- MarsRL: Advancing Multi-Agent Reasoning System via Reinforcement Learning with Agentic Pipeline Parallelism☆18Nov 18, 2025Updated 10 months ago
- ☆15Apr 17, 2026Updated 5 months ago
- MemOCR: an OCR-driven visual memory agent.☆35May 17, 2026Updated 4 months ago
- [NeurIPS 2025] The implementation of paper "On Reasoning Strength Planning in Large Reasoning Models"☆33Jul 6, 2025Updated last year
- EVA: Efficient Reinforcement Learning for End-to-End Video Agent☆26May 6, 2026Updated 4 months ago
- [ACL'26 Oral] AgentOCR is a token-efficient framework that compresses multi-turn agent history by rendering it into images and adopting R…☆50Mar 1, 2026Updated 7 months ago
- On Path to Multimodal Generalist: General-Level and General-Bench☆22Jul 11, 2025Updated last year
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- Will Pre-Training Ever End? A First Step Toward Next-Generation Foundation MLLMs via Self-Improving Systematic Cognition☆31May 14, 2025Updated last year
- An approach to utomatically generating browser environment with verifiable tasks☆72Mar 24, 2026Updated 6 months ago
- ☆17Apr 11, 2025Updated last year
- Official implementation of the ΔBelief-RL method.☆31Feb 28, 2026Updated 7 months ago
- ConfAgents: A Conformal-Guided Multi-Agent Framework for Cost-Efficient Medical Diagnosis☆15Jul 22, 2026Updated 2 months ago
- DLLM-Searcher has been accepted by SIGIR 2026! 🥳☆34Jan 23, 2026Updated 8 months ago
- [EMNLP 2026] MDM-Prime-v2: Binary Encoding and Index Shuffling Enable Scaling of Diffusion Language Models☆30Sep 11, 2026Updated 3 weeks ago
- [ACL 2025 Findings] GenS: Generative Frame Sampler for Long Video Understanding☆22Aug 21, 2025Updated last year
- 上朝式科研:AI-powered research workflow showcase☆134Apr 7, 2026Updated 5 months ago
- Proton VPN Special Offer - Get 70% off • AdSpecial partner offer. Trusted by over 100 million users worldwide. Tested, Approved and Recommended by Experts.
- SearchEyes: Towards Frontier Multimodal Deep Search Intelligence via Search World Simulation. A typed knowledge graph unifies data synthe…☆26Jul 8, 2026Updated 2 months ago
- Implementation for Bridging Semantic and Kinematic Conditions with Diffusion-based Discrete Motion Tokenizer.☆38Mar 20, 2026Updated 6 months ago
- The official codebase for "Experiential Reinforcement Learning" - https://arxiv.org/pdf/2602.13949v1☆79Jul 2, 2026Updated 3 months ago
- ☆21May 1, 2026Updated 5 months ago
- Reproduction code for paper "MineExplorer: Evaluating Open-World Exploration of MLLM Agents in Minecraft"☆25Jun 12, 2026Updated 3 months ago
- Cross-Attention Guided Loss-Based Deep Dual-Branch Fusion Network for Liver Tumor Classification☆16Sep 26, 2024Updated 2 years ago
- This repository includes code for our paper: ResRL: Boosting LLM Reasoning via Negative Sample Projection Residual Reinforcement Learning…☆17May 2, 2026Updated 5 months ago
- ☆46Jan 16, 2026Updated 8 months ago
- MICCAI 2023: Radiomics-Informed Deep Learning for Classification of Atrial Fibrillation Sub-Types from Left-Atrium CT Volumes☆15Jul 23, 2023Updated 3 years ago
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- [ICLR 2026] The implementation of paper "AlphaSteer: Learning Refusal Steering with Principled Null-Space Constraint"☆67Nov 20, 2025Updated 10 months ago
- UniDoc-RL: Unified Document Understanding with Reinforcement Learning☆18May 21, 2026Updated 4 months ago
- ☆42Aug 11, 2026Updated last month
- Dynamic dual-granularity skill bank for agentic RL, jointly evolving policy and skills to improve long-horizon decision making in agentic…☆70Apr 1, 2026Updated 6 months ago
- ☆19May 25, 2026Updated 4 months ago
- ☆16Jun 17, 2026Updated 3 months ago
- In-Context Reinforcement Learning for Tool Use in Large Language Models☆48Mar 26, 2026Updated 6 months ago