RETROAGENT: From Solving to Evolving via Retrospective Dual Intrinsic Feedback
☆27Mar 30, 2026Updated 4 months ago
Alternatives and similar repositories for RetroAgent
Users that are interested in RetroAgent are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Co-evolving policy actors and experience extractors for efficient experience-driven agent RL☆51May 12, 2026Updated 2 months ago
- ☆23Apr 5, 2026Updated 4 months ago
- [arxiv: 2604.14142] From P(y|x) to P(y): Investigating Reinforcement Learning in Pre-train Space☆17Apr 16, 2026Updated 3 months ago
- Official Repo for DAC-RL: Training LLMs for Divide-and-Conquer Reasoning Elevates Test-Time Scalability☆16Feb 26, 2026Updated 5 months ago
- ☆61Apr 9, 2026Updated 3 months ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- A Massive Multi-Discipline Lecture Understanding Benchmark☆34Apr 20, 2026Updated 3 months ago
- Build coherent and visually polished multimodal webpages with hierarchical planning, AIGC tools, and iterative reflection.☆16May 17, 2026Updated 2 months ago
- MarsRL: Advancing Multi-Agent Reasoning System via Reinforcement Learning with Agentic Pipeline Parallelism☆18Nov 18, 2025Updated 8 months ago
- ☆15Apr 17, 2026Updated 3 months ago
- MemOCR: an OCR-driven visual memory agent.☆33May 17, 2026Updated 2 months ago
- [NeurIPS 2025] The implementation of paper "On Reasoning Strength Planning in Large Reasoning Models"☆31Jul 6, 2025Updated last year
- ☆14Apr 22, 2024Updated 2 years ago
- EVA: Efficient Reinforcement Learning for End-to-End Video Agent☆26May 6, 2026Updated 2 months ago
- [ACL'26 Oral] AgentOCR is a token-efficient framework that compresses multi-turn agent history by rendering it into images and adopting R…☆42Mar 1, 2026Updated 5 months ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- On Path to Multimodal Generalist: General-Level and General-Bench☆21Jul 11, 2025Updated last year
- An approach to utomatically generating browser environment with verifiable tasks☆64Mar 24, 2026Updated 4 months ago
- ☆17Apr 11, 2025Updated last year
- Official implementation of the ΔBelief-RL method.☆31Feb 28, 2026Updated 5 months ago
- ConfAgents: A Conformal-Guided Multi-Agent Framework for Cost-Efficient Medical Diagnosis☆15Jul 22, 2026Updated 2 weeks ago
- Will Pre-Training Ever End? A First Step Toward Next-Generation Foundation MLLMs via Self-Improving Systematic Cognition☆31May 14, 2025Updated last year
- DLLM-Searcher has been accepted by SIGIR 2026! 🥳☆33Jan 23, 2026Updated 6 months ago
- ☆50Feb 4, 2026Updated 6 months ago
- MDM-Prime-v2: Binary Encoding and Index Shuffling Enable Scaling of Diffusion Language Models☆27May 23, 2026Updated 2 months ago
- Open source password manager - Proton Pass • AdSecurely store, share, and autofill your credentials with Proton Pass, the end-to-end encrypted password manager trusted by millions.
- [ACL 2025 Findings] GenS: Generative Frame Sampler for Long Video Understanding☆22Aug 21, 2025Updated 11 months ago
- 上朝式科 研:AI-powered research workflow showcase☆130Apr 7, 2026Updated 3 months ago
- SearchEyes: Towards Frontier Multimodal Deep Search Intelligence via Search World Simulation. A typed knowledge graph unifies data synthe…☆22Jul 8, 2026Updated 3 weeks ago
- The official codebase for "Experiential Reinforcement Learning" - https://arxiv.org/pdf/2602.13949v1☆76Jul 2, 2026Updated last month
- Implementation for Bridging Semantic and Kinematic Conditions with Diffusion-based Discrete Motion Tokenizer.☆36Mar 20, 2026Updated 4 months ago
- ☆19Aug 28, 2025Updated 11 months ago
- ☆21May 1, 2026Updated 3 months ago
- Reproduction code for paper "MineExplorer: Evaluating Open-World Exploration of MLLM Agents in Minecraft"☆18Jun 12, 2026Updated last month
- An implementation of parameter server framework in PyTorch RPC.☆12Nov 12, 2021Updated 4 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- Cross-Attention Guided Loss-Based Deep Dual-Branch Fusion Network for Liver Tumor Classification☆16Sep 26, 2024Updated last year
- This repository includes code for our paper: ResRL: Boosting LLM Reasoning via Negative Sample Projection Residual Reinforcement Learning…☆15May 2, 2026Updated 3 months ago
- MICCAI 2023: Radiomics-Informed Deep Learning for Classification of Atrial Fibrillation Sub-Types from Left-Atrium CT Volumes☆15Jul 23, 2023Updated 3 years ago
- [ICLR 2026] The implementation of paper "AlphaSteer: Learning Refusal Steering with Principled Null-Space Constraint"☆62Nov 20, 2025Updated 8 months ago
- 🌟 SwarmAgent: A framework for simulating social group dynamics using multi-agent collaboration, aiding insights into collective behavior…☆13Dec 5, 2023Updated 2 years ago
- UniDoc-RL: Unified Document Understanding with Reinforcement Learning☆16May 21, 2026Updated 2 months ago
- ☆40Oct 2, 2024Updated last year