☆15Apr 17, 2026Updated 3 months ago
Alternatives and similar repositories for AgentV-RL
Users that are interested in AgentV-RL are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- RETROAGENT: From Solving to Evolving via Retrospective Dual Intrinsic Feedback☆26Mar 30, 2026Updated 3 months ago
- Self-Hinting Language Models Enhance Reinforcement Learning☆26Mar 28, 2026Updated 3 months ago
- [ICLR 2026] M2-Miner: Multi-Agent Enhanced MCTS for Mobile GUI Agent Data Mining☆55Apr 22, 2026Updated 2 months ago
- Codes for the paper "BAPO: Stabilizing Off-Policy Reinforcement Learning for LLMs via Balanced Policy Optimization with Adaptive Clipping…☆94Jan 29, 2026Updated 5 months ago
- Code for the paper "Stack Attention: Improving the Ability of Transformers to Model Hierarchical Patterns"☆18Mar 15, 2024Updated 2 years ago
- Serverless GPU API endpoints on Runpod - Get Bonus Credits • AdSkip the infrastructure headaches. Auto-scaling, pay-as-you-go, no-ops approach lets you focus on innovating your application.
- The official implementation of InfoRM [NeurIPS 2024].☆16Oct 25, 2025Updated 8 months ago
- A Teacher–Student Cooperation Framework to Synthesize Student-Consistent SFT Data☆33May 1, 2026Updated 2 months ago
- [ACL 2025] Agentic Reward Modeling: Integrating Human Preferences with Verifiable Correctness Signals for Reliable Reward Systems☆134Jun 11, 2025Updated last year
- [ICLR2026] "Co-rewarding: Stable Self-supervised RL for Eliciting Reasoning in Large Language Models"☆30Feb 4, 2026Updated 5 months ago
- ☆26Jun 10, 2025Updated last year
- WikiVideo: Article Generation from Multiple Videos☆15Nov 14, 2025Updated 8 months ago
- Beyond Basic RAG, Empowering Real-Time Deep Research