The official codebase for our paper, FLEX: Continuous Agent Evolution via Forward Learning from Experience.
☆85Jun 9, 2026Updated 2 months ago
Alternatives and similar repositories for FLEX
Users that are interested in FLEX are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- ☆36May 11, 2026Updated 3 months ago
- This is the repository for paper "CREATOR: Tool Creation for Disentangling Abstract and Concrete Reasoning of Large Language Models"☆31Oct 8, 2023Updated 2 years ago
- [ICML 2024] Learning with Complementary Labels Revisited: The Selected-Completely-at-Random Setting Is More Practical☆13May 12, 2024Updated 2 years ago
- MemGen: Weaving Generative Latent Memory for Self-Evolving Agents☆408Jun 10, 2026Updated 2 months ago
- [ICML'26] Scaling Long-Horizon LLM Agent via Context-Folding☆183May 18, 2026Updated 2 months ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- ☆15Aug 23, 2025Updated 11 months ago
- [ICLR 2025 Spotlight] Realistic Evaluation of Deep Partial-Label Learning Algorithms☆15Feb 2, 2025Updated last year
- [AAAI 2026] Relation-R1: Progressively Cognitive Chain-of-Thought Guided Reinforcement Learning for Unified Relation Comprehension☆20Mar 6, 2026Updated 5 months ago
- An AI agent memory framework that converts an agent’s own interaction traces—both successes and failures—into reusable, high-level reason…☆62Feb 9, 2026Updated 6 months ago
- The official implementation of the paper "Mem-α: Learning Memory Construction via Reinforcement Learning"☆222Dec 25, 2025Updated 7 months ago
- [CVPR 2026] This is the official PyTorch implementation of "MoDES: Accelerating Mixture-of-Experts Multimodal Large Language Models via D…☆32Mar 16, 2026Updated 4 months ago
- [ICML 2026] Meta Context Engineering via Agentic Skill Evolution☆156May 4, 2026Updated 3 months ago
- Code Release for the 2023 NeurIPS Paper How does GPT-2 compute greater-than?: Interpreting mathematical abilities in a pre-trained langua…☆17Dec 6, 2024Updated last year
- [NeurIPS 2025] TTRL: Test-Time Reinforcement Learning☆1,110Apr 15, 2026Updated 4 months ago
- Deploy open-source AI quickly and easily - Special Bonus Offer • AdRunpod Hub is built for open source. One-click deployment and autoscaling endpoints without provisioning your own infrastructure.
- Implementation of Reinforcement Pre-Training (RPT) for Language Models - ArXiv:2506.08007☆21Jul 19, 2025Updated last year
- SkillGrad: Optimizing Agent Skills Like Gradient Descent☆26May 28, 2026Updated 2 months ago
- [ICLR 2023] Is the Performance of My Deep Network Too Good to Be True? A Direct Approach to Estimating the Bayes Error in Binary Classifi…☆24Aug 12, 2025Updated last year
- Synthetic Video hallucination and Mitigation☆25Sep 21, 2025Updated 10 months ago
- ☆16Apr 26, 2021Updated 5 years ago
- Github Repo for Reinforced Reasoning for Embodied Planning☆19Aug 16, 2025Updated 11 months ago
- [ICML 2026] XSkill: Continual Learning from Experience and Skills in Multimodal Agents☆251May 13, 2026Updated 3 months ago
- ☆26Oct 9, 2025Updated 10 months ago
- Agent KB: Leveraging Cross-Domain Experience for Agentic Problem Solving☆451Aug 19, 2025Updated 11 months ago
- Deploy open-source AI quickly and easily - Special Bonus Offer • AdRunpod Hub is built for open source. One-click deployment and autoscaling endpoints without provisioning your own infrastructure.
- The code of our AAAI'20 paper "GraphER: Token-Centric Entity Resolution with Graph Convolutional Neural Networks"☆11Aug 10, 2020Updated 6 years ago
- Official code repo for our work "Native Visual Understanding: Resolving Resolution Dilemmas in Vision-Language Models"☆55Jun 17, 2025Updated last year
- ☆523Aug 5, 2026Updated last week
- A Collection of Papers about Memory for Language Agents☆635Jul 23, 2026Updated 3 weeks ago
- Official implementation of Language Models as Compilers: Simulating the Execution Of Pseudocode Improves Algorithmic Reasoning in Languag…☆23Apr 8, 2024Updated 2 years ago
- [ICLR 2026] Agentic Reinforced Policy Optimization (ARPO)☆1,108Jul 13, 2026Updated last month
- ☆12Oct 17, 2024Updated last year
- CL-bench: A Benchmark for Context Learning☆576May 12, 2026Updated 3 months ago
- 本仓库是基于 Gazebo、ArduPilot/MAVROS、D435i RGB-D 深度相机与 YOLO 的无人机城市仿真及识别-预测-避障系统,包含基础包的仿真环境、移动行人与车辆场景、目标检测与拓展包的轨迹预测、安全降落点生成、仿真飞控状态机、真实飞控轻量程序以及录像…☆24Jun 24, 2026Updated last month
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- ☆15Feb 10, 2025Updated last year
- A MemAgent framework that can be extrapolated to 3.5M, along with a training framework for RL training of any agent workflow.☆1,093May 12, 2026Updated 3 months ago
- ☆73Feb 1, 2026Updated 6 months ago
- [arXiv:2605.19952] "Rethinking How to Remember: Beyond Atomic Facts in Lifelong LLM Agent Memory"☆17May 20, 2026Updated 2 months ago
- An approach to utomatically generating browser environment with verifiable tasks☆67Mar 24, 2026Updated 4 months ago
- ☆13Mar 29, 2026Updated 4 months ago
- ☆1,285Oct 15, 2025Updated 10 months ago