☆41May 12, 2026Updated 2 months ago
Alternatives and similar repositories for SkillFlow
Users that are interested in SkillFlow are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- VCR-Bench: A Comprehensive Evaluation Framework for Video Chain-of-Thought Reasoning☆37May 9, 2026Updated 2 months ago
- [EMNLP 2025] Official Implement of "CRITICTOOL: Evaluating Self-Critique Capabilities of Large Language Models in Tool-Calling Error Scen…☆18Sep 2, 2025Updated 11 months ago
- ☆79Apr 12, 2026Updated 3 months ago
- ☆47Apr 9, 2025Updated last year
- Structured Decomposition and Conditional Skill Orchestration for Complex Image Generation☆32Updated this week
- Managed Database hosting by DigitalOcean • AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- ☆47Apr 8, 2026Updated 3 months ago
- Official Repo of "Flow-OPD: On-Policy Distillation for Flow Matching Models"☆269Jun 24, 2026Updated last month
- Any research. One Claw. 🦞 From any materials to research with fully autonomous & skill-driven researcher.☆440May 9, 2026Updated 2 months ago
- A Bi-Level Multi-Agent LLM System for Internet-Scale Information Search and Extraction☆40May 1, 2026Updated 3 months ago
- ☆38Jan 1, 2026Updated 7 months ago
- Agent Skill Evaluation and Evolution: Frameworks and Benchmarks☆27Jul 15, 2026Updated 2 weeks ago
- Official code repo of Video-Browser: Towards Agentic Open-web Video Browsing☆28Jan 19, 2026Updated 6 months ago
- A Self-Evolving Framework Through Executable Subagent Accumulation and Reuse☆60Jul 10, 2026Updated 3 weeks ago
- ☆18Apr 1, 2025Updated last year
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- [ICML 2026] Multimodal deep-research MLLM and benchmark. The first long-horizon multimodal deep-research MLLM, extending the number of re…☆659Jun 8, 2026Updated last month
- Terrarium: Multi-turn data engine for evaluating and optimizing LLM agents in living environments.☆54Jul 14, 2026Updated 3 weeks ago
- Automating Sub-Agent Creation for Agentic Orchestration☆153May 25, 2026Updated 2 months ago
- ☆23May 7, 2026Updated 2 months ago
- Resources for our paper: "Agent-R: Training Language Model Agents to Reflect via Iterative Self-Training"☆174Oct 20, 2025Updated 9 months ago
- This respository is used for time reasoning task for mult-session dialogue system.☆17Feb 7, 2026Updated 5 months ago
- Dynamic dual-granularity skill bank for agentic RL, jointly evolving policy and skills to improve long-horizon decision making in agentic…☆71Apr 1, 2026Updated 4 months ago
- [ICML 2026] ScalingAR: Scaling Confidence for Autoregressive Image Generation☆22May 5, 2026Updated 2 months ago
- ConfAgents: A Conformal-Guided Multi-Agent Framework for Cost-Efficient Medical Diagnosis☆15Jul 22, 2026Updated last week
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- Scaf-GRPO: Scaffolded Group Relative Policy Optimization for Enhancing LLM Reasoning☆24Feb 8, 2026Updated 5 months ago
- [AACL2025] Code for paper Chain-of-Query: Unleashing the Power of LLMs in SQL-Aided Table Understanding via Multi-Agent Collaboration☆21Jan 19, 2026Updated 6 months ago
- DLER: Doing Length pEnalty Right - Incentivizing More Intelligence per Token via Reinforcement Learning☆18Nov 11, 2025Updated 8 months ago
- Official implementation of Vector-ICL: In-context Learning with Continuous Vector Representations (ICLR 2025)☆24Jun 2, 2025Updated last year
- Model-based Hindsight Experience Replay☆10Jun 8, 2022Updated 4 years ago
- ☆53Feb 25, 2026Updated 5 months ago
- MarsRL: Advancing Multi-Agent Reasoning System via Reinforcement Learning with Agentic Pipeline Parallelism☆18Nov 18, 2025Updated 8 months ago
- Takagi-Sugeno-Kang Fuzzy System☆19May 20, 2020Updated 6 years ago
- ☆18Apr 8, 2025Updated last year
- Deploy open-source AI quickly and easily - Special Bonus Offer • AdRunpod Hub is built for open source. One-click deployment and autoscaling endpoints without provisioning your own infrastructure.
- [KDD 2025] AgentGen: Enhancing Planning Abilities for Large Language Model based Agent via Environment and Task Generation☆34Nov 18, 2025Updated 8 months ago
- ☆136Mar 22, 2025Updated last year
- Official Repo for DAC-RL: Training LLMs for Divide-and-Conquer Reasoning Elevates Test-Time Scalability☆16Feb 26, 2026Updated 5 months ago
- TrustAgent: Towards Safe and Trustworthy LLM-based Agents☆58Feb 7, 2025Updated last year
- ☆27Mar 17, 2026Updated 4 months ago
- [ICLR 2026] Code for "Group Critical-token Policy Optimization for Autoregressive Image Generation"☆58Dec 4, 2025Updated 8 months ago
- [ACL'25] Code for "Aligning Large Language Models to Follow Instructions and Hallucinate Less via Effective Data Filtering"☆21Jul 23, 2025Updated last year