☆41May 12, 2026Updated 3 months ago
Alternatives and similar repositories for SkillFlow
Users that are interested in SkillFlow are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- [EMNLP 2025] Official Implement of "CRITICTOOL: Evaluating Self-Critique Capabilities of Large Language Models in Tool-Calling Error Scen…☆18Sep 2, 2025Updated 11 months ago
- ☆79Apr 12, 2026Updated 4 months ago
- V2P-Bench: Evaluating Video-Language Understanding with Visual Prompts for Better Human-Model Interaction☆38May 9, 2026Updated 3 months ago
- Structured Decomposition and Conditional Skill Orchestration for Complex Image Generation☆32Jul 30, 2026Updated 3 weeks ago
- ☆48Apr 8, 2026Updated 4 months ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- Official Repo of "Flow-OPD: On-Policy Distillation for Flow Matching Models"☆284Jun 24, 2026Updated last month
- A Bi-Level Multi-Agent LLM System for Internet-Scale Information Search and Extraction☆41Aug 8, 2026Updated 2 weeks ago
- ☆38Jan 1, 2026Updated 7 months ago
- Agent Skill Evaluation and Evolution: Frameworks and Benchmarks☆28Jul 15, 2026Updated last month
- A Self-Evolving Framework Through Executable Subagent Accumulation and Reuse☆62Jul 10, 2026Updated last month
- ☆18Apr 1, 2025Updated last year
- [COLM'26] SkillLearnBench is the first benchmark for evaluating continual learning methods that automatically generate agent skills.☆77Jul 9, 2026Updated last month
- [ICML 2026 & EMNLP 2026] Multimodal deep-research MLLM and benchmark. The first long-horizon multimodal deep-research MLLM, extending the…☆677Aug 8, 2026Updated 2 weeks ago
- Terrarium: Multi-turn data engine for evaluating and optimizing LLM agents in living environments.☆59Jul 14, 2026Updated last month
- Serverless GPU API endpoints on Runpod - Get Bonus Credits • AdSkip the infrastructure headaches. Auto-scaling, pay-as-you-go, no-ops approach lets you focus on innovating your application.
- Automating Sub-Agent Creation for Agentic Orchestration☆157May 25, 2026Updated 2 months ago
- ☆24May 7, 2026Updated 3 months ago
- Resources for our paper: "Agent-R: Training Language Model Agents to Reflect via Iterative Self-Training"☆175Oct 20, 2025Updated 10 months ago
- This respository is used for time reasoning task for mult-session dialogue system.☆18Feb 7, 2026Updated 6 months ago
- [ICML 2026] ScalingAR: Scaling Confidence for Autoregressive Image Generation☆22May 5, 2026Updated 3 months ago
- ConfAgents: A Conformal-Guided Multi-Agent Framework for Cost-Efficient Medical Diagnosis☆15Jul 22, 2026Updated last month
- Scaf-GRPO: Scaffolded Group Relative Policy Optimization for Enhancing LLM Reasoning☆25Feb 8, 2026Updated 6 months ago
- [AACL2025] Code for paper Chain-of-Query: Unleashing the Power of LLMs in SQL-Aided Table Understanding via Multi-Agent Collaboration☆21Jan 19, 2026Updated 7 months ago
- DLER: Doing Length pEnalty Right - Incentivizing More Intelligence per Token via Reinforcement Learning☆18Nov 11, 2025Updated 9 months ago
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- Official implementation of Vector-ICL: In-context Learning with Continuous Vector Representations (ICLR 2025)☆24Jun 2, 2025Updated last year
- Model-based Hindsight Experience Replay☆10Jun 8, 2022Updated 4 years ago
- ☆53Feb 25, 2026Updated 5 months ago
- MarsRL: Advancing Multi-Agent Reasoning System via Reinforcement Learning with Agentic Pipeline Parallelism☆18Nov 18, 2025Updated 9 months ago
- ☆18Apr 8, 2025Updated last year
- [KDD 2025] AgentGen: Enhancing Planning Abilities for Large Language Model based Agent via Environment and Task Generation☆35Nov 18, 2025Updated 9 months ago
- ☆136Mar 22, 2025Updated last year
- Official Repo for DAC-RL: Training LLMs for Divide-and-Conquer Reasoning Elevates Test-Time Scalability☆16Feb 26, 2026Updated 5 months ago
- TrustAgent: Towards Safe and Trustworthy LLM-based Agents☆60Feb 7, 2025Updated last year
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- 2025寒假 豆包MarsCode & 字节跳动青训营-抖音商场后端Java实现☆15Mar 5, 2025Updated last year
- Task Aware Downscaling for efficient storing and accurate reconstruction in image and video domain☆12Jul 25, 2024Updated 2 years ago
- [ICLR 2026] Code for "Group Critical-token Policy Optimization for Autoregressive Image Generation"☆58Dec 4, 2025Updated 8 months ago
- [ACL'25] Code for "Aligning Large Language Models to Follow Instructions and Hallucinate Less via Effective Data Filtering"☆21Jul 23, 2025Updated last year
- ☆37May 11, 2026Updated 3 months ago
- CoEvoSkills: Self-Evolving Agent Skills via Co-Evolutionary Verification — COLM 2026☆60Updated this week
- flexible super-resolution☆15Mar 9, 2022Updated 4 years ago