ScoreFlow: Mastering LLM Agent Workflows via Score-based Preference Optimization
☆97May 22, 2025Updated last year
Alternatives and similar repositories for ScoreFlow
Users that are interested in ScoreFlow are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- [COLM2025] "Weak-for-Strong: Training Weak Meta-Agent to Harness Strong Executors"☆57Oct 6, 2025Updated 9 months ago
- [NeurIPS 2025 Spotlight] Co-Evolving LLM Coder and Unit Tester via Reinforcement Learning☆167Sep 19, 2025Updated 10 months ago
- Official implementation of the paper "MusicInfuser: Making Video Diffusion Listen and Dance" (CVPR`26)☆85May 3, 2026Updated 2 months ago
- Codes for our paper "AgentMonitor: A Plug-and-Play Framework for Predictive and Secure Multi-Agent Systems"☆13Dec 13, 2024Updated last year
- ☆17Jun 10, 2025Updated last year
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- Async MCP server with Minimax API integration for image generation and text-to-speech☆50Jan 29, 2026Updated 5 months ago
- [ICLR 2026] LongRLVR: Long-Context Reinforcement Learning Requires Verifiable Context Rewards.☆19Mar 16, 2026Updated 4 months ago
- AutoLibra: Metric Induction for Agents from Open-Ended Human Feedback☆19Apr 23, 2026Updated 2 months ago
- [TMLR 2025] Jigsaw-R1: A Study of Rule-based Visual Reinforcement Learning with Jigsaw Puzzles☆15Oct 17, 2025Updated 9 months ago
- [preprint] Think Longer to Explore Deeper: Learn to Explore In-Context via Length-Incentivized Reinforcement Learning☆19Feb 18, 2026Updated 5 months ago
- ☆61Dec 12, 2025Updated 7 months ago
- MT3:多任务多音轨音乐转录的 Gradio 演示。(全中文汉化)☆12Mar 24, 2025Updated last year
- ☆24Apr 15, 2025Updated last year
- Official code repo for our work "Native Visual Understanding: Resolving Resolution Dilemmas in Vision-Language Models"☆54Jun 17, 2025Updated last year
- Managed Database hosting by DigitalOcean • AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- [NeurIPS 2025 Spotlight] LLM post-training suite — featuring ReasonFlux, ReasonFlux-PRM, and ReasonFlux-Coder.☆540Sep 27, 2025Updated 9 months ago
- [ICLR 2025] Benchmarking Agentic Workflow Generation☆155Feb 19, 2025Updated last year
- [EMNLP 2024] Introducing Filtered Direct Preference Optimization (fDPO) that enhances language model alignment with human preferences by …☆16Nov 27, 2024Updated last year
- Kyutai with an "eye"☆254Mar 26, 2025Updated last year
- The original Shared Recurrent Memory Transformer implementation☆36Jul 11, 2025Updated last year
- ☆145May 6, 2025Updated last year
- [ICML'25 Oral] Multi-agent Architecture Search via Agentic Supernet☆279Nov 13, 2025Updated 8 months ago
- Official source code repository for paper BubbleRAG.☆16Jun 1, 2026Updated last month
- 🔥🔥🔥 ICLR 2025 Oral. Automating Agentic Workflow Generation.☆554Dec 25, 2025Updated 6 months ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- ☆52Nov 9, 2025Updated 8 months ago
- Official implementation of MAS-GPT: Training LLMs to Build LLM-based Multi-Agent Systems☆81Jun 26, 2025Updated last year
- ☆37May 29, 2025Updated last year
- aigc evals☆10Dec 2, 2023Updated 2 years ago
- [ICLR 2026] Official code for TraceRL: Revolutionizing post-training for Diffusion LLMs, powering the SOTA TraDo series.☆512Jan 28, 2026Updated 5 months ago
- ☆80Nov 19, 2024Updated last year
- An open platform for enhancing the capability of LLMs in workflow orchestration.☆195Mar 11, 2025Updated last year
- ☆52Feb 12, 2025Updated last year
- ☆18Jun 9, 2026Updated last month
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- Official Implementation of "Learning to Refuse: Towards Mitigating Privacy Risks in LLMs"☆10Dec 13, 2024Updated last year
- PiFlow: Principle-aware Scientific Discovery with Multi-Agent Collaboration☆48Jan 7, 2026Updated 6 months ago
- This code is part of a workshop conducted on how to build your own Document AI Agent using Open Source LLMs☆16May 8, 2025Updated last year
- A deep research framework☆33Apr 21, 2026Updated 3 months ago
- Look Back to Reason Forward: Revisitable Memory for Long-Context LLM Agents☆42Apr 13, 2026Updated 3 months ago
- ☆32May 15, 2022Updated 4 years ago
- Code for ACL25-findings. An LLM-based agent simulation framework that simulates human behavior and generates dynamic, text-based social g…☆96Mar 15, 2026Updated 4 months ago