ScoreFlow: Mastering LLM Agent Workflows via Score-based Preference Optimization
☆98May 22, 2025Updated last year
Alternatives and similar repositories for ScoreFlow
Users that are interested in ScoreFlow are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- [COLM2025] "Weak-for-Strong: Training Weak Meta-Agent to Harness Strong Executors"☆57Oct 6, 2025Updated 10 months ago
- [NeurIPS 2025 Spotlight] Co-Evolving LLM Coder and Unit Tester via Reinforcement Learning☆166Sep 19, 2025Updated 11 months ago
- Codes for our paper "AgentMonitor: A Plug-and-Play Framework for Predictive and Secure Multi-Agent Systems"☆14Dec 13, 2024Updated last year
- ☆17Jun 10, 2025Updated last year
- Code for Expert Supervised Reinforcement Learning☆10Apr 7, 2021Updated 5 years ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- Async MCP server with Minimax API integration for image generation and text-to-speech☆50Jan 29, 2026Updated 7 months ago
- autogen 中文文档☆10Nov 7, 2023Updated 2 years ago
- [TMLR 2025] Jigsaw-R1: A Study of Rule-based Visual Reinforcement Learning with Jigsaw Puzzles☆15Oct 17, 2025Updated 10 months ago
- The official Implementation for TKDE paper "Individual and Structural Graph Information Bottlenecks for Out-of-Distribution Generalizatio…☆14Aug 6, 2023Updated 3 years ago
- [preprint] Think Longer to Explore Deeper: Learn to Explore In-Context via Length-Incentivized Reinforcement Learning☆19Feb 18, 2026Updated 6 months ago
- ☆61Dec 12, 2025Updated 8 months ago
- Official code repo for our work "Native Visual Understanding: Resolving Resolution Dilemmas in Vision-Language Models"☆54Jun 17, 2025Updated last year
- [NeurIPS 2025 Spotlight] LLM post-training suite — featuring ReasonFlux, ReasonFlux-PRM, and ReasonFlux-Coder.☆542Sep 27, 2025Updated 11 months ago
- Implementation of the Decrypto benchmark for multi-agent reasoning and theory of mind.☆23Jan 19, 2026Updated 7 months ago
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- Cross-library augmentation toolbox supporting 300 operators over 8 libraries + AI transforms☆12Jan 11, 2022Updated 4 years ago
- ☆144May 6, 2025Updated last year
- [EMNLP 2024] Introducing Filtered Direct Preference Optimization (fDPO) that enhances language model alignment with human preferences by …☆16Nov 27, 2024Updated last year
- Agent-based implementation of RAG, incorporating AI agents into the RAG pipeline to orchestrate its components and perform additional act…☆20Feb 20, 2025Updated last year
- The original Shared Recurrent Memory Transformer implementation☆36Updated this week
- [ICML'25 Oral] Multi-agent Architecture Search via Agentic Supernet☆282Nov 13, 2025Updated 9 months ago
- Official source code repository for paper BubbleRAG.☆17Jun 1, 2026Updated 2 months ago
- ☆54Nov 9, 2025Updated 9 months ago
- 🔥🔥🔥 ICLR 2025 Oral. Automating Agentic Workflow Generation.☆583Dec 25, 2025Updated 8 months ago
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- Official implementation of MAS-GPT: Training LLMs to Build LLM-based Multi-Agent Systems☆81Jun 26, 2025Updated last year
- ☆37May 29, 2025Updated last year
- ☆29May 27, 2026Updated 3 months ago
- [ICLR 2026] Official code for TraceRL: Revolutionizing post-training for Diffusion LLMs, powering the SOTA TraDo series.☆519Jan 28, 2026Updated 7 months ago
- A simple template for theoretical computer science assignments☆12Sep 6, 2023Updated 2 years ago
- An open platform for enhancing the capability of LLMs in workflow orchestration.☆195Mar 11, 2025Updated last year
- Offline Policy Evaluation via Adaptive Weighting with Data from Contextual Bandits☆11Oct 21, 2024Updated last year
- ☆54Feb 12, 2025Updated last year
- PiFlow: Principle-aware Scientific Discovery with Multi-Agent Collaboration☆51Jan 7, 2026Updated 7 months ago
- Simple, predictable pricing with DigitalOcean hosting • AdAlways know what you'll pay with monthly caps and flat pricing. Enterprise-grade infrastructure trusted by 600k+ customers.
- This code is part of a workshop conducted on how to build your own Document AI Agent using Open Source LLMs☆16May 8, 2025Updated last year
- Official Implementation of "Learning to Refuse: Towards Mitigating Privacy Risks in LLMs"☆10Dec 13, 2024Updated last year
- Look Back to Reason Forward: Revisitable Memory for Long-Context LLM Agents☆45Apr 13, 2026Updated 4 months ago
- [ICML 2026 Spotlight] Critique-GRPO: Advancing LLM Reasoning with Natural Language and Numerical Feedback☆74Jun 3, 2026Updated 2 months ago
- Beyond KV Caching: Shared Attention for Efficient LLMs☆20Jul 19, 2024Updated 2 years ago
- [ICLR 2026 Oral & ICML 2026] Generative Universal Verifier as Multimodal Meta-Reasoner☆65May 29, 2026Updated 3 months ago
- ☆90Sep 11, 2024Updated last year