OpenJudge: A Unified Framework for Holistic Evaluation and Quality Rewards
☆789Aug 3, 2026Updated last week
Alternatives and similar repositories for OpenJudge
Users that are interested in OpenJudge are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- FlowLLM: Build LLM applications with ease.☆34Jun 27, 2026Updated last month
- A benchmark for evaluating LLM × harness performance.☆99Aug 3, 2026Updated last week
- LLM-powered MCP server for building financial deep-research agents, integrating web search, Crawl4AI scraping, and entity extraction into…☆24Feb 11, 2026Updated 6 months ago
- Trinity-RFT is a general-purpose, flexible and scalable framework designed for reinforcement fine-tuning (RFT) of large language models (…☆683Updated this week
- A platform for running AI agents on realtime sport data and make predictions about game outcomes.☆48Jul 27, 2026Updated 2 weeks ago
- Managed Database hosting by DigitalOcean • AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- ReMe: Memory Management Kit for Agents - Remember Me, Refine Me.☆3,315Updated this week
- AgentEvolver: Towards Efficient Self-Evolving Agent System☆1,533Apr 1, 2026Updated 4 months ago
- A curated collection of skills around AgentScope ecosystem and CoPaw applications.☆120Apr 15, 2026Updated 4 months ago
- ☆252Updated this week
- Multi-tenant fine-tuning for LLMs with Tinker-compatible API☆64Updated this week
- A curated list of resources (surveys, papers, benchmarks, and opensource projects) on Rubrics☆107Aug 8, 2026Updated last week
- verl/HybridFlow: A Flexible and Efficient RL Post-Training Framework☆22,970Updated this week
- A production-ready runtime framework for agent apps with secure tool sandboxing, Agent-as-a-Service APIs, scalable deployment, full-stack…☆856Jun 4, 2026Updated 2 months ago
- A development-oriented visualization toolkit☆637Jun 15, 2026Updated 2 months ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- Use PEFT or Full-parameter to CPT/SFT/DPO/GRPO 600+ LLMs (Qwen3.6, DeepSeek-V4, GLM-5.1, InternLM3, Llama4, ...) and 300+ MLLMs (Qwen3-VL…☆15,192Updated this week
- A collection of ready-to-use Python sample agents built with AgentScope and AgentScope Runtime, covering use cases from CLI tools to full…☆338Apr 10, 2026Updated 4 months ago
- An Efficient and User-Friendly Scaling Library for Reinforcement Learning with Large Language Models☆3,362Updated this week
- Cutting-edge platform for LLM agent tuning. Deliver RL tuning with flexibility, reliability, speed, multi-agent optimization and realtime…☆232Updated this week
- EasyR1: An Efficient, Scalable, Multi-Modality RL Training Framework based on veRL☆5,116Jul 30, 2026Updated 2 weeks ago
- An Easy-to-use, Scalable and High-performance Agentic RL Framework based on Ray (PPO & DAPO & REINFORCE++ & VLM & TIS & vLLM & Ray & Asy…☆9,914Updated this week
- Deep Research as Rubric for Reinforcement Learning☆25Jun 30, 2026Updated last month
- A streamlined and customizable framework for efficient large model (LLM, VLM, AIGC) evaluation and performance benchmarking.☆3,240Updated this week
- Open Rubric System: Scaling Reinforcement Learning with Pairwise Adaptive Rubric☆19Mar 5, 2026Updated 5 months ago
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- Tongyi Deep Research, the Leading Open-source Deep Research Agent☆19,831Feb 27, 2026Updated 5 months ago
- slime is an LLM post-training framework for RL Scaling.☆8,027Updated this week
- Data processing for and with foundation models! 🍎 🍋 🌽 ➡️ ➡️🍸 🍹 🍷☆6,887Updated this week
- A set of examples based on verl for end-to-end RL training recipes.☆324Aug 5, 2026Updated last week
- Build and run agents you can see, understand and trust.☆28,966Updated this week
- Scaling Preference Data Curation via Human-AI Synergy☆153Jul 3, 2025Updated last year
- Search-R1: An Efficient, Scalable RL Training Framework for Reasoning & Search Engine Calling interleaved LLM based on veRL☆5,295Nov 13, 2025Updated 9 months ago
- Train your Agent model via our easy and efficient framework☆1,778Dec 5, 2025Updated 8 months ago
- Democratizing Reinforcement Learning for LLMs☆5,784Updated this week
- End-to-end encrypted email - Proton Mail • AdSpecial offer: 40% Off Yearly / 80% Off First Month. All Proton services are open source and independently audited for security.
- The RL Bridge for LLM-based Agent Applications. Made Simple & Flexible.☆5,667Updated this week
- RAGEN leverages reinforcement learning to train LLM reasoning agents in interactive, stochastic environments.☆2,770Jul 24, 2026Updated 3 weeks ago
- The official code repo of paper "Chasing the Tail: Effective Rubric-based Reward Modeling for Large Language Model Post-Training"☆30Feb 20, 2026Updated 5 months ago
- ☆29Jan 31, 2026Updated 6 months ago
- [ICLR'26] RM-R1: Unleashing the Reasoning Potential of Reward Models☆168Jun 26, 2025Updated last year
- TBD☆68Mar 13, 2026Updated 5 months ago
- ReSearch: Learning to Reason with Search for LLMs via Reinforcement Learning & ReCall: Learning to Reason with Tool Call for LLMs via Rei…☆1,428May 16, 2025Updated last year