OpenJudge: A Unified Framework for Holistic Evaluation and Quality Rewards
☆816Aug 3, 2026Updated last month
Alternatives and similar repositories for OpenJudge
Users that are interested in OpenJudge are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- FlowLLM: Build LLM applications with ease.☆34Jun 27, 2026Updated 2 months ago
- A benchmark for evaluating LLM × harness performance.☆109Aug 3, 2026Updated last month
- LLM-powered MCP server for building financial deep-research agents, integrating web search, Crawl4AI scraping, and entity extraction into…☆24Feb 11, 2026Updated 6 months ago
- Trinity-RFT is a general-purpose, flexible and scalable framework designed for reinforcement fine-tuning (RFT) of large language models (…☆698Aug 13, 2026Updated 3 weeks ago
- A platform for running AI agents on realtime sport data and make predictions about game outcomes.☆49Jul 27, 2026Updated last month
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- ReMe: Memory Management Kit for Agents - Remember Me, Refine Me.☆3,409Updated this week
- AgentEvolver: Towards Efficient Self-Evolving Agent System☆1,553Apr 1, 2026Updated 5 months ago
- A curated collection of skills around AgentScope ecosystem and CoPaw applications.☆123Apr 15, 2026Updated 4 months ago
- ☆264Aug 10, 2026Updated 3 weeks ago
- A curated list of resources (surveys, papers, benchmarks, and opensource projects) on Rubrics☆111Aug 18, 2026Updated 2 weeks ago
- verl/HybridFlow: A Flexible and Efficient RL Post-Training Framework☆23,290Updated this week
- A production-ready runtime framework for agent apps with secure tool sandboxing, Agent-as-a-Service APIs, scalable deployment, full-stack…☆863Jun 4, 2026Updated 3 months ago
- A development-oriented visualization toolkit☆647Updated this week
- Use PEFT or Full-parameter to CPT/SFT/DPO/GRPO 600+ LLMs (Qwen3.6, DeepSeek-V4, GLM-5.1, InternLM3, Llama4, ...) and 300+ MLLMs (Qwen3-VL…☆15,514Updated this week
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- A collection of ready-to-use Python sample agents built with AgentScope and AgentScope Runtime, covering use cases from CLI tools to full…☆342Apr 10, 2026Updated 4 months ago
- An Efficient and User-Friendly Scaling Library for Reinforcement Learning with Large Language Models☆3,382Updated this week
- Cutting-edge platform for LLM agent tuning. Deliver RL tuning with flexibility, reliability, speed, multi-agent optimization and realtime…☆233Aug 25, 2026Updated last week
- EasyR1: An Efficient, Scalable, Multi-Modality RL Training Framework based on veRL☆5,148Updated this week
- An Easy-to-use, Scalable and High-performance Agentic RL Framework based on Ray (PPO & DAPO & REINFORCE++ & VLM & TIS & vLLM & Ray & Asy…☆9,976Aug 13, 2026Updated 3 weeks ago
- Deep Research as Rubric for Reinforcement Learning☆25Jun 30, 2026Updated 2 months ago
- A streamlined and customizable framework for efficient large model (LLM, VLM, AIGC) evaluation and performance benchmarking.☆3,372Updated this week
- Open Rubric System: Scaling Reinforcement Learning with Pairwise Adaptive Rubric☆22Mar 5, 2026Updated 6 months ago
- Tongyi Deep Research, the Leading Open-source Deep Research Agent☆19,908Feb 27, 2026Updated 6 months ago
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- slime is an LLM post-training framework for RL Scaling.☆8,381Updated this week
- Data processing for and with foundation models! 🍎 🍋 🌽 ➡️ ➡️🍸 🍹 🍷☆6,991Updated this week
- A set of examples based on verl for end-to-end RL training recipes.☆330Updated this week
- Build and run agents you can see, understand and trust.☆30,667Updated this week
- Scaling Preference Data Curation via Human-AI Synergy☆153Jul 3, 2025Updated last year
- Search-R1: An Efficient, Scalable RL Training Framework for Reasoning & Search Engine Calling interleaved LLM based on veRL☆5,366Nov 13, 2025Updated 9 months ago
- Train your Agent model via our easy and efficient framework☆1,780Dec 5, 2025Updated 9 months ago
- Democratizing Reinforcement Learning for LLMs☆5,814Aug 24, 2026Updated last week
- The RL Bridge for LLM-based Agent Applications. Made Simple & Flexible.☆5,723Updated this week
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- Agent RL framework for LLM agents: multi-turn reinforcement learning with StarPO and reasoning-collapse diagnostics☆2,789Aug 23, 2026Updated last week
- The official code repo of paper "Chasing the Tail: Effective Rubric-based Reward Modeling for Large Language Model Post-Training"☆30Feb 20, 2026Updated 6 months ago
- ☆29Jan 31, 2026Updated 7 months ago
- [ICLR'26] RM-R1: Unleashing the Reasoning Potential of Reward Models☆171Jun 26, 2025Updated last year
- TBD☆69Mar 13, 2026Updated 5 months ago
- ReSearch: Learning to Reason with Search for LLMs via Reinforcement Learning & ReCall: Learning to Reason with Tool Call for LLMs via Rei…☆1,432May 16, 2025Updated last year
- HeartBench is an evaluation benchmark for the psychological and social sciences field, designed to transcend traditional knowledge and re…☆51Jan 7, 2026Updated 7 months ago