OpenJudge: A Unified Framework for Holistic Evaluation and Quality Rewards
☆747Jul 9, 2026Updated 2 weeks ago
Alternatives and similar repositories for OpenJudge
Users that are interested in OpenJudge are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- FlowLLM: Build LLM applications with ease.☆34Jun 27, 2026Updated last month
- A benchmark for evaluating LLM × harness performance.☆95Updated this week
- LLM-powered MCP server for building financial deep-research agents, integrating web search, Crawl4AI scraping, and entity extraction into…☆24Feb 11, 2026Updated 5 months ago
- Trinity-RFT is a general-purpose, flexible and scalable framework designed for reinforcement fine-tuning (RFT) of large language models (…☆672Updated this week
- A platform for running AI agents on realtime sport data and make predictions about game outcomes.☆43Jul 15, 2026Updated last week
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- ReMe: Memory Management Kit for Agents - Remember Me, Refine Me.☆3,218Updated this week
- AgentEvolver: Towards Efficient Self-Evolving Agent System☆1,504Apr 1, 2026Updated 3 months ago
- A curated collection of skills around AgentScope ecosystem and CoPaw applications.☆117Apr 15, 2026Updated 3 months ago
- ☆241Updated this week
- Multi-tenant fine-tuning for LLMs with Tinker-compatible API☆61Updated this week
- A curated list of resources (surveys, papers, benchmarks, and opensource projects) on Rubrics☆101Jul 13, 2026Updated 2 weeks ago
- A production-ready runtime framework for agent apps with secure tool sandboxing, Agent-as-a-Service APIs, scalable deployment, full-stack…☆842Jun 4, 2026Updated last month
- verl/HybridFlow: A Flexible and Efficient RL Post-Training Framework☆22,667Updated this week
- A development-oriented visualization toolkit☆609Jun 15, 2026Updated last month
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- Use PEFT or Full-parameter to CPT/SFT/DPO/GRPO 600+ LLMs (Qwen3.6, DeepSeek-V4, GLM-5.1, InternLM3, Llama4, ...) and 300+ MLLMs (Qwen3-VL…☆14,952Updated this week
- A collection of ready-to-use Python sample agents built with AgentScope and AgentScope Runtime, covering use cases from CLI tools to full…☆332Apr 10, 2026Updated 3 months ago
- An Efficient and User-Friendly Scaling Library for Reinforcement Learning with Large Language Models☆3,327Updated this week
- Cutting-edge platform for LLM agent tuning. Deliver RL tuning with flexibility, reliability, speed, multi-agent optimization and realtime…☆229Jul 2, 2026Updated 3 weeks ago
- EasyR1: An Efficient, Scalable, Multi-Modality RL Training Framework based on veRL☆5,082Updated this week
- An Easy-to-use, Scalable and High-performance Agentic RL Framework based on Ray (PPO & DAPO & REINFORCE++ & VLM & TIS & vLLM & Ray & Asy…☆9,853Jul 14, 2026Updated last week
- Deep Research as Rubric for Reinforcement Learning☆23Jun 30, 2026Updated 3 weeks ago
- A streamlined and customizable framework for efficient large model (LLM, VLM, AIGC) evaluation and performance benchmarking.☆3,144Updated this week
- Open Rubric System: Scaling Reinforcement Learning with Pairwise Adaptive Rubric☆19Mar 5, 2026Updated 4 months ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- Tongyi Deep Research, the Leading Open-source Deep Research Agent☆19,730Feb 27, 2026Updated 4 months ago
- slime is an LLM post-training framework for RL Scaling.☆7,645Updated this week
- Data processing for and with foundation models! 🍎 🍋 🌽 ➡️ ➡️🍸 🍹 🍷☆6,777Updated this week
- A set of examples based on verl for end-to-end RL training recipes.☆312Updated this week
- Build and run agents you can see, understand and trust.☆28,282Updated this week
- Search-R1: An Efficient, Scalable RL Training Framework for Reasoning & Search Engine Calling interleaved LLM based on veRL☆5,156Nov 13, 2025Updated 8 months ago
- Train your Agent model via our easy and efficient framework☆1,773Dec 5, 2025Updated 7 months ago
- Democratizing Reinforcement Learning for LLMs☆5,732Updated this week
- The RL Bridge for LLM-based Agent Applications. Made Simple & Flexible.☆5,604Updated this week
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- RAGEN leverages reinforcement learning to train LLM reasoning agents in interactive, stochastic environments.☆2,757Updated this week
- The official code repo of paper "Chasing the Tail: Effective Rubric-based Reward Modeling for Large Language Model Post-Training"☆30Feb 20, 2026Updated 5 months ago
- ☆29Jan 31, 2026Updated 5 months ago
- [ICLR'26] RM-R1: Unleashing the Reasoning Potential of Reward Models☆167Jun 26, 2025Updated last year
- Scaling Preference Data Curation via Human-AI Synergy☆152Jul 3, 2025Updated last year
- TBD☆65Mar 13, 2026Updated 4 months ago
- ReSearch: Learning to Reason with Search for LLMs via Reinforcement Learning & ReCall: Learning to Reason with Tool Call for LLMs via Rei…☆1,427May 16, 2025Updated last year