OpenJudge: A Unified Framework for Holistic Evaluation and Quality Rewards
☆852Sep 11, 2026Updated 2 weeks ago
Alternatives and similar repositories for OpenJudge
Users that are interested in OpenJudge are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- FlowLLM: Build LLM applications with ease.☆35Jun 27, 2026Updated 2 months ago
- A benchmark for evaluating LLM × harness performance.☆111Aug 3, 2026Updated last month
- LLM-powered MCP server for building financial deep-research agents, integrating web search, Crawl4AI scraping, and entity extraction into…☆26Feb 11, 2026Updated 7 months ago
- Trinity-RFT is a general-purpose, flexible and scalable framework designed for reinforcement fine-tuning (RFT) of large language models (…☆703Updated this week
- A platform for running AI agents on realtime sport data and make predictions about game outcomes.☆50Jul 27, 2026Updated last month
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- A curated collection of skills around AgentScope ecosystem and CoPaw applications.☆126Sep 10, 2026Updated 2 weeks ago
- ReMe: Memory Management Kit for Agents - Remember Me, Refine Me.☆3,509Updated this week
- AgentEvolver: Towards Efficient Self-Evolving Agent System☆1,574Apr 1, 2026Updated 5 months ago
- ☆267Aug 10, 2026Updated last month
- Multi-tenant fine-tuning for LLMs with Tinker-compatible API☆70Sep 2, 2026Updated 3 weeks ago
- A curated list of resources (surveys, papers, benchmarks, and opensource projects) on Rubrics☆120Sep 12, 2026Updated last week
- verl/HybridFlow: A Flexible and Efficient RL Post-Training Framework☆23,601Updated this week
- A production-ready runtime framework for agent apps with secure tool sandboxing, Agent-as-a-Service APIs, scalable deployment, full-stack…☆872Jun 4, 2026Updated 3 months ago
- A development-oriented visualization toolkit☆652Sep 2, 2026Updated 3 weeks ago
- Proton VPN Special Offer - Get 70% off • AdSpecial partner offer. Trusted by over 100 million users worldwide. Tested, Approved and Recommended by Experts.
- Use PEFT or Full-parameter to CPT/SFT/DPO/GRPO 600+ LLMs (Qwen3.8, DeepSeek-V4, GLM-5.1, InternLM3, Llama4, ...) and 300+ MLLMs (Qwen3-VL…☆15,723Updated this week
- A collection of ready-to-use Python sample agents built with AgentScope and AgentScope Runtime, covering use cases from CLI tools to full…☆348Apr 10, 2026Updated 5 months ago
- An Efficient and User-Friendly Scaling Library for Reinforcement Learning with Large Language Models☆3,407Updated this week
- Cutting-edge platform for LLM agent tuning. Deliver RL tuning with flexibility, reliability, speed, multi-agent optimization and realtime…☆237Aug 25, 2026Updated last month
- EasyR1: An Efficient, Scalable, Multi-Modality RL Training Framework based on veRL☆5,170Updated this week
- An Easy-to-use, Scalable and High-performance Agentic RL Framework based on Ray (PPO & DAPO & REINFORCE++ & VLM & TIS & vLLM & Ray & Asy…☆10,042Sep 17, 2026Updated last week
- Deep Research as Rubric for Reinforcement Learning☆25Jun 30, 2026Updated 2 months ago
- A streamlined and customizable framework for efficient large model (LLM, VLM, AIGC) evaluation and performance benchmarking.☆3,467Updated this week
- Open Rubric System: Scaling Reinforcement Learning with Pairwise Adaptive Rubric☆27Mar 5, 2026Updated 6 months ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- Tongyi Deep Research, the Leading Open-source Deep Research Agent☆19,988Feb 27, 2026Updated 6 months ago
- slime is an LLM post-training framework for RL Scaling.☆8,527Updated this week
- Data processing for and with foundation models! 🍎 🍋 🌽 ➡️ ➡️🍸 🍹 🍷☆7,092Updated this week
- Build and run agents you can see, understand and trust.☆32,301Updated this week
- A set of examples based on verl for end-to-end RL training recipes.☆334Updated this week
- Train your Agent model via our easy and efficient framework☆1,783Dec 5, 2025Updated 9 months ago
- Scaling Preference Data Curation via Human-AI Synergy☆153Jul 3, 2025Updated last year
- Search-R1: An Efficient, Scalable RL Training Framework for Reasoning & Search Engine Calling interleaved LLM based on veRL☆5,446Nov 13, 2025Updated 10 months ago
- Agent RL framework for LLM agents: multi-turn reinforcement learning with StarPO and reasoning-collapse diagnostics☆2,806Aug 23, 2026Updated last month
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- Democratizing Reinforcement Learning for LLMs☆5,836Sep 12, 2026Updated last week
- The RL Bridge for LLM-based Agent Applications. Made Simple & Flexible.☆5,793Updated this week
- The official code repo of paper "Chasing the Tail: Effective Rubric-based Reward Modeling for Large Language Model Post-Training"☆32Feb 20, 2026Updated 7 months ago
- ☆29Jan 31, 2026Updated 7 months ago
- [ICLR'26] RM-R1: Unleashing the Reasoning Potential of Reward Models☆171Jun 26, 2025Updated last year
- TBD☆69Mar 13, 2026Updated 6 months ago
- HeartBench is an evaluation benchmark for the psychological and social sciences field, designed to transcend traditional knowledge and re…☆52Jan 7, 2026Updated 8 months ago