Official repository for AutoRule: Reasoning Chain-of-thought Extracted Rule-based Rewards Improve Preference Learning
☆17Jul 24, 2025Updated last year
Alternatives and similar repositories for AutoRule
Users that are interested in AutoRule are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- The codebase and some introductions of FineMed.☆31Sep 11, 2025Updated last year
- [ICLR 2024] "Data Distillation Can Be Like Vodka: Distilling More Times For Better Quality" by Xuxi Chen*, Yu Yang*, Zhangyang Wang, Baha…☆15May 18, 2024Updated 2 years ago
- Demo of google-map-react and Next.js☆10May 18, 2018Updated 8 years ago
- SCoRe: Training Language Models to Self-Correct via Reinforcement Learning☆16May 14, 2026Updated 4 months ago
- Implementation of "PAG: Multi-Turn Reinforced LLM Self-Correction with Policy as Generative Verifier"☆17Jun 27, 2025Updated last year
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- WisdoMentor - Series: A LLM for undergraduates | 博导智言(辅助大学生 学习)☆13May 9, 2024Updated 2 years ago
- 基于Llama3,通过进一步CPT,SFT,ORPO得到的中文版Llama3☆16Apr 24, 2024Updated 2 years ago
- [ICLR 2026] Skill-Targeted Adaptive Training☆28Mar 12, 2026Updated 6 months ago
- R3: Robust Rubric-Agnostic Reward Models☆24Jul 12, 2025Updated last year
- Open Rubric System: Scaling Reinforcement Learning with Pairwise Adaptive Rubric☆25Mar 5, 2026Updated 6 months ago
- MLLM @ Game☆17May 12, 2025Updated last year
- ACL 2026☆27Nov 19, 2025Updated 10 months ago
- ChatGPT-Client is a ChatGPT Client with Offical OpenAI API.☆11May 30, 2024Updated 2 years ago
- ☆16Feb 15, 2026Updated 7 months ago
- GPUs on demand by Runpod - Special Offer Available • AdRun AI, ML, and HPC workloads on powerful cloud GPUs—without limits or wasted spend. Deploy GPUs in under a minute and pay by the second.
- ☆15Oct 4, 2024Updated last year
- 数据预处理——插值法填补缺失值,并且标记填充位置☆10Apr 19, 2019Updated 7 years ago
- A simple adaboost code using decision stumps as weak classifiers☆11Nov 1, 2012Updated 13 years ago
- Code repository for ICLR 2026 paper "ResearchRubrics: A Benchmark of Prompts and Rubrics For Evaluating Deep Research Agents" (https://ww…☆31Feb 10, 2026Updated 7 months ago
- Indonesian law dataset containing section annotation of court decision documents☆19Jul 7, 2022Updated 4 years ago
- [CVPR 2025] LoRA Recycle: Unlocking Tuning-Free Few-Shot Adaptability in Visual Foundation Models by Recycling Pre-Tuned LoRAs☆15Jun 20, 2025Updated last year
- ☆17Jan 14, 2026Updated 8 months ago
- [NeurIPS 2024 D&B Track] DACO: Towards Application-Driven and Comprehensive Data Analysis via Code Generation☆14Mar 5, 2025Updated last year
- PathPiece tokenizer☆14Nov 10, 2024Updated last year
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- This is the repo for constructing a comprehensive and rigorous evaluation framework for LLM calibration.☆14Apr 9, 2024Updated 2 years ago
- CaMML:Context-Aware MultiModal Learner for Large Models (ACL 2024 SAC Award)☆15May 21, 2025Updated last year
- AAAI-22 paper: Synthetic Disinformation Attacks on Automated Fact Verification Systems☆12Feb 23, 2022Updated 4 years ago
- ☆14Apr 22, 2024Updated 2 years ago
- Code and data from the paper 'Human Feedback is not Gold Standard'☆20Aug 9, 2026Updated last month
- ☆12Jun 30, 2024Updated 2 years ago
- The official code repo of paper "Chasing the Tail: Effective Rubric-based Reward Modeling for Large Language Model Post-Training"☆32Feb 20, 2026Updated 6 months ago
- ☆12May 6, 2022Updated 4 years ago
- Awesome_CV的中文版本,clone本项目到overleaf即可轻松愉快编写自己的CV☆18May 24, 2024Updated 2 years ago
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- Deep Data Research. Seek More, See Beyond.☆16Feb 6, 2026Updated 7 months ago
- Diverse Demonstrations Improve In-context Compositional Generalization☆13Jul 7, 2023Updated 3 years ago
- ☆29Jan 31, 2026Updated 7 months ago
- Code for "Unlearning Traces the Influential Training Data of Language Models"☆13Jun 13, 2024Updated 2 years ago
- Enhancing Complex Question Answering over Knowledge Graphs through Evidence Pattern Retrieval, WWW 2024☆15Oct 22, 2024Updated last year
- Fairer Preferences Elicit Improved Human-Aligned Large Language Model Judgments (Zhou et al., EMNLP 2024)☆14Oct 3, 2024Updated last year
- Official Code Repository for [AutoScale📈: Scale-Aware Data Mixing for Pre-Training LLMs] Published as a conference paper at **COLM 2025*…☆14Aug 8, 2025Updated last year