☆29Jul 16, 2024Updated 2 years ago
Alternatives and similar repositories for CPO
Users that are interested in CPO are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- [ACL'24] Beyond One-Preference-Fits-All Alignment: Multi-Objective Direct Preference Optimization☆101Aug 20, 2024Updated 2 years ago
- Code for our EMNLP 2019 paper titled "Sentence-Level Content Planning and Style Specification for Neural Text Generation"☆17May 4, 2020Updated 6 years ago
- Official code for "Decoding-Time Language Model Alignment with Multiple Objectives".☆30Oct 30, 2024Updated last year
- ☆33Aug 21, 2025Updated last year
- Official code for the paper: DRA-GRPO: Exploring Diversity-Aware Reward Adjustment for R1-Zero-Like Training of Large Language Models☆24Jan 6, 2026Updated 7 months ago
- Bare Metal GPUs on DigitalOcean Gradient AI • AdPurpose-built for serious AI teams training foundational models, running large-scale inference, and pushing the boundaries of what's possible.
- ☆17Oct 9, 2025Updated 10 months ago
- This is the code repo for the paper "RAG-DDR: Optimizing Retrieval-Augmented Generation Using Differentiable Data Rewards".☆23Oct 28, 2024Updated last year
- ☆34Aug 5, 2023Updated 3 years ago
- 哈尔滨工业大学 Typst 论文模板 | 本仓库是 universal-hit-thesis 的镜像,为维护模板早期版本在 Typst Universe 上的链接而设置,请移步至 HITSZ OSA 的仓库☆12Mar 8, 2025Updated last year
- [IROS 2024] "ComTraQ-MPC: Meta-Trained DQN-MPC Integration for Trajectory Tracking with Limited Active Localization Updates" by Gokul Put…☆13Apr 10, 2025Updated last year
- Benchmarking Social Intelligence of Language Agents through Interactive Scenarios☆12Jan 4, 2025Updated last year
- ☆18Mar 23, 2025Updated last year
- Official Repository of "[ICLR26] TRAPO: A Semi-Supervised Reinforcement Learning Framework for Boosting LLM Reasoning"☆29Feb 6, 2026Updated 6 months ago
- PathPiece tokenizer☆14Nov 10, 2024Updated last year
- Proton VPN Special Offer - Get 70% off • AdSpecial partner offer. Trusted by over 100 million users worldwide. Tested, Approved and Recommended by Experts.
- GreenLIT: Using GPT-J with Multi-Task Learning to Create New Screenplays☆16Nov 27, 2022Updated 3 years ago
- ☆15Mar 12, 2022Updated 4 years ago
- Code and data repository for "The Mirage of Model Editing: Revisiting Evaluation in the Wild"☆18Aug 27, 2025Updated last year
- Code to reproduce results from "Invertible generative models for inverse problems: mitigating representation error and dataset bias"☆21Jul 9, 2020Updated 6 years ago
- Implementation of the paper "Improving the Accuracy-Robustness Trade-off of Classifiers via Adaptive Smoothing".☆10Feb 6, 2024Updated 2 years ago
- ☆10Mar 4, 2024Updated 2 years ago
- [COLING 2025] Official repo of paper: "Not Aligned" is Not "Malicious": Being Careful about Hallucinations of Large Language Models' Jail…☆12Jul 26, 2024Updated 2 years ago
- Learning MLPs to replace GNN☆10Jun 3, 2023Updated 3 years ago
- ☆29May 13, 2025Updated last year
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- https://footprints.baulab.info☆17Oct 4, 2024Updated last year
- [AAAI2021] Automated Model Design and Benchmarking of 3D Deep Learning Models for COVID-19 Detection with Chest CT Scans☆13Mar 30, 2024Updated 2 years ago
- Code, data, and pretrained models for the paper "Generating Wikipedia Article Sections from Diverse Data Sources"☆21Feb 5, 2021Updated 5 years ago
- ☆42Nov 21, 2023Updated 2 years ago
- [NLPCC 2021] Shared Task on AutoIE2: Sub-Event Identification☆14Jul 19, 2021Updated 5 years ago
- The code of “Improving Weak-to-Strong Generalization with Scalable Oversight and Ensemble Learning”☆17Feb 26, 2024Updated 2 years ago
- Opensource code for ICML 2026 poster☆16Nov 26, 2025Updated 9 months ago
- DGL implementation of GRAND(Graph Random Neural Network, NeurIPS 2020)☆18Mar 19, 2021Updated 5 years ago
- [NeurIPS 2024] Official code of $\beta$-DPO: Direct Preference Optimization with Dynamic $\beta$☆51Oct 23, 2024Updated last year
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- CodeUltraFeedback: aligning large language models to coding preferences (TOSEM 2025)☆76Jun 25, 2024Updated 2 years ago
- Learning to Generate STRUCTURED Output with Schema Reinforcement Learning☆26Mar 2, 2025Updated last year
- Secure and Scalable Federated Learning using Serverless Computing☆14Jan 31, 2024Updated 2 years ago
- [EMNLP 2024 Main] Official repository of paper "SLANG: New Concept Comprehension of Large Language Models"☆14Oct 27, 2024Updated last year
- Codebase for Math Neurosurgery: Isolating LLMs' Math Reasoning Abilities Using Only Forward Passes☆24Jun 15, 2025Updated last year
- Out-of-distribution generalization benchmarks for image recognition models☆14Apr 5, 2020Updated 6 years ago
- Investigating and Defending Shortcut Learning in Personalized Diffusion Models☆14Nov 19, 2024Updated last year