DPO, but faster 🚀
☆51Dec 6, 2024Updated last year
Alternatives and similar repositories for dpo-prefix-sharing
Users that are interested in dpo-prefix-sharing are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- This repository contains the joint use of CPO and SimPO method for better reference-free preference learning methods.☆59Aug 13, 2024Updated 2 years ago
- Training and evaluation code for the paper "Headless Language Models: Learning without Predicting with Contrastive Weight Tying" (https:/…☆30Apr 17, 2024Updated 2 years ago
- rabitq rust implementation☆11May 14, 2026Updated 3 months ago
- torch_remat fine-grained activation checkpointing API☆16Aug 19, 2026Updated 2 weeks ago
- My hybrid TTS network that combines, VALL-E, VoiceBox, SpeechFlow, Seamless and TortoiseTTS into one☆26Aug 5, 2024Updated 2 years ago
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- ☆25Aug 20, 2025Updated last year
- Official code of *Virgo: A Preliminary Exploration on Reproducing o1-like MLLM*☆20May 27, 2025Updated last year
- ☆15Feb 5, 2026Updated 7 months ago
- MUX-PLMs: Pretraining LMs with Data Multiplexing☆15Jan 29, 2023Updated 3 years ago
- ☆49Nov 10, 2023Updated 2 years ago
- ☆38Jul 16, 2025Updated last year
- Latent Large Language Models☆19Aug 24, 2024Updated 2 years ago
- Impact of typos and common misspellings on LLM task performance.☆25Mar 22, 2024Updated 2 years ago
- Nearest Neighbor Normalization (EMNLP 2024)☆21Nov 1, 2024Updated last year
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- ☆16Jul 29, 2025Updated last year
- NISQA - Non-Intrusive Speech Quality and TTS Naturalness Assessment☆16Apr 13, 2022Updated 4 years ago
- Experimental repository for research implementation of NoLoCo.☆31Mar 26, 2026Updated 5 months ago
- ☆16Feb 6, 2024Updated 2 years ago
- ☆56Apr 30, 2025Updated last year
- Official implementation of "GPT or BERT: why not both?"☆65Jul 28, 2025Updated last year
- ☆23Jan 23, 2026Updated 7 months ago
- ☆18Apr 23, 2025Updated last year
- ☆83Jun 8, 2026Updated 2 months ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- ☆116Jan 21, 2025Updated last year
- Investigating the generalization behavior of LM probes trained to predict truth labels: (1) from one annotator to another, and (2) from e…☆34May 23, 2024Updated 2 years ago
- Official code for our paper, "LoRA-Pro: Are Low-Rank Adapters Properly Optimized? "☆149Apr 8, 2025Updated last year
- HPLT Analytics☆16Jul 22, 2026Updated last month
- ☆36Mar 12, 2025Updated last year
- Contains my experiments with the `big_vision` repo to train ViTs on ImageNet-1k.☆22Jan 16, 2023Updated 3 years ago
- Bridge Megatron-Core to Hugging Face/Reinforcement Learning☆231Jun 15, 2026Updated 2 months ago
- Official implementation of ECCV24 paper: POA☆24Aug 8, 2024Updated 2 years ago
- Large language models (LLMs) made easy, EasyLM is a one stop solution for pre-training, finetuning, evaluating and serving LLMs in JAX/Fl…☆78Aug 17, 2024Updated 2 years ago
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- Optimizing Anytime Reasoning via Budget Relative Policy Optimization☆54Jul 15, 2025Updated last year
- ☆20Nov 4, 2025Updated 10 months ago
- Kinetics: Rethinking Test-Time Scaling Laws☆86Jul 11, 2025Updated last year
- Landing repository for the paper "Predicting the Order of Upcoming Tokens Improves Language Modeling"☆48May 13, 2026Updated 3 months ago
- [ICLR 2026] RuleReasoner: Reinforced Rule-based Reasoning via Domain-aware Dynamic Sampling☆42Feb 25, 2026Updated 6 months ago
- [ACL 2025] Outlier-Safe Pre-Training for Robust 4-Bit Quantization of Large Language Models☆39Nov 4, 2025Updated 10 months ago
- [ICLR2025] γ -MOD: Mixture-of-Depth Adaptation for Multimodal Large Language Models☆45Oct 28, 2025Updated 10 months ago