Source code to accompany research paper on training multi token prediction language models using self-distillation.
☆42Feb 21, 2026Updated 6 months ago
Alternatives and similar repositories for mtp-lm
Users that are interested in mtp-lm are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- ☆14Mar 2, 2025Updated last year
- Code for "DynaGuard: A Dynamic Guardrail Model With User-Defined Policies."☆23Nov 3, 2025Updated 9 months ago
- (ECCV 2026): Official code for Flash-BoN: Instant Drafts for Inference-Time Scaling in Diffusion Models☆21Jul 9, 2026Updated last month
- ☆20Nov 4, 2025Updated 9 months ago
- Learning from Mixed Rollouts: Logit Fusion as a Bridge Between Imitation and Exploration☆18Feb 24, 2026Updated 6 months ago
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- Code for Dayal Kalra's research internship on scalable curvature measures for neural networks.☆29Feb 3, 2026Updated 6 months ago
- Code for the paper Multi-Stream LLMs: Unblocking Language Models with Parallel Streams of Thoughts, Inputs and Outputs☆70Jun 23, 2026Updated 2 months ago
- ☆31May 21, 2026Updated 3 months ago
- Official implementation of "Learning To Draft: Adaptive Speculative Decoding with Reinforcement Learning" (ICLR 2026)☆22Mar 1, 2026Updated 6 months ago
- [ICML 2026] Hybrid Policy Distillation (HPD) is a practical distillation framework for reasoning-oriented language models. This repositor…☆24Apr 24, 2026Updated 4 months ago
- 📄 Forgetting in Language Models: Capacity, Optimization, and Self-Generated Replay☆25Jul 17, 2026Updated last month
- Gemstones: A Model Suite for Multi-Faceted Scaling Laws (NeurIPS 2025)☆35Sep 28, 2025Updated 11 months ago
- SCT: An Efficient Self-Supervised Cross-View Training For Sentence Embedding (TACL)☆16Jul 27, 2024Updated 2 years ago
- ☆66Jul 3, 2026Updated last month
- Managed Database hosting by DigitalOcean • AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- PyTorch Implementation of Zero-Shot Vision Encoder Grafting via LLM Surrogates [ICCV'25]☆54Jul 10, 2025Updated last year
- [ICML2026] Reproduce Kimi K1.5/K2 RL algorithm and rollout system☆21Apr 9, 2026Updated 4 months ago
- Cross-GPU KV Cache Marketplace☆27Nov 12, 2025Updated 9 months ago
- Official Implementation of MARS☆30Apr 21, 2026Updated 4 months ago
- ☆18Sep 6, 2025Updated 11 months ago
- [ICLR 2026] Official code for BézierFlow: Learning Bézier Stochastic Interpolant Schedulers for Few-Step Generation☆23Apr 13, 2026Updated 4 months ago
- A Teacher–Student Cooperation Framework to Synthesize Student-Consistent SFT Data☆37May 1, 2026Updated 4 months ago
- A selective knowledge distillation algorithm for efficient speculative decoders☆39Nov 27, 2025Updated 9 months ago
- [NeurIPS 2024] Goldfish Loss: Mitigating Memorization in Generative LLMs☆98Nov 17, 2024Updated last year
- Managed Database hosting by DigitalOcean • AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- ☆22Feb 10, 2025Updated last year
- Accelerating RL for LLM Reasoning with Optimal Advantage Regression☆41May 30, 2025Updated last year
- PyTorch Implementation of Image Generation with a Sphere Encoder☆46May 20, 2026Updated 3 months ago
- PyTorch implementation for "Generative Modeling on Manifolds Through Mixture of Riemannian Diffusion Processes" (ICML 2024).☆13Jul 21, 2024Updated 2 years ago
- Auditing agents for fine-tuning safety☆22Oct 21, 2025Updated 10 months ago
- On the Quartic Invariant of Odd Degree Binary Forms — paper, Lean formalization, and computational verification☆16Apr 16, 2026Updated 4 months ago
- ProAct is a framework designed to enable Large Language Model (LLM) agents to perform accurate, multi-turn lookahead reasoning in interac…☆18Feb 11, 2026Updated 6 months ago
- Inverse Scaling in Test-Time Compute☆26Dec 3, 2025Updated 8 months ago
- [EMNLP 2026] MDM-Prime-v2: Binary Encoding and Index Shuffling Enable Scaling of Diffusion Language Models☆30May 23, 2026Updated 3 months ago
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- ☆46Jan 30, 2026Updated 7 months ago
- An MLX implementation of Meta AI's ESM-2 protein language model☆16Aug 16, 2025Updated last year
- ☆19May 25, 2026Updated 3 months ago
- [ICML 2026] Official implementation for paper: Learning Self-Correction in Vision–Language Models via Rollout Augmentation☆16Jun 4, 2026Updated 2 months ago
- The official implementation of NOSA☆20Jun 11, 2026Updated 2 months ago
- [ICML 2026] Code for V1: Unifying Generation and Self-Verification for Parallel Reasoners.☆39Mar 5, 2026Updated 5 months ago
- This repository contains code for the paper "Better Estimation of the KL Divergence Between Language Models"☆19May 30, 2025Updated last year