[ICML 2026] d3LLM: Ultra-Fast Diffusion LLM π
β154May 1, 2026Updated 4 months ago
Alternatives and similar repositories for d3LLM
Users that are interested in d3LLM are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- [ICLR 2026] dParallel: Learnable Parallel Decoding for dLLMsβ67Apr 12, 2026Updated 5 months ago
- Official implementation of "Fast-dLLM: Training-free Acceleration of Diffusion LLM by Enabling KV Cache and Parallel Decoding"β1,090May 30, 2026Updated 3 months ago
- [ICLR 2026] Official code for TraceRL: Revolutionizing post-training for Diffusion LLMs, powering the SOTA TraDo series.β521Jan 28, 2026Updated 7 months ago
- [ICLR 2026] Discrete Diffusion Forcing (D2F): dLLMs Can Do Faster-Than-AR Inferenceβ262Feb 3, 2026Updated 7 months ago
- LoPA: Scaling dLLM Inference via Lookahead Parallel Decodingβ40Aug 3, 2026Updated last month
- GPU virtual machines on DigitalOcean Gradient AI β’ AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- dInfer: An Efficient Inference Framework for Diffusion Language Modelsβ482Feb 11, 2026Updated 7 months ago
- [ICML 2026] Jacobi Forcing: Fast and Accurate Diffusion-style Decodingβ120Feb 20, 2026Updated 7 months ago
- SDAR (Synergy of Diffusion and AutoRegression), a large diffusion language modelοΌ1.7B, 4B, 8B, 30BοΌβ528Jul 29, 2026Updated last month
- JetSpec: Breaking the Scaling Ceiling of Speculative Decoding with Causal Parallel Tree Draftingβ227Updated this week
- Free Draft-and-Verification: Toward Lossless Parallel Decoding for Diffusion Large Language Modelsβ24May 19, 2026Updated 4 months ago
- Easy and Efficient dLLM Fine-Tuningβ270Sep 10, 2026Updated last week
- Efficient Long-context Language Model Training by Core Attention Disaggregationβ105Apr 7, 2026Updated 5 months ago
- Flexible and Pluggable Serving Engine for Diffusion LLMsβ156Jul 13, 2026Updated 2 months ago
- [NeurIPS'25] dKV-Cache: The Cache for Diffusion Language Modelsβ136May 22, 2025Updated last year
- AI Agents on DigitalOcean Gradient AI Platform β’ AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- [ICLR 2026] Learning to Parallel: Accelerating Diffusion Large Language Models via Learnable Parallel Decodingβ35Jan 27, 2026Updated 7 months ago
- Official PyTorch implementation of the paper "Accelerating Diffusion Large Language Models with SlowFast Sampling: The Three Golden Princβ¦β46Jul 18, 2025Updated last year
- [ICML 2026 Outstanding Paper] Minimalist RL for Diffusion LLMs. 89.1% on GSM8K.β280Jul 6, 2026Updated 2 months ago
- [Accepted By EMNLP 2026 Main Conference] Sequential Diffusion Language Model (SDLM) enhances pre-trained autoregressive language models bβ¦β99Dec 27, 2025Updated 8 months ago
- [NeurIPS 2025] Scaling Speculative Decoding with Lookahead Reasoningβ69Oct 31, 2025Updated 10 months ago
- DMax: Aggressive Parallel Decoding for dLLMsβ132Jul 5, 2026Updated 2 months ago
- CDLM: Consistency Diffusion Language Models for Faster Samplingβ41Nov 25, 2025Updated 9 months ago
- A post-training framework for diffusion language models with supervised fine-tuning and reinforcement learningβ165Mar 30, 2026Updated 5 months ago
- LLaDA2.0 is the diffusion language model series developed by InclusionAI team, Ant Group.β528Jul 22, 2026Updated 2 months ago
- Deploy on Railway without the complexity - Free Credits Offer β’ AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- [ICML 2026] Official repository for the paper "Light Forcing: Accelerating Autoregressive Video Diffusion via Sparse Attention"β52Aug 5, 2026Updated last month
- [ICLR 2026] Official PyTorch implementation for "ReFusion: A Diffusion Large Language Model with Parallel Autoregressive Decoding"β64Dec 26, 2025Updated 8 months ago
- Official Implementation for the paper "d1: Scaling Reasoning in Diffusion Large Language Models via Reinforcement Learning"β457Aug 30, 2026Updated 3 weeks ago
- Code for paper "SPG Sandwiched Policy Gradient for Masked Diffusion Language Models"β66Oct 29, 2025Updated 10 months ago
- DLLM-Searcher has been accepted by SIGIR 2026! π₯³β34Jan 23, 2026Updated 8 months ago
- [ICLR 2026] ParallelBench: Understanding the Tradeoffs of Parallel Decoding in Diffusion LLMsβ47Mar 27, 2026Updated 5 months ago
- Official PyTorch implementation for "Principled RL for Diffusion LLMs Emerges from a Sequence-Level Perspective"β40Jan 25, 2026Updated 7 months ago
- dLLM: Simple Diffusion Language Modelingβ2,693Jul 17, 2026Updated 2 months ago
- Open diffusion language model for code generation β releasing pretraining, evaluation, inference, and checkpoints.β656Jul 20, 2026Updated 2 months ago
- Bare Metal GPUs on DigitalOcean Gradient AI β’ AdPurpose-built for serious AI teams training foundational models, running large-scale inference, and pushing the boundaries of what's possible.
- A lightweight Inference Engine built for block diffusion modelsβ47Apr 12, 2026Updated 5 months ago
- Stable-DiffCoder is a family of lightweight open-source code DLLMs(diffusion large language models) comprising base and instruct models, β¦β85Mar 9, 2026Updated 6 months ago
- Advancing Block Diffusion Language Models for Test-Time Scalingβ16Feb 14, 2026Updated 7 months ago
- Diffusion Language Models For Code Infilling Beyond Fixed-size Canvasβ119Feb 3, 2026Updated 7 months ago
- Official PyTorch implementation of the paper "dLLM-Cache: Accelerating Diffusion Large Language Models with Adaptive Caching" (dLLM-Cacheβ¦β214May 1, 2026Updated 4 months ago
- Official implementation of "Diffusion Language Models Know the Answer Before Decoding"β61Apr 28, 2026Updated 4 months ago
- Vortex: Programmable Sparse Attention for Agents as Algorithm Designersβ69Aug 22, 2026Updated last month