[ICML 2026][Ultra Powerful Few-Step Diffusion RL] TDM-R1: Reinforcing Few-Step Diffusion Models with Non-Differentiable Reward
☆116May 25, 2026Updated last month
Alternatives and similar repositories for TDM-R1
Users that are interested in TDM-R1 are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- [ICCV 2025][Few-Step Student Surpasses Teacher Diffusion] Learning Few-Step Diffusion Models by Trajectory Distribution Matching☆99Mar 16, 2026Updated 4 months ago
- [ECCV 2026] Official Code of "Distribution Matching Distillation Meets Reinforcement Learning"☆285Feb 1, 2026Updated 5 months ago
- Official Repo of "D-OPSD: On-Policy Self-Distillation for Continuously Tuning Step-Distilled Diffusion Models"☆289May 22, 2026Updated 2 months ago
- Continuous-Time Distribution Matching for Few-Step Diffusion Distillation👏☆147May 11, 2026Updated 2 months ago
- paper collection: alignment of diffusion models☆29Mar 6, 2026Updated 4 months ago
- Virtual machines for every use case on DigitalOcean • AdGet dependable uptime with 99.99% SLA, simple security tools, and predictable monthly pricing with DigitalOcean's virtual machines, called Droplets.
- [ICML 2026] Official implementation for "DirectEdit: Step-Level Accurate Inversion for Flow-Based Image Editing".☆28May 5, 2026Updated 2 months ago
- [ICLR 2026][Ultra Fast&Powerful Diffusion RL] Reinforcing Diffusion Models by Direct Group Preference Optimization☆85May 26, 2026Updated last month
- ArcFlow: Unleashing 2-Step Text-to-Image Generation via High-Precision Non-Linear Flow Distillation☆128May 20, 2026Updated 2 months ago
- [ICLR 2025] You Only Sample Once: Taming One-Step Text-To-Image Synthesis by Self-Cooperative Diffusion GANs☆71Mar 7, 2026Updated 4 months ago
- 🚀 [ICLR 2026] SenseFlow: Scaling Distribution Matching for Flow-based Text-to-Image Distillation☆113Mar 14, 2026Updated 4 months ago
- [arXiv 2026] This is the official PyTorch implementation of "MeanFlowNFT: Bringing Forward-Process RL to Average-Velocity Generators".☆62Updated this week
- [NeurIPS 2025] Reward-Instruct: A Reward-Centric Approach to Fast Photo-Realistic Image Generation☆35Oct 24, 2025Updated 9 months ago
- [ICML 2026] The offical code of Diversity-Preserved Distribution Matching Distillation for Fast Visual Synthesis☆87Jun 2, 2026Updated last month
- Official implementation of Pref-GRPO: Pairwise Preference Reward-based GRPO for Stable Text-to-Image Reinforcement Learning☆274Feb 10, 2026Updated 5 months ago
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- Official implementation of AsymFlow, pi-Flow, GMFlow☆452Jul 14, 2026Updated last week
- [arXiv 2026] This is the official PyTorch implementation of "RTDMD: Reinforcing Few-step Generators via Reward-Tilted Distribution Matchi…☆41Jun 6, 2026Updated last month
- Official implementation of HPSv3: Towards Wide-Spectrum Human Preference Score (ICCV2025)☆326Dec 5, 2025Updated 7 months ago
- ☆23Jun 29, 2026Updated 3 weeks ago
- Piecewise Sparse Attention Is Wiser for Efficient Diffusion Transformers☆34Jul 1, 2026Updated 3 weeks ago
- Edit-R1: Reinforce Image Editing with Diffusion Negative-Aware Finetuning and MLLM Implicit Feedback☆295Jan 24, 2026Updated 6 months ago
- Self-Evolving Image Generation Agents via Tool-Orchestrated Visual Experience Distillation☆77May 22, 2026Updated 2 months ago
- [ICML'26] Code and website for Self-Flow: Self-Supervised Flow Matching for Scalable Multi-Modal Synthesis☆535May 23, 2026Updated 2 months ago
- [ICLR 2026 Oral] DiffusionNFT: Online Diffusion Reinforcement with Forward Process☆979Feb 10, 2026Updated 5 months ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- [ICLR 2026] Taming large-scale few-step training with self-adversarial flows! 👏🏻☆536Feb 24, 2026Updated 5 months ago
- [SIGGRAPH Asia 2026] DiffusionOPD: A Unified Perspective of On-Policy Distillation in Diffusion Models☆140Updated this week
- A unified framework for easy reinforcement learning in Flow-Matching models☆636Jul 12, 2026Updated last week
- FDFO: Finite Difference Flow Optimization☆115Apr 27, 2026Updated 2 months ago
- CoTj (Chain-of-Trajectories) upgrades diffusion models from fixed System-1 denoising schedules to System-2 style, condition-adaptive traj…☆23Mar 24, 2026Updated 4 months ago
- Official implementation of Tuna-2: Pixel Embeddings Beat Vision Encoders for Unified Understanding and Generation☆738Updated this week
- rCM & Causal-rCM: Leading and Unified Algorithms/Infrastructures for Bidirectional/Autoregressive Video Diffusion Distillation at Scale☆770Jun 25, 2026Updated 3 weeks ago
- Official Implementation of SAGE-GRPO:Manifold-Aware Exploration for Reinforcement Learning in Video Generation☆126Apr 2, 2026Updated 3 months ago
- CVPR 2026 Highlight: Pluggable Pruning with Contiguous Layer Distillation for Diffusion Transformers☆86Apr 9, 2026Updated 3 months ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- Flow Map OPD for AnyStep Video Diffusion☆398May 23, 2026Updated 2 months ago
- [NeurIPS 2025] An official implementation of Flow-GRPO: Training Flow Matching Models via Online RL☆2,428May 7, 2026Updated 2 months ago
- [ICML 2026] Official PyTorch implementation of paper “CoCoEdit: Content-Consistent Image Editing via Region Regularized Reinforcement Lea…☆26Jun 14, 2026Updated last month
- Official codes for the paper "GARDO: Reinforcing Diffusion Models without Reward Hacking"☆61May 3, 2026Updated 2 months ago
- ☆79Jul 3, 2026Updated 3 weeks ago
- [NeurIPS 2025 D&B🔥] ImgEdit: A Unified Image Editing Dataset and Benchmark☆327Nov 5, 2025Updated 8 months ago
- PiD: Fast and High-Resolution Latent Decoding with Pixel Diffusion☆976Updated this week