Codebase of 'From Denoising to Refining: A Corrective Framework for Vision-Language Diffusion Model'
☆45Jun 27, 2026Updated 3 weeks ago
Alternatives and similar repositories for ReDiff
Users that are interested in ReDiff are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- A unified framework for controllable caption generation across images, videos, and audio. Supports multi-modal inputs and customizable ca…☆54Jul 24, 2025Updated 11 months ago
- [ICLR 2025] ChartMimic: Evaluating LMM’s Cross-Modal Reasoning Capability via Chart-to-Code Generation☆132Dec 19, 2025Updated 7 months ago
- [ICML 2026] The offical code of Diversity-Preserved Distribution Matching Distillation for Fast Visual Synthesis☆87Jun 2, 2026Updated last month
- [2025-TMLR] A Survey on the Honesty of Large Language Models☆66Dec 8, 2024Updated last year
- [CVPR 2026] See Less, See Right: Bi-directional Perceptual Shaping For Multimodal Reasoning☆21Jun 28, 2026Updated 3 weeks ago
- End-to-end encrypted email - Proton Mail • AdSpecial offer: 40% Off Yearly / 80% Off First Month. All Proton services are open source and independently audited for security.
- ☆92May 8, 2026Updated 2 months ago
- Free Draft-and-Verification: Toward Lossless Parallel Decoding for Diffusion Large Language Models☆23May 19, 2026Updated 2 months ago
- [ICML 2026] Smaller Models are Natural Explorers for Policy-Level Diversity in GRPO☆19Jun 15, 2026Updated last month
- The official implementation of dLLM-Var☆35Nov 6, 2025Updated 8 months ago
- [NeurIPS'25] The official code of "PeRL: Permutation-Enhanced Reinforcement Learning for Interleaved Vision-Language Reasoning"☆30Mar 30, 2026Updated 3 months ago
- [NLPCC 2021] Shared Task on AutoIE2: Sub-Event Identification☆14Jul 19, 2021Updated 5 years ago
- [EMNLP 2023] Question Answering as Programming for Solving Time-Sensitive Questions☆12Dec 18, 2023Updated 2 years ago
- Tracking the latest and greatest research papers on diffusion large language models.☆32Mar 13, 2026Updated 4 months ago
- [ICCV 2025] Prompt-A-Video☆24Feb 2, 2025Updated last year
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- ☆34Feb 12, 2026Updated 5 months ago
- The code for paper "LLM-Neo: Parameter Efficient Knowledge Distillation for Large Language Models"☆15Mar 2, 2025Updated last year
- [ICLR 2026 Oral & ICML 2026] Generative Universal Verifier as Multimodal Meta-Reasoner☆64May 29, 2026Updated last month
- [ICLR 2026] dParallel: Learnable Parallel Decoding for dLLMs☆65Apr 12, 2026Updated 3 months ago
- [ICLR'26] Official PyTorch implementation of "Time Is a Feature: Exploiting Temporal Dynamics in Diffusion Language Models".☆66Mar 5, 2026Updated 4 months ago
- ☆63Jun 25, 2024Updated 2 years ago
- [ICASSP 2022] Official PyTorch Implementation for "Attention Probe: Vision Transformer Distillation in the Wild" (ICASSP 2022)☆11Jan 23, 2022Updated 4 years ago
- [ECCV'24] Official Implementation of Autoregressive Visual Entity Recognizer.☆14Mar 2, 2024Updated 2 years ago
- [EMNLP 2024] A Multi-level Hallucination Diagnostic Benchmark for Tool-Augmented Large Language Models.☆22Sep 23, 2024Updated last year
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- RePO: Replay-Enhanced Policy Optimization☆23Jun 12, 2025Updated last year
- [ICLR 2026] High-Fidelity Visual Reasoning on Structured Images☆29Updated this week
- 基于paddlex目标检测的工业场景下违规使用手机识别。☆12Jun 11, 2022Updated 4 years ago
- Official Implementation of LaViDa: :A Large Diffusion Language Model for Multimodal Understanding☆227Dec 17, 2025Updated 7 months ago
- Project for "LaSagnA: Language-based Segmentation Assistant for Complex Queries".☆63Apr 29, 2024Updated 2 years ago
- ☆15Updated this week
- Official PyTorch implementation for "Principled RL for Diffusion LLMs Emerges from a Sequence-Level Perspective"☆39Jan 25, 2026Updated 5 months ago
- [NeurIPS 2023] Assessor360: Multi-sequence Network for Blind Omnidirectional Image Quality Assessment☆38Oct 11, 2023Updated 2 years ago
- [NeurIPS 2025] Wan-Move: Motion-controllable Video Generation via Latent Trajectory Guidance☆644Jan 5, 2026Updated 6 months ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- The official implementation of MaskGRPO: Consolidating Reinforcement Learning for Multimodal Discrete Diffusion Models. (ICLR 2026, arxiv…☆19Jan 27, 2026Updated 5 months ago
- Memorize-and-Generate: Towards Long-Term Consistency in Real-Time Video Generation☆18Mar 20, 2026Updated 4 months ago
- ☆29May 24, 2024Updated 2 years ago
- ☆38Oct 11, 2022Updated 3 years ago
- A Benchmark for Evaluating MLLMs' Geometry Performance on Long-Step Problems Requiring Auxiliary Lines☆38Apr 27, 2026Updated 2 months ago
- Reflect-DiT: Inference-Time Scaling for Text-to-Image Diffusion Transformers via In-Context Reflection☆56Aug 16, 2025Updated 11 months ago
- [EMNLP 2025 Oral] Official codebase for Seeing More, Saying More: Lightweight Language Experts are Dynamic Video Token Compressors.☆18Sep 7, 2025Updated 10 months ago