☆56May 16, 2026Updated 4 months ago
Alternatives and similar repositories for LLaDA-o
Users that are interested in LLaDA-o are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- ☆24Dec 1, 2025Updated 9 months ago
- [ICLR 26] Context Tokens are Anchors: Understanding the Repeat Curse in dMLLMs from an Information Flow Perspective☆31Mar 6, 2026Updated 6 months ago
- Methods and code for extending the context length of diffusion language models☆56Dec 7, 2025Updated 9 months ago
- Learning 1D Causal Visual Representation with De-focus Attention Networks☆35Jun 7, 2024Updated 2 years ago
- Official PyTorch implementation for "Effective and Efficient Masked Image Generation Models"☆35Apr 8, 2025Updated last year
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- [NeurIPS 2025 Spotlight] Implementation of "KLASS: KL-Guided Fast Inference in Masked Diffusion Models"☆34Jan 3, 2026Updated 8 months ago
- Automatically update arXiv papers about SOT & VLT, Multi-modal Learning, LLM and Video Understanding using Github Actions.☆48Updated this week
- Fast Diffusion-Based Counterfactuals for Shortcut Removal and Generation (ECCV 2024 ORAL)☆17Sep 3, 2024Updated 2 years ago
- [ICME 2023] FlowText: Synthesizing Realistic Scene Text Video with Optical Flow Estimation☆13May 13, 2023Updated 3 years ago
- An imaginary extension of rotary position embeddings for long-context language models☆33Dec 9, 2025Updated 9 months ago
- ☆10Aug 22, 2023Updated 3 years ago
- ✨✨[ICML 2026] Omni-Diffusion: Unified Multimodal Understanding and Generation with Masked Discrete Diffusion☆155Mar 12, 2026Updated 6 months ago
- [ICRA 26] C^2ROPE: Causal Continuous Rotary Positional Encoding for 3D Large Multimodal-Models Reasoning☆28Feb 13, 2026Updated 7 months ago
- [ICLR 2026] AdaBlock-dLLM: Semantic-Aware Diffusion LLM Inference via Adaptive Block Size☆17Jan 28, 2026Updated 7 months ago
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- ☆17Mar 9, 2026Updated 6 months ago
- Official code for paper "Reasoning Fails Where Step Flow Breaks" (ACL 2026)☆19Apr 19, 2026Updated 5 months ago
- [ICLR 2025] A Comprehensive Framework for Developing and Evaluating Multimodal Role-Playing Agents☆101Feb 2, 2026Updated 7 months ago
- Implementation of Em_Garde: a proposal-retrieval framework for streaming video understanding☆33Jun 24, 2026Updated 2 months ago
- This is the official code for MolReasoner: Toward Effective and Interpretable Reasoning for Molecular LLMs☆46Aug 28, 2025Updated last year
- Baselines for Model-Based Optimization installation fixes and compatible with newer AMPERE+ GPUs (e.g. 3090)☆11Apr 30, 2023Updated 3 years ago
- RobuQ: Pushing DiTS to W1.58A2 via Robust Activation Quantization☆17Jun 28, 2026Updated 2 months ago
- Implementation for the paper "Unified Multimodal Model with Unlikelihood Training for Visual Dialog"☆13May 12, 2023Updated 3 years ago
- [ICLR 2025] Code&Data for the paper "Super(ficial)-alignment: Strong Models May Deceive Weak Models in Weak-to-Strong Generalization"☆15Jun 21, 2024Updated 2 years ago
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- [ICLR'26] Official PyTorch implementation of "Time Is a Feature: Exploiting Temporal Dynamics in Diffusion Language Models".☆66Mar 5, 2026Updated 6 months ago
- Official implementation of "Fast-dLLM: Training-free Acceleration of Diffusion LLM by Enabling KV Cache and Parallel Decoding"☆1,088May 30, 2026Updated 3 months ago
- [ICLR 2026] Official repository of "Beyond Fixed: Training-Free Variable-Length Denoising for Diffusion Large Language Models"☆175Feb 16, 2026Updated 7 months ago
- Source code for paper "VD-PCR: Improving Visual Dialog with Pronoun Coreference Resolution"☆10Nov 1, 2022Updated 3 years ago
- Implementaiton of "DiLM: Distilling Dataset into Language Model for Text-level Dataset Distillation" (accepted by NAACL2024 Findings)".☆28Feb 10, 2025Updated last year
- dInfer: An Efficient Inference Framework for Diffusion Language Models☆481Feb 11, 2026Updated 7 months ago
- [NeurIPS 2025] Official implementation for our paper "Scaling Diffusion Transformers Efficiently via μP".☆100Nov 2, 2025Updated 10 months ago
- ☆15Sep 1, 2025Updated last year
- [Preprint] Efficient Generative Model Training via Embedded Representation Warmup☆36Oct 15, 2025Updated 11 months ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- A post-training framework for diffusion language models with supervised fine-tuning and reinforcement learning☆165Mar 30, 2026Updated 5 months ago
- 3D-MolT5: Leveraging Discrete Structural Information for Molecule-Text Modeling (ICLR 2025)☆20Oct 23, 2025Updated 10 months ago
- Seeing Far and Clearly: Mitigating Hallucinations in MLLMs with Attention Causal Decoding (CVPR 2025 Oral)☆43Nov 28, 2025Updated 9 months ago
- ☆16Apr 28, 2023Updated 3 years ago
- Official repository for paper "DeepCritic: Deliberate Critique with Large Language Models"☆41Jun 24, 2025Updated last year
- Official PyTorch implementation for ICLR2025 paper "Scaling up Masked Diffusion Models on Text"☆387Dec 22, 2024Updated last year
- [ICLR2025] γ -MOD: Mixture-of-Depth Adaptation for Multimodal Large Language Models☆46Oct 28, 2025Updated 10 months ago