[ICML 2026 Spotlight] Code for miXed Discrete Diffusion Language Model
☆29Mar 16, 2026Updated 4 months ago
Alternatives and similar repositories for XDLM
Users that are interested in XDLM are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Official repo of From Indoor to Open World: Revealing the Spatial Reasoning Gap in MLLMs☆24Jun 23, 2026Updated last month
- Implementation of paper "CC-Diff: Enhancing Contextual Coherence in Remote Sensing Image Synthesis"☆28Dec 19, 2025Updated 7 months ago
- [ICLR 2026] Geometric-Mean Policy Optimization☆104Jan 26, 2026Updated 6 months ago
- The official implementation of MaskGRPO: Consolidating Reinforcement Learning for Multimodal Discrete Diffusion Models. (ICLR 2026, arxiv…☆19Jan 27, 2026Updated 6 months ago
- official repo for `thinking with images through-self-calling`☆26Dec 28, 2025Updated 7 months ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- GlimpRouter: Efficient Collaborative Inference by Glimpsing One Token of Thoughts☆16Apr 24, 2026Updated 3 months ago
- [CVPR 2026] LongVideo-R1: Smart Navigation for Low-cost Long Video Understanding☆51Jul 7, 2026Updated last month
- [ICCV 2023] Generative Prompt Model for Weakly Supervised Object Localization☆57Nov 10, 2023Updated 2 years ago
- Decorrelate Irrelevant, Purify Relevant: Overcome Textual Spurious Correlations from a Feature Perspective☆11Nov 16, 2022Updated 3 years ago
- [AAAI2025] ChatterBox: Multi-round Multimodal Referring and Grounding, Multimodal, Multi-round dialogues☆62May 2, 2025Updated last year
- ☆20Dec 14, 2024Updated last year
- [CVPR2025] Hybrid-Level Instruction Injection for Video Token Compression in Multi-modal Large Language Models☆21Apr 30, 2025Updated last year
- [NeurIPS 2024] Artemis: Towards Referential Understanding in Complex Videos☆27Apr 8, 2025Updated last year
- [ICLR 2026] dParallel: Learnable Parallel Decoding for dLLMs☆66Apr 12, 2026Updated 3 months ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- CVPR2024, Semantic-aware SAM for Point-Prompted Instance Segmentation☆38Jan 20, 2025Updated last year
- ☆18Mar 26, 2026Updated 4 months ago
- ☆76Mar 1, 2023Updated 3 years ago
- 2D-TPE: Two-Dimensional Positional Encoding Enhances Table Understanding for Large Language Models (WWW 2025)☆10Apr 15, 2025Updated last year
- ☆39Updated this week
- CANDI: Continuous and Discrete Diffusion☆28Oct 27, 2025Updated 9 months ago
- ☆25Dec 16, 2025Updated 7 months ago
- [ECCV 2024] ControlCap: Controllable Region-level Captioning☆81Oct 25, 2024Updated last year
- The first continuous diffusion language model that rivals discrete counterparts on standard language modeling benchmarks like LM1B and Op…☆88Jun 14, 2026Updated last month
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- Official InfiniBench: A Benchmark for Large Multi-Modal Models in Long-Form Movies and TV Shows☆20Nov 4, 2025Updated 9 months ago
- ☆15Jul 20, 2026Updated 3 weeks ago
- Official Repo for DAC-RL: Training LLMs for Divide-and-Conquer Reasoning Elevates Test-Time Scalability☆16Feb 26, 2026Updated 5 months ago
- Revealing and unlocking the context boundary of reward models☆21May 10, 2026Updated 3 months ago
- [CVPR2025] VideoICL: Confidence-based Iterative In-context Learning for Out-of-Distribution Video Understanding☆24Mar 24, 2025Updated last year
- ☆22Jan 2, 2026Updated 7 months ago
- ☆25Dec 13, 2024Updated last year
- [ICLR 2025] Large (Vision) Language Models are Unsupervised In-Context Learners☆22Jun 6, 2025Updated last year
- [ICLR 2025] LongPO: Long Context Self-Evolution of Large Language Models through Short-to-Long Preference Optimization☆44Feb 27, 2025Updated last year
- Virtual machines for every use case on DigitalOcean • AdGet dependable uptime with 99.99% SLA, simple security tools, and predictable monthly pricing with DigitalOcean's virtual machines, called Droplets.
- Cross-Self KV Cache Pruning for Efficient Vision-Language Inference☆10Dec 15, 2024Updated last year
- Remasking Discrete Diffusion Models with Inference-Time Scaling☆77Feb 7, 2026Updated 6 months ago
- All-in-one benchmarking platform for evaluating LLM.☆15Nov 12, 2025Updated 8 months ago
- [ICML 2026] Official Implementation of Prism: Efficient Test-Time Scaling via Hierarchical Search and Self-Verification for Discrete Diff…☆22Mar 4, 2026Updated 5 months ago
- (CVPR2023/TPAMI2024) Integrally Pre-Trained Transformer Pyramid Networks -- A Hierarchical Vision Transformer for Masked Image Modeling☆216Jul 28, 2024Updated 2 years ago
- ☆27Feb 18, 2024Updated 2 years ago
- The official data and code for EMNLP 2023 main conference paper: CRT-QA: A Dataset of Complex Reasoning Question Answering over Tabular D…☆13May 19, 2025Updated last year