[ICLR 2026] π» Uniform Discrete Diffusion with Metric Path for Video Generation
β126May 20, 2026Updated 4 months ago
Alternatives and similar repositories for URSA
Users that are interested in URSA are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- [CVPR2026] Exploring Spatial Intelligence from a Generative Perspectiveβ32Jun 3, 2026Updated 3 months ago
- [ICML 2026 Spotlight] UDM-GRPO: Stable and Efficient Group Relative Policy Optimization for Uniform Discrete Diffusion Modelsβ29Sep 9, 2026Updated 2 weeks ago
- [ACL'26] EvoToken-DLM (Beyond Hard Masks: Progressive Token Evolution for Diffusion Language)β49Apr 7, 2026Updated 5 months ago
- [ICLR 2025] Autoregressive Video Generation without Vector Quantizationβ660Oct 29, 2025Updated 10 months ago
- β35Apr 10, 2026Updated 5 months ago
- Deploy to Railway using AI coding agents - Free Credits Offer β’ AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- β21Jun 13, 2026Updated 3 months ago
- [NeurIPS 2025] Unveiling Chain of Step Reasoning for Vision-Language Models with Fine-grained Rewardsβ18Oct 6, 2025Updated 11 months ago
- Native Multimodal Models are World Learnersβ1,557Dec 30, 2025Updated 8 months ago
- Unsupervised Learning of Generalizable Robot Motion from Compact State Representationβ45Jun 10, 2026Updated 3 months ago
- [ICLR'26] Official PyTorch implementation of "Time Is a Feature: Exploiting Temporal Dynamics in Diffusion Language Models".β66Mar 5, 2026Updated 6 months ago
- Official implementation of paper "VMoBA: Mixture-of-Block Attention for Video Diffusion Models"β65Jul 1, 2025Updated last year
- Repo of HawkLlama.β16Jan 2, 2025Updated last year
- TVRBench: Target Viewpoint Reproduction Benchmark for Active Spatial Intelligenceβ29Jun 2, 2026Updated 3 months ago
- One-shot and Few-shot 3D Editing without Per-Scene Optimizationβ175Aug 21, 2025Updated last year
- Managed Kubernetes at scale on DigitalOcean β’ AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- [ICCV'25] Unified Open-World Segmentation with Multi-Modal Promptsβ16Jun 16, 2026Updated 3 months ago
- [NeurIPS 2025 Oral]InfinityβοΈ: Uniο¬ed Spacetime AutoRegressive Modeling for Visual Generationβ787Apr 16, 2026Updated 5 months ago
- [NeurIPS 2025 Spotlight] FUDOKI: Discrete Flow-based Unified Understanding and Generation via Kinetic-Optimal Velocitiesβ78Dec 21, 2025Updated 9 months ago
- [NeurIPS'24] A Simple Image Segmentation Framework via In-Context Examplesβ68Oct 29, 2024Updated last year
- [CVPR2026] Eliciting Complex Spatial Reasoning in MLLMs through Wide-Baseline Matchingβ19Jun 4, 2026Updated 3 months ago
- [ICML2026] ACTIVE-O3: Empowering Multimodal Large Language Models with Active Perception via GRPOβ84Apr 30, 2026Updated 4 months ago
- [NeurIPS 2025] Training-Free Efficient Video Generation via Dynamic Token Carvingβ289Aug 4, 2025Updated last year
- β36Oct 21, 2022Updated 3 years ago
- [ICLR 2026] LayerSync: Self-aligning Intermediate Layersβ25Mar 21, 2026Updated 6 months ago
- Deploy open-source AI quickly and easily - Special Bonus Offer β’ AdRunpod Hub is built for open source. One-click deployment and autoscaling endpoints without provisioning your own infrastructure.
- [ICLR 2026] Official Repo for Rolling Forcing: Autoregressive Long Video Diffusion in Real Timeβ460Oct 31, 2025Updated 10 months ago
- JoVA: Unified Multimodal Learning for Joint Video-Audio Generationβ34Dec 22, 2025Updated 9 months ago
- [NeurIPS'24] Unleashing the Potential of the Diffusion Model in Few-shot Semantic Segmentation (Diffews)β51Apr 14, 2025Updated last year
- [ICLR 2026 Oral] DiffusionNFT: Online Diffusion Reinforcement with Forward Processβ1,069Feb 10, 2026Updated 7 months ago
- DenseFusion-1M: Merging Vision Experts for Comprehensive Multimodal Perceptionβ161Dec 6, 2024Updated last year
- [ACL 2026 Findings, ICCV 2025 Workshop Outstanding Paper Award] VChain: Chain-of-Visual-Thought for Reasoning in Video Generationβ121Apr 8, 2026Updated 5 months ago
- [CVPR 2026 Highlight] Reward Forcing: Efficient Streaming Video Generation with Rewarded Distribution Matching Distillationβ363Dec 15, 2025Updated 9 months ago
- Scaling Text-to-Image Diffusion Transformers with Representation Autoencodersβ263Feb 13, 2026Updated 7 months ago
- Official implementation of the paper "TOKENTRIM: INFERENCE-TIME TOKEN PRUNING FOR AUTOREGRESSIVE LONG VIDEO GENERATION"β15Feb 8, 2026Updated 7 months ago
- Managed hosting for WordPress and PHP on Cloudways β’ AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- [ACL 2026] Official implementation of "Less is More: Improving LLM Reasoning with Minimal Test-Time Intervention"β41Apr 18, 2026Updated 5 months ago
- [NeurIPS 2025] An official implementation of Flow-GRPO: Training Flow Matching Models via Online RLβ2,531May 7, 2026Updated 4 months ago
- β57May 6, 2026Updated 4 months ago
- (CVPR 2025) From Slow Bidirectional to Fast Autoregressive Video Diffusion Modelsβ1,441Aug 7, 2025Updated last year
- Official repo for UAE [ECCV 2026]β214Jul 28, 2026Updated last month
- β151Sep 11, 2026Updated last week
- β39Mar 5, 2026Updated 6 months ago