[ICLR 2026] π» Uniform Discrete Diffusion with Metric Path for Video Generation
β123May 20, 2026Updated 2 months ago
Alternatives and similar repositories for URSA
Users that are interested in URSA are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- [CVPR2026] Exploring Spatial Intelligence from a Generative Perspectiveβ30Jun 3, 2026Updated last month
- [ICML 2026 Spotlight] UDM-GRPO: Stable and Efficient Group Relative Policy Optimization for Uniform Discrete Diffusion Modelsβ27May 1, 2026Updated 2 months ago
- [ACL'26] EvoToken-DLM (Beyond Hard Masks: Progressive Token Evolution for Diffusion Language)β48Apr 7, 2026Updated 3 months ago
- [ICLR 2025] Autoregressive Video Generation without Vector Quantizationβ656Oct 29, 2025Updated 8 months ago
- β35Apr 10, 2026Updated 3 months ago
- Bare Metal GPUs on DigitalOcean Gradient AI β’ AdPurpose-built for serious AI teams training foundational models, running large-scale inference, and pushing the boundaries of what's possible.
- β18Jun 13, 2026Updated last month
- [NeurIPS 2025] Unveiling Chain of Step Reasoning for Vision-Language Models with Fine-grained Rewardsβ18Oct 6, 2025Updated 9 months ago
- Native Multimodal Models are World Learnersβ1,537Dec 30, 2025Updated 6 months ago
- Unsupervised Learning of Generalizable Robot Motion from Compact State Representationβ40Jun 10, 2026Updated last month
- [ICLR'26] Official PyTorch implementation of "Time Is a Feature: Exploiting Temporal Dynamics in Diffusion Language Models".β66Mar 5, 2026Updated 4 months ago
- Official implementation of paper "VMoBA: Mixture-of-Block Attention for Video Diffusion Models"β64Jul 1, 2025Updated last year
- Repo of HawkLlama.β16Jan 2, 2025Updated last year
- TVRBench: Target Viewpoint Reproduction Benchmark for Active Spatial Intelligenceβ25Jun 2, 2026Updated last month
- One-shot and Few-shot 3D Editing without Per-Scene Optimizationβ175Aug 21, 2025Updated 11 months ago
- Deploy to Railway using AI coding agents - Free Credits Offer β’ AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- [ICCV'25] Unified Open-World Segmentation with Multi-Modal Promptsβ16Jun 16, 2026Updated last month
- [NeurIPS 2025 Oral]InfinityβοΈ: Uniο¬ed Spacetime AutoRegressive Modeling for Visual Generationβ773Apr 16, 2026Updated 3 months ago
- [NeurIPS 2025 Spotlight] FUDOKI: Discrete Flow-based Unified Understanding and Generation via Kinetic-Optimal Velocitiesβ77Dec 21, 2025Updated 7 months ago
- [NeurIPS'24] A Simple Image Segmentation Framework via In-Context Examplesβ68Oct 29, 2024Updated last year
- [CVPR2026] Eliciting Complex Spatial Reasoning in MLLMs through Wide-Baseline Matchingβ19Jun 4, 2026Updated last month
- [ICML2026] ACTIVE-O3: Empowering Multimodal Large Language Models with Active Perception via GRPOβ83Apr 30, 2026Updated 2 months ago
- [NeurIPS 2025] Training-Free Efficient Video Generation via Dynamic Token Carvingβ287Aug 4, 2025Updated 11 months ago
- β36Oct 21, 2022Updated 3 years ago
- [ICLR 2026] LayerSync: Self-aligning Intermediate Layersβ22Mar 21, 2026Updated 4 months ago
- Managed hosting for WordPress and PHP on Cloudways β’ AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- JoVA: Unified Multimodal Learning for Joint Video-Audio Generationβ33Dec 22, 2025Updated 7 months ago
- [ICLR 2026] Official Repo for Rolling Forcing: Autoregressive Long Video Diffusion in Real Timeβ444Oct 31, 2025Updated 8 months ago
- [NeurIPS'24] Unleashing the Potential of the Diffusion Model in Few-shot Semantic Segmentation (Diffews)β51Apr 14, 2025Updated last year
- [ICLR 2026 Oral] DiffusionNFT: Online Diffusion Reinforcement with Forward Processβ983Feb 10, 2026Updated 5 months ago
- DenseFusion-1M: Merging Vision Experts for Comprehensive Multimodal Perceptionβ159Dec 6, 2024Updated last year
- [ACL 2026 Findings, ICCV 2025 Workshop Outstanding Paper Award] VChain: Chain-of-Visual-Thought for Reasoning in Video Generationβ120Apr 8, 2026Updated 3 months ago
- [CVPR 2026 Highlight] Reward Forcing: Efficient Streaming Video Generation with Rewarded Distribution Matching Distillationβ352Dec 15, 2025Updated 7 months ago
- Scaling Text-to-Image Diffusion Transformers with Representation Autoencodersβ255Feb 13, 2026Updated 5 months ago
- Official implementation of the paper "TOKENTRIM: INFERENCE-TIME TOKEN PRUNING FOR AUTOREGRESSIVE LONG VIDEO GENERATION"β15Feb 8, 2026Updated 5 months ago
- Deploy to Railway using AI coding agents - Free Credits Offer β’ AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- [ACL 2026] Official implementation of "Less is More: Improving LLM Reasoning with Minimal Test-Time Intervention"β41Apr 18, 2026Updated 3 months ago
- [NeurIPS 2025] An official implementation of Flow-GRPO: Training Flow Matching Models via Online RLβ2,430May 7, 2026Updated 2 months ago
- β47May 6, 2026Updated 2 months ago
- Official repo for UAEβ207Jun 21, 2026Updated last month
- (CVPR 2025) From Slow Bidirectional to Fast Autoregressive Video Diffusion Modelsβ1,408Aug 7, 2025Updated 11 months ago
- β145Nov 8, 2025Updated 8 months ago
- β39Mar 5, 2026Updated 4 months ago