Test-time Scaling for VAR models
☆33Sep 19, 2025Updated 10 months ago
Alternatives and similar repositories for TTS-VAR
Users that are interested in TTS-VAR are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- DreamVideo-Omni: Omni-Motion Controlled Multi-Subject Video Customization with Latent Identity Reinforcement Learning☆17May 27, 2026Updated 2 months ago
- [ICLR2026] The official code of "Routing Matters in MoE: Scaling Diffusion Transformers with Explicit Routing Guidance"☆50Mar 23, 2026Updated 4 months ago
- [ICCV2025] The official code of "DreamRelation: Relation-Centric Video Customization"☆27Feb 4, 2026Updated 6 months ago
- Image Tokenizer Needs Post-Training☆24Oct 4, 2025Updated 10 months ago
- UniClawBench project page: https://uniclawbench.github.io/☆38Jul 28, 2026Updated 2 weeks ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- [ICML 2026] Official Implementation of Prism: Efficient Test-Time Scaling via Hierarchical Search and Self-Verification for Discrete Diff…☆22Mar 4, 2026Updated 5 months ago
- [SIGGRAPH Asia 2025] DiffCamera: Arbitrary Refocusing on Images☆16Jan 26, 2026Updated 6 months ago
- ☆17Jul 30, 2024Updated 2 years ago
- [NeurIPS 2025] Inference-Time Text-to-Video Alignment with Diffusion Latent Beam Search☆18Aug 4, 2026Updated last week
- [CVPR2026 Findings] VHS: Verifier on Hidden States, an efficient inference-time scaling verification framework for DiT-based image genera…☆16Mar 25, 2026Updated 4 months ago
- The official repo of "MACRO: Advancing Multi-Reference Image Generation with Structured Long-Context Data"☆67Mar 27, 2026Updated 4 months ago
- Blended Point Cloud Diffusion for Localized Text-guided Shape Editing☆17Jul 2, 2026Updated last month
- [ICML 2026] Smaller Models are Natural Explorers for Policy-Level Diversity in GRPO☆19Jun 15, 2026Updated last month
- This project is the official implementation of 'DreamOmni3: Scribble-based Editing and Generation''☆40Aug 8, 2026Updated last week
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- [CVPR 2024] "Towards Robust Audiovisual Segmentation in Complex Environments with Quantization-based Semantic Decomposition"☆12Feb 27, 2024Updated 2 years ago
- [Arxiv 2025] Official PyTorch implementation of DiffMoE, TC-DiT, EC-DiT and Dense DiT☆176Oct 21, 2025Updated 9 months ago
- 🔥 Official impl. of "DetailFlow: 1D Coarse-to-Fine Autoregressive Image Generation via Next-Detail Prediction"☆172Jul 10, 2025Updated last year
- Official repository of Text-Image Conditioned 3D Generation (TIGON, CVPR 2026)☆29Jul 23, 2026Updated 3 weeks ago
- ☆41Mar 17, 2026Updated 4 months ago
- [NeurIPS 2025] Wan-Move: Motion-controllable Video Generation via Latent Trajectory Guidance☆653Jan 5, 2026Updated 7 months ago
- [NeurIPS 2025] IEAP: Image Editing As Programs with Diffusion Models☆118Sep 27, 2025Updated 10 months ago
- AD-TUNING: An Adaptive CHILD-TUNING Approach to Efficient Hyperparameter Optimization of Child Networks for Speech Processing Tasks in th…☆11Feb 23, 2024Updated 2 years ago
- [CVPR25] IAR☆18Jun 13, 2025Updated last year
- Deploy open-source AI quickly and easily - Special Bonus Offer • AdRunpod Hub is built for open source. One-click deployment and autoscaling endpoints without provisioning your own infrastructure.
- [ICML2026 Spotlight] UniPercept: Towards Unified Perceptual-Level Image Understanding across Aesthetics, Quality, Structure, and Texture☆165Jul 13, 2026Updated last month
- ☆83Oct 18, 2025Updated 9 months ago
- NextFlow🚀: Unified Sequential Modeling Activates Multimodal Understanding and Generation☆331Jan 9, 2026Updated 7 months ago
- The code for "VISTA: Enhancing Long-Duration and High-Resolution Video Understanding by VIdeo SpatioTemporal Augmentation" [CVPR2025]☆20Feb 27, 2025Updated last year
- [CVPR2026 Highlight] Cubic Discrete Diffusion: Discrete Visual Generation on High-Dimensional Representation Tokens https://arxiv.org/abs…☆63Apr 10, 2026Updated 4 months ago
- ☆15May 13, 2025Updated last year
- ☆24May 23, 2025Updated last year
- [ICLR 2026] M2-Miner: Multi-Agent Enhanced MCTS for Mobile GUI Agent Data Mining☆55Apr 22, 2026Updated 3 months ago
- [ICML 2026] What Does Vision Tool-Use Reinforcement Learning Really Learn? Disentangling Tool-Induced and Intrinsic Effects for Crop-and-…☆23May 15, 2026Updated 2 months ago
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- A unified framework for controllable caption generation across images, videos, and audio. Supports multi-modal inputs and customizable ca…☆54Jul 24, 2025Updated last year
- [ICLR2026] AliTok: Towards Sequence Modeling Alignment between Tokenizer and Autoregressive Model☆56Oct 12, 2025Updated 10 months ago
- Official implementation of TSSR☆16Mar 5, 2026Updated 5 months ago
- UniEval: Unified Holistic Evaluation for Unified Multimodal Understanding and Generation☆25May 16, 2025Updated last year
- ☆19Jun 26, 2025Updated last year
- (ICCV2025) Official repository of paper "ViSpeak: Visual Instruction Feedback in Streaming Videos"☆54Jul 1, 2025Updated last year
- Data-free knowledge distillation using Gaussian noise (NeurIPS paper)☆15Mar 24, 2023Updated 3 years ago