Test-time Scaling for VAR models
☆33Sep 19, 2025Updated 10 months ago
Alternatives and similar repositories for TTS-VAR
Users that are interested in TTS-VAR are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- DreamVideo-Omni: Omni-Motion Controlled Multi-Subject Video Customization with Latent Identity Reinforcement Learning☆16May 27, 2026Updated last month
- [ICLR2026] The official code of "Routing Matters in MoE: Scaling Diffusion Transformers with Explicit Routing Guidance"☆46Mar 23, 2026Updated 4 months ago
- [ICCV2025] The official code of "DreamRelation: Relation-Centric Video Customization"☆27Feb 4, 2026Updated 5 months ago
- Image Tokenizer Needs Post-Training☆24Oct 4, 2025Updated 9 months ago
- UniClawBench project page: https://uniclawbench.github.io/☆37Updated this week
- Deploy open-source AI quickly and easily - Special Bonus Offer • AdRunpod Hub is built for open source. One-click deployment and autoscaling endpoints without provisioning your own infrastructure.
- [ICLR 2026] | MMSU: A Massive Multi-task Spoken Language Understanding and Reasoning Benchmark☆17Feb 12, 2026Updated 5 months ago
- [ICML 2026] Official Implementation of Prism: Efficient Test-Time Scaling via Hierarchical Search and Self-Verification for Discrete Diff…☆22Mar 4, 2026Updated 4 months ago
- [SIGGRAPH Asia 2025] DiffCamera: Arbitrary Refocusing on Images☆16Jan 26, 2026Updated 5 months ago
- ☆19May 15, 2026Updated 2 months ago
- [NeurIPS 2025] Inference-Time Text-to-Video Alignment with Diffusion Latent Beam Search☆18Feb 24, 2026Updated 5 months ago
- [CVPR2026 Findings] VHS: Verifier on Hidden States, an efficient inference-time scaling verification framework for DiT-based image genera…☆16Mar 25, 2026Updated 4 months ago
- The official repo of "MACRO: Advancing Multi-Reference Image Generation with Structured Long-Context Data"☆67Mar 27, 2026Updated 3 months ago
- Blended Point Cloud Diffusion for Localized Text-guided Shape Editing☆17Jul 2, 2026Updated 3 weeks ago
- Code for Draft Attention☆103May 22, 2025Updated last year
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- [ICML 2026] Smaller Models are Natural Explorers for Policy-Level Diversity in GRPO☆19Jun 15, 2026Updated last month
- This project is the official implementation of 'DreamOmni3: Scribble-based Editing and Generation''☆40Dec 30, 2025Updated 6 months ago
- ☆32Mar 24, 2023Updated 3 years ago
- [CVPR 2024] "Towards Robust Audiovisual Segmentation in Complex Environments with Quantization-based Semantic Decomposition"☆12Feb 27, 2024Updated 2 years ago
- [NeurIPS 2025] Efficient Reasoning Vision Language Models☆460Sep 18, 2025Updated 10 months ago
- 🔥 Official impl. of "DetailFlow: 1D Coarse-to-Fine Autoregressive Image Generation via Next-Detail Prediction"☆170Jul 10, 2025Updated last year
- ☆12Feb 23, 2022Updated 4 years ago
- Official repository of Text-Image Conditioned 3D Generation (TIGON, CVPR 2026)☆27Updated this week
- ☆40Mar 17, 2026Updated 4 months ago
- End-to-end encrypted email - Proton Mail • AdSpecial offer: 40% Off Yearly / 80% Off First Month. All Proton services are open source and independently audited for security.
- [NeurIPS 2025] IEAP: Image Editing As Programs with Diffusion Models☆118Sep 27, 2025Updated 9 months ago
- MICRO 2023 Evaluation Artifact for TeAAL☆11Oct 26, 2023Updated 2 years ago
- [ICML2026 Spotlight] UniPercept: Towards Unified Perceptual-Level Image Understanding across Aesthetics, Quality, Structure, and Texture☆157Jul 13, 2026Updated last week
- [ICLR 2026] Official repo of paper "Reconstruction Alignment Improves Unified Multimodal Models". Unlocking the Massive Zero-shot Potenti…☆411May 23, 2026Updated 2 months ago
- ☆83Oct 18, 2025Updated 9 months ago
- Official code for WACV 2024 paper, "Annotation-free Audio-Visual Segmentation"☆38Oct 11, 2024Updated last year
- T2I-ReasonBench: Benchmarking Reasoning-Informed Text-to-Image Generation☆37Sep 16, 2025Updated 10 months ago
- NextFlow🚀: Unified Sequential Modeling Activates Multimodal Understanding and Generation☆331Jan 9, 2026Updated 6 months ago
- The code for "VISTA: Enhancing Long-Duration and High-Resolution Video Understanding by VIdeo SpatioTemporal Augmentation" [CVPR2025]☆20Feb 27, 2025Updated last year
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- [CVPR2026 Highlight] Cubic Discrete Diffusion: Discrete Visual Generation on High-Dimensional Representation Tokens https://arxiv.org/abs…☆63Apr 10, 2026Updated 3 months ago
- ☆15May 13, 2025Updated last year
- ☆24May 23, 2025Updated last year
- [ICLR 2026] M2-Miner: Multi-Agent Enhanced MCTS for Mobile GUI Agent Data Mining☆55Apr 22, 2026Updated 3 months ago
- [ICML 2026] What Does Vision Tool-Use Reinforcement Learning Really Learn? Disentangling Tool-Induced and Intrinsic Effects for Crop-and-…☆22May 15, 2026Updated 2 months ago
- A unified framework for controllable caption generation across images, videos, and audio. Supports multi-modal inputs and customizable ca…☆54Jul 24, 2025Updated last year
- [ICLR2026] AliTok: Towards Sequence Modeling Alignment between Tokenizer and Autoregressive Model☆56Oct 12, 2025Updated 9 months ago