Latest Advances on Autoregressive Visual Models.π
β29Mar 15, 2025Updated last year
Alternatives and similar repositories for Awesome-Visual-Autoregressive-Model
Users that are interested in Awesome-Visual-Autoregressive-Model are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- CoDi:Subject-Consistent and Pose-Diverse Text-to-Image Generationβ38Aug 1, 2025Updated last year
- L2P: Unlocking Latent Potential for Pixel Generationβ39May 22, 2026Updated 3 months ago
- TextCrafter: Accurately Rendering Multiple Texts in Complex Visual Scenesβ97Nov 26, 2025Updated 9 months ago
- This is the official repository of UltraHR-100K.β45Nov 21, 2025Updated 9 months ago
- Code for "L2P: Unlocking Latent Potential for Pixel Generation"β191Jul 11, 2026Updated last month
- Managed hosting for WordPress and PHP on Cloudways β’ AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- [CVPR 2025] InstanceCap: Improving Text-to-Video Generation via Instance-aware Structured Caption πβ45Jul 5, 2025Updated last year
- [AAAI 2025] Exploiting Multimodal Spatial-temporal Patterns for Video Object Trackingβ118May 18, 2025Updated last year
- The MIG benchmark of CVPR2024 MIGCβ15Mar 3, 2024Updated 2 years ago
- β18Jul 9, 2024Updated 2 years ago
- Explicit Context Reasoning with Supervision for Visual Tracking (ACM MM 25)β18Jul 20, 2025Updated last year
- [ICCV 2025] Region-Aware Text-to-Image Generation via Hard Binding and Soft Refinement π₯β622Dec 12, 2025Updated 8 months ago
- [ICLR 2026] MotionSight's official code implementation.β48Apr 24, 2026Updated 4 months ago
- a collection of awesome autoregressive visual generation modelsβ82Apr 17, 2025Updated last year
- A collection of diffusion models based on FLUX/DiT for image/video generation, editing, reconstruction, inpainting .etc.β86Jun 20, 2025Updated last year
- 1-Click AI Models by DigitalOcean Gradient β’ AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- (NeurIPS 2025 D&B Track) OverLayBench: A Benchmark for Layout-to-Image Generation with Dense Overlapsβ27May 4, 2026Updated 4 months ago
- Notion LifeOS PARA system β agent skill for Claude Code, OpenClaw, Codex and moreβ23Mar 24, 2026Updated 5 months ago
- A curated list of resources focused on Visual AutoRegressive Modeling, makes GPT-style AR models surpass diffusion transformers in image β¦β42Mar 2, 2025Updated last year
- β28Mar 7, 2025Updated last year
- Official implementation for ICDAR 2024 Oral paper "ICAL: Implicit Character-Aided Learning for Enhanced Handwritten Mathematical Expressiβ¦β29Aug 16, 2024Updated 2 years ago
- SDXL API provides a seamless interface for image generation and retrieval using Stable Diffusion XL integrated with Cloudflare AI Workersβ¦β14Feb 29, 2024Updated 2 years ago
- The implementation of Decoupling Layout from Glyph in Online Chinese Handwriting Generation (ICLR 2025)β25May 26, 2025Updated last year
- An extension for ComfyUI to add IPAdapter nodes for clip vision model with different input size.β18Feb 9, 2025Updated last year
- β‘οΈQwen-Image 4.8xπ speedup with Hybrid Acceleration for low VRAM GPUsβ17Oct 24, 2025Updated 10 months ago
- Managed Database hosting by DigitalOcean β’ AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- Source code of the article "Non Euclidean Sliced Optimal Transort Sampling" published at Eurographics 2024, authors : Baptiste GENEST, Niβ¦β13Aug 28, 2024Updated 2 years ago
- [CVPR 2025] Official code of "From Zero to Detail: Deconstructing Ultra-High-Definition Image Restoration from Progressive Spectral Perspβ¦β60Apr 16, 2026Updated 4 months ago
- [ICCV 2025] Official implementation for Describe, Don't Dictate: Semantic Image Editing with Natural Language Intentβ15Nov 4, 2025Updated 10 months ago
- [TAI 2025] Official implementation of TAI-accepted paper: ShadowMaskFormer: Mask Augmented Patch Embedding for Shadow Removalβ15May 8, 2025Updated last year
- [`CVPR 2024`] Official code repository for " 'Previously On ...' From Recaps to Story Summarization". https://arxiv.org/abs/2405.11487β14Feb 21, 2025Updated last year
- In OLHWDB ,you can find the ptts files, this code can help you get the information of the pttsβ11Mar 8, 2022Updated 4 years ago
- β13Apr 5, 2020Updated 6 years ago
- A framework for camera-controllable image editing using unified geometric guidance and video models.β66Aug 20, 2026Updated 2 weeks ago
- [ICLR 2025] Causal Graphical Models for Vision-Language Compositional Understandingβ10Apr 15, 2025Updated last year
- Managed Kubernetes at scale on DigitalOcean β’ AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- β23May 7, 2024Updated 2 years ago
- [ICCV 2025 Highlight] "Edit360: 2D Image Edits to 3D Assets from Any Angle"β21Feb 4, 2026Updated 7 months ago
- Generative Expressive Conversational Speech Synthesis (Accepted by MM'2024)β78Nov 1, 2024Updated last year
- The diffusion model is simple to implementβ17Oct 10, 2022Updated 3 years ago
- Source code for "A Compact Representation of Measured BRDFs Using Neural Processes" (TOG 2021)β15Nov 25, 2024Updated last year
- Official model implementation and benchmark evaluation repository of <AnyEdit: Unified High-Quality Image Edit with Any Idea>β34Jul 18, 2025Updated last year
- The official code of OneActor: Consistent Subject Generation via Cluster-Conditioned Guidance (NeurIPS 2024)β17Dec 23, 2024Updated last year