[CVPR 2025] π₯ Official impl. of "TokenFlow: Unified Image Tokenizer for Multimodal Understanding and Generation".
β464Aug 8, 2025Updated last year
Alternatives and similar repositories for TokenFlow
Users that are interested in TokenFlow are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- [NeurIPS 2025 Spotlight] A Unified Tokenizer for Visual Generation and Understandingβ530Nov 14, 2025Updated 8 months ago
- [ICLR 2025] VILA-U: a Unified Foundation Model Integrating Visual Understanding and Generationβ425Apr 25, 2025Updated last year
- SEED-Voken: A Series of Powerful Visual Tokenizersβ1,020Nov 25, 2025Updated 8 months ago
- Autoregressive Model Beats Diffusion: π¦ Llama for Scalable Image Generationβ1,965Aug 15, 2024Updated last year
- NextFlowπ: Unified Sequential Modeling Activates Multimodal Understanding and Generationβ331Jan 9, 2026Updated 7 months ago
- Virtual machines for every use case on DigitalOcean β’ AdGet dependable uptime with 99.99% SLA, simple security tools, and predictable monthly pricing with DigitalOcean's virtual machines, called Droplets.
- High-performance Image Tokenizers for VAR and ARβ307Apr 25, 2025Updated last year
- [ICLR & NeurIPS 2025] Repository for Show-o series, One Single Transformer to Unify Multimodal Understanding and Generation.β1,970Jan 8, 2026Updated 7 months ago
- π This is a repository for organizing papers, codes and other resources related to unified multimodal models.β828Oct 10, 2025Updated 9 months ago
- This repo contains the code for 1D tokenizer and generatorβ1,172Mar 20, 2025Updated last year
- [ICCV 2025] Official repo for "GigaTok: Scaling Visual Tokenizers to 3 Billion Parameters for Autoregressive Image Generation"β204Jan 7, 2026Updated 7 months ago
- [TMLR 2025π₯] A survey for the autoregressive models in vision.β805May 5, 2026Updated 3 months ago
- [ICLR 2025] Autoregressive Video Generation without Vector Quantizationβ658Oct 29, 2025Updated 9 months ago
- [CVPR 2025 Oral]Infinity β : Scaling Bitwise AutoRegressive Modeling for High-Resolution Image Synthesisβ1,587Apr 16, 2026Updated 3 months ago
- Official implementation of BLIP3o-Seriesβ1,664Nov 29, 2025Updated 8 months ago
- Deploy on Railway without the complexity - Free Credits Offer β’ AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- PyTorch implementation of MAR+DiffLoss https://arxiv.org/abs/2406.11838β1,946Feb 20, 2026Updated 5 months ago
- Pytorch implementation for the paper titled "SimpleAR: Pushing the Frontier of Autoregressive Visual Generation"β431Jun 20, 2025Updated last year
- [CVPR2025 Highlight] PAR: Parallelized Autoregressive Visual Generation. https://yuqingwang1029.github.io/PAR-projectβ186Mar 20, 2025Updated last year
- [NeurIPS 2025] Vision as a Dialect: Unifying Visual Understanding and Generation via Text-Aligned Representationsβ202Sep 18, 2025Updated 10 months ago
- [ICCV2025]Code Release of Harmonizing Visual Representations for Unified Multimodal Understanding and Generationβ192May 21, 2025Updated last year
- Native Multimodal Models are World Learnersβ1,543Dec 30, 2025Updated 7 months ago
- [CVPR 2025 Oral] Reconstruction vs. Generation: Taming Optimization Dilemma in Latent Diffusion Modelsβ1,520Dec 16, 2025Updated 7 months ago
- Official PyTorch Implementation of "Diffusion Transformers with Representation Autoencoders"β1,988Feb 25, 2026Updated 5 months ago
- This is a repo to track the latest autoregressive visual generation papers.β430Jun 25, 2025Updated last year
- Deploy to Railway using AI coding agents - Free Credits Offer β’ AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- EVE Series: Encoder-Free Vision-Language Models from BAAIβ376Jul 24, 2025Updated last year
- FlexTok: Resampling Images into 1D Token Sequences of Flexible Lengthβ328Jun 2, 2025Updated last year
- Open-source unified multimodal modelβ6,141May 4, 2026Updated 3 months ago
- (Accepted by IJCV) Liquid: Language Models are Scalable and Unified Multi-modal Generatorsβ642Jun 1, 2026Updated 2 months ago
- Official PyTorch Implementation of "Latent Denoising Makes Good Visual Tokenizers"β196Feb 24, 2026Updated 5 months ago
- This repository includes the official implementation of our paper "Beyond Next-Token: Next-X Prediction for Autoregressive Visual Generatβ¦β252Oct 12, 2025Updated 9 months ago
- [ICLR'25 Oral] Representation Alignment for Generation: Training Diffusion Transformers Is Easier Than You Thinkβ1,695Mar 16, 2025Updated last year
- [ECCV 2026] Towards Scalable Pre-training of Visual Tokenizers for Generationβ498Apr 15, 2026Updated 3 months ago
- Next-Token Prediction is All You Needβ2,436Jan 12, 2026Updated 6 months ago
- Wordpress hosting with auto-scaling - Free Trial Offer β’ AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- [ICLR 2025] ControlAR: Controllable Image Generation with Autoregressive Modelsβ329Aug 2, 2026Updated last week
- π₯ Official impl. of "DetailFlow: 1D Coarse-to-Fine Autoregressive Image Generation via Next-Detail Prediction"β171Jul 10, 2025Updated last year
- Official Implementation of "Lumina-mGPT: Illuminate Flexible Photorealistic Text-to-Image Generation with Multimodal Generative Pretrainiβ¦β648Oct 16, 2025Updated 9 months ago
- Selftok: Discrete Visual Tokens of Autoregression, by Diffusion, and for Reasoningβ238May 30, 2025Updated last year
- β323May 29, 2025Updated last year
- β189Jun 27, 2025Updated last year
- Official inference code and LongText-Bench benchmark for our paper X-Omni (https://arxiv.org/pdf/2507.22058).β428Aug 26, 2025Updated 11 months ago