[ICLR2026] AliTok: Towards Sequence Modeling Alignment between Tokenizer and Autoregressive Model
☆56Oct 12, 2025Updated 9 months ago
Alternatives and similar repositories for alitok
Users that are interested in alitok are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- [ICCV2025] TokenBridge: Bridging Continuous and Discrete Tokens for Autoregressive Visual Generation. https://yuqingwang1029.github.io/To…☆158Jul 24, 2025Updated last year
- FACM: Flow-Anchored Consistency Models☆147Aug 6, 2025Updated 11 months ago
- Implementation of the paper "MaskBit: Embedding-free Image Generation from Bit Tokens"☆94Apr 10, 2025Updated last year
- [NeurIPS 2025] ViewPoint: Panoramic Video Generation with Pretrained Diffusion Models☆34Jul 1, 2025Updated last year
- [ICCV 2025] Official repo for "GigaTok: Scaling Visual Tokenizers to 3 Billion Parameters for Autoregressive Image Generation"☆204Jan 7, 2026Updated 6 months ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- [Preprint] UCGM: Unified Continuous Generative Models☆185May 27, 2025Updated last year
- Image and video Tokenizer/VAE selection guide, text and face reconstruction evaluation.☆152Jun 11, 2026Updated last month
- FlexTok: Resampling Images into 1D Token Sequences of Flexible Length☆322Jun 2, 2025Updated last year
- ☆34Mar 4, 2025Updated last year
- Official Implementation for the paper: A Variational Framework for Improving Naturalness in Generative Spoken Language Models☆24Jun 18, 2025Updated last year
- Implementation of the paper "Variable Bitrate Residual Vector Quantization for Audio Coding"☆11Apr 10, 2025Updated last year
- ☆321May 29, 2025Updated last year
- ☆53Jun 13, 2025Updated last year
- [Preprint] GMem: A Modular Approach for Ultra-Efficient Generative Models☆43Mar 11, 2025Updated last year
- End-to-end encrypted email - Proton Mail • AdSpecial offer: 40% Off Yearly / 80% Off First Month. All Proton services are open source and independently audited for security.
- Inference for the STFT-VAE continuous audio codec (24kHz, 3.125Hz latent)☆43Jul 12, 2026Updated 2 weeks ago
- [Arxiv'25] MGVQ: Could VQ-VAE Beat VAE? A Generalizable Tokenizer with Multi-group Quantization☆55Sep 16, 2025Updated 10 months ago
- [ICLR 2026 Oral] Locality-aware Parallel Decoding for Efficient Autoregressive Image Generation☆104May 8, 2026Updated 2 months ago
- This repo contains the code for 1D tokenizer and generator☆1,167Mar 20, 2025Updated last year
- High-performance Image Tokenizers for VAR and AR☆307Apr 25, 2025Updated last year
- 🔥 Official impl. of "DetailFlow: 1D Coarse-to-Fine Autoregressive Image Generation via Next-Detail Prediction"☆170Jul 10, 2025Updated last year
- This repository includes the official implementation of our paper "Beyond Next-Token: Next-X Prediction for Autoregressive Visual Generat…☆251Oct 12, 2025Updated 9 months ago
- ☆27Jun 22, 2026Updated last month
- Codebase for InfoTok: Adaptive Discrete Video Tokenizer via Information-Theoretic Compression☆53Mar 18, 2026Updated 4 months ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- [NeurIPS 2025 Spotlight] A Unified Tokenizer for Visual Generation and Understanding☆529Nov 14, 2025Updated 8 months ago
- 5Hz Deep-Compression Speech VAE for AR-Diffusion and CALMs☆57Nov 19, 2025Updated 8 months ago
- This repository provides the official implementation of VTBench, a benchmark designed to evaluate the performance of visual tokenizers (V…☆35Jul 30, 2025Updated 11 months ago
- Official Implementation of "UniFlow: A Unified Pixel Flow Tokenizer for Visual Understanding and Generation"☆143Oct 17, 2025Updated 9 months ago
- [ICLR2026] WeTok: Powerful Discrete Tokenization for High-Fidelity Visual Reconstruction☆70Sep 3, 2025Updated 10 months ago
- Official repository for “Duo-Tok: Dual-Track Semantic Music Tokenizer for Vocal–Accompaniment Generation.”☆32Nov 26, 2025Updated 8 months ago
- SEED-Voken: A Series of Powerful Visual Tokenizers☆1,018Nov 25, 2025Updated 8 months ago
- the official repo for "D-AR: Diffusion via Autoregressive Models"☆138Jan 29, 2026Updated 5 months ago
- Official Implementation for Diffusion Models Without Classifier-free Guidance☆175Feb 18, 2025Updated last year
- GPUs on demand by Runpod - Special Offer Available • AdRun AI, ML, and HPC workloads on powerful cloud GPUs—without limits or wasted spend. Deploy GPUs in under a minute and pay by the second.
- ICLR 2026-MVAR: Visual Autoregressive Modeling with Scale and Spatial Markovian Conditioning☆38Apr 17, 2026Updated 3 months ago
- ☆15Apr 16, 2026Updated 3 months ago
- ☆31Jul 16, 2025Updated last year
- [Arxiv'25] DINO-Tok: Adapting DINO for Visual Tokenizers☆40Apr 11, 2026Updated 3 months ago
- Variable Bitrate Residual Vector Quantization for Audio Coding☆54May 1, 2025Updated last year
- Official inference code and LongText-Bench benchmark for our paper X-Omni (https://arxiv.org/pdf/2507.22058).☆426Aug 26, 2025Updated 11 months ago
- ☆145Jun 28, 2024Updated 2 years ago