[ICLR2026] AliTok: Towards Sequence Modeling Alignment between Tokenizer and Autoregressive Model
☆56Oct 12, 2025Updated 11 months ago
Alternatives and similar repositories for alitok
Users that are interested in alitok are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- [ICCV2025] TokenBridge: Bridging Continuous and Discrete Tokens for Autoregressive Visual Generation. https://yuqingwang1029.github.io/To…☆161Jul 24, 2025Updated last year
- FACM: Flow-Anchored Consistency Models☆147Aug 6, 2025Updated last year
- Implementation of the paper "MaskBit: Embedding-free Image Generation from Bit Tokens"☆94Apr 10, 2025Updated last year
- [NeurIPS 2025] ViewPoint: Panoramic Video Generation with Pretrained Diffusion Models☆35Jul 1, 2025Updated last year
- [ICCV 2025] Official repo for "GigaTok: Scaling Visual Tokenizers to 3 Billion Parameters for Autoregressive Image Generation"☆206Jan 7, 2026Updated 8 months ago
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- [Preprint] UCGM: Unified Continuous Generative Models☆188May 27, 2025Updated last year
- Image and video Tokenizer/VAE selection guide, text and face reconstruction evaluation.☆152Jun 11, 2026Updated 3 months ago
- FlexTok: Resampling Images into 1D Token Sequences of Flexible Length☆334Sep 11, 2026Updated last week
- ☆36Mar 4, 2025Updated last year
- Official Implementation for the paper: A Variational Framework for Improving Naturalness in Generative Spoken Language Models☆24Jun 18, 2025Updated last year
- Implementation of the paper "Variable Bitrate Residual Vector Quantization for Audio Coding"☆11Apr 10, 2025Updated last year
- ☆323May 29, 2025Updated last year
- ☆53Jun 13, 2025Updated last year
- [Preprint] GMem: A Modular Approach for Ultra-Efficient Generative Models☆43Mar 11, 2025Updated last year
- End-to-end encrypted email - Proton Mail • AdSpecial offer: 40% Off Yearly / 80% Off First Month. All Proton services are open source and independently audited for security.
- Inference for the STFT-VAE continuous audio codec (24kHz, 3.125Hz latent)☆47Jul 12, 2026Updated 2 months ago
- [Arxiv'25] MGVQ: Could VQ-VAE Beat VAE? A Generalizable Tokenizer with Multi-group Quantization☆55Sep 16, 2025Updated last year
- [ICLR 2026 Oral] Locality-aware Parallel Decoding for Efficient Autoregressive Image Generation☆104May 8, 2026Updated 4 months ago
- This repo contains the code for 1D tokenizer and generator☆1,176Mar 20, 2025Updated last year
- High-performance Image Tokenizers for VAR and AR☆307Apr 25, 2025Updated last year
- 🔥 Official impl. of "DetailFlow: 1D Coarse-to-Fine Autoregressive Image Generation via Next-Detail Prediction"☆174Jul 10, 2025Updated last year
- This repository includes the official implementation of our paper "Beyond Next-Token: Next-X Prediction for Autoregressive Visual Generat…☆251Oct 12, 2025Updated 11 months ago
- ☆28Jun 22, 2026Updated 3 months ago
- Codebase for InfoTok: Adaptive Discrete Video Tokenizer via Information-Theoretic Compression☆61Mar 18, 2026Updated 6 months ago
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- [NeurIPS 2025 Spotlight] A Unified Tokenizer for Visual Generation and Understanding☆533Nov 14, 2025Updated 10 months ago
- 5Hz Deep-Compression Speech VAE for AR-Diffusion and CALMs☆57Nov 19, 2025Updated 10 months ago
- This repository provides the official implementation of VTBench, a benchmark designed to evaluate the performance of visual tokenizers (V…☆36Jul 30, 2025Updated last year
- Official Implementation of "UniFlow: A Unified Pixel Flow Tokenizer for Visual Understanding and Generation"☆143Oct 17, 2025Updated 11 months ago
- [ICLR2026] WeTok: Powerful Discrete Tokenization for High-Fidelity Visual Reconstruction☆71Sep 3, 2025Updated last year
- Official repository for “Duo-Tok: Dual-Track Semantic Music Tokenizer for Vocal–Accompaniment Generation.”☆32Nov 26, 2025Updated 9 months ago
- SEED-Voken: A Series of Powerful Visual Tokenizers☆1,022Nov 25, 2025Updated 10 months ago
- the official repo for "D-AR: Diffusion via Autoregressive Models"☆138Jan 29, 2026Updated 7 months ago
- Official Implementation for Diffusion Models Without Classifier-free Guidance☆175Feb 18, 2025Updated last year
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- ICLR 2026-MVAR: Visual Autoregressive Modeling with Scale and Spatial Markovian Conditioning☆39Apr 17, 2026Updated 5 months ago
- ☆15Apr 16, 2026Updated 5 months ago
- ☆31Jul 16, 2025Updated last year
- [Arxiv'25] DINO-Tok: Adapting DINO for Visual Tokenizers☆44Apr 11, 2026Updated 5 months ago
- Variable Bitrate Residual Vector Quantization for Audio Coding☆56May 1, 2025Updated last year
- Official inference code and LongText-Bench benchmark for our paper X-Omni (https://arxiv.org/pdf/2507.22058).☆428Aug 26, 2025Updated last year
- ☆84Oct 18, 2025Updated 11 months ago