[ICLR2026] AliTok: Towards Sequence Modeling Alignment between Tokenizer and Autoregressive Model
☆56Oct 12, 2025Updated 10 months ago
Alternatives and similar repositories for alitok
Users that are interested in alitok are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- [ICCV2025] TokenBridge: Bridging Continuous and Discrete Tokens for Autoregressive Visual Generation. https://yuqingwang1029.github.io/To…☆158Jul 24, 2025Updated last year
- FACM: Flow-Anchored Consistency Models☆147Aug 6, 2025Updated last year
- Implementation of the paper "MaskBit: Embedding-free Image Generation from Bit Tokens"☆94Apr 10, 2025Updated last year
- [NeurIPS 2025] ViewPoint: Panoramic Video Generation with Pretrained Diffusion Models☆34Jul 1, 2025Updated last year
- [ICCV 2025] Official repo for "GigaTok: Scaling Visual Tokenizers to 3 Billion Parameters for Autoregressive Image Generation"☆204Jan 7, 2026Updated 7 months ago
- Managed Database hosting by DigitalOcean • AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- [Preprint] UCGM: Unified Continuous Generative Models☆187May 27, 2025Updated last year
- Image and video Tokenizer/VAE selection guide, text and face reconstruction evaluation.☆152Jun 11, 2026Updated 2 months ago
- FlexTok: Resampling Images into 1D Token Sequences of Flexible Length☆328Jun 2, 2025Updated last year
- ☆35Mar 4, 2025Updated last year
- Official Implementation for the paper: A Variational Framework for Improving Naturalness in Generative Spoken Language Models☆24Jun 18, 2025Updated last year
- Implementation of the paper "Variable Bitrate Residual Vector Quantization for Audio Coding"☆11Apr 10, 2025Updated last year
- ☆323May 29, 2025Updated last year
- ☆53Jun 13, 2025Updated last year
- [Preprint] GMem: A Modular Approach for Ultra-Efficient Generative Models☆43Mar 11, 2025Updated last year
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- Inference for the STFT-VAE continuous audio codec (24kHz, 3.125Hz latent)☆47Jul 12, 2026Updated last month
- [Arxiv'25] MGVQ: Could VQ-VAE Beat VAE? A Generalizable Tokenizer with Multi-group Quantization☆55Sep 16, 2025Updated 10 months ago
- [ICLR 2026 Oral] Locality-aware Parallel Decoding for Efficient Autoregressive Image Generation☆104May 8, 2026Updated 3 months ago
- This repo contains the code for 1D tokenizer and generator☆1,172Mar 20, 2025Updated last year
- High-performance Image Tokenizers for VAR and AR☆306Apr 25, 2025Updated last year
- 🔥 Official impl. of "DetailFlow: 1D Coarse-to-Fine Autoregressive Image Generation via Next-Detail Prediction"☆172Jul 10, 2025Updated last year
- This repository includes the official implementation of our paper "Beyond Next-Token: Next-X Prediction for Autoregressive Visual Generat…☆251Oct 12, 2025Updated 10 months ago
- ☆28Jun 22, 2026Updated last month
- Codebase for InfoTok: Adaptive Discrete Video Tokenizer via Information-Theoretic Compression☆55Mar 18, 2026Updated 4 months ago
- Deploy open-source AI quickly and easily - Special Bonus Offer • AdRunpod Hub is built for open source. One-click deployment and autoscaling endpoints without provisioning your own infrastructure.
- [NeurIPS 2025 Spotlight] A Unified Tokenizer for Visual Generation and Understanding☆530Nov 14, 2025Updated 9 months ago
- 5Hz Deep-Compression Speech VAE for AR-Diffusion and CALMs☆57Nov 19, 2025Updated 8 months ago
- This repository provides the official implementation of VTBench, a benchmark designed to evaluate the performance of visual tokenizers (V…☆35Jul 30, 2025Updated last year
- Official Implementation of "UniFlow: A Unified Pixel Flow Tokenizer for Visual Understanding and Generation"☆143Oct 17, 2025Updated 9 months ago
- [ICLR2026] WeTok: Powerful Discrete Tokenization for High-Fidelity Visual Reconstruction☆71Sep 3, 2025Updated 11 months ago
- Official repository for “Duo-Tok: Dual-Track Semantic Music Tokenizer for Vocal–Accompaniment Generation.”☆32Nov 26, 2025Updated 8 months ago
- SEED-Voken: A Series of Powerful Visual Tokenizers☆1,020Nov 25, 2025Updated 8 months ago
- the official repo for "D-AR: Diffusion via Autoregressive Models"☆138Jan 29, 2026Updated 6 months ago
- Official Implementation for Diffusion Models Without Classifier-free Guidance☆175Feb 18, 2025Updated last year
- Proton VPN Special Offer - Get 70% off • AdSpecial partner offer. Trusted by over 100 million users worldwide. Tested, Approved and Recommended by Experts.
- ICLR 2026-MVAR: Visual Autoregressive Modeling with Scale and Spatial Markovian Conditioning☆38Apr 17, 2026Updated 3 months ago
- ☆15Apr 16, 2026Updated 3 months ago
- ☆31Jul 16, 2025Updated last year
- [Arxiv'25] DINO-Tok: Adapting DINO for Visual Tokenizers☆43Apr 11, 2026Updated 4 months ago
- Variable Bitrate Residual Vector Quantization for Audio Coding☆55May 1, 2025Updated last year
- Official inference code and LongText-Bench benchmark for our paper X-Omni (https://arxiv.org/pdf/2507.22058).☆428Aug 26, 2025Updated 11 months ago
- ☆145Jun 28, 2024Updated 2 years ago