Codebase for InfoTok: Adaptive Discrete Video Tokenizer via Information-Theoretic Compression
☆55Mar 18, 2026Updated 4 months ago
Alternatives and similar repositories for InfoTok
Users that are interested in InfoTok are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- [ICCV 2025] CHORDS: Diffusion Sampling Accelerator with Multi-core Hierarchical ODE Solvers☆17Mar 3, 2026Updated 5 months ago
- [CVPR 2026] Official repo for "EVATok: Adaptive Length Video Tokenization for Efficient Visual Autoregressive Generation"☆63Mar 13, 2026Updated 4 months ago
- ☆13Apr 28, 2025Updated last year
- [ICLR2026] AliTok: Towards Sequence Modeling Alignment between Tokenizer and Autoregressive Model☆56Oct 12, 2025Updated 9 months ago
- FlexTok: Resampling Images into 1D Token Sequences of Flexible Length☆328Jun 2, 2025Updated last year
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- YODA: Yet Another One-step Diffusion-based Video Compression☆21Jul 26, 2026Updated 2 weeks ago
- Official repository for Color Equivariant Convolutional Networks.☆10Nov 16, 2023Updated 2 years ago
- [ICCV 2025] Official repo for "GigaTok: Scaling Visual Tokenizers to 3 Billion Parameters for Autoregressive Image Generation"☆204Jan 7, 2026Updated 7 months ago
- DC-VideoGen: Efficient Video Generation with Deep Compression Video Autoencoder☆191Oct 5, 2025Updated 10 months ago
- [CVPR2026 Highlight] Cubic Discrete Diffusion: Discrete Visual Generation on High-Dimensional Representation Tokens https://arxiv.org/abs…☆63Apr 10, 2026Updated 4 months ago
- Single-stage End-to-End Training for Tokenization and Generation☆118Mar 24, 2026Updated 4 months ago
- ☆30Sep 25, 2024Updated last year
- [ICCV2025] TokenBridge: Bridging Continuous and Discrete Tokens for Autoregressive Visual Generation. https://yuqingwang1029.github.io/To…☆158Jul 24, 2025Updated last year
- Official Implementation of Paper: FVQ: Scalable Training for Vector-Quantized Networks with 100% Codebook Utilization (ICLR2026)☆26Jan 30, 2026Updated 6 months ago
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- paper collection: alignment of diffusion models☆29Mar 6, 2026Updated 5 months ago
- ☆47Mar 22, 2024Updated 2 years ago
- [ICML2026] Auto-Regressive Long Video Generation via 2-Bit KV-Cache Quantization☆61Jul 26, 2026Updated 2 weeks ago
- The repository of paper "Implicit-explicit Integrated Representations for Multi-view Video Compression"☆15Apr 27, 2025Updated last year
- [CVPR-2026] DiverseDiT: Towards Diverse Representation Learning in Diffusion Transformers☆22Updated this week
- [CVPR 2026] Adaptive Spectral Feature Forecasting for Diffusion Sampling Acceleration☆130Apr 30, 2026Updated 3 months ago
- Learning An Effective Transformer for Remote Sensing Satellite Image Dehazing☆12Sep 25, 2023Updated 2 years ago
- [CVPR2026] LeapAlign: Post-Training Flow Matching Models at Any Generation Step by Building Two-Step Trajectories☆55Jun 13, 2026Updated last month
- [ICLR 2026] Official implementation of DiCache: Let Diffusion Model Determine Its Own Cache☆62Jan 26, 2026Updated 6 months ago
- Serverless GPU API endpoints on Runpod - Get Bonus Credits • AdSkip the infrastructure headaches. Auto-scaling, pay-as-you-go, no-ops approach lets you focus on innovating your application.
- ☆37Mar 24, 2026Updated 4 months ago
- A curated list of papers and resources for text-to-image evaluation.☆30Sep 6, 2023Updated 2 years ago
- ☆146Nov 8, 2025Updated 9 months ago
- Visual Generation Tuning☆101Apr 16, 2026Updated 3 months ago
- Official Implementation of MultiWorld: Scalable Multi-Agent Multi-View Video World Models☆250May 12, 2026Updated 2 months ago
- Differentiable Hierarchical Visual Tokenization☆45Nov 26, 2025Updated 8 months ago
- [EMNLP 2025 Outstanding Paper Award] Official repo for DiscoSG: Towards Discourse-Level Text Scene Graph Parsing through Iterative Graph …☆22Nov 16, 2025Updated 8 months ago
- ☆26Jul 13, 2026Updated 3 weeks ago
- [CVPR 2026 Highlight] Reward Forcing: Efficient Streaming Video Generation with Rewarded Distribution Matching Distillation☆352Dec 15, 2025Updated 7 months ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- Geo-Align: Video Generation Alignment via Metric Geometry Reward☆34May 25, 2026Updated 2 months ago
- [CVPR'26] AdapTok: Learning Adaptive and Temporally Causal Video Tokenization in a 1D Latent Space☆30Mar 15, 2026Updated 4 months ago
- Official Pytorch implementation of MuCo: Multi-turn Contrastive Learning for Multimodal Embedding Model (CVPR 2026)☆15Apr 16, 2026Updated 3 months ago
- [ICLR 2024] Official pytorch implementation of "Denoising Task Routing for Diffusion Models"☆25Feb 19, 2024Updated 2 years ago
- Implementation of "VQ-HPS: Human Pose and Shape Estimation in a Vector-Quantized Latent Space" - ECCV 2024☆14Mar 24, 2025Updated last year
- [ICLR 2026] 🐻 Uniform Discrete Diffusion with Metric Path for Video Generation☆124May 20, 2026Updated 2 months ago
- Official implementation of Tuna-2: Pixel Embeddings Beat Vision Encoders for Unified Understanding and Generation☆745Jul 22, 2026Updated 2 weeks ago