📖 This is a repository for organizing papers, codes and other resources related to visual tokenizers.
☆17Jul 7, 2026Updated last week
Alternatives and similar repositories for Awesome-Visual-Tokenizers
Users that are interested in Awesome-Visual-Tokenizers are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- ☆11Dec 15, 2025Updated 7 months ago
- HITSZ 面向对象的软件构造实践课程项目(实现飞机大战的安卓版)☆17Nov 12, 2023Updated 2 years ago
- ☆15Mar 22, 2026Updated 3 months ago
- Repo from the "Learning with limited labeled data" seminar @ Uni of Tuebingen. A collection of notes, notebooks and slideshows to underst…☆17Apr 13, 2023Updated 3 years ago
- Provably (and non-vacuously) bounding test error of deep neural networks under distribution shift with unlabeled test data.☆10Feb 27, 2024Updated 2 years ago
- Bare Metal GPUs on DigitalOcean Gradient AI • AdPurpose-built for serious AI teams training foundational models, running large-scale inference, and pushing the boundaries of what's possible.
- [ICLR'26] SPRINT: Sparse-Dense Residual Fusion for Efficient Diffusion Transformers☆16Mar 19, 2026Updated 4 months ago
- The application of large pre-trained vision model DINOv2 from MetaAI for feature points matching, and a ViT decoder used for Auto Encoder☆18Apr 27, 2023Updated 3 years ago
- Code for DVD A Diagnostic Dataset for Multi-step Reasoning in Video Grounded Dialogue☆14Oct 12, 2021Updated 4 years ago
- [NeurIPS 2025] Official repo of "Martian World Model: Controllable Video Synthesis with Physically Accurate 3D Reconstructions"☆20Aug 6, 2025Updated 11 months ago
- A small project that uses Discrete Denoising Diffusion Probabilistic Models (D3PMs), a generative model for discrete data that builds upo…☆16Aug 10, 2024Updated last year
- 大学Latex答辩模版,当前包含川大、哈工大、中科大。☆11Jul 22, 2024Updated last year
- Collect papers and codes about VQGAN in various Computer Vision tasks☆10Dec 20, 2022Updated 3 years ago
- Grounding Language Models for Compositional and Spatial Reasoning☆18Oct 26, 2022Updated 3 years ago
- ☆28Oct 7, 2025Updated 9 months ago
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- Code release for "Generative Modeling of Weights: Generalization or Memorization?"☆23Apr 9, 2026Updated 3 months ago
- Official PyTorch Implementation of "Latent Denoising Makes Good Visual Tokenizers"☆195Feb 24, 2026Updated 4 months ago
- Med-DANet Series (ECCV 2022 & WACV 2024)☆13Jan 2, 2024Updated 2 years ago
- [ICLR 2026 Oral & ICML 2026] Generative Universal Verifier as Multimodal Meta-Reasoner☆64May 29, 2026Updated last month
- Official PyTorch implementation of the paper "Enhancing Vision-Language Pre-Training with Jointly Learned Questioner and Dense Captioner"☆15Aug 9, 2023Updated 2 years ago
- ☆16May 28, 2026Updated last month
- ☆13Oct 12, 2020Updated 5 years ago
- Official code for the MICCAI 2025 paper "Semantically Consistent Discrete Diffusion for 3D Biological Graph Generation"☆19Jul 7, 2025Updated last year
- PyTorch implementation of "PatchGame: Learning to Signal Mid-level Patches in Referential Games" to appear in NeurIPS 2021☆24Jun 4, 2021Updated 5 years ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- ☆78Jul 3, 2024Updated 2 years ago
- Official Implementation of "Latent Posterior-Mean Rectified Flow for Higher-Fidelity Perceptual Face Restoration"☆20Jul 15, 2025Updated last year
- UniVesselSeg official repository☆18Jan 26, 2026Updated 5 months ago
- [ICLR 2026] Many-for-Many: Unify the Training of Multiple Video and Image Generation and Manipulation Tasks☆32Feb 5, 2026Updated 5 months ago
- Self-Teaching Autoencoder learning reconstructions through latent agreement, not pixel loss.☆29May 25, 2026Updated last month
- official implementation of the paper "Delving into Latent Spectral Biasing of Video VAEs for Superior Diffusability".☆71Dec 25, 2025Updated 6 months ago
- [CVPR-2026] DiverseDiT: Towards Diverse Representation Learning in Diffusion Transformers☆20Mar 26, 2026Updated 3 months ago
- [NAACL 2025] Source code for MMEvalPro, a more trustworthy and efficient benchmark for evaluating LMMs☆25Sep 26, 2024Updated last year
- [ICCV 2025] MobileViCLIP: An Efficient Video-Text Model for Mobile Devices☆24Dec 11, 2025Updated 7 months ago
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- Paper Reading of IMCC groups.☆18Oct 22, 2025Updated 8 months ago
- Research4CSBeginners☆36Apr 11, 2026Updated 3 months ago
- Existing literature about training-data analysis.☆17Dec 17, 2021Updated 4 years ago
- [ICML 2026]☆17Jul 4, 2026Updated 2 weeks ago
- This repo holds the official code for the paper "FreMIM: Fourier Transform Meets Masked Image Modeling for Medical Image Segmentation".☆24Jan 2, 2024Updated 2 years ago
- Official code for CVPR 2024 paper, "SC-Tune: Unleashing Self-Consistent Referential Comprehension in Large Vision Language Models"☆16Apr 22, 2024Updated 2 years ago
- PyTorch Implementation: "Optimizing the Latent Space of Generative Networks"☆23Mar 12, 2019Updated 7 years ago