[CVPR'25] MergeVQ: A Unified Framework for Visual Generation and Representation with Token Merging and Quantization
☆51Jul 22, 2025Updated last year
Alternatives and similar repositories for MergeVQ
Users that are interested in MergeVQ are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Implementation of RankE: End-to-End Discrete Text-to-Image Post-Training via Rank-Consistent Alignment☆22May 27, 2026Updated 2 months ago
- Envision: Benchmarking Unified Understanding & Generation for Causal World Process Insights☆32Jan 9, 2026Updated 7 months ago
- ScalingOpt - Optimization Community☆105Jun 1, 2026Updated 2 months ago
- Does Understanding Inform Generation in Unified Multimodal Models? From Analysis to Path Forward☆60Nov 27, 2025Updated 8 months ago
- CAIRI Supervised, Semi- and Self-Supervised Visual Representation Learning Toolbox and Benchmark☆658Oct 15, 2025Updated 9 months ago
- End-to-end encrypted cloud storage - Proton Drive • AdSpecial offer: 40% Off Yearly / 80% Off First Month. Protect your most important files, photos, and documents from prying eyes.
- About Official PyTorch(MMCV) implementation of “SUMix: Mixup with Semantic and Uncertain Information” (ECCV 2024)☆12Sep 2, 2024Updated last year
- The official implementation of the ECCV'24 paper MC-CoT: Boosting the Power of Small Multimodal Reasoning Models to Match Larger Models w…☆26May 19, 2024Updated 2 years ago
- [CPAL 2026 oral] Offical implementation of "ROSE: Reordered SparseGPT for More Accurate One-Shot Large Language Models Pruning”☆17Jul 31, 2026Updated last week
- SEED-Voken: A Series of Powerful Visual Tokenizers☆1,020Nov 25, 2025Updated 8 months ago
- [NeurIPS 2024] Stabilize the Latent Space for Image Autoregressive Modeling: A Unified Perspective☆78Oct 31, 2024Updated last year
- Small Drafts, Big Verdict: Information-Intensive Visual Reasoning via Speculation (ICLR 2026)☆21Apr 27, 2026Updated 3 months ago
- ☆15Sep 1, 2025Updated 11 months ago
- ☆22Feb 24, 2026Updated 5 months ago
- (NeurIPS 2025) Vision Foundation Models as Effective Visual Tokenizers for Autoregressive Image Generation☆77May 21, 2026Updated 2 months ago
- Serverless GPU API endpoints on Runpod - Get Bonus Credits • AdSkip the infrastructure headaches. Auto-scaling, pay-as-you-go, no-ops approach lets you focus on innovating your application.
- [NeurIPS 2024] Image Understanding Makes for A Good Tokenizer for Image Generation☆21Dec 17, 2024Updated last year
- A Novel Semantic Segmentation Network using Enhanced Boundaries in Cluttered Scenes (WACV 2025)☆12Aug 11, 2025Updated last year
- Code repository for the paper "MrT5: Dynamic Token Merging for Efficient Byte-level Language Models."☆59Sep 25, 2025Updated 10 months ago
- [ICLR 2024] MogaNet: Efficient Multi-order Gated Aggregation Network☆273Jun 28, 2025Updated last year
- This repo focuses on supervised and self-supervised bio-sequence representation learning☆22Oct 11, 2023Updated 2 years ago
- [ICLR 2025] Official Pytorch Implementation of "Mix-LN: Unleashing the Power of Deeper Layers by Combining Pre-LN and Post-LN" by Pengxia…☆30Jul 24, 2025Updated last year
- [ICLR 2026] MergeMix: A Unified Augmentation Paradigm for Visual and Multi-Modal Understanding☆22Feb 27, 2026Updated 5 months ago
- The official implement of paper 《DaMo: Data Mixing Optimizer in Fine-tuning Multimodal LLMs for Mobile Phone Agents》☆30Oct 23, 2025Updated 9 months ago
- ☆56Nov 26, 2024Updated last year
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- [ICLR2026] WeTok: Powerful Discrete Tokenization for High-Fidelity Visual Reconstruction☆71Sep 3, 2025Updated 11 months ago
- ☆14Sep 22, 2025Updated 10 months ago
- High-performance Image Tokenizers for VAR and AR☆306Apr 25, 2025Updated last year
- 📐 [CVPR 2026] GGBench: A Geometric Generative Reasoning Benchmark for Unified Multimodal Models☆18Apr 1, 2026Updated 4 months ago
- [AAAI'26] Steering One-Step Diffusion Model with Fidelity-Rich Decoder for Fast Image Compression☆19Dec 21, 2025Updated 7 months ago
- ☆36Mar 12, 2025Updated last year
- [ICCV 2025] Official repo for "GigaTok: Scaling Visual Tokenizers to 3 Billion Parameters for Autoregressive Image Generation"☆204Jan 7, 2026Updated 7 months ago
- ☆49Feb 9, 2026Updated 6 months ago
- ☆17May 28, 2026Updated 2 months ago
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- Code release for "Generative Modeling of Weights: Generalization or Memorization?"☆23Apr 9, 2026Updated 4 months ago
- Official implementation of "Universal Deep Image Compression via Content-Adaptive Optimization with Adapters" presented at WACV 23 (https…☆31Sep 17, 2023Updated 2 years ago
- AAAI23-Directed Acyclic Graph Structure Learning from Dynamic Graphs☆12Nov 25, 2022Updated 3 years ago
- SearchEyes: Towards Frontier Multimodal Deep Search Intelligence via Search World Simulation. A typed knowledge graph unifies data synthe…☆23Jul 8, 2026Updated last month
- [EMNLP 2025 Findings] MEXA: Towards General Multimodal Reasoning with Dynamic Multi-Expert Aggregation☆15Aug 22, 2025Updated 11 months ago
- PEA-Diffusion: Parameter-Efficient Adapter with Knowledge Distillation in non-English Text-to-Image Generation☆37Oct 28, 2024Updated last year
- This repository categorizes the papers about masked image modeling according to their main contributions. The classification is based on …☆27May 3, 2025Updated last year