☆59Sep 28, 2023Updated 2 years ago
Alternatives and similar repositories for mixed-resolution-vit
Users that are interested in mixed-resolution-vit are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Official implementation of Scaling Laws in Patchification: An Image Is Worth 50,176 Tokens And More☆25Feb 25, 2025Updated last year
- CVPR2023: Vector Quantization with Self-Attention for Quality-Independent Representation Learning.☆15May 17, 2024Updated 2 years ago
- ☆22May 12, 2026Updated 3 months ago
- Unconditional music synthesis using a diffusion model in the STFT domain☆12May 31, 2022Updated 4 years ago
- [NAACL 2022] TreeMix: Compositional Constituency-based Data Augmentation for Natural Language Understanding☆10Jul 15, 2023Updated 3 years ago
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- ☆17Jul 4, 2021Updated 5 years ago
- Official implementation for Wavelet Feature Maps Compression for Image-to-Image CNNs, NeurIPS 2022.☆37Oct 12, 2022Updated 3 years ago
- Code release for "Generative Modeling of Weights: Generalization or Memorization?"☆23Apr 9, 2026Updated 4 months ago
- [EMNLP'2023 Findings] MoqaGPT, for zero-shot multimodal question answering with LLMs☆13Dec 28, 2024Updated last year
- ☆10Feb 12, 2024Updated 2 years ago
- [EMNLP'2024 Findings] Explore generated documents for enhanced IR with LLMs. We enhance BM25 to surpass strong dense retriever on many da…☆14Mar 28, 2025Updated last year
- RecConv: Efficient Recursive Convolutions for Multi-Frequency Representations☆19Oct 13, 2025Updated 10 months ago
- Moving-camera background model (a CVPR '20 paper)☆16Oct 22, 2020Updated 5 years ago
- ☆23Jun 13, 2022Updated 4 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- Bilingual Singing Voice Synthesis☆18Mar 25, 2024Updated 2 years ago
- PyTorch reimplementation of FlexiViT: One Model for All Patch Sizes☆70May 5, 2024Updated 2 years ago
- ☆15Nov 23, 2023Updated 2 years ago
- towhee+elasticsearch实现本地以图搜图☆11Apr 23, 2023Updated 3 years ago
- Official PyTorch Implementation of Exploring Stochastic Autoregressive Image Modeling for Visual Representation, Accepted by AAAI 2023.☆15Jul 3, 2023Updated 3 years ago
- ☆46Oct 27, 2023Updated 2 years ago
- 🔊 A comprehensive list of open-source datasets for voice and sound computing (50+ datasets).☆20Apr 1, 2021Updated 5 years ago
- [CVPR 2025] Official Pytorch Code for Spatial Transport Optimization by Repositioning Attention Map for Training-Free Text-to-Image Synth…☆15Jun 21, 2025Updated last year
- ☆10Jun 14, 2022Updated 4 years ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- ☆17Jun 28, 2024Updated 2 years ago
- [ICCV23] Official implementation of eP-ALM: Efficient Perceptual Augmentation of Language Models.☆27Oct 27, 2023Updated 2 years ago
- LaVIT: Empower the Large Language Model to Understand and Generate Visual Content☆604Oct 6, 2024Updated last year
- [TMI 2024] Harvard Glaucoma Fairness (Harvard-GF): A Retinal Nerve Disease Dataset for Fairness Learning and Fair Identity Normalization☆10Apr 9, 2024Updated 2 years ago
- Code for the paper "MULTI-BAND MASKING FOR WAVEFORM-BASED SINGING VOICE SEPARATION" that was accepted on EUSIPCO2022☆15Jun 18, 2022Updated 4 years ago
- Official PyTorch implementation of Agglomerative Token Clustering presented at ECCV 2024☆20Sep 19, 2024Updated last year
- ☆32Jun 18, 2025Updated last year
- Backprop with Low-Precision Activations☆11Oct 28, 2019Updated 6 years ago
- [NeurIPS 2024] SeeClear: This repo is the official implementation of "SeeClear: Semantic Distillation Enhances Pixel Condensation for Vid…☆18Oct 8, 2024Updated last year
- Managed Database hosting by DigitalOcean • AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- A machine learning library focused on deep learning☆11Jun 2, 2015Updated 11 years ago
- Code for WACV 2021 Paper "Meta Module Network for Compositional Visual Reasoning"☆43May 13, 2021Updated 5 years ago
- ☆12Nov 16, 2020Updated 5 years ago
- Implementation of the ALI-G algorithm (PyTorch, Tensorflow)☆23Mar 7, 2021Updated 5 years ago
- An imbalanced dataset sampler for PyTorch.☆11Jan 20, 2022Updated 4 years ago
- Majesty Diffusion by @Dango233 and @apolinario (@multimodalart)☆25Jul 26, 2022Updated 4 years ago
- [EMNLP 2024] Preserving Multi-Modal Capabilities of Pre-trained VLMs for Improving Vision-Linguistic Compositionality☆24Oct 8, 2024Updated last year