[ECCV 2024] Official implementation of the paper "Towards Latent Masked Image Modeling for Self-Supervised Visual Representation Learning"
☆30Mar 5, 2025Updated last year
Alternatives and similar repositories for LatentMIM
Users that are interested in LatentMIM are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Official codebase for "Unveiling the Power of Audio-Visual Early Fusion Transformers with Dense Interactions through Masked Modeling".☆43Aug 2, 2024Updated last year
- [ICCV 2025] Code for "Flow Stochastic Segmentation Networks"☆17Jun 9, 2026Updated last month
- Official implementation of "Positional-encoding Image Prior" (PIP)☆18Mar 1, 2023Updated 3 years ago
- Code for WACV 2023 paper "Out-of-distribution Detection via Frequency-regularized Generative Models" by Mu Cai and Yixuan Li☆11May 1, 2023Updated 3 years ago
- The Spacetime of Diffusion Models: An Information Geometry Perspective (ICLR 2026 Oral)☆48Feb 21, 2026Updated 5 months ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- [CVPR2026] ChangeBridge: Spatiotemporal Image Generation with Multimodal Controls for Remote Sensing☆17Jun 7, 2026Updated last month
- [AAAI 2026] WDT-MD: Wavelet Diffusion Transformers for Microaneurysm Detection in Fundus Images☆16Mar 25, 2026Updated 3 months ago
- Solution to the AMOS-MM challenge☆16Sep 13, 2025Updated 10 months ago
- [ICML 2024] GeoReasoner: Geo-localization with Reasoning in Street Views using a Large Vision-Language Model☆74Feb 1, 2026Updated 5 months ago
- Offical implementation of work 6 DoF Localization of Text Descriptions in Large-Scale Scenes with Gaussian Representation☆19Feb 5, 2025Updated last year
- Code for the paper: F. Ragusa, G. M. Farinella, A. Furnari. StillFast: An End-to-End Approach for Short-Term Object Interaction Anticipat…☆13Apr 11, 2023Updated 3 years ago
- [ECCV-2024] Transferable Targeted Adversarial Attack, CLIP models, Generative adversarial network, Multi-target attacks☆39Apr 23, 2025Updated last year
- The official repo for "Stepping Stones: A Progressive Training Strategy for Audio-Visual Semantic Segmentation", ECCV 2024☆18Oct 11, 2024Updated last year
- [ECCV 2024] Brain-ID: Learning Contrast-agnostic Anatomical Representations for Brain Imaging☆38Jan 31, 2026Updated 5 months ago
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- FunnyBirds: A Synthetic Vision Dataset for a Part-Based Analysis of Explainable AI Methods (ICCV 2023)☆17Apr 8, 2024Updated 2 years ago
- Checkpoints, logs and source code for AAAI-23 paper 'Data-Efficient Image Quality Assessment with Attention-Panel Decoder'☆39Apr 3, 2024Updated 2 years ago
- PyTorch implementation of RiT: Vanilla Diffusion Transformers Suffice in Representation Space☆26May 23, 2026Updated last month
- [ECCV'24 Oral] SPVLoc estimates 6D camera pose by matching images to semantic 3D models of indoor scenes, without scene-specific training…☆45Mar 4, 2026Updated 4 months ago
- [MedIA 2026] Hi-End-MAE: Hierarchical encoder-driven masked autoencoders are stronger vision learners for medical image segmentation☆33Feb 16, 2026Updated 5 months ago
- Unofficial version of LaneExtraction☆13Oct 12, 2022Updated 3 years ago
- Self-supervised algorithm for learning representations from ego-centric video data. Code is tested on EPIC-Kitchens-100 and Ego4D in PyTo…☆13Oct 23, 2022Updated 3 years ago
- [CVPR 2026] Scale Space Diffusion☆30Jul 9, 2026Updated last week
- [ECCV 2024] Official Pytorch Implementation of A Comprehensive Study of Multimodal Large Language Models for Image Quality Assessment☆94Jul 20, 2024Updated 2 years ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- Official implementation of "Goal Force: Teaching Video Models To Accomplish Physics-Conditioned Goals" (CVPR 2026)☆41Feb 25, 2026Updated 4 months ago
- Official repo for Contrastive Diffusion Loss☆14Dec 12, 2024Updated last year
- Offical Code for TBSNet(AAAI 2024)☆14Feb 17, 2024Updated 2 years ago
- Code for the paper "AverNet: All-in-one Video Restoration for Time-varying Unknown Degradations" (NeurIPS 2024)☆36Oct 29, 2024Updated last year
- ☆16Dec 29, 2025Updated 6 months ago
- ☆16Dec 29, 2025Updated 6 months ago
- Semantic Self-adaptation: Enhancing Generalization with a Single Sample (TMLR 2023)☆18Jul 21, 2023Updated 3 years ago
- ☆15Oct 14, 2025Updated 9 months ago
- [CVPR 2024] Diversity-aware Channel Pruning for StyleGAN Compression☆26Jul 23, 2025Updated 11 months ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- Downstream semantic segmentation evaluation of DGInStyle.☆25Apr 1, 2024Updated 2 years ago
- ☆20Jun 4, 2026Updated last month
- Source code for Regional Homogeneity: Towards Learning Transferable Universal Adversarial Perturbations Against Defenses (ECCV 2020)☆42Apr 2, 2019Updated 7 years ago
- Midjourney X Instant Collage -- Collage Template + Grid + Quality Style☆12May 25, 2025Updated last year
- Boosting Unsupervised Semantic Segmentation with Principal Mask Proposals (TMLR 2024)☆19Nov 27, 2024Updated last year
- [ECCV 2026] Official code of "Representation Alignment for Just Image Transformers is not Easier than You Think"☆47Jun 18, 2026Updated last month
- Casande-RL☆11May 9, 2023Updated 3 years ago