[ECCV 2024] Official implementation of the paper "Towards Latent Masked Image Modeling for Self-Supervised Visual Representation Learning"
☆31Mar 5, 2025Updated last year
Alternatives and similar repositories for LatentMIM
Users that are interested in LatentMIM are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- [ECCV 2024 Oral] Audio-Synchronized Visual Animation☆60Mar 15, 2026Updated 4 months ago
- Official Codebase of "A Unified Audio-Visual Learning Framework for Localization, Separation, and Recognition" (ICML 2023)☆12Jun 1, 2023Updated 3 years ago
- [NeurIPS 2024] Activating Self-Attention for Multi-Scene Absolute Pose Regression☆14Feb 24, 2025Updated last year
- [ICCV 2025] Code for "Flow Stochastic Segmentation Networks"☆18Jun 9, 2026Updated 2 months ago
- Official codebase for "Unveiling the Power of Audio-Visual Early Fusion Transformers with Dense Interactions through Masked Modeling".☆43Aug 2, 2024Updated 2 years ago
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- [CVPR2026] ChangeBridge: Spatiotemporal Image Generation with Multimodal Controls for Remote Sensing☆23Jun 7, 2026Updated 2 months ago
- Official Codebase of "A Closer Look at Weakly-Supervised Audio-Visual Source Localization" (NeurIPS 2022)☆22Dec 6, 2022Updated 3 years ago
- ☆22Aug 8, 2024Updated 2 years ago
- [AAAI 2026] WDT-MD: Wavelet Diffusion Transformers for Microaneurysm Detection in Fundus Images☆16Mar 25, 2026Updated 4 months ago
- Solution to the AMOS-MM challenge☆16Sep 13, 2025Updated 11 months ago
- [ICML 2024] GeoReasoner: Geo-localization with Reasoning in Street Views using a Large Vision-Language Model☆75Feb 1, 2026Updated 6 months ago
- ☆22Jul 3, 2025Updated last year
- Offical implementation of work 6 DoF Localization of Text Descriptions in Large-Scale Scenes with Gaussian Representation☆19Feb 5, 2025Updated last year
- [ECCV-2024] Transferable Targeted Adversarial Attack, CLIP models, Generative adversarial network, Multi-target attacks☆39Apr 23, 2025Updated last year
- Bare Metal GPUs on DigitalOcean Gradient AI • AdPurpose-built for serious AI teams training foundational models, running large-scale inference, and pushing the boundaries of what's possible.
- (BMVC 2022--Oral) Official repository for "Adversarial Pixel Restoration as a Pretext Task for Transferable Perturbations" …☆35Jan 8, 2023Updated 3 years ago
- Official PyTorch Repository of "Minority-Oriented Vicinity Expansion with Attentive Aggregation for Video Long-Tailed Recognition" (AAAI …☆13Jul 27, 2023Updated 3 years ago
- The official repo for "Stepping Stones: A Progressive Training Strategy for Audio-Visual Semantic Segmentation", ECCV 2024☆18Oct 11, 2024Updated last year
- Adapters Strike Back (CVPR 2024)☆43Jul 24, 2024Updated 2 years ago
- Code for the paper "AMEGO: Active Memory from long EGOcentric videos" published at ECCV 2024☆45Dec 7, 2024Updated last year
- Checkpoints, logs and source code for AAAI-23 paper 'Data-Efficient Image Quality Assessment with Attention-Panel Decoder'☆39Apr 3, 2024Updated 2 years ago
- Repository for "Enhanced Super-Resolution Training via Mimicked Alignment for Real-World Scenes", ACCV 2024☆16Dec 2, 2024Updated last year
- PyTorch implementation of RiT: Vanilla Diffusion Transformers Suffice in Representation Space☆27May 23, 2026Updated 2 months ago
- [MedIA 2026] Hi-End-MAE: Hierarchical encoder-driven masked autoencoders are stronger vision learners for medical image segmentation☆33Feb 16, 2026Updated 5 months ago
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- Self-supervised algorithm for learning representations from ego-centric video data. Code is tested on EPIC-Kitchens-100 and Ego4D in PyTo…☆13Oct 23, 2022Updated 3 years ago
- [ECCV 2024] Official Pytorch Implementation of A Comprehensive Study of Multimodal Large Language Models for Image Quality Assessment☆94Jul 20, 2024Updated 2 years ago
- [CVPR 2026] Scale Space Diffusion☆31Jul 9, 2026Updated last month
- Official repo for Contrastive Diffusion Loss☆14Dec 12, 2024Updated last year
- Offical Code for TBSNet(AAAI 2024)☆14Feb 17, 2024Updated 2 years ago
- ☆47Oct 5, 2025Updated 10 months ago
- ☆16Dec 29, 2025Updated 7 months ago
- ☆16Dec 29, 2025Updated 7 months ago
- Semantic Self-adaptation: Enhancing Generalization with a Single Sample (TMLR 2023)☆18Jul 21, 2023Updated 3 years ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- Material de la charla "The bad guys in AI - atacando sistemas de machine learning"☆16Nov 22, 2022Updated 3 years ago
- This is the official pytorch implementation of "Progressive Proxy Anchor Propagation for Unsupervised Semantic Segmentation" (ECCV 2024).☆18Aug 7, 2024Updated 2 years ago
- Downstream semantic segmentation evaluation of DGInStyle.☆25Apr 1, 2024Updated 2 years ago
- ☆32Jun 1, 2023Updated 3 years ago
- [ECCV 2026] Official code of "Representation Alignment for Just Image Transformers is not Easier than You Think"☆48Jun 18, 2026Updated last month
- Casande-RL☆11May 9, 2023Updated 3 years ago
- Repository of PanoVQA (CVPR 2026)☆27Mar 21, 2026Updated 4 months ago