[ICLR 23 oral] The Modality Focusing Hypothesis: Towards Understanding Crossmodal Knowledge Distillation
☆44Jul 10, 2023Updated 3 years ago
Alternatives and similar repositories for MFH
Users that are interested in MFH are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Code and data release for the paper "Seeing the Arrow of Time in Large Multimodal Models"☆16Oct 2, 2025Updated 11 months ago
- [CVPR 2023] Diversity-Aware Meta Visual Prompting☆84Nov 30, 2023Updated 2 years ago
- Papers about Hallucination in Multi-Modal Large Language Models (MLLMs)☆103Nov 21, 2024Updated last year
- Official Codebase of "A Unified Audio-Visual Learning Framework for Localization, Separation, and Recognition" (ICML 2023)☆12Jun 1, 2023Updated 3 years ago
- ☆14Jul 5, 2026Updated last month
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- The source codes and results of Efficient Wavelet Boost Learning-Based Multi-stage Progressive Refinement Network for Underwater Image En…☆41May 24, 2022Updated 4 years ago
- A python implement for Certifiable Robust Multi-modal Training☆21Jun 21, 2025Updated last year
- Accepted by TMM 2022☆20Aug 18, 2022Updated 4 years ago
- [ICCV 2023] Audio-Visual Class-Incremental Learning☆38Sep 29, 2024Updated last year
- Experiments and data for the paper "When and why vision-language models behave like bags-of-words, and what to do about it?" Oral @ ICLR …☆295Jun 7, 2023Updated 3 years ago
- [ECCV 2022] Tackling Long-Tailed Category Distribution Under Domain Shifts☆25Nov 29, 2022Updated 3 years ago
- [2022 TPAMI] Contrastive Positive Sample Propagation along the Audio-Visual Event Line☆32Mar 6, 2023Updated 3 years ago
- Foundation models based medical image analysis☆238Jul 29, 2026Updated last month
- Hyper-AdaC: Adaptive clustering-based hypergraph representation of whole slide images for survival analysis☆16Nov 28, 2022Updated 3 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- ☆36Jul 25, 2024Updated 2 years ago
- ☆13Jan 8, 2020Updated 6 years ago
- I2M2: Jointly Modeling Inter- & Intra-Modality Dependencies for Multi-modal Learning (NeurIPS 2024)☆21Oct 30, 2024Updated last year
- ☆22Nov 24, 2022Updated 3 years ago
- [ACM-MM'24 Oral] PASSION: Towards Effective Incomplete Multi-Modal Medical Image Segmentation with Imbalanced Missing Rates☆36Jun 4, 2025Updated last year
- Official PyTorch Implementation of the Longhorn Deep State Space Model☆57Dec 4, 2024Updated last year
- [MM 2023 Oral] Online Distillation-enhanced Multi-modal Transformer for Sequential Recommendation☆17Jan 10, 2024Updated 2 years ago
- [CVPR 2024 Highlight] OPERA: Alleviating Hallucination in Multi-Modal Large Language Models via Over-Trust Penalty and Retrospection-Allo…☆415Aug 24, 2024Updated 2 years ago
- Multimodal Variational Auto-encoder based Audio-Visual Segmentation [ICCV2023].☆20Sep 19, 2024Updated last year
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- [AAAI 2024] ConceptBed Evaluations for Personalized Text-to-Image Diffusion Models☆25Jun 1, 2023Updated 3 years ago
- an official PyTorch implementation of the paper "Partial Network Cloning", CVPR 2023☆13Mar 21, 2023Updated 3 years ago
- Deployed a facial emotion recognition using neural network model which predicts the emotion from faces in images, videos and live feed fr…☆12May 2, 2021Updated 5 years ago
- Building an efficient music recommendation system which determines the emotion of user using Facial Recognition techniques.☆13Jul 23, 2021Updated 5 years ago
- only contain face detect 、5/81 points 、face recognization models☆10Jul 9, 2020Updated 6 years ago
- This repository contains the source code related to the paper Compressed Volumetric Heatmaps for Multi-Person 3D Pose Estimation☆11Jun 23, 2020Updated 6 years ago
- Code for the paper: "SuS-X: Training-Free Name-Only Transfer of Vision-Language Models" [ICCV'23]☆104Aug 22, 2023Updated 3 years ago
- ☆22Jul 23, 2026Updated last month
- Mind the Gap: Understanding the Modality Gap in Multi-modal Contrastive Representation Learning☆177Sep 26, 2022Updated 3 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- The code repository for ICML24 paper "Tabular Insights, Visual Impacts: Transferring Expertise from Tables to Images"☆23Mar 11, 2025Updated last year
- Harmonic-NAS: Hardware-Aware Multimodal Neural Architecture Search on Resource-constrained Devices (ACML 2023)☆16May 7, 2024Updated 2 years ago
- rPPG-based Biometric Authentication☆11Jun 3, 2025Updated last year
- BenchX: A Unified Benchmark Framework for Medical Vision-Language Pretraining on Chest X-Rays☆51Dec 27, 2025Updated 8 months ago
- A Spatial–Temporal Video Quality Assessment Method via Comprehensive HVS Simulation☆17Jan 13, 2024Updated 2 years ago
- Test-time adaptation via Nearest neighbor information (TAST), submitted to ICLR'23☆24Jul 11, 2023Updated 3 years ago
- ICML2024-ReconBoost: Boosting Can Achieve Modality Reconcilement☆30May 2, 2025Updated last year