Official implementation of "MST-Distill: Mixture of Specialized Teachers for Cross-Modal Knowledge Distillation" (ACM MM 2025)
☆35Mar 5, 2026Updated 5 months ago
Alternatives and similar repositories for MST-Distill
Users that are interested in MST-Distill are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- MoIIE: Mixture of Intra- and Inter-Modality Experts for Large Vision Language Models☆20Aug 14, 2025Updated last year
- SynthRL: Scaling Visual Reasoning with Verifiable Data Synthesis☆71Jul 24, 2025Updated last year
- Graph in Graph Neural Network (https://arxiv.org/abs/2407.00696)☆16Sep 12, 2024Updated last year
- Pixels, Patterns, but no Poetry: To See the World like Humans☆18Aug 11, 2025Updated last year
- [ICMR 2025] Official Repository for The Paper, Let Network Decide What to Learn: Symbolic Music Understanding Model Based on Large-scale …☆19Aug 17, 2025Updated last year
- Managed Database hosting by DigitalOcean • AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- An official pytorch implementation of the paper: [MV-Adapter: Multimodal Video Transfer Learning for Video Text Retrieval].☆14Jul 27, 2024Updated 2 years ago
- ☆15Oct 13, 2025Updated 10 months ago
- ☆13Apr 2, 2025Updated last year
- The official github repo for MixEval-X, the first any-to-any, real-world benchmark.☆17Feb 15, 2025Updated last year
- ☆26Oct 4, 2024Updated last year
- ☆14Feb 13, 2025Updated last year
- ☆14Jul 13, 2024Updated 2 years ago
- A unified framework for vision-language environments with Gymnasium-compatible interface☆38Mar 17, 2026Updated 5 months ago
- ☆24Nov 4, 2025Updated 9 months ago
- Managed Database hosting by DigitalOcean • AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- ☆15Dec 31, 2024Updated last year
- ☆12Apr 26, 2022Updated 4 years ago
- ☆16Apr 27, 2025Updated last year
- Official implementation of SBNet as described in "Single-branch Network for Multimodal Training".☆13Aug 28, 2023Updated 3 years ago
- [NeurIPS 2024 Spotlight] Code for the paper "Flex-MoE: Modeling Arbitrary Modality Combination via the Flexible Mixture-of-Experts"☆85Jun 9, 2025Updated last year
- This is a summary of research on noisy correspondence. There may be omissions. If anything is missing please get in touch with us. Our em…☆87May 24, 2026Updated 3 months ago
- The official github repo for "Training Optimal Large Diffusion Language Models", the first-ever large-scale diffusion language models sca…☆46Nov 6, 2025Updated 9 months ago
- [CVPR 2024] Official repository of ST_GT☆10Sep 15, 2024Updated last year
- Enhancing Recipe Retrieval with Foundation Models: A Data Augmentation Perspective☆15Oct 22, 2024Updated last year
- Serverless GPU API endpoints on Runpod - Get Bonus Credits • AdSkip the infrastructure headaches. Auto-scaling, pay-as-you-go, no-ops approach lets you focus on innovating your application.
- This github contains the implementation of the method proposed in MDGNN_BS paper☆13May 9, 2024Updated 2 years ago
- Official implementation of RMoE (Layerwise Recurrent Router for Mixture-of-Experts)☆33Aug 4, 2024Updated 2 years ago
- ☆12Mar 28, 2024Updated 2 years ago
- ☆50Jan 15, 2026Updated 7 months ago
- This is the source code of our paper PALT in EMNLP2022.☆11Nov 19, 2022Updated 3 years ago
- [VLDB'23] SUREL+ is a novel set-based computation framework for scalable subgraph-based graph representation learning.☆17Apr 10, 2025Updated last year
- [Neural Networks 2025]Text-guided Image Restoration and Semantic Enhancement for Text-to-Image Person Retrieval☆12Dec 24, 2024Updated last year
- ☆12Oct 30, 2025Updated 10 months ago
- Community-aware Graph Transformer (CGT) is a novel Graph Transformer model that utilizes community structures to address node degree bias…☆15Aug 27, 2025Updated last year
- GPUs on demand by Runpod - Special Offer Available • AdRun AI, ML, and HPC workloads on powerful cloud GPUs—without limits or wasted spend. Deploy GPUs in under a minute and pay by the second.
- [IEEE Transactions on Information Forensics and Security'25] Pytorch implementation of CAMeL: Cross-modality Adaptive Meta-Learning for T…☆17Jan 5, 2026Updated 7 months ago
- Official PyTorch implementation for Hypersphere-Based Remote Sensing Cross-Modal Text–Image Retrieval via Curriculum Learning.☆16Aug 10, 2024Updated 2 years ago
- TECHS: Temporal Logical Graph Networks for Explainable Extrapolation Reasoning☆10Jan 16, 2024Updated 2 years ago
- ☆16Jul 18, 2024Updated 2 years ago
- The code of Fine-Grained Visual-Language Alignment for Remote Sensing Image-Text Retrieval(IEEE Transactions on Geoscience and Remote Sen…☆15Jun 30, 2025Updated last year
- Think Before You Move: Latent Motion Reasoning for Text-to-Motion Generation☆18Jan 4, 2026Updated 7 months ago
- About [AAAI 2025] Official repository of paper titled "DM-Adapter: Domain-Aware Mixture-of-Adapters for Text-Based Person Retrieval"☆17Jul 29, 2026Updated last month