MoIIE: Mixture of Intra- and Inter-Modality Experts for Large Vision Language Models
☆20Aug 14, 2025Updated 11 months ago
Alternatives and similar repositories for MoIIE
Users that are interested in MoIIE are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Follow Me: Conversation Planning for Target-driven Recommendation Dialogue Systems☆12Aug 1, 2023Updated 3 years ago
- Official implementation of "MST-Distill: Mixture of Specialized Teachers for Cross-Modal Knowledge Distillation" (ACM MM 2025)☆35Mar 5, 2026Updated 5 months ago
- Explicit Context Reasoning with Supervision for Visual Tracking (ACM MM 25)☆18Jul 20, 2025Updated last year
- The official implementation of the paper Collaborating Vision, Depth, and Thermal Signals for Multi-Modal Tracking: Dataset and Algorithm☆16Jul 29, 2025Updated last year
- ☆54Dec 9, 2024Updated last year
- Virtual machines for every use case on DigitalOcean • AdGet dependable uptime with 99.99% SLA, simple security tools, and predictable monthly pricing with DigitalOcean's virtual machines, called Droplets.
- Code and dataset for paper "SpatialBench: Benchmarking Multimodal Large Language Models for Spatial Cognition"☆19Mar 17, 2026Updated 4 months ago
- Modality-missing RGBT Tracking: Invertible Prompt Learning and High-quality Benchmarks (IJCV2024))☆27Mar 13, 2026Updated 4 months ago
- ☆17Mar 25, 2025Updated last year
- This repository provides a comprehensive research platform for Unaligned RGBT object tracking☆29Jul 24, 2026Updated 2 weeks ago
- [NeurIPS‘25] Vid-SME: Membership Inference Attacks against Large Video Understanding Models☆26Jun 20, 2026Updated last month
- The official implementation for the CVPR'2025 paper Dynamic Updates for Language Adaptation in Visual-Language Tracking☆44Mar 27, 2025Updated last year
- ☆30Apr 3, 2024Updated 2 years ago
- Learning Low-rank and Sparse Discriminative Correlation Filters for Coarse-to-Fine Visual Object Tracking☆10Apr 15, 2021Updated 5 years ago
- [ICLR 2026] M2-Miner: Multi-Agent Enhanced MCTS for Mobile GUI Agent Data Mining☆55Apr 22, 2026Updated 3 months ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- [CVPR2026] SpikeTrack: A Spike-driven Framework for Efficient Visual Tracking☆41Apr 8, 2026Updated 4 months ago
- Official repository of the ACL 2024 paper "Rethinking Task-Oriented Dialogue Systems: From Complex Modularity to Zero-Shot Autonomous Age…☆20May 28, 2024Updated 2 years ago
- The official implementation for the paper [Towards Unified Token Learning for Vision-Language Tracking].☆24Dec 13, 2023Updated 2 years ago
- (2025' IJCV) This is the offical implementation for the paper titled "FusionBooster: A Unified Image Fusion Boosting Paradigm".☆15Jul 23, 2025Updated last year
- ☆32Jan 16, 2025Updated last year
- (2021' TIM) This is the official implementation for the paper titled "UNIFusion: A Lightweight Unified Image Fusion Network".☆11Apr 12, 2023Updated 3 years ago
- Predicting the onset of Alzheimer's Disease using MRI & PET scans.☆11Dec 18, 2018Updated 7 years ago
- (AAAI23) Riemannian Local Mechanism for SPD Neural Networks☆11Mar 11, 2024Updated 2 years ago
- Official implementation of RMoE (Layerwise Recurrent Router for Mixture-of-Experts)☆33Aug 4, 2024Updated 2 years ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- Learning Deep Multi-Level Similarity for Thermal Infrared Object Tracking☆11Mar 18, 2021Updated 5 years ago
- OneVOS: Unifying Video Object Segmentation with All-in-One Transformer Framework☆13Feb 27, 2025Updated last year
- ☆18Apr 18, 2025Updated last year
- [CVPR'25] AVF-MAE++ : Scaling Affective Video Facial Masked Autoencoders via Efficient Audio-Visual Self-Supervised Learning☆22Jun 11, 2026Updated 2 months ago
- ☆14Nov 14, 2023Updated 2 years ago
- [ICCV‘25] Official implementation of paper "Towards Performance Consistency in Multi-Level Model Collaboration"☆45Oct 23, 2025Updated 9 months ago
- [TNNLS 2024] Refocus the Attention for Parameter-Efficient Thermal Infrared Object Tracking☆10Jun 20, 2025Updated last year
- DEP-Former: Multimodal Depression Recognition Based on Facial Expressions and Audio Features via Emotional Changes☆19Sep 18, 2024Updated last year
- Mariana Pro colorscheme from Sublime Text ported to Vim☆10Aug 2, 2025Updated last year
- Bare Metal GPUs on DigitalOcean Gradient AI • AdPurpose-built for serious AI teams training foundational models, running large-scale inference, and pushing the boundaries of what's possible.
- (CVPR24) Riemannian Multinomial Logistics Regression for SPD Neural Networks☆15Feb 5, 2025Updated last year
- R1-Track: Direct Application of MLLMs to Visual Object Tracking via Reinforcement Learning.☆66May 14, 2025Updated last year
- A PyTorch implementation of PTSA-MCTS from [Accelerating Monte Carlo Tree Search with Probability Tree State Abstraction].☆16Oct 21, 2023Updated 2 years ago
- 基于 Anatole 开发的 Halo 博客主题 Knarc☆13Apr 6, 2023Updated 3 years ago
- Official implementation of Bifrost-1: Bridging Multimodal LLMs and Diffusion Models with Patch-level CLIP Latents (NeurIPS 2025)☆47Nov 24, 2025Updated 8 months ago
- https://huggingface.co/datasets/multimodal-reasoning-lab/Zebra-CoT☆137Jan 30, 2026Updated 6 months ago
- ☆29Mar 30, 2025Updated last year