[TMLR 2022] High-Modality Multimodal Transformer
☆116Nov 2, 2024Updated last year
Alternatives and similar repositories for HighMMT
Users that are interested in HighMMT are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- ☆11Aug 20, 2024Updated 2 years ago
- [NeurIPS 2021] Multiscale Benchmarks for Multimodal Representation Learning☆639Jan 27, 2024Updated 2 years ago
- ☆21Mar 15, 2023Updated 3 years ago
- [ICLR 2023] MultiViz: Towards Visualizing and Understanding Multimodal Models☆100Aug 22, 2024Updated 2 years ago
- Holistic evaluation of multimodal foundation models☆48Aug 11, 2024Updated 2 years ago
- Serverless GPU API endpoints on Runpod - Get Bonus Credits • AdSkip the infrastructure headaches. Auto-scaling, pay-as-you-go, no-ops approach lets you focus on innovating your application.
- ☆18Jul 10, 2024Updated 2 years ago
- Intepretability method to find what navigation agents learn☆19Jun 16, 2022Updated 4 years ago
- The repo for "Balanced Multimodal Learning via On-the-fly Gradient Modulation", CVPR 2022 (ORAL)☆323Sep 22, 2025Updated last year
- Official code for the paper: "Metadata Archaeology"☆19May 10, 2023Updated 3 years ago
- [NeurIPS 2023] Factorized Contrastive Learning: Going Beyond Multi-view Redundancy☆78Nov 13, 2023Updated 2 years ago
- ActMAD: Activation Matching to Align Distributions for Test-Time-Training (CVPR 2023)☆21Jun 27, 2023Updated 3 years ago
- Code and datasets for "Text encoders are performance bottlenecks in contrastive vision-language models". Coming soon!☆11May 24, 2023Updated 3 years ago
- ☆14May 31, 2022Updated 4 years ago
- [ICCV 2021] Multimodal Knowledge Expansion☆10Aug 28, 2021Updated 5 years ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- Code repository of the paper "CITRIS: Causal Identifiability from Temporal Intervened Sequences" and "iCITRIS: Causal Representation Lear…☆63Jun 16, 2023Updated 3 years ago
- This repository contains the code of our paper 'Skip \n: A simple method to reduce hallucination in Large Vision-Language Models'.☆15Feb 12, 2024Updated 2 years ago
- Implementation for "DeltaPhi: Learning Physical Trajectory Residual for PDE Solving"☆13Jun 17, 2024Updated 2 years ago
- Lottery Tickets in Evolutionary Optimization (Lange & Sprekeler, ICML 2023)☆17Jun 2, 2023Updated 3 years ago
- ☆26May 8, 2022Updated 4 years ago
- This is an implementation of the paper "Are We Done with Object-Centric Learning?"☆14Jun 21, 2026Updated 3 months ago
- ☆22Jun 4, 2025Updated last year
- SMILE: A Multimodal Dataset for Understanding Laughter☆13Jun 15, 2023Updated 3 years ago
- Official Codebase of "A Unified Audio-Visual Learning Framework for Localization, Separation, and Recognition" (ICML 2023)☆12Jun 1, 2023Updated 3 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- PyTorch codes for "LST: Ladder Side-Tuning for Parameter and Memory Efficient Transfer Learning"☆241Jan 20, 2023Updated 3 years ago
- ✨✨The Curse of Multi-Modalities (CMM): Evaluating Hallucinations of Large Multimodal Models across Language, Visual, and Audio☆55Jul 11, 2025Updated last year
- ☆13Jul 20, 2024Updated 2 years ago
- AgentHive provides the primitives and helpers for a seamless usage of robohive within TorchRL.☆36Jan 12, 2024Updated 2 years ago
- ☆19Jan 30, 2023Updated 3 years ago
- The repo for "Enhancing Multi-modal Cooperation via Sample-level Modality Valuation", CVPR 2024☆62Nov 5, 2024Updated last year
- The Social-IQ 2.0 Challenge Release for the Artificial Social Intelligence Workshop at ICCV '23☆39Oct 13, 2023Updated 2 years ago
- Counterfactual Evaluation and Learning for Interactive Systems: Foundations, Implementations, and Recent Advances☆12Aug 14, 2022Updated 4 years ago
- [CVPR 2024] VCoder: Versatile Vision Encoders for Multimodal Large Language Models☆279Apr 17, 2024Updated 2 years ago
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- Repository for the PopulAtion Parameter Averaging (PAPA) paper☆31Apr 11, 2024Updated 2 years ago
- Official code for the paper: [ICCV2023] Sound Localization from Motion: Jointly Learning Sound Direction and Camera Rotation☆44Jul 16, 2026Updated 2 months ago
- Source code for the paper "Policy Architectures for Compositional Generalization in Control"☆30May 19, 2022Updated 4 years ago
- EARL: Environment for Autonomous Reinforcement Learning☆38Nov 24, 2022Updated 3 years ago
- ☆161Jun 13, 2022Updated 4 years ago
- Tutorials for doing scientific machine learning (SciML) and high-performance differential equation solving with open source software.☆22Dec 8, 2025Updated 9 months ago
- Beta-VAE, Conditional-VAE, Total Correlation-VAE, FactorVAE, Relevance Factor-VAE, Multi-Level VAE, (Soft)-IntroVAE (Beta-Version), LVAE,…☆17Aug 19, 2025Updated last year