Official repository for the MMFM challenge
☆26Jun 18, 2024Updated 2 years ago
Alternatives and similar repositories for MMFM-Challenge
Users that are interested in MMFM-Challenge are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Meta-Prompting for Automating Zero-shot Visual Recognition with LLMs (ECCV 2024)☆20Jul 15, 2024Updated 2 years ago
- Repository for the paper: dense and aligned captions (dac) promote compositional reasoning in vl models☆28Nov 29, 2023Updated 2 years ago
- Repository for the paper: Teaching Structured Vision & Language Concepts to Vision & Language Models☆47Sep 25, 2023Updated 2 years ago
- ☆22Mar 14, 2024Updated 2 years ago
- ☆14May 25, 2021Updated 5 years ago
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- Personalized Lip Reading: Adapting to Your Unique Lip Movements with Vision and Language (AAAI 2025)☆24Jun 29, 2026Updated last month
- ☆11May 24, 2024Updated 2 years ago
- [CVPR25] Official Implementation of CAV-MAE Sync☆31Apr 5, 2026Updated 3 months ago
- An experiment to see if chatgpt can improve the output of the stanford alpaca dataset☆12Mar 29, 2023Updated 3 years ago
- how to build up Knowledge graph☆13Nov 16, 2021Updated 4 years ago
- ☆10Feb 7, 2022Updated 4 years ago
- [CVPR 2024] Retrieval-Augmented Image Captioning with External Visual-Name Memory for Open-World Comprehension☆64Apr 8, 2024Updated 2 years ago
- Fine tuning Mistral-7b with PEFT(Parameter Efficient Fine-Tuning) and LoRA(Low-Rank Adaptation) on Puffin Dataset(multi-turn conversation…☆12Nov 23, 2023Updated 2 years ago
- ☆12Mar 12, 2023Updated 3 years ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- ☆13Oct 17, 2024Updated last year
- X-MIC: Cross-Modal Instance Conditioning for Egocentric Action Generalization, CVPR 2024☆11Nov 7, 2024Updated last year
- ☆22Jun 4, 2025Updated last year
- [ACL2025] Unsolvable Problem Detection: Robust Understanding Evaluation for Large Multimodal Models☆82Mar 6, 2026Updated 4 months ago
- C++17 implementation of einops for libtorch - clear and reliable tensor manipulations with einstein-like notation☆12Oct 16, 2023Updated 2 years ago
- LLMGeo: Benchmarking Large Language Models on Image Geolocation In-the-wild☆16Oct 31, 2024Updated last year
- CaMML:Context-Aware MultiModal Learner for Large Models (ACL 2024 SAC Award)☆15May 21, 2025Updated last year
- Refactor your code with local LLM in VSCode☆13Mar 14, 2024Updated 2 years ago
- 🤡 An up-to-date & curated list of awesome KBQA papers, methods & resources.☆10Jul 14, 2022Updated 4 years ago
- Deploy open-source AI quickly and easily - Special Bonus Offer • AdRunpod Hub is built for open source. One-click deployment and autoscaling endpoints without provisioning your own infrastructure.
- WildLife Documentary Dataset☆14Jun 19, 2017Updated 9 years ago
- Official implementation of AdaMML. https://arxiv.org/abs/2105.05165.☆52Apr 14, 2022Updated 4 years ago
- Code for the C2KD paper (ICASSP 2023)☆20May 15, 2023Updated 3 years ago
- text-only training or language-free training for multimodal tasks (image/audio/video caption, retrieval, text2image)☆13Oct 15, 2024Updated last year
- ☆13May 9, 2023Updated 3 years ago
- ☆16Jan 3, 2023Updated 3 years ago
- Pytorch version of DeCEMBERT: Learning from Noisy Instructional Videos via Dense Captions and Entropy Minimization (NAACL 2021)☆17Jan 12, 2023Updated 3 years ago
- ☆27Aug 28, 2023Updated 2 years ago
- ☆14Jul 7, 2024Updated 2 years ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- Make machine learning simpler with Galaxy☆12Jul 16, 2024Updated 2 years ago
- A Comprehensive Benchmark and Toolkit for Evaluating Video-based Large Language Models!☆140Dec 31, 2023Updated 2 years ago
- We propose the Flowmind2digital method and the hdFlowmind dataset in this paper.☆14Nov 17, 2025Updated 8 months ago
- Accepted to ICLR 2025. MetaMetrics is a calibrated meta-metric designed to evaluate generation tasks across different modalities aligned …☆15Dec 30, 2024Updated last year
- Python Library to evaluate VLM models' robustness across diverse benchmarks☆227Jun 30, 2026Updated 3 weeks ago
- Code for paper: "Region Proposals for Saliency Map Refinement for Weakly-supervised Disease Localisation and Classification"☆14Jun 29, 2021Updated 5 years ago
- ☆15Feb 24, 2023Updated 3 years ago