[ACL 2026 Oral] Official implementation of LaMI: Augmenting Large Language Models via Late Multi-Image Fusion
β20Jul 4, 2026Updated 3 months ago
Alternatives and similar repositories for LaMI
Users that are interested in LaMI are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Official implementation of "Describing Sets of Images with Textual-PCA".β16Feb 13, 2023Updated 3 years ago
- The official code for the SALMonπ£ benchmark (ICASSP 2025 - Oral)β50Aug 15, 2025Updated last year
- β19Jan 8, 2025Updated last year
- [InterSpeech 2023] The official PyTorch implementation of: "AudioToken: Adaptation of Text-Conditioned Diffusion Models for Audio-to-Imagβ¦β90May 18, 2026Updated 4 months ago
- This repo contains the official PyTorch implementation of "Analyzing Discrete Self Supervised Speech Representation For Spoken Language Mβ¦β21Jan 3, 2023Updated 3 years ago
- Managed hosting for WordPress and PHP on Cloudways β’ AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- Official PyTorch Implementation for the "Unsupervised Model Tree Heritage Recovery" paper (ICLR 2025).β62Jul 1, 2025Updated last year
- [AAAI 2025] Official Implementation for "Click2Mask: Local Editing with Dynamic Mask Generation" Paper.β21Jan 22, 2026Updated 8 months ago
- This repo contains the official PyTorch implementation of "A Systematic Comparison of Phonetic Aware Techniques for Speech Enhancement" (β¦β28Aug 8, 2022Updated 4 years ago
- The official implementation of "A Language Modeling Approach to Diacritic-Free Hebrew TTS"β112Jun 12, 2025Updated last year
- CaMML:Context-Aware MultiModal Learner for Large Models (ACL 2024 SAC Award)β15May 21, 2025Updated last year
- β48Jul 7, 2025Updated last year
- β16Sep 6, 2024Updated 2 years ago
- Code for our paper: "Where's Waldo: Diffusion Features For Personalized Segmentation and Retrieval".β14Feb 26, 2025Updated last year
- Official implementation of the pipeline presented in I hear your true colors: Image Guided Audio Generationβ125Jan 18, 2023Updated 3 years ago
- Managed Database hosting by DigitalOcean β’ AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- [ICML 2024] Official Repository for the paper "Transformers Get Stable: An End-to-End Signal Propagation Theory for Language Models"β11Jul 19, 2024Updated 2 years ago
- β16Jul 23, 2024Updated 2 years ago
- Official PyTorch Implementation for the "Recovering the Pre-Fine-Tuning Weights of Generative Models" paper (ICML 2024).β86Apr 15, 2025Updated last year
- An official implementation of ProbeGenβ14Oct 20, 2024Updated last year
- Official This-Is-My Dataset published in CVPR 2023β16Jul 18, 2024Updated 2 years ago
- β13Jul 10, 2024Updated 2 years ago
- Official implementation of "DGD: Dynamic 3D Gaussians Distillation".β70Aug 16, 2024Updated 2 years ago
- Official PyTorch implementation of the paper: "Deep Audio Waveform Prior" (Interspeech 2022) https://arxiv.org/abs/2207.10441β12Oct 25, 2022Updated 3 years ago
- β20Oct 1, 2026Updated last week
- GPU virtual machines on DigitalOcean Gradient AI β’ AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- Automatic Measurement of Vowel Duration for Consonant Vowel Consonant (CVC) sound files (JASA 2016)β14Feb 25, 2017Updated 9 years ago
- [TPAMI 2026] Enhancing MMDiT-Based Text-to-Image Models for Similar Subject Generationβ15Mar 7, 2026Updated 7 months ago
- β16Jun 14, 2024Updated 2 years ago
- Official implemention of "Make It Count: Text-to-Image Generation with an Accurate Number of Objects" (CVPR 2025)β97Mar 12, 2025Updated last year
- Official repository for NAST: Noise Aware Speech Tokenization for Speech Language Models (Interspeech 2024) https://arxiv.org/abs/2406.11β¦β47Jul 2, 2024Updated 2 years ago
- Official implementation of MARS: Mixture of Auto-Regressive Models for Fine-grained Text-to-image Synthesisβ86Jul 16, 2024Updated 2 years ago
- [ICLR 2025] Adaptive prompt tailored pruning of T2I diffusion models.β15Feb 1, 2025Updated last year
- [AAAI 2024] The official PyTorch implementation of "Diverse and Aligned Audio-to-Video Generation via Text-to-Video Model Adaptation"β130May 18, 2026Updated 4 months ago
- An official PyTorch implementation for CLIPPRβ30Jul 22, 2023Updated 3 years ago
- Wordpress hosting with auto-scaling - Free Trial Offer β’ AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- The official repo of continuous speculative decodingβ36Mar 28, 2025Updated last year
- LMM for VQA, tcsvt versionβ10Jul 19, 2024Updated 2 years ago
- Official code for GlobalSplat: Efficient Feed-Forward 3D Gaussian Splatting via Global Scene Tokensβ48Jun 30, 2026Updated 3 months ago
- Official PyTorch Implementation for the "A Deep Inverse-Mapping Model for a Flapping Robotic Wing" Paper (ICLR 2025)β23Dec 16, 2025Updated 9 months ago
- AV-Link: Temporally-Aligned Diffusion Features for Cross-Modal Audio-Video Generationβ17Aug 3, 2025Updated last year
- Code for our papers : "Generating images of rare concepts using pre-trained diffusion models" (AAAI 24) and "Norm-guided latent space expβ¦β88Dec 27, 2023Updated 2 years ago
- A spoken version of the textual story cloze benchmarkβ22Aug 6, 2023Updated 3 years ago