Multi-Modal Language Modeling with Image, Audio and Text Integration, included multi-images and multi-audio in a single multiturn.
☆18Feb 20, 2024Updated 2 years ago
Alternatives and similar repositories for multimodal-LLM
Users that are interested in multimodal-LLM are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Repository contains code to fine-tune WhisperASR model☆23Dec 16, 2022Updated 3 years ago
- ☆13Mar 21, 2023Updated 3 years ago
- A minimal re-implementation of orthogonal fine-tuning (OFT), a diffusion method, for LLMs. Based on nanoGPT and minLoRA.☆14Nov 17, 2023Updated 2 years ago
- Video scrubbing with WebCodecs☆16Nov 4, 2025Updated 10 months ago
- This project is from the Airbnb Recruitment Challenge on Kaggle. The challenge is to solve a multi-class classification problem of predic…☆11Feb 22, 2022Updated 4 years ago
- Serverless GPU API endpoints on Runpod - Get Bonus Credits • AdSkip the infrastructure headaches. Auto-scaling, pay-as-you-go, no-ops approach lets you focus on innovating your application.
- ☆12May 23, 2023Updated 3 years ago
- ☆16Dec 21, 2025Updated 9 months ago
- An implement of SPEECHSPLIT☆15Sep 12, 2020Updated 6 years ago
- fast opus bindings for node and browsers☆15Feb 11, 2024Updated 2 years ago
- code for training and using chess embeddings models☆14Jun 9, 2024Updated 2 years ago
- 😜Constrative Learning of Sentence Embedding using LoRA (EECS487 final project)☆13Apr 19, 2023Updated 3 years ago
- Official implementation for "Think Before You Segment: High-Quality Reasoning Segmentation with GPT Chain of Thoughts"☆22Jun 28, 2025Updated last year
- [ICLR 2022 Spotlight] Multi-Stage Episodic Control for Strategic Exploration in Text Games☆16Feb 8, 2026Updated 7 months ago
- ☆16Nov 24, 2025Updated 10 months ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- ☆51Jul 3, 2026Updated 2 months ago
- Syntexmex plugin for blender☆16Mar 28, 2020Updated 6 years ago
- An Official Repo of CVPR '20 "MSeg: A Composite Dataset for Multi-Domain Segmentation"☆16Aug 23, 2020Updated 6 years ago
- An SVM model for multi-class classification of Thyroid data.☆11Dec 9, 2019Updated 6 years ago
- ☆23Sep 2, 2024Updated 2 years ago
- Docker image for WhisperX by Max Bain☆13Sep 24, 2025Updated last year
- Forked UnrealEnginePython to add support for UE 5.1-5.3☆24Sep 14, 2023Updated 3 years ago
- ☆16May 1, 2021Updated 5 years ago
- Dialogue Planning via Brownian Bridge Stochastic Process for Goal-directed Proactive Dialogue (ACL Findings 2023)☆21Nov 10, 2025Updated 10 months ago
- GPUs on demand by Runpod - Special Offer Available • AdRun AI, ML, and HPC workloads on powerful cloud GPUs—without limits or wasted spend. Deploy GPUs in under a minute and pay by the second.
- A japanese Doujinshi about mne-python. PDF, 181p.☆18Jan 30, 2026Updated 7 months ago
- A Workshop on EEG with Drone Control Demos☆18Dec 16, 2024Updated last year
- Brain Computer Interface (BCI) with Neurosky Mindwave Mobile 2 that enables anyone to use computer, mobilephone etc. with his/her thought…☆24Oct 3, 2019Updated 6 years ago
- ☆10Jul 22, 2015Updated 11 years ago
- Contains a refined version of a vector trace of the Laughing Man logo from the anime "Ghost In The Shell - Stand Alone Complex" I did in …☆20Jan 8, 2018Updated 8 years ago
- This app uses OpenAI's LLM model to answer questions about your PDF file. Upload your PDF file and ask questions about it. The app will r…☆13May 13, 2025Updated last year
- ☆15Aug 9, 2024Updated 2 years ago
- Official PyTorch implementation for "MMS-LLaMA: Efficient LLM-based Audio-Visual Speech Recognition with Minimal Multimodal Speech Tokens…☆48Jun 12, 2025Updated last year
- ☆12Jan 27, 2025Updated last year
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- Simple LLM-enabled document Q&A app built using Langchain and Streamlit☆10Dec 4, 2024Updated last year
- ☆37Jul 24, 2026Updated 2 months ago
- Game-based AI Platforms☆26Jun 27, 2024Updated 2 years ago
- Homework 1 for Introduction to Operating Systems, Fall 2015☆10Oct 26, 2015Updated 10 years ago
- Directly applying advancements in transfer learning from BERT results in poor accuracy in domain-specific areas like law because of a wor…☆10May 7, 2023Updated 3 years ago
- Audio-Visual Speech Recognition☆25Jul 7, 2025Updated last year
- ☆14Apr 10, 2023Updated 3 years ago