gpt-omni / mini-omni

open-source multimodal large language model that can hear, talk while thinking. Featuring real-time end-to-end speech input and streaming audio output conversational capabilities.
3,066Updated 2 months ago

Alternatives and similar repositories for mini-omni:

Users that are interested in mini-omni are comparing it to the libraries listed below