Service for testing out the new Qwen2.5 omni model
☆62Apr 30, 2025Updated last year
Alternatives and similar repositories for qwen2.5_omni_chat
Users that are interested in qwen2.5_omni_chat are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- *NIX SHELL with Local AI/LLM integration☆26Feb 26, 2025Updated last year
- Yet another frontend for LLM, written using .NET and WinUI 3☆11Sep 14, 2025Updated last year
- Your personal and private AI☆54Apr 3, 2025Updated last year
- An fully autonomous agent that accesses the browser and performs tasks.☆18Aug 17, 2026Updated last month
- Adding a multi-text multi-speaker script (diffe) that is based on a script from asiff00 on issue 61 for Sesame: A Conversational Speech G…☆26Mar 28, 2025Updated last year
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- Yet Another (LLM) Web UI, made with Gemini☆12Dec 25, 2024Updated last year
- ☆18Jul 1, 2025Updated last year
- Run Orpheus 3B Locally with Gradio UI, Standalone App☆25Apr 1, 2025Updated last year
- Local-first RAG application for technical documentation and research papers☆28Jun 26, 2026Updated 3 months ago
- Orpheus Chat WebUI☆79Mar 27, 2025Updated last year
- LLamaHTML is a simple html file to communicate with a running llamacpp llama-server☆25Aug 5, 2025Updated last year
- A Windows tool to query various LLM AIs. Supports branched conversations, history and summaries among others.☆36May 11, 2026Updated 4 months ago
- FastAPI template made for AI☆42Apr 22, 2026Updated 5 months ago
- MCP server for searching npm packages☆16Feb 20, 2026Updated 7 months ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- CompChomper is a framework for measuring how LLMs perform at code completion.☆21Apr 29, 2025Updated last year
- Lightweight Llama 3 8B Inference Engine in CUDA C☆52Mar 21, 2025Updated last year
- ☆21Jan 25, 2025Updated last year
- Mistral Vibe rewritten in Rust by Devstral 2☆21Dec 23, 2025Updated 9 months ago
- Developer tools to debug and build realtime voice agents. Supports multiple models.☆49Aug 5, 2025Updated last year
- A Conversational Speech Generation Model with Gradio UI and OpenAI compatible API. UI and API support CUDA, MLX and CPU devices.☆215May 9, 2025Updated last year
- Inspect LLM's logprobs and perplexity over a piece of text, or compare two LLMs (like a git diff)☆18Aug 3, 2026Updated last month
- Offline LLM chatbot with personalized memory — works on CPU with multi-session memory support.☆22Jan 10, 2026Updated 8 months ago
- Agentic BYOK Browser-Based Website Builder☆60Updated this week
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- ☆101Mar 28, 2026Updated 5 months ago
- ☆59Oct 10, 2025Updated 11 months ago
- Personal voice assistant, with voice interruption and Twilio support☆18Feb 24, 2025Updated last year
- OpenAI compatible TTS for Sesame CSM:1b & dia:1.6b - Voice Cloning from File/YT☆438Sep 26, 2025Updated last year
- A repository to store helpful information and emerging insights in regard to LLMs☆21Oct 27, 2023Updated 2 years ago
- deep hermes, but decides how to respond based on its OWN decision, no need for system prompts.☆45Apr 1, 2025Updated last year
- ☆97Nov 6, 2024Updated last year
- ☆42Aug 2, 2025Updated last year
- BROKEN REPO. DO NOT USE UNDER ANY CIRCUMSTANCES☆20Updated this week
- Serverless GPU API endpoints on Runpod - Get Bonus Credits • AdSkip the infrastructure headaches. Auto-scaling, pay-as-you-go, no-ops approach lets you focus on innovating your application.
- Lightweight C inference for Qwen3 GGUF. Multiturn prefix caching & batch processing.☆26Sep 1, 2025Updated last year
- Realtime tts reading of large textfiles by your favourite voice. +Translation via LLM (Python script)☆52Oct 18, 2024Updated last year
- Holly - host your own AI coding agent inside docker container. Keep your system safe.☆23Jul 13, 2026Updated 2 months ago
- ☆190Jun 22, 2025Updated last year
- Protocol for Augmented Memory of Project Artifacts (MCP compatible) - extended☆23Jan 24, 2026Updated 8 months ago
- Compose, manage, and run MCP servers as Docker containers. With a Unified API gateway built in.☆58Oct 9, 2025Updated 11 months ago
- 🔥 Alternative to Ollama — multi-model serving with sub-ms model switching · CPU-only 20B inference for Edge AI · llama.cpp + stablediffu…☆40Aug 22, 2026Updated last month