A high-throughput and memory-efficient inference and serving engine for LLMs - Optimized for AMD gfx906 GPUs, e.g. Radeon VII / MI50 / MI60
☆91Jun 23, 2026Updated 3 months ago
Alternatives and similar repositories for vllm-gfx906-mobydick
Users that are interested in vllm-gfx906-mobydick are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- ML software (llama.cpp, ComfyUI, vLLM) builds for AMD gfx906 GPUs, e.g. Radeon VII / MI50 / MI60☆347Updated this week
- llama.cpp-gfx906☆145Aug 23, 2026Updated last month
- ☆28May 12, 2026Updated 4 months ago
- vLLM for AMD gfx906 GPUs, e.g. Radeon VII / MI50 / MI60☆434Feb 20, 2026Updated 7 months ago
- triton for AMD gfx906 GPUs, e.g. Radeon VII / MI50 / MI60☆48Dec 8, 2025Updated 9 months ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- A database of knowledge around inference & training on GFX906 GPUs https://skyne98.github.io/wiki-gfx906/☆17Feb 21, 2026Updated 7 months ago
- Profile-guided GPU kernel optimizer for AMD/RDNA3. Auto-tunes llama.cpp MMVQ kernels per model shape. 2x decode speedup on 7900 XTX.☆69Sep 5, 2026Updated 3 weeks ago
- FORK of VLLM for AMD MI25/50/60. A high-throughput and memory-efficient inference and serving engine for LLMs☆70May 4, 2025Updated last year
- Advanced interoperability middleware for GPGPU acceleration. Facilitates cross vendor hardware abstraction and API translation for parall…☆103Sep 8, 2026Updated 3 weeks ago
- Random AI notes for working with local models or playing around with random machine learning bits.☆64Jun 7, 2026Updated 3 months ago
- The main repository for building Pascal-compatible versions of ML applications and libraries.☆219Aug 23, 2025Updated last year
- This is HTTPS/HTTP Server Library for ESP32, WT32_ETH01, ESP32 + LwIP W5500, ESP32 + LwIP W6100, ESP32 + LwIP ENC28J60. In the future, th…☆23Jan 10, 2023Updated 3 years ago
- A FastAPI-based TTS server that provides OpenAI-compatible API endpoints using KittenTTS☆17Aug 16, 2025Updated last year
- RDNA-native LLM inference engine in Rust.☆650Updated this week
- Deploy open-source AI quickly and easily - Special Bonus Offer • AdRunpod Hub is built for open source. One-click deployment and autoscaling endpoints without provisioning your own infrastructure.
- GPU-accelerated voice assistant — local LLM, fine-tuned Whisper, Kokoro TTS, AMD ROCm☆25Apr 2, 2026Updated 6 months ago
- Triton for AMD MI25/50/60. Development repository for the Triton language and compiler☆36Dec 15, 2025Updated 9 months ago
- A fork of vLLM enabling Pascal architecture GPUs☆37Feb 21, 2025Updated last year
- Simple node proxy for llama-server that enables MCP use☆19May 10, 2025Updated last year
- A conversational voice-to-voice open-weights LLM-powered assistant designed to run on high-end consumer or workstation class hardware☆78Sep 20, 2026Updated last week
- Local AI text-to-speech with voice cloning and voice design, powered by GGML. C++17 port of OmniVoice (k2-fsa/OmniVoice). 646 languages, …☆182Updated this week
- A llamacpp wrapper to manage and monitor your llama server instance over a web ui.☆29Jun 16, 2026Updated 3 months ago
- Smart OpenAI‑compatible proxy for llama.cpp: manages slots, saves/restores KV cache to disk, routes requests by prefix similarity, and pr…☆55Sep 16, 2026Updated 2 weeks ago
- Fast LLM swapping with sleep/wake support, compatible with vllm, llama.cpp, etc. llama-swap fork.☆57Apr 5, 2026Updated 5 months ago
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- Open API and Wyoming wrapper around Chatterbox☆32Aug 18, 2026Updated last month
- Enable true multi gpu capability in Comfy UI using XDiT XFuser and FSDP managed by Ray☆457Updated this week
- A simple pam account module to process HBAC rules stored on an IPA server☆10May 14, 2018Updated 8 years ago
- Random llm scripts☆41Sep 6, 2026Updated 3 weeks ago
- bluetooth gyroscopic mouse like a wii remote control☆15Oct 17, 2021Updated 4 years ago
- Guidances for Test setup of 16 AMD MI50 32GB (for Deepseek v3.2)☆29May 11, 2026Updated 4 months ago
- coreless esp32 controlled drone☆11Mar 17, 2023Updated 3 years ago
- A premium RAG-based AI Assistant built with React and FastAPI. Features efficient document indexing and high-accuracy retrieval-augmented…☆20Sep 8, 2026Updated 3 weeks ago
- Local ears and mouth for your LLM — offline, private, safe, free & open source. Voice-enable any local LLM stack: Claude Code, OpenCode, …☆45Sep 12, 2026Updated 3 weeks ago
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- Openvpn client in a docker container.☆11Nov 5, 2024Updated last year
- A command line tool for translation of Flutter ARB files☆16Jul 3, 2026Updated 3 months ago
- Sherpa-onnx-tts-stt source for homeassisstant addon with Kroko Onnx Streaming STT integration.☆32Dec 18, 2025Updated 9 months ago
- ☆43May 4, 2026Updated 4 months ago
- ☆59Oct 10, 2025Updated 11 months ago
- ESP32-S2 1.54" x 1.54" TFT Display Board☆11Mar 22, 2022Updated 4 years ago
- Various Cobbler config files☆14Jan 23, 2013Updated 13 years ago