A high-throughput and memory-efficient inference and serving engine for LLMs - Optimized for AMD gfx906 GPUs, e.g. Radeon VII / MI50 / MI60
☆77Jun 23, 2026Updated last month
Alternatives and similar repositories for vllm-gfx906-mobydick
Users that are interested in vllm-gfx906-mobydick are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- ML software (llama.cpp, ComfyUI, vLLM) builds for AMD gfx906 GPUs, e.g. Radeon VII / MI50 / MI60☆295Updated this week
- llama.cpp-gfx906☆140Mar 22, 2026Updated 4 months ago
- vLLM for AMD gfx906 GPUs, e.g. Radeon VII / MI50 / MI60☆433Feb 20, 2026Updated 5 months ago
- triton for AMD gfx906 GPUs, e.g. Radeon VII / MI50 / MI60☆48Dec 8, 2025Updated 7 months ago
- LLM inference in C/C++, but for GFX906!☆20Jul 15, 2026Updated 2 weeks ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- A database of knowledge around inference & training on GFX906 GPUs https://skyne98.github.io/wiki-gfx906/☆15Feb 21, 2026Updated 5 months ago
- vLLM for AMD gfx906 GPUs, e.g. Radeon VII / MI50 / MI60☆25Jul 2, 2026Updated last month
- Profile-guided GPU kernel optimizer for AMD/RDNA3. Auto-tunes llama.cpp MMVQ kernels per model shape. 2x decode speedup on 7900 XTX.☆64Jul 7, 2026Updated 3 weeks ago
- FORK of VLLM for AMD MI25/50/60. A high-throughput and memory-efficient inference and serving engine for LLMs☆71May 4, 2025Updated last year
- Advanced interoperability middleware for GPGPU acceleration. Facilitates cross vendor hardware abstraction and API translation for parall…☆98May 23, 2026Updated 2 months ago
- ☆13Dec 26, 2022Updated 3 years ago
- Random AI notes for working with local models or playing around with random machine learning bits.☆61Jun 7, 2026Updated last month
- The main repository for building Pascal-compatible versions of ML applications and libraries.☆215Aug 23, 2025Updated 11 months ago
- This is HTTPS/HTTP Server Library for ESP32, WT32_ETH01, ESP32 + LwIP W5500, ESP32 + LwIP W6100, ESP32 + LwIP ENC28J60. In the future, th…☆23Jan 10, 2023Updated 3 years ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- A FastAPI-based TTS server that provides OpenAI-compatible API endpoints using KittenTTS☆17Aug 16, 2025Updated 11 months ago
- Implementation of MambaFormer in Pytorch ++ Zeta from the paper: "Can Mamba Learn How to Learn? A Comparative Study on In-Context Learnin…☆23Jul 27, 2026Updated last week
- RDNA-native LLM inference engine in Rust.☆496Updated this week
- A PyTorch native platform for training generative AI models☆17Jun 30, 2026Updated last month
- GPU-accelerated voice assistant — local LLM, fine-tuned Whisper, Kokoro TTS, AMD ROCm☆18Apr 2, 2026Updated 4 months ago
- Proxy for OpenAI☆16Sep 2, 2025Updated 11 months ago
- A conversational voice-to-voice open-weights LLM-powered assistant designed to run on high-end consumer or workstation class hardware☆75Jul 6, 2026Updated 3 weeks ago
- A llamacpp wrapper to manage and monitor your llama server instance over a web ui.☆21Jun 16, 2026Updated last month
- Local AI text-to-speech with voice cloning and voice design, powered by GGML. C++17 port of OmniVoice (k2-fsa/OmniVoice). 646 languages, …☆150Jul 21, 2026Updated last week
- Serverless GPU API endpoints on Runpod - Get Bonus Credits • AdSkip the infrastructure headaches. Auto-scaling, pay-as-you-go, no-ops approach lets you focus on innovating your application.
- Smart OpenAI‑compatible proxy for llama.cpp: manages slots, saves/restores KV cache to disk, routes requests by prefix similarity, and pr…☆49Nov 14, 2025Updated 8 months ago
- Fast LLM swapping with sleep/wake support, compatible with vllm, llama.cpp, etc. llama-swap fork.☆51Apr 5, 2026Updated 3 months ago
- Enable true multi gpu capability in Comfy UI using XDiT XFuser and FSDP managed by Ray☆378Updated this week
- Random llm scripts☆41May 26, 2026Updated 2 months ago
- High fidelity neural audio codec for TTS models☆36Dec 22, 2025Updated 7 months ago
- bluetooth gyroscopic mouse like a wii remote control☆14Oct 17, 2021Updated 4 years ago
- Guidances for Test setup of 16 AMD MI50 32GB (for Deepseek v3.2)☆27May 11, 2026Updated 2 months ago
- AI plays Doom — pit Vision Language Models against demons and each other. Solo scenarios, deathmatch arena, 1-4 agents with any OpenAI-co…☆20Mar 12, 2026Updated 4 months ago
- ESP8266 PID controller for sous vide using MAX6675☆10May 13, 2017Updated 9 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- coreless esp32 controlled drone☆11Mar 17, 2023Updated 3 years ago
- A premium RAG-based AI Assistant built with React and FastAPI. Features efficient document indexing and high-accuracy retrieval-augmented…☆18Jun 3, 2026Updated 2 months ago
- Openvpn client in a docker container.☆11Nov 5, 2024Updated last year
- The HIP Environment and ROCm Kit - A lightweight open source build system for HIP and ROCm☆1,183Updated this week
- Sherpa-onnx-tts-stt source for homeassisstant addon with Kroko Onnx Streaming STT integration.☆30Dec 18, 2025Updated 7 months ago
- Structured local memory storage and retrieval for LLM agents☆15May 19, 2026Updated 2 months ago
- ☆44May 4, 2026Updated 3 months ago