Linux & Powershell scripts to easily set up and run the Qwen 3.5 series locally on Windows and Linux with llama.cpp.
☆98Aug 14, 2026Updated this week
Alternatives and similar repositories for local-qwen3-coder-env
Users that are interested in local-qwen3-coder-env are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- A PowerShell script to fully automate the setup of `llama.cpp` on Windows. It installs all prerequisites, including the correct CUDA Tool…☆17May 12, 2026Updated 3 months ago
- ik_llama.cpp's Thireus fork with release builds for macOS/Windows/Ubuntu CPU, Vulkan and CUDA☆170Updated this week
- llama.cpp/ik_llama.cpp launcher: loads big MoE models across mismatched multi-GPU rigs by exact VRAM math.☆264Updated this week
- ☆44May 4, 2026Updated 3 months ago
- a fast and lightweight distributed background task processing framework with seamless scheduling.☆15Mar 30, 2026Updated 4 months ago
- GPUs on demand by Runpod - Special Offer Available • AdRun AI, ML, and HPC workloads on powerful cloud GPUs—without limits or wasted spend. Deploy GPUs in under a minute and pay by the second.
- this is an easy way to make ai podcast useing ai loccaly like ollama and the tts of piper☆16Feb 17, 2026Updated 6 months ago
- A fully local document intelligence system that allows users to build a persistent private knowledge base from documents and query it usi…☆17Apr 15, 2026Updated 4 months ago
- Practical local LLM recipes and benchmarks for RTX 5060 Ti setups☆118Updated this week
- llama.cpp fork with additional SOTA quants and improved performance☆3,056Updated this week
- Memory hygiene skills for Hermes Agent — dreaming (3-phase consolidation)☆44Jun 3, 2026Updated 2 months ago
- Messy repo filled with messy tests about hardware and LLMs. Built for me, public for you.☆46Updated this week
- Produce your own Dynamic 3.0 Quants and achieve optimum accuracy & SOTA quantization performance! Input a target size and the toolchain w…☆149Updated this week
- Thireus's fork of llama.cpp with Cuda 12.8 and 13.3 release builds and Windows patch for loading more .gguf shards + llama-sweep-bench☆30Updated this week
- llama-swap + a minimal ollama compatible api☆61May 26, 2026Updated 2 months ago
- Virtual machines for every use case on DigitalOcean • AdGet dependable uptime with 99.99% SLA, simple security tools, and predictable monthly pricing with DigitalOcean's virtual machines, called Droplets.
- Python-based cryptocurrency trading bot that functions on the Coinbase Pro exchange. This bot currently uses RSI values and 24HR Volume i…☆10May 16, 2022Updated 4 years ago
- Smart home assistant powered by an SLM☆19May 16, 2026Updated 3 months ago
- Llama.cpp runner/swapper and proxy that emulates LMStudio / Ollama backends☆60Aug 21, 2025Updated 11 months ago
- Recursive self-correcting intelligence framework☆17Nov 13, 2025Updated 9 months ago
- A harness optimized to smaller LLMs☆2,444Updated this week
- Smart OpenAI‑compatible proxy for llama.cpp: manages slots, saves/restores KV cache to disk, routes requests by prefix similarity, and pr…☆51Nov 14, 2025Updated 9 months ago
- Custom Template for checking the availiability of an entity.☆13Oct 31, 2025Updated 9 months ago
- Kon is a minimal coding agent (and also a highly opinionated one)☆344Aug 1, 2026Updated 2 weeks ago
- An open-source AI agent that lives in your terminal, privacy oriented, with no telemetry.☆78Updated this week
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- Scripts and tools for optimizing quantizations in llama.cpp with GGUF imatrices.☆19Jan 10, 2025Updated last year
- GPU-accelerated voice assistant — local LLM, fine-tuned Whisper, Kokoro TTS, AMD ROCm☆21Apr 2, 2026Updated 4 months ago
- ☆16Dec 16, 2024Updated last year
- Deploying full-stack on-prem deep research agent that can be run entirely on a local machine for $0!☆34Nov 8, 2025Updated 9 months ago
- An OpenAI-compatible ASR/STT API server powered by Meta's omnilingual-asr model. Supports real-time streaming via WebSocket and batch tra…☆19Jan 2, 2026Updated 7 months ago
- LLM inference in C/C++☆141Updated this week
- Multi-modal benchmark for measuring sensitive-information redaction in human recording data☆25Jun 6, 2026Updated 2 months ago
- ☆22Aug 5, 2026Updated last week
- ☆21Jul 4, 2025Updated last year
- Bare Metal GPUs on DigitalOcean Gradient AI • AdPurpose-built for serious AI teams training foundational models, running large-scale inference, and pushing the boundaries of what's possible.
- Run autoresearch on any NVIDIA GPUs (Works on 2-4GB+ Cards)☆25Mar 22, 2026Updated 4 months ago
- ☆12Sep 9, 2024Updated last year
- ☆19Nov 9, 2025Updated 9 months ago
- KVarN, KV cache precision tail, low-bit quants in llama.cpp for longer context of better precision in the same VRAM☆886Updated this week
- ROCm/AMD GPU benchmark suite for llama.cpp, whisper.cpp, PyTorch☆21Feb 4, 2026Updated 6 months ago
- A micro-services based system for indexing cryptocurrencies☆16Dec 7, 2022Updated 3 years ago
- ☆14Oct 9, 2023Updated 2 years ago