Linux & Powershell scripts to easily set up and run the Qwen 3.5 series locally on Windows and Linux with llama.cpp.
☆107Aug 22, 2026Updated 3 weeks ago
Alternatives and similar repositories for local-qwen3-coder-env
Users that are interested in local-qwen3-coder-env are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- A PowerShell script to fully automate the setup of `llama.cpp` on Windows. It installs all prerequisites, including the correct CUDA Tool…☆21May 12, 2026Updated 4 months ago
- ik_llama.cpp's Thireus fork with release builds for macOS/Windows/Ubuntu CPU, Vulkan and CUDA☆176Updated this week
- llama.cpp/ik_llama.cpp launcher: loads big MoE models across mismatched multi-GPU rigs by exact VRAM math.☆271Updated this week
- ☆43May 4, 2026Updated 4 months ago
- a fast and lightweight distributed background task processing framework with seamless scheduling.☆15Mar 30, 2026Updated 5 months ago
- Virtual machines for every use case on DigitalOcean • AdGet dependable uptime with 99.99% SLA, simple security tools, and predictable monthly pricing with DigitalOcean's virtual machines, called Droplets.
- this is an easy way to make ai podcast useing ai loccaly like ollama and the tts of piper☆16Feb 17, 2026Updated 6 months ago
- Practical local LLM recipes and benchmarks for RTX 5060 Ti setups☆163Aug 19, 2026Updated 3 weeks ago
- llama.cpp fork with additional SOTA quants and improved performance☆3,217Updated this week
- Memory hygiene skills for Hermes Agent — dreaming (3-phase consolidation)☆46Jun 3, 2026Updated 3 months ago
- Thireus's fork of llama.cpp with Cuda 12.8 and 13.3 release builds and Windows patch for loading more .gguf shards + llama-sweep-bench☆30Updated this week
- llama-swap + a minimal ollama compatible api☆64Updated this week
- Produce your own Dynamic 3.0 Quants and achieve optimum accuracy & SOTA quantization performance! Input a target size and the toolchain w…☆164Updated this week
- Single-file installer for a club-3090 webserver providing an admin control panel, a reverse proxy that automatically routes requests to t…☆25Jul 8, 2026Updated 2 months ago
- Llama.cpp runner/swapper and proxy that emulates LMStudio / Ollama backends☆61Aug 21, 2025Updated last year
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- Smart home assistant powered by an SLM☆20Updated this week
- Recursive self-correcting intelligence framework☆17Nov 13, 2025Updated 9 months ago
- Copilot with deepseek and more...☆13Mar 7, 2025Updated last year
- A harness optimized to smaller LLMs☆2,579Aug 29, 2026Updated 2 weeks ago
- Smart OpenAI‑compatible proxy for llama.cpp: manages slots, saves/restores KV cache to disk, routes requests by prefix similarity, and pr…☆53Nov 14, 2025Updated 9 months ago
- Custom Template for checking the availiability of an entity.☆13Oct 31, 2025Updated 10 months ago
- Kon is a minimal coding agent (and also a highly opinionated one)☆353Aug 1, 2026Updated last month
- An electron Wrapper for Open-Interpreter for the lablab.ai hackathon☆12Oct 14, 2023Updated 2 years ago
- An open-source AI agent that lives in your terminal, privacy oriented, with no telemetry.☆79Updated this week
- Deploy open-source AI quickly and easily - Special Bonus Offer • AdRunpod Hub is built for open source. One-click deployment and autoscaling endpoints without provisioning your own infrastructure.
- Scripts and tools for optimizing quantizations in llama.cpp with GGUF imatrices.☆19Jan 10, 2025Updated last year
- Automate faster with n8n templates. Connect apps like Gmail and Slack with ready-to-use, AI-powered workflows. Save time and boost produc…☆15May 13, 2025Updated last year
- AMD APU compatible Ollama. Get up and running with OpenAI gpt-oss, DeepSeek-R1, Gemma 3 and other models.☆16Aug 18, 2026Updated 3 weeks ago
- Compare tables within or across databases☆15Jun 4, 2026Updated 3 months ago
- GPU-accelerated voice assistant — local LLM, fine-tuned Whisper, Kokoro TTS, AMD ROCm☆25Apr 2, 2026Updated 5 months ago
- ☆16Dec 16, 2024Updated last year
- An OpenAI-compatible ASR/STT API server powered by Meta's omnilingual-asr model. Supports real-time streaming via WebSocket and batch tra…☆21Jan 2, 2026Updated 8 months ago
- One-click Qwen3.6-27B inference on Windows. 158 tok/s on RTX 5090, 72 tok/s on RTX 3090. Native, no WSL, no Docker, no telemetry.☆227May 14, 2026Updated 3 months ago
- Multi-modal benchmark for measuring sensitive-information redaction in human recording data☆25Jun 6, 2026Updated 3 months ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- ☆21Jul 4, 2025Updated last year
- ☆12Aug 3, 2024Updated 2 years ago
- Run autoresearch on any NVIDIA GPUs (Works on 2-4GB+ Cards)☆25Mar 22, 2026Updated 5 months ago
- ☆20Mar 17, 2026Updated 5 months ago
- ☆19Nov 9, 2025Updated 10 months ago
- KVarN, KV cache precision tail, low-bit quants in llama.cpp for longer context of better precision in the same VRAM☆1,073Updated this week
- ROCm/AMD GPU benchmark suite for llama.cpp, whisper.cpp, PyTorch☆21Feb 4, 2026Updated 7 months ago