Reproducible llama.cpp configs + per-category quality benches for Qwen3.6-27B on a single RTX 4090. Winners, dead ends, and the silent-corruption bug.
☆23Apr 26, 2026Updated 3 months ago
Alternatives and similar repositories for qwen36-4090-recipes
Users that are interested in qwen36-4090-recipes are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- An agentic runtime that enables secure, extensible and configurable AI automation from any model☆17Jul 26, 2026Updated last week
- SNDR Core Engine (Genesis) — vLLM runtime patch-overlay for Qwen3.6 + Gemma4 on consumer NVIDIA (Ampere sm_86, 2× A5000/3090). Qwen3.6-35…☆131Updated this week
- Repository for the Q-Filters method (https://arxiv.org/pdf/2503.02812)☆34Mar 7, 2025Updated last year
- ☆118Apr 28, 2026Updated 3 months ago
- Use OpenCode / Claude Code For Free☆49Apr 29, 2026Updated 3 months ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- Claude Code skill that turns any topic, conversation, or file into a single-file, beautifully typeset HTML document - trigger with /paper☆24Jul 6, 2026Updated 3 weeks ago
- ☆10Oct 24, 2024Updated last year
- GPU overclocking utility for Blackwell RTX 50-series on Linux☆22May 26, 2026Updated 2 months ago
- Tiny evaluation of leading LLMs on competitive programming problems☆14Apr 10, 2026Updated 3 months ago
- ☆13Jun 29, 2024Updated 2 years ago
- Kiwix ZIM-to-vector RAG system for local, offline LLM knowledge retrieval☆29Mar 24, 2026Updated 4 months ago
- Mixed-vendor GPU inference cluster manager with speculative decoding☆31Jul 2, 2026Updated last month
- A Python implementation of an agent swarm system that works with local LLM servers. The system allows you to create multiple agents that …☆14Nov 20, 2024Updated last year
- Open API and Wyoming wrapper around Chatterbox☆28Jan 2, 2026Updated 7 months ago
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- Community recipes for serving LLMs on RTX 3090/4090/5090 CUDA gpus. Multi-engine (vLLM, llama.cpp, ik_llama) and model-agnostic. Currentl…☆1,868Updated this week
- Phoenix CLI for humans and agents☆18Updated this week
- Python package for extractive NLP using the OpenAI API☆17Aug 28, 2024Updated last year
- ☆17Feb 1, 2024Updated 2 years ago
- Distil SN97 — Competitive Model Distillation on Bittensor☆35May 20, 2026Updated 2 months ago
- ☆15Apr 26, 2025Updated last year
- Optimizing Causal LMs through GRPO with weighted reward functions and automated hyperparameter tuning using Optuna☆60Oct 18, 2025Updated 9 months ago
- Using modal.com to process FineWeb-edu data☆20Apr 11, 2026Updated 3 months ago
- Adding a multi-text multi-speaker script (diffe) that is based on a script from asiff00 on issue 61 for Sesame: A Conversational Speech G…☆26Mar 28, 2025Updated last year
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- X Developer Challenge☆12Apr 25, 2024Updated 2 years ago
- How we built Witan - four months of engineering an LLM spreadsheet agent☆97May 15, 2026Updated 2 months ago
- PyTorch version of TensorFlow without a PhD☆10Feb 26, 2017Updated 9 years ago
- LINQ for JavaScript library, which allows to work with arrays in a more easy way and focus on business logic.☆11Dec 19, 2016Updated 9 years ago
- A streaming local chatbot☆34Jul 3, 2025Updated last year
- Use Hermes-2-Pro-Mistral-7B function calling with your OpenAI API compatible code.☆18May 7, 2024Updated 2 years ago
- AI Agents Workshop with Red Hat AI☆13Feb 26, 2025Updated last year
- Deep learning framework; image classification; Nature Food publication☆11Mar 28, 2024Updated 2 years ago
- ☆17Feb 23, 2026Updated 5 months ago
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- A Ray-based LLM server compatible with OpenAI API☆12Mar 12, 2024Updated 2 years ago
- prinzbench is a private benchmark that ranks LLMs based on their ability to conduct legal research and analysis and locate obscure public…☆127Jul 18, 2026Updated 2 weeks ago
- ☆20Apr 27, 2025Updated last year
- Ubuntu Server edition: automated setup script for Intel Arc Pro B70 GPU LLM inference server with vLLM tensor parallelism. 140 tok/s on 2…☆31Apr 26, 2026Updated 3 months ago
- Client Code Examples, Use Cases and Benchmarks for Enterprise h2oGPTe RAG-Based GenAI Platform☆90Sep 9, 2025Updated 10 months ago
- General plug-and-play inference library for Recursive Language Models (RLMs), supporting various sandboxes.☆17Apr 26, 2026Updated 3 months ago
- a set of scripts to easily convert all training data from huggingface into alpaca instruct or sharegpt format, which should allow for eas…☆20Mar 14, 2025Updated last year