Reproducible llama.cpp configs + per-category quality benches for Qwen3.6-27B on a single RTX 4090. Winners, dead ends, and the silent-corruption bug.
☆24Apr 26, 2026Updated 3 months ago
Alternatives and similar repositories for qwen36-4090-recipes
Users that are interested in qwen36-4090-recipes are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- An agentic runtime that enables secure, extensible and configurable AI automation from any model☆17Aug 15, 2026Updated last week
- Repository for the Q-Filters method (https://arxiv.org/pdf/2503.02812)☆34Mar 7, 2025Updated last year
- The htop for LLM inference see exactly where every GB of VRAM goes and get measured quantization savings.☆70Aug 11, 2026Updated last week
- Using experimental methods to merge large language models☆11Jul 11, 2026Updated last month
- ☆119Apr 28, 2026Updated 3 months ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- Lightweight, ergonomic Solana JSON-RPC wasm client☆17Aug 16, 2026Updated last week
- Use OpenCode / Claude Code For Free☆49Apr 29, 2026Updated 3 months ago
- ☆10Oct 24, 2024Updated last year
- A Python post-processor for 3D printer G-code files, implementing advanced flow and temperature smoothing for improved print quality and …☆28Jun 27, 2026Updated last month
- GPU overclocking utility for Blackwell RTX 50-series on Linux☆22May 26, 2026Updated 2 months ago
- Tiny evaluation of leading LLMs on competitive programming problems☆14Apr 10, 2026Updated 4 months ago
- ☆13Jun 29, 2024Updated 2 years ago
- A Python implementation of an agent swarm system that works with local LLM servers. The system allows you to create multiple agents that …☆14Nov 20, 2024Updated last year
- Open API and Wyoming wrapper around Chatterbox☆31Updated this week
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- Community recipes for serving LLMs on RTX 3090/4090/5090 CUDA gpus. Multi-engine (vLLM, llama.cpp, ik_llama) and model-agnostic. Currentl…☆2,068Updated this week
- Python package for extractive NLP using the OpenAI API☆17Aug 28, 2024Updated last year
- Distil SN97 — Competitive Model Distillation on Bittensor☆35May 20, 2026Updated 3 months ago
- ☆15Apr 26, 2025Updated last year
- A benchmarking harness for coding agents.☆16Jul 31, 2026Updated 3 weeks ago
- My Gen AI research☆11Jun 3, 2024Updated 2 years ago
- Optimizing Causal LMs through GRPO with weighted reward functions and automated hyperparameter tuning using Optuna☆60Oct 18, 2025Updated 10 months ago
- Using modal.com to process FineWeb-edu data☆20Apr 11, 2026Updated 4 months ago
- Adding a multi-text multi-speaker script (diffe) that is based on a script from asiff00 on issue 61 for Sesame: A Conversational Speech G…☆26Mar 28, 2025Updated last year
- Managed Database hosting by DigitalOcean • AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- X Developer Challenge☆12Apr 25, 2024Updated 2 years ago
- PyTorch version of TensorFlow without a PhD☆10Feb 26, 2017Updated 9 years ago
- LINQ for JavaScript library, which allows to work with arrays in a more easy way and focus on business logic.☆11Dec 19, 2016Updated 9 years ago
- alternative way to calculating self attention☆18May 25, 2024Updated 2 years ago
- A streaming local chatbot☆34Jul 3, 2025Updated last year
- Factory's Droid actions☆52Updated this week
- ☆17Feb 23, 2026Updated 6 months ago
- Add ability to interrupt own message☆14Apr 21, 2024Updated 2 years ago
- A Ray-based LLM server compatible with OpenAI API☆12Mar 12, 2024Updated 2 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- prinzbench is a private benchmark that ranks LLMs based on their ability to conduct legal research and analysis and locate obscure public…☆128Jul 18, 2026Updated last month
- ☆20Apr 27, 2025Updated last year
- Using DSPy to optimize Chat-to-SQL☆15Nov 17, 2025Updated 9 months ago
- Ubuntu Server edition: automated setup script for Intel Arc Pro B70 GPU LLM inference server with vLLM tensor parallelism. 140 tok/s on 2…☆33Apr 26, 2026Updated 3 months ago
- LLM inference in C/C++☆55Updated this week
- 📄 Nano JSX Template using Isomorphic JSX.☆14Oct 7, 2022Updated 3 years ago
- Client Code Examples, Use Cases and Benchmarks for Enterprise h2oGPTe RAG-Based GenAI Platform☆90Sep 9, 2025Updated 11 months ago