☆52Jun 11, 2023Updated 3 years ago
Alternatives and similar repositories for servereless-runpod-ggml
Users that are interested in servereless-runpod-ggml are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- ☆50Oct 10, 2023Updated 2 years ago
- Deploy your GGML models to HuggingFace Spaces with Docker and gradio☆37Jun 6, 2023Updated 3 years ago
- Utensil's LLM Playground (2023)☆10May 24, 2026Updated 3 months ago
- GPT4MAX is a free AI chatbot app built with Next.js, the Vercel AI SDK, and OpenAI GPT-4 Turbo.☆18May 10, 2024Updated 2 years ago
- TheBloke's Dockerfiles☆310Mar 8, 2024Updated 2 years ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- QLoRA: Efficient Finetuning of Quantized LLMs☆78Apr 10, 2024Updated 2 years ago
- An open platform for training, serving, and evaluating large language models. Release repo for Vicuna and FastChat-T5.☆11May 26, 2023Updated 3 years ago
- The Runpod worker template for serving our large language model endpoints. Powered by vLLM.☆466Sep 11, 2026Updated last week
- Set of tools for creating backups, compaction and restoration of Apache Kafka® Clusters☆23Jan 20, 2026Updated 8 months ago
- ☆40May 14, 2025Updated last year
- ☆16Feb 15, 2023Updated 3 years ago
- ☆17Updated this week
- Scapytain is a web application that enables you to store, organise and run test campaigns on top of Scapy.☆20Jun 19, 2018Updated 8 years ago
- ComfyUI wrapper for Kokoro-onnx☆37Jan 19, 2025Updated last year
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- Convenient wrapper for fine-tuning and inference of Large Language Models (LLMs) with several quantization techniques (GTPQ, bitsandbytes…☆143Oct 17, 2023Updated 2 years ago
- Convert your README into a Website☆15Oct 24, 2010Updated 15 years ago
- Host LLM via text-generation-inference☆16Dec 5, 2023Updated 2 years ago
- Automates Flatcam generation of G-code for my (and maybe your) PCB milling process☆10Jun 16, 2023Updated 3 years ago
- This repo helps to transform text into a better form for lora training☆12Apr 9, 2023Updated 3 years ago
- An RAG (retrieval augmented generation) app which iterates through a PDF document and can answer user's questions based on the document u…☆16Mar 23, 2025Updated last year
- A more memory-efficient rewrite of the HF transformers implementation of Llama for use with quantized weights.☆63Oct 13, 2023Updated 2 years ago
- ☆28Sep 4, 2023Updated 3 years ago
- The official API used by all Turing AI services☆14Mar 2, 2024Updated 2 years ago
- Managed Database hosting by DigitalOcean • AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- Fuzzing Infrastructure with k8s & cephfs☆12Jul 23, 2020Updated 6 years ago
- ☆18Sep 3, 2026Updated 2 weeks ago
- Install and configure Kubectl for the specified Oracle Engine for Kubernetes (OKE) cluster☆20Dec 9, 2024Updated last year
- ☆34Jun 13, 2024Updated 2 years ago
- ☆17Mar 14, 2017Updated 9 years ago
- Docker that queries Starlinks Dishy and serves you the response on it's own web server☆12Nov 6, 2021Updated 4 years ago
- Implementation of StyleTTS for Mandarin☆11Jun 22, 2023Updated 3 years ago
- ☆14Oct 31, 2023Updated 2 years ago
- StyleTTS2 + Vocos as a Decoder☆13Jul 31, 2026Updated last month
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- ☆10Jan 5, 2015Updated 11 years ago
- KServe community docs for contributions and process☆18Aug 19, 2026Updated last month
- ComfyUI API Workflow Dependency Graph☆17Jul 30, 2024Updated 2 years ago
- An offline CPU-first low-resource chat application to perform RAG on your corpus of data. Powered by OpenChat and CTranslate2.☆15Sep 12, 2026Updated last week
- G2pw's inference speed is accelerated by about 8-10 times. Change loop generated predictive data to only once and model loop prediction b…☆14Dec 30, 2023Updated 2 years ago
- Awesome Curated | Contructive Developmental Theory: Adult Development, Dialectical Thought Form Framework, Immunity to Change, etc☆13May 9, 2025Updated last year
- ☆19Aug 1, 2024Updated 2 years ago