A guidance compatibility layer for llama-cpp-python
☆37Sep 11, 2023Updated 3 years ago
Alternatives and similar repositories for llama-cpp-guidance
Users that are interested in llama-cpp-guidance are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Plug n Play GBNF Compiler for llama.cpp☆33Nov 8, 2023Updated 2 years ago
- "Pacha" TUI (Text User Interface) is a JavaScript application that utilizes the "blessed" library. It serves as a frontend for llama.cpp …☆38Aug 3, 2023Updated 3 years ago
- Karpathy's llama2.c transpiled to MLX for Apple Silicon☆14Dec 28, 2023Updated 2 years ago
- Evaluating practical performance of local multi-turn conversational LLMs.☆20Aug 1, 2025Updated last year
- ☆62Jan 21, 2024Updated 2 years ago
- Deploy open-source AI quickly and easily - Special Bonus Offer • AdRunpod Hub is built for open source. One-click deployment and autoscaling endpoints without provisioning your own infrastructure.
- Agent-friendly shell and session runtime built on top of Jido VFS☆20Updated this week
- The llama-cpp-agent framework is a tool designed for easy interaction with Large Language Models (LLMs). Allowing users to chat with LLM …☆659Mar 9, 2026Updated 6 months ago
- Code for paper: "QuIP: 2-Bit Quantization of Large Language Models With Guarantees" adapted for Llama models☆40Aug 4, 2023Updated 3 years ago
- After my server ui improvements were successfully merged, consider this repo a playground for experimenting, tinkering and hacking around…☆53Aug 18, 2024Updated 2 years ago
- Minimal R wrapper for llama.cpp☆56Jun 5, 2023Updated 3 years ago
- CI scripts designed to build a Pascal-compatible version of vLLM.☆13Aug 10, 2024Updated 2 years ago
- Sources and instructions for building an Intel(r) Edison-based monitoring system witih motion detection and cloud/social connection☆20Aug 20, 2017Updated 9 years ago
- ☆10May 16, 2023Updated 3 years ago
- PyTorch implementation of R2D2 (Recurrent Reply Distributed DQN)☆13Nov 14, 2019Updated 6 years ago
- Serverless GPU API endpoints on Runpod - Get Bonus Credits • AdSkip the infrastructure headaches. Auto-scaling, pay-as-you-go, no-ops approach lets you focus on innovating your application.
- Make codex, gemini-cli, or any other coding agent into an RLM.☆20Feb 27, 2026Updated 7 months ago
- Local LLaMAs/Models in VSCode☆52Jun 5, 2023Updated 3 years ago
- A Tauri2 plugin for embedding a terminal in your application☆22Jul 8, 2026Updated 2 months ago
- A survey on machine learning for combinatorial optimization.☆12Dec 27, 2021Updated 4 years ago
- ☆14Jun 11, 2021Updated 5 years ago
- A simple experiment on letting two local LLM have a conversation about anything!☆112Jul 3, 2024Updated 2 years ago
- An llm wrapper for OpenAI☆13Dec 14, 2024Updated last year
- Experiments with open source LLMs☆75Aug 21, 2026Updated last month
- Neural agents evolved with Neataptic to seek targets☆21Aug 27, 2017Updated 9 years ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- ☆19Aug 31, 2022Updated 4 years ago
- A tool for generating function arguments and choosing what function to call with local LLMs☆435Mar 12, 2024Updated 2 years ago
- Summarization with Pointer-Generator Networks☆15Sep 1, 2020Updated 6 years ago
- ☆10Dec 19, 2024Updated last year
- Instant Neural Graphics Primitives from scratch, zero dependencies. Learning by doing.☆10Aug 18, 2023Updated 3 years ago
- Copy a bunch of files into your clipboard to provide context for LLMs☆115Feb 8, 2026Updated 7 months ago
- ☆13Nov 16, 2022Updated 3 years ago
- 33B Chinese LLM, DPO QLORA, 100K context, AirLLM 70B inference with single 4GB GPU☆14May 5, 2024Updated 2 years ago
- 📽 Python package to live stream ML-Agents training process from Google Colab to Twitch/YouTube server.☆14Mar 27, 2022Updated 4 years ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- Source code for Pathfinding in Stochastic Environments paper.☆15Oct 27, 2022Updated 3 years ago
- JAX implementation of GPTQ quantization algorithm☆10Jul 19, 2023Updated 3 years ago
- ☆166Jun 1, 2023Updated 3 years ago
- Improve Devcontainer Creation☆17Sep 3, 2026Updated 3 weeks ago
- Clean RL implementation using MLX☆34Mar 8, 2024Updated 2 years ago
- An Autonomous LLM Agent that runs on Wizcoder-15B☆333Oct 21, 2024Updated last year
- mlx implementations of various transformers, speedups, training☆34Dec 14, 2023Updated 2 years ago