Run inference on MPT-30B using CPU
☆574Jun 30, 2023Updated 3 years ago
Alternatives and similar repositories for mpt-30B-inference
Users that are interested in mpt-30B-inference are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Run inference on replit-3B code instruct model using CPU☆161Jul 5, 2023Updated 3 years ago
- Python bindings for the Transformer models implemented in C/C++ using GGML library.☆1,886Jan 28, 2024Updated 2 years ago
- Chat with your data privately using MPT-30b☆183Jun 29, 2023Updated 3 years ago
- LLM training code for Databricks foundation models☆4,443Mar 25, 2026Updated 5 months ago
- This is the official code for MobileSAM project that makes SAM lightweight for mobile applications and beyond!☆5,865May 5, 2026Updated 4 months ago
- Open source password manager - Proton Pass • AdSecurely store, share, and autofill your credentials with Proton Pass, the end-to-end encrypted password manager trusted by millions.
- Running Llama 2 and other Open-Source LLMs on CPU Inference Locally for Document Q&A☆971Nov 6, 2023Updated 2 years ago
- LLMs build upon Evol Insturct: WizardLM, WizardCoder, WizardMath☆9,481Jun 7, 2025Updated last year
- A more memory-efficient rewrite of the HF transformers implementation of Llama for use with quantized weights.☆2,943Sep 30, 2023Updated 2 years ago
- CodeTF: One-stop Transformer Library for State-of-the-art Code LLM☆1,480May 1, 2025Updated last year
- Run any Llama 2 locally with gradio UI on GPU or CPU from anywhere (Linux/Windows/Mac). Use `llama2-wrapper` as your local llama2 backend…☆1,936Mar 22, 2024Updated 2 years ago
- LLM as a Chatbot Service☆3,318Nov 20, 2023Updated 2 years ago
- Salesforce open-source LLMs with 8k sequence length.☆726Jun 2, 2026Updated 3 months ago
- ☆2,544Jan 7, 2025Updated last year
- Large Language Model Text Generation Inference☆10,890Mar 21, 2026Updated 5 months ago
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- Locally hosted tool that connects documents to LLMs for summarization and querying, with a simple GUI.☆797Aug 1, 2023Updated 3 years ago
- OpenLLaMA, a permissively licensed open source reproduction of Meta AI’s LLaMA 7B trained on the RedPajama dataset☆7,529Jul 16, 2023Updated 3 years ago
- prompt2model - Generate Deployable Models from Natural Language Instructions☆2,018Dec 29, 2024Updated last year
- ☆133Nov 24, 2023Updated 2 years ago
- H2O LLM Studio - a framework and no-code GUI for fine-tuning LLMs. Documentation: https://docs.h2o.ai/h2o-llmstudio/☆5,177Updated this week
- LongLLaMA is a large language model capable of handling long contexts. It is based on OpenLLaMA and fine-tuned with the Focused Transform…☆1,465Nov 7, 2023Updated 2 years ago
- LLMs custom-chatbots console ⚡☆5,231Feb 27, 2024Updated 2 years ago
- LOMO: LOw-Memory Optimization☆993Jul 2, 2024Updated 2 years ago
- Cross-Platform, GPU Accelerated Whisper 🏎️☆1,794Feb 27, 2024Updated 2 years ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- FlagAI (Fast LArge-scale General AI models) is a fast, easy-to-use and extensible toolkit for large-scale model.☆3,869Jul 13, 2026Updated last month
- An Open-source Toolkit for LLM Development☆2,800Jan 13, 2025Updated last year
- Gorilla: Training and Evaluating LLMs for Function Calls (Tool Calls)☆13,017Apr 13, 2026Updated 4 months ago
- QLoRA: Efficient Finetuning of Quantized LLMs☆11,009Jun 10, 2024Updated 2 years ago
- Plug in and Play Implementation of Tree of Thoughts: Deliberate Problem Solving with Large Language Models that Elevates Model Reasoning …☆4,590Jul 29, 2025Updated last year
- LLaMA v2 Chatbot☆1,410Aug 27, 2023Updated 3 years ago
- Explore large language models in 512MB of RAM☆1,192Feb 19, 2026Updated 6 months ago
- Run evaluation on LLMs using human-eval benchmark☆431Sep 12, 2023Updated 2 years ago
- ☆1,064May 29, 2023Updated 3 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- Scale LLM Engine public repository☆837Aug 27, 2026Updated last week
- An open source implementation of OpenAI's ChatGPT Code interpreter☆3,536Mar 20, 2024Updated 2 years ago
- 🌸 Run LLMs at home, BitTorrent-style. Fine-tuning and inference up to 10x faster than offloading☆10,542Sep 7, 2024Updated 2 years ago
- Universal LLM Deployment Engine with ML Compilation☆23,140Aug 17, 2026Updated 3 weeks ago
- Run any open-source LLMs, such as DeepSeek and Llama, as OpenAI compatible API endpoint in the cloud.☆12,528Aug 31, 2026Updated last week
- [ICLR 2024] Fine-tuning LLaMA to follow Instructions within 1 Hour and 1.2M Parameters☆5,914Mar 14, 2024Updated 2 years ago
- Superagent protects your AI applications against prompt injections, data leaks, and harmful outputs. Embed safety directly into your app …☆6,733Aug 25, 2026Updated 2 weeks ago