Inference on CPU code for LLaMA models
☆136Mar 19, 2023Updated 3 years ago
Alternatives and similar repositories for llama-cpu
Users that are interested in llama-cpu are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Chat with Meta's LLaMA models at home made easy☆839Apr 2, 2023Updated 3 years ago
- Inference code for LLaMA models☆46Mar 5, 2023Updated 3 years ago
- Fork of Facebooks LLaMa model to run on CPU☆766Mar 6, 2023Updated 3 years ago
- LLaMA implementation for HuggingFace Transformers☆38Mar 17, 2023Updated 3 years ago
- ☆16Jun 6, 2023Updated 3 years ago
- Bare Metal GPUs on DigitalOcean Gradient AI • AdPurpose-built for serious AI teams training foundational models, running large-scale inference, and pushing the boundaries of what's possible.
- ☆453Oct 15, 2023Updated 2 years ago
- Example of Alpaca-LoRA with llama index.☆31Mar 30, 2023Updated 3 years ago
- Inference code for LLaMA models☆35Mar 7, 2023Updated 3 years ago
- Open ChatGLM Eyes to See the World☆13Mar 30, 2023Updated 3 years ago
- Simple and clean Python implementation of TextRank as per seminal paper by Rada Mihalcea and Paul Tarau. This implementation performs bot…☆12Jan 26, 2021Updated 5 years ago
- ☆26Mar 5, 2023Updated 3 years ago
- A simple package for leveraging Falcon 180B and the HF ecosystem's tools, including training/inference scripts, safetensors, integrations…☆12Mar 11, 2024Updated 2 years ago
- Connect to a GitHub Action runner via SSH through an ngrok tunnel.☆15Dec 5, 2024Updated last year
- Forces DeepSeek R1 models to engage in extended reasoning by intercepting early termination tokens.☆19Feb 12, 2025Updated last year
- End-to-end encrypted email - Proton Mail • AdSpecial offer: 40% Off Yearly / 80% Off First Month. All Proton services are open source and independently audited for security.
- 4 bits quantization of LLaMA using GPTQ☆3,071Jul 13, 2024Updated 2 years ago
- This project is based on the original code of the inteoryx / twitter-video-dl project, which allows users to download Twitter videos as M…☆14Nov 13, 2025Updated 10 months ago
- llmon-py is a multimodal webui for Llama 3-8B.☆16Jul 1, 2024Updated 2 years ago
- Quantized inference code for LLaMA models☆1,034Mar 17, 2023Updated 3 years ago
- Quillan-Ronin - a sovereign mini Software 3.0 runtime on Universal BitNet 1.58-bit logic and a 9B EGGROLL Swarm. v5.4.0-oni features a 34…☆27Updated this week
- Waffer-thin FlaskGPT on Vercel.☆12Jun 1, 2023Updated 3 years ago
- Pascal library for cross-compiler consistent and exact conversion between double-precision floating point number values and strings☆13Updated this week
- A Tensorflow2.x implementation of EfficientDet☆13Mar 4, 2021Updated 5 years ago
- Synthèses vocale piper oobabooga☆14Feb 24, 2024Updated 2 years ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- Data and code for APPDIA: A Discourse-aware Transformer-based Style Transfer Model for Offensive Social Media Conversations (COLING 2022)…☆13Sep 8, 2022Updated 4 years ago
- Python package for Geometric / Clifford Algebra with Pytorch.☆16Jun 2, 2026Updated 4 months ago
- Fast transforms over finite fields☆13Jul 17, 2021Updated 5 years ago
- Agent Zero: The Real "Q" and "Strawberry" of Advanced AI - Combining Quantum-Inspired Algorithms, Complex Mathematical Reasoning, and Eth…☆15Jul 4, 2026Updated 3 months ago
- Exploring the Limitations of Large Language Models on Multi-Hop Queries☆33Mar 2, 2025Updated last year
- Open LLaMA Eyes to See the World☆175Apr 16, 2023Updated 3 years ago
- For converting LLM datasets from one format into another.☆22Nov 12, 2025Updated 10 months ago
- Script for Debian based systems to easily install Matrix-Synapse☆14Oct 14, 2024Updated last year
- Fine-Tuning and Evaluating a Falcon 7B Model for generating HTML code from input prompts.☆11Jan 6, 2024Updated 2 years ago
- Managed Database hosting by DigitalOcean • AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- ☆21Dec 14, 2024Updated last year
- Inference code for facebook LLaMA models with Wrapyfi support☆128Mar 16, 2023Updated 3 years ago
- High-speed download of LLaMA, Facebook's 65B parameter GPT model☆4,107Jun 28, 2023Updated 3 years ago
- ☆10Aug 18, 2022Updated 4 years ago
- Running large language models on a single GPU for throughput-oriented scenarios.☆9,347Oct 28, 2024Updated last year
- ☆12Nov 9, 2018Updated 7 years ago
- Hacked on extensions to the excellent PasDoc☆17Jul 5, 2017Updated 9 years ago