Inference on CPU code for LLaMA models
☆136Mar 19, 2023Updated 3 years ago
Alternatives and similar repositories for llama-cpu
Users that are interested in llama-cpu are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Chat with Meta's LLaMA models at home made easy☆840Apr 2, 2023Updated 3 years ago
- Inference code for LLaMA models☆46Mar 5, 2023Updated 3 years ago
- Fork of Facebooks LLaMa model to run on CPU☆766Mar 6, 2023Updated 3 years ago
- LLaMA implementation for HuggingFace Transformers☆38Mar 17, 2023Updated 3 years ago
- Inference code for LLaMA models with Gradio Interface and rolling generation like ChatGPT☆48Mar 15, 2023Updated 3 years ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- Searching in-memory corpus with Corpus Query Language (CQL)☆19Dec 2, 2024Updated last year
- ☆454Oct 15, 2023Updated 2 years ago
- Example of Alpaca-LoRA with llama index.☆31Mar 30, 2023Updated 3 years ago
- A repository dedicated to exploring and analyzing the latest developments in Generative AI☆27Apr 1, 2025Updated last year
- The zhong [|] Chinese grammars☆15Jul 24, 2026Updated last month
- ☆17Jan 5, 2023Updated 3 years ago
- Corpus Annotation Graph builder (CAG) is an architectural framework that employs the build-and-annotate pattern for creating a graph.☆14Dec 7, 2023Updated 2 years ago
- Python wrapper for the CWB to extract concordances and score frequency lists☆22Aug 3, 2026Updated last month
- Connect to a GitHub Action runner via SSH through an ngrok tunnel.☆15Dec 5, 2024Updated last year
- Virtual machines for every use case on DigitalOcean • AdGet dependable uptime with 99.99% SLA, simple security tools, and predictable monthly pricing with DigitalOcean's virtual machines, called Droplets.
- NLP-helper for OCR-ed pages in PAGE XML format☆10Dec 6, 2024Updated last year
- Forces DeepSeek R1 models to engage in extended reasoning by intercepting early termination tokens.☆19Feb 12, 2025Updated last year
- A simple extension that uses Bark Text-to-Speech for audio output☆10Nov 20, 2023Updated 2 years ago
- 4 bits quantization of LLaMA using GPTQ☆3,070Jul 13, 2024Updated 2 years ago
- Keyword spaCy is a spaCy pipeline component for extracting keywords from text using cosine similarity.☆14Dec 7, 2023Updated 2 years ago
- Easy to use hybrid index for semantic + keyword search.☆17Jul 19, 2023Updated 3 years ago
- This repository provides German documentation relating to the text recognition and transcription platform eScriptorium. The documentation…☆16Updated this week
- llmon-py is a multimodal webui for Llama 3-8B.☆16Jul 1, 2024Updated 2 years ago
- [WIP] Download full text pdf and supplemental materials for each PubMed IDs.☆14Apr 5, 2017Updated 9 years ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- Quantized inference code for LLaMA models☆1,035Mar 17, 2023Updated 3 years ago
- Automated Semantic Analysis of Discourse Markers☆11May 30, 2022Updated 4 years ago
- A repository to store helpful information and emerging insights in regard to LLMs☆21Oct 27, 2023Updated 2 years ago
- Waffer-thin FlaskGPT on Vercel.☆12Jun 1, 2023Updated 3 years ago
- Synthèses vocale piper oobabooga☆14Feb 24, 2024Updated 2 years ago
- Agent Zero: The Real "Q" and "Strawberry" of Advanced AI - Combining Quantum-Inspired Algorithms, Complex Mathematical Reasoning, and Eth…☆15Jul 4, 2026Updated 2 months ago
- Wikibase extension that allows defining RDF mappings for Wikibase Entities☆17Aug 13, 2026Updated last month
- Open LLaMA Eyes to See the World☆175Apr 16, 2023Updated 3 years ago
- For converting LLM datasets from one format into another.☆22Nov 12, 2025Updated 10 months ago
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- Fine-Tuning and Evaluating a Falcon 7B Model for generating HTML code from input prompts.☆11Jan 6, 2024Updated 2 years ago
- This repository hosts the dataset for the paper Computer Science Named Entity Recognition in the Open Research Knowledge Graph☆21Jan 8, 2024Updated 2 years ago
- ☆17Jul 29, 2024Updated 2 years ago
- llama-4bit-colab☆63Mar 18, 2023Updated 3 years ago
- Easily convert HuggingFace models to GGUF-format for llama.cpp☆21Jul 27, 2024Updated 2 years ago
- Inference code for facebook LLaMA models with Wrapyfi support☆128Mar 16, 2023Updated 3 years ago
- Implementation for Phenotype prediction from single-cell RNA-seq data using attention-based neural networks (Bioinformatics 2024).☆13Jul 15, 2024Updated 2 years ago