QLoRA: Efficient Finetuning of Quantized LLMs
β11,030Jun 10, 2024Updated 2 years ago
Alternatives and similar repositories for qlora
Users that are interested in qlora are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- π€ PEFT: State-of-the-art Parameter-Efficient Fine-Tuning.β21,749Updated this week
- Accessible large language models via k-bit quantization for PyTorch.β8,510Sep 7, 2026Updated 3 weeks ago
- An open platform for training, serving, and evaluating large language models. Release repo for Vicuna and Chatbot Arena.β39,555May 1, 2026Updated 5 months ago
- Instruct-tune LLaMA on consumer hardwareβ18,900Jul 29, 2024Updated 2 years ago
- LLMs build upon Evol Insturct: WizardLM, WizardCoder, WizardMathβ9,482Jun 7, 2025Updated last year
- Deploy on Railway without the complexity - Free Credits Offer β’ AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- Code and documentation to train Stanford's Alpaca models, and generate the data.β30,234Jul 17, 2024Updated 2 years ago
- [ICLR 2024] Fine-tuning LLaMA to follow Instructions within 1 Hour and 1.2M Parametersβ5,912Mar 14, 2024Updated 2 years ago
- Fast and memory-efficient exact attentionβ25,068Updated this week
- An easy-to-use LLMs quantization package with user-friendly apis, based on GPTQ algorithm.β5,067Apr 11, 2025Updated last year
- Train transformer language models with reinforcement learning.β19,443Updated this week
- DeepSpeed is a deep learning optimization library that makes distributed training and inference easy, efficient, and effective.β43,172Updated this week
- Large Language Model Text Generation Inferenceβ10,884Mar 21, 2026Updated 6 months ago
- OpenLLaMA, a permissively licensed open source reproduction of Meta AIβs LLaMA 7B trained on the RedPajama datasetβ7,525Jul 16, 2023Updated 3 years ago
- Inference code for Llama modelsβ59,619Jan 26, 2025Updated last year
- Serverless GPU API endpoints on Runpod - Get Bonus Credits β’ AdSkip the infrastructure headaches. Auto-scaling, pay-as-you-go, no-ops approach lets you focus on innovating your application.
- Code for loralib, an implementation of "LoRA: Low-Rank Adaptation of Large Language Models"β13,827Dec 17, 2024Updated last year
- LLM training code for Databricks foundation modelsβ4,448Mar 25, 2026Updated 6 months ago
- 4 bits quantization of LLaMA using GPTQβ3,069Jul 13, 2024Updated 2 years ago
- [NeurIPS'23 Oral] Visual Instruction Tuning (LLaVA) built towards GPT-4V level capabilities and beyond.β25,057Aug 12, 2024Updated 2 years ago
- [ICLR 2024] Efficient Streaming Language Models with Attention Sinksβ7,267Jul 11, 2024Updated 2 years ago
- A high-throughput and memory-efficient inference and serving engine for LLMsβ93,111Updated this week
- Gorilla: Training and Evaluating LLMs for Function Calls (Tool Calls)β13,044Apr 13, 2026Updated 5 months ago
- RWKV (pronounced RwaKuv) is an RNN with great LLM performance, which can also be directly trained like a GPT transformer (parallelizable)β¦β14,738Updated this week
- Large-scale Self-supervised Pre-training Across Tasks, Languages, and Modalitiesβ22,226Sep 21, 2026Updated last week
- GPUs on demand by Runpod - Special Offer Available β’ AdRun AI, ML, and HPC workloads on powerful cloud GPUsβwithout limits or wasted spend. Deploy GPUs in under a minute and pay by the second.
- Universal LLM Deployment Engine with ML Compilationβ23,200Updated this week
- Implementation of the LLaMA language model based on nanoGPT. Supports flash attention, Int8 and GPTQ 4bit quantization, LoRA and LLaMA-Adβ¦β6,082Jul 1, 2025Updated last year
- [MLSys 2024 Best Paper Award] AWQ: Activation-aware Weight Quantization for LLM Compression and Accelerationβ3,645Jul 17, 2025Updated last year
- LlamaIndex is the document processing platform for AIβ52,395Updated this week
- Making large AI models cheaper, faster and more accessibleβ41,441Updated this week
- Go ahead and axolotl questionsβ12,513Updated this week
- A guidance language for controlling large language models.β21,786May 21, 2026Updated 4 months ago
- StableLM: Stability AI Language Modelsβ15,672Apr 8, 2024Updated 2 years ago
- An Extensible Toolkit for Finetuning and Inference of Large Foundation Models. Large Models for All.β8,486Aug 10, 2026Updated last month
- Deploy on Railway without the complexity - Free Credits Offer β’ AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- Instruction Tuning with GPT-4β4,331Jun 11, 2023Updated 3 years ago
- Code for the ICLR 2023 paper "GPTQ: Accurate Post-training Quantization of Generative Pretrained Transformers".β2,381Mar 27, 2024Updated 2 years ago
- The RedPajama-Data repository contains code for preparing large datasets for training large language models.β4,993Jun 3, 2026Updated 4 months ago
- Code and documents of LongLoRA and LongAlpaca (ICLR 2024 Oral)β2,686Aug 14, 2024Updated 2 years ago
- Aligning pretrained language models with instruction data generated by themselves.β4,612Mar 27, 2023Updated 3 years ago
- Open-sourced codes for MiniGPT-4 and MiniGPT-v2 (https://minigpt-4.github.io, https://minigpt-v2.github.io/)β25,605Sep 2, 2024Updated 2 years ago
- Tools for merging pretrained large language models.β7,386Sep 12, 2026Updated 3 weeks ago