QLoRA: Efficient Finetuning of Quantized LLMs
☆79Apr 10, 2024Updated 2 years ago
Alternatives and similar repositories for qlora
Users that are interested in qlora are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- A bagel, with everything.☆326Apr 11, 2024Updated 2 years ago
- Customizable implementation of the self-instruct paper.☆1,051Mar 7, 2024Updated 2 years ago
- This is the Mixture-of-Agents (MoA) concept, adapted from the original work by TogetherAI. My version is tailored for local model usage a…☆12Jun 25, 2024Updated 2 years ago
- ☆75Sep 5, 2023Updated 2 years ago
- ☆28Aug 30, 2023Updated 2 years ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- QLoRA with Enhanced Multi GPU Support☆38Aug 8, 2023Updated 3 years ago
- Modified Stanford-Alpaca Trainer for Training Replit's Code Model☆47Jun 1, 2023Updated 3 years ago
- Code for the arXiv preprint "The Unreasonable Effectiveness of Easy Training Data"☆48Jan 17, 2024Updated 2 years ago
- Code for fine-tuning Platypus fam LLMs using LoRA☆625Feb 4, 2024Updated 2 years ago
- ☆16Feb 21, 2026Updated 5 months ago
- Demonstration that finetuning RoPE model on larger sequences than the pre-trained model adapts the model context limit☆62Jun 21, 2023Updated 3 years ago
- ☆166Jun 1, 2023Updated 3 years ago
- A Python library for efficient and flexible cycle-consistency training of transformer models via iteratie back-translation. Memory and co…☆11Jan 13, 2025Updated last year
- Inference Llama 2 in one file of pure Python☆426Nov 21, 2025Updated 8 months ago
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- Traing PRO extension for oobabooga WebUI - recent dev version☆52Aug 7, 2025Updated last year
- A collection of modular datasets generated by GPT-4, General-Instruct - Roleplay-Instruct - Code-Instruct - and Toolformer☆1,666Sep 15, 2023Updated 2 years ago
- Merge Transformers language models by use of gradient parameters.☆215Aug 8, 2024Updated 2 years ago
- YaRN: Efficient Context Window Extension of Large Language Models☆1,768Apr 17, 2024Updated 2 years ago
- Các thí nghiệm liên quan tới LLMs cho tiếng Việt (insprised by Physics of LLMs Series)☆11Oct 21, 2024Updated last year
- Train Llama Loras Easily☆30Aug 3, 2023Updated 3 years ago
- Learn & build: Always available expertise powered by AI☆14Jul 10, 2023Updated 3 years ago
- ☆131Oct 1, 2024Updated last year
- Full finetuning of large language models without large memory requirements☆92Sep 22, 2025Updated 10 months ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- Inference code for LLaMA models☆20Apr 3, 2025Updated last year
- ☆79Dec 26, 2023Updated 2 years ago
- code for training and using chess embeddings models☆14Jun 9, 2024Updated 2 years ago
- Parameter-Efficient Sparsity Crafting From Dense to Mixture-of-Experts for Instruction Tuning on General Tasks (EMNLP'24)☆144Sep 20, 2024Updated last year
- Experiments on speculative sampling with Llama models☆129Jun 8, 2023Updated 3 years ago
- Training LLMs with QLoRA + FSDP☆1,550Nov 9, 2024Updated last year
- Analyzing LLM Alignment via Token distribution shift☆17Jan 26, 2024Updated 2 years ago
- reimagine the implementation of C-3PO droid voice synthesizer and multilingual translation and communication capabilities with the latest…☆12Mar 6, 2024Updated 2 years ago
- FuseAI Project☆600Jan 25, 2025Updated last year
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- extension for text WebUI☆20Aug 7, 2025Updated last year
- A more memory-efficient rewrite of the HF transformers implementation of Llama for use with quantized weights.☆2,937Sep 30, 2023Updated 2 years ago
- QLoRA: Efficient Finetuning of Quantized LLMs☆10,991Jun 10, 2024Updated 2 years ago
- ☆40Mar 25, 2023Updated 3 years ago
- Exploring finetuning public checkpoints on filter 8K sequences on Pile☆116Mar 22, 2023Updated 3 years ago
- High-performance tokenized language data-loader for Python C++ extension☆15Jul 22, 2024Updated 2 years ago
- AI shorts generator☆26Apr 18, 2024Updated 2 years ago