QLoRA: Efficient Finetuning of Quantized LLMs
☆79Apr 10, 2024Updated 2 years ago
Alternatives and similar repositories for qlora
Users that are interested in qlora are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- A bagel, with everything.☆326Apr 11, 2024Updated 2 years ago
- Customizable implementation of the self-instruct paper.☆1,051Mar 7, 2024Updated 2 years ago
- This is the Mixture-of-Agents (MoA) concept, adapted from the original work by TogetherAI. My version is tailored for local model usage a…☆12Jun 25, 2024Updated 2 years ago
- ☆75Sep 5, 2023Updated 2 years ago
- ☆28Aug 30, 2023Updated 2 years ago
- Proton VPN Special Offer - Get 70% off • AdSpecial partner offer. Trusted by over 100 million users worldwide. Tested, Approved and Recommended by Experts.
- QLoRA with Enhanced Multi GPU Support☆38Aug 8, 2023Updated 2 years ago
- ☆21Oct 6, 2023Updated 2 years ago
- Modified Stanford-Alpaca Trainer for Training Replit's Code Model☆47Jun 1, 2023Updated 3 years ago
- Code for the arXiv preprint "The Unreasonable Effectiveness of Easy Training Data"☆48Jan 17, 2024Updated 2 years ago
- Code for fine-tuning Platypus fam LLMs using LoRA☆625Feb 4, 2024Updated 2 years ago
- ☆16Feb 21, 2026Updated 5 months ago
- Demonstration that finetuning RoPE model on larger sequences than the pre-trained model adapts the model context limit☆62Jun 21, 2023Updated 3 years ago
- ☆166Jun 1, 2023Updated 3 years ago
- Traing PRO extension for oobabooga WebUI - recent dev version☆53Aug 7, 2025Updated 11 months ago
- End-to-end encrypted email - Proton Mail • AdSpecial offer: 40% Off Yearly / 80% Off First Month. All Proton services are open source and independently audited for security.
- A collection of modular datasets generated by GPT-4, General-Instruct - Roleplay-Instruct - Code-Instruct - and Toolformer☆1,668Sep 15, 2023Updated 2 years ago
- Merge Transformers language models by use of gradient parameters.☆215Aug 8, 2024Updated last year
- YaRN: Efficient Context Window Extension of Large Language Models☆1,740Apr 17, 2024Updated 2 years ago
- Các thí nghiệm liên quan tới LLMs cho tiếng Việt (insprised by Physics of LLMs Series)☆11Oct 21, 2024Updated last year
- Train Llama Loras Easily☆30Aug 3, 2023Updated 2 years ago
- Learn & build: Always available expertise powered by AI☆14Jul 10, 2023Updated 3 years ago
- ☆131Oct 1, 2024Updated last year
- OpenLine Protocol (OLP) — typed claim/evidence graphs for agentic systems with a 5-number digest. One-command FastAPI demo; auditable, re…☆15Feb 14, 2026Updated 5 months ago
- Full finetuning of large language models without large memory requirements☆92Sep 22, 2025Updated 10 months ago
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- ☆78Dec 26, 2023Updated 2 years ago
- code for training and using chess embeddings models☆14Jun 9, 2024Updated 2 years ago
- Parameter-Efficient Sparsity Crafting From Dense to Mixture-of-Experts for Instruction Tuning on General Tasks (EMNLP'24)☆145Sep 20, 2024Updated last year
- Experiments on speculative sampling with Llama models☆129Jun 8, 2023Updated 3 years ago
- Training LLMs with QLoRA + FSDP☆1,550Nov 9, 2024Updated last year
- Analyzing LLM Alignment via Token distribution shift☆17Jan 26, 2024Updated 2 years ago
- reimagine the implementation of C-3PO droid voice synthesizer and multilingual translation and communication capabilities with the latest…☆12Mar 6, 2024Updated 2 years ago
- FuseAI Project☆600Jan 25, 2025Updated last year
- extension for text WebUI☆20Aug 7, 2025Updated 11 months ago
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- A more memory-efficient rewrite of the HF transformers implementation of Llama for use with quantized weights.☆2,934Sep 30, 2023Updated 2 years ago
- ☆40Mar 25, 2023Updated 3 years ago
- Exploring finetuning public checkpoints on filter 8K sequences on Pile☆116Mar 22, 2023Updated 3 years ago
- High-performance tokenized language data-loader for Python C++ extension☆15Jul 22, 2024Updated 2 years ago
- AI shorts generator☆26Apr 18, 2024Updated 2 years ago
- Text WebUI extension to add clever Notebooks to Chat mode☆148Aug 7, 2025Updated 11 months ago
- This is the oficial repository for "Safer-Instruct: Aligning Language Models with Automated Preference Data"☆17Feb 22, 2024Updated 2 years ago