Implementation of DoRA
☆310Jun 7, 2024Updated 2 years ago
Alternatives and similar repositories for dora
Users that are interested in dora are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- LoRA and DoRA from Scratch Implementations☆231Mar 5, 2024Updated 2 years ago
- Official implementation of "DoRA: Weight-Decomposed Low-Rank Adaptation"☆121Apr 28, 2024Updated 2 years ago
- ☆234Jun 24, 2024Updated 2 years ago
- ☆206Dec 5, 2024Updated last year
- [EMNLP 2023, Main Conference] Sparse Low-rank Adaptation of Pre-trained Language Models☆87Mar 5, 2024Updated 2 years ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- [ICML2024 (Oral)] Official PyTorch implementation of DoRA: Weight-Decomposed Low-Rank Adaptation☆999Mar 24, 2026Updated 5 months ago
- Training LLMs with QLoRA + FSDP☆1,552Nov 9, 2024Updated last year
- ☆279Oct 31, 2023Updated 2 years ago
- GaLore: Memory-Efficient LLM Training by Gradient Low-Rank Projection☆1,703Oct 28, 2024Updated last year
- Official code for ReLoRA from the paper Stack More Layers Differently: High-Rank Training Through Low-Rank Updates☆475Apr 21, 2024Updated 2 years ago
- Codebase for Merging Language Models (ICML 2024)☆874May 5, 2024Updated 2 years ago
- The official implementation of Self-Play Fine-Tuning (SPIN)☆1,255May 8, 2024Updated 2 years ago
- MoRA: High-Rank Updating for Parameter-Efficient Fine-Tuning☆361Aug 7, 2024Updated 2 years ago
- Cascade Speculative Drafting☆33Apr 2, 2024Updated 2 years ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- ☆235Jun 11, 2024Updated 2 years ago
- Implementation of Spectral State Space Models☆16Feb 23, 2024Updated 2 years ago
- A pytorch quantization backend for optimum☆1,053Aug 25, 2026Updated 3 weeks ago
- This is a repository for "PMET: Precise Model Editing in a Transformer"☆58Sep 28, 2023Updated 2 years ago
- ☆23Dec 18, 2023Updated 2 years ago
- Stanford NLP Python library for Representation Finetuning (ReFT)☆1,586Mar 5, 2026Updated 6 months ago
- Low-Rank adapter extraction for fine-tuned transformers models☆181May 2, 2024Updated 2 years ago
- Accessible large language models via k-bit quantization for PyTorch.☆8,487Sep 7, 2026Updated last week
- ☆179Jul 22, 2024Updated 2 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- PyTorch compiler that accelerates training and inference. Get built-in optimizations for performance, memory, parallelism, and easily wri…☆1,473Updated this week
- PyTorch native post-training library☆5,812Sep 9, 2026Updated last week
- S-LoRA: Serving Thousands of Concurrent LoRA Adapters☆1,925Jan 21, 2024Updated 2 years ago
- Token Omission Via Attention☆129Oct 13, 2024Updated last year
- Let's create synthetic textbooks together :)☆74Jan 29, 2024Updated 2 years ago
- Implementation of the training framework proposed in Self-Rewarding Language Model, from MetaAI☆1,411Apr 11, 2024Updated 2 years ago
- Utilities for PyTorch distributed☆26Feb 27, 2025Updated last year
- The Truth Is In There: Improving Reasoning in Language Models with Layer-Selective Rank Reduction☆396Jul 9, 2024Updated 2 years ago
- AdaLoRA: Adaptive Budget Allocation for Parameter-Efficient Fine-Tuning (ICLR 2023).☆395Jun 1, 2023Updated 3 years ago
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- Official repository of NEFTune: Noisy Embeddings Improves Instruction Finetuning☆412May 17, 2024Updated 2 years ago
- A python package made to generate sequences (greedy and beam-search) from Pytorch (not necessarily HF transformers) models.☆19Dec 12, 2025Updated 9 months ago
- Repo for Rho-1: Token-level Data Selection & Selective Pretraining of LLMs.☆472Apr 18, 2024Updated 2 years ago
- Merge safetensor files using the technique described in "Language Models are Super Mario: Absorbing Abilities from Homologous Models as a…☆83Oct 17, 2024Updated last year
- Code and documents of LongLoRA and LongAlpaca (ICLR 2024 Oral)☆2,688Aug 14, 2024Updated 2 years ago
- Implementation of 💍 Ring Attention, from Liu et al. at Berkeley AI, in Pytorch☆545May 16, 2025Updated last year
- Schedule-Free Optimization in PyTorch☆2,323Jul 28, 2026Updated last month