5X faster 60% less memory QLoRA finetuning
☆21May 28, 2024Updated 2 years ago
Alternatives and similar repositories for unsloth
Users that are interested in unsloth are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Parameter-Efficient Sparsity Crafting From Dense to Mixture-of-Experts for Instruction Tuning on General Tasks☆31May 22, 2024Updated 2 years ago
- Parameter-Efficient Sparsity Crafting From Dense to Mixture-of-Experts for Instruction Tuning on General Tasks (EMNLP'24)☆144Sep 20, 2024Updated last year
- Low-Rank adapter extraction for fine-tuned transformers models☆181May 2, 2024Updated 2 years ago
- Curriculum training of instruction-following LLMs with Unsloth☆14Dec 15, 2025Updated 8 months ago
- ☆69May 26, 2024Updated 2 years ago
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- My Gen AI research☆11Jun 3, 2024Updated 2 years ago
- LLM-Training-API: Including Embeddings & ReRankers, mergekit, LaserRMT☆27Feb 18, 2024Updated 2 years ago
- ☆29Apr 29, 2024Updated 2 years ago
- This is our own implementation of 'Layer Selective Rank Reduction'☆240May 26, 2024Updated 2 years ago
- A library for simplifying training with multi gpu setups in the HuggingFace / PyTorch ecosystem.☆16Jun 10, 2026Updated 2 months ago
- ☆10Oct 18, 2023Updated 2 years ago
- Repository for the Q-Filters method (https://arxiv.org/pdf/2503.02812)☆34Mar 7, 2025Updated last year
- ☆26Mar 25, 2025Updated last year
- Diapositivas, notebooks y material de charlas, talleres y el grupo de estudio☆12Apr 24, 2024Updated 2 years ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- FastAPI WebSocket server for the OpenVoice text-to-speech model.☆12Jun 6, 2024Updated 2 years ago
- Creates CMM script that can directly executed on Kaggle from easy merge script☆14Mar 6, 2026Updated 5 months ago
- Optimizing Causal LMs through GRPO with weighted reward functions and automated hyperparameter tuning using Optuna☆60Oct 18, 2025Updated 10 months ago
- Porting of espressif/arduino-esp32 example to M5Stack CoreS3 (GC0308)☆11Nov 30, 2023Updated 2 years ago
- Official implementation of the ICML 2024 paper RoSA (Robust Adaptation)☆46May 20, 2026Updated 3 months ago
- A collection of autogen skills for use with locally run models☆14Feb 29, 2024Updated 2 years ago
- Jeroen Cottaar's work for the Kaggle Geophysical Waveform Inversion competition (2nd place)☆14Aug 11, 2025Updated last year
- Build your own custom knowledge base from various sources such as youtube videos transcripts, tweets, articles, videos and audios. Uses G…☆13Dec 15, 2023Updated 2 years ago
- ☆16Dec 11, 2023Updated 2 years ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- ☆14Feb 7, 2024Updated 2 years ago
- MicroPython viper documentation and examples☆16Apr 19, 2024Updated 2 years ago
- ☆25Dec 13, 2024Updated last year
- Code for PHATGOOSE introduced in "Learning to Route Among Specialized Experts for Zero-Shot Generalization"☆93Feb 27, 2024Updated 2 years ago
- CI scripts designed to build a Pascal-compatible version of vLLM.☆13Aug 10, 2024Updated 2 years ago
- Simple LLM inference server☆20Jun 13, 2024Updated 2 years ago
- ☆16Mar 14, 2024Updated 2 years ago
- ☆13Feb 10, 2021Updated 5 years ago
- Vite + Mantine + Vanilla extract template☆11Jul 23, 2026Updated last month
- Virtual machines for every use case on DigitalOcean • AdGet dependable uptime with 99.99% SLA, simple security tools, and predictable monthly pricing with DigitalOcean's virtual machines, called Droplets.
- A copy of the DirectX Headers from MinGW-64.☆14Sep 7, 2023Updated 2 years ago
- ☆32Aug 27, 2024Updated 2 years ago
- A Model Agnostic function to directly remove specified layers from the LLM☆10May 23, 2024Updated 2 years ago
- the small distributed language model toolkit; fine-tune state-of-the-art LLMs anywhere, rapidly☆33Oct 19, 2024Updated last year
- [ICML 2024] Temporal Spiking Neural Networks with Synaptic Delay for Graph Reasoning☆11Jun 1, 2024Updated 2 years ago
- Basic library for spatial audio SOFA files☆12Sep 29, 2020Updated 5 years ago
- ☆10Oct 24, 2024Updated last year