flow-merge is a powerful Python library that enables seamless merging of multiple transformer-based language models using the most popular merge methods such as model soups, SLERP, ties-MERGING or DARE.
☆21Feb 12, 2025Updated last year
Alternatives and similar repositories for flow-merge
Users that are interested in flow-merge are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Unofficial Implementation of Evolutionary Model Merging☆42Mar 28, 2024Updated 2 years ago
- LLM-Training-API: Including Embeddings & ReRankers, mergekit, LaserRMT☆27Feb 18, 2024Updated 2 years ago
- DELLA-Merging: Reducing Interference in Model Merging through Magnitude-Based Sampling☆37Jul 12, 2024Updated 2 years ago
- A quick Crew AI tutorial☆23May 9, 2024Updated 2 years ago
- Source code for the paper "Positional Attention: Expressivity and Learnability of Algorithmic Computation"☆14May 26, 2025Updated last year
- GPUs on demand by Runpod - Special Offer Available • AdRun AI, ML, and HPC workloads on powerful cloud GPUs—without limits or wasted spend. Deploy GPUs in under a minute and pay by the second.
- ☆14Jun 6, 2023Updated 3 years ago
- Automated Identification of Redundant Layer Blocks for Pruning in Large Language Models☆267Apr 23, 2024Updated 2 years ago
- Plug in & Play Pytorch Implementation of the paper: "Evolutionary Optimization of Model Merging Recipes" by Sakana AI☆33Nov 11, 2024Updated last year
- ☆10Jan 20, 2025Updated last year
- ☆11Jun 24, 2022Updated 4 years ago
- configuration files on a Mac☆22Jul 25, 2026Updated 2 weeks ago
- Batch Implementation for Kokoro for enhanced on-device performance☆29May 10, 2025Updated last year
- ☆10Feb 12, 2024Updated 2 years ago
- A block pruning framework for LLMs.☆28May 17, 2025Updated last year
- Serverless GPU API endpoints on Runpod - Get Bonus Credits • AdSkip the infrastructure headaches. Auto-scaling, pay-as-you-go, no-ops approach lets you focus on innovating your application.
- Implementation of LPLR algorithm for matrix compression☆33Nov 21, 2023Updated 2 years ago
- A library for easily merging multiple LLM experts, and efficiently train the merged LLM.☆516Aug 26, 2024Updated last year
- OpenRewrite recipes for Python.☆33Jan 13, 2026Updated 6 months ago
- Official Code For Dual Grained Quantization: Efficient Fine-Grained Quantization for LLM☆14Dec 27, 2023Updated 2 years ago
- Efficient and Effective Weight-Ensembling Mixture of Experts for Multi-Task Model Merging. Arxiv, 2024.☆16Oct 28, 2024Updated last year
- A toolset for easy formant extraction and visualization from wav files and TTS models☆33Sep 2, 2022Updated 3 years ago
- ☆11Oct 22, 2024Updated last year
- Official implementation of "NESS: Node Embeddings from Static Subgraphs"☆20May 30, 2023Updated 3 years ago
- The official code repo and data hub of top_nsigma sampling strategy for LLMs.☆26Feb 11, 2025Updated last year
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- The official code for "Advancing Multimodal Large Language Models with Quantization-Aware Scale Learning for Efficient Adaptation" | [MM2…☆14Dec 7, 2024Updated last year
- Materials for paper "Are Large Language Models Temporally Grounded?"☆14Nov 16, 2023Updated 2 years ago
- Cuda implemenation of flash-kmeans, 2x faster☆23May 8, 2026Updated 3 months ago
- Conditional DDPM for characterizing radio sources from dirty images. (autumn 2023)☆11Nov 30, 2023Updated 2 years ago
- The repository for our paper: Neighboring Perturbations of Knowledge Editing on Large Language Models☆16May 4, 2024Updated 2 years ago
- GitHub Repository Naming Suggestion Service: AI powered naming suggestions for your GitHub repositories☆17Jan 9, 2023Updated 3 years ago
- ☆15Sep 24, 2023Updated 2 years ago
- code for EMNLP 2024 paper: Interpreting Arithmetic Mechanism in Large Language Models through Comparative Neuron Analysis☆12Nov 17, 2024Updated last year
- Medical Mixture of Experts LLM using Mergekit.☆22Mar 6, 2024Updated 2 years ago
- Virtual machines for every use case on DigitalOcean • AdGet dependable uptime with 99.99% SLA, simple security tools, and predictable monthly pricing with DigitalOcean's virtual machines, called Droplets.
- A QT GUI for large language models☆40Dec 27, 2023Updated 2 years ago
- [ECCV24] MixDQ: Memory-Efficient Few-Step Text-to-Image Diffusion Models with Metric-Decoupled Mixed Precision Quantization☆14Nov 27, 2024Updated last year
- [EMNLP 2024] Quantize LLM to extremely low-bit, and finetune the quantized LLMs☆16Jul 18, 2024Updated 2 years ago
- ☆12Dec 27, 2024Updated last year
- Qt GUI for LLM assisted co-writing☆12Jul 28, 2024Updated 2 years ago
- Native progress bars from macOS☆18Nov 5, 2024Updated last year
- Code for the paper "Function-Space Learning Rates"☆23Jun 3, 2025Updated last year