flow-merge is a powerful Python library that enables seamless merging of multiple transformer-based language models using the most popular merge methods such as model soups, SLERP, ties-MERGING or DARE.
☆21Feb 12, 2025Updated last year
Alternatives and similar repositories for flow-merge
Users that are interested in flow-merge are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Code for evaluating with Flow-Judge-v0.1 - an open-source, lightweight (3.8B) language model optimized for LLM system evaluations. Crafte…☆86Oct 29, 2024Updated last year
- Unofficial Implementation of Evolutionary Model Merging☆42Mar 28, 2024Updated 2 years ago
- Source code for the paper "Positional Attention: Expressivity and Learnability of Algorithmic Computation"☆14May 26, 2025Updated last year
- Automated Identification of Redundant Layer Blocks for Pruning in Large Language Models☆266Apr 23, 2024Updated 2 years ago
- Spherical Merge Pytorch/HF format Language Models with minimal feature loss.☆153Sep 10, 2023Updated 2 years ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- ☆11Jun 24, 2022Updated 4 years ago
- Official code for "Algorithmic Capabilities of Random Transformers" (NeurIPS 2024)☆15Sep 28, 2024Updated last year
- ☆10Feb 12, 2024Updated 2 years ago
- A QT GUI for large language models☆40Dec 27, 2023Updated 2 years ago
- A block pruning framework for LLMs.☆28May 17, 2025Updated last year
- Implementation of LPLR algorithm for matrix compression☆33Nov 21, 2023Updated 2 years ago
- A library for easily merging multiple LLM experts, and efficiently train the merged LLM.☆518Aug 26, 2024Updated last year
- Official Code For Dual Grained Quantization: Efficient Fine-Grained Quantization for LLM☆14Dec 27, 2023Updated 2 years ago
- Efficient and Effective Weight-Ensembling Mixture of Experts for Multi-Task Model Merging. Arxiv, 2024.☆16Oct 28, 2024Updated last year
- GPUs on demand by Runpod - Special Offer Available • AdRun AI, ML, and HPC workloads on powerful cloud GPUs—without limits or wasted spend. Deploy GPUs in under a minute and pay by the second.
- ☆11Oct 22, 2024Updated last year
- Code for paper "Concrete Subspace Learning based Interference Elimination for Multi-task Model Fusion"☆14Mar 28, 2024Updated 2 years ago
- The official code repo and data hub of top_nsigma sampling strategy for LLMs.☆26Feb 11, 2025Updated last year
- code for EMNLP 2024 paper: How do Large Language Models Learn In-Context? Query and Key Matrices of In-Context Heads are Two Towers for M…☆13Nov 17, 2024Updated last year
- The official code for "Advancing Multimodal Large Language Models with Quantization-Aware Scale Learning for Efficient Adaptation" | [MM2…☆14Dec 7, 2024Updated last year
- Materials for paper "Are Large Language Models Temporally Grounded?"☆14Nov 16, 2023Updated 2 years ago
- ☆14Mar 31, 2024Updated 2 years ago
- Cuda implemenation of flash-kmeans, 2x faster☆23May 8, 2026Updated 2 months ago
- The repository for our paper: Neighboring Perturbations of Knowledge Editing on Large Language Models☆16May 4, 2024Updated 2 years ago
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- Comparing LanceDB and Elasticsearch for full-text search and vector search performance☆30Feb 8, 2026Updated 5 months ago
- Medical Mixture of Experts LLM using Mergekit.☆22Mar 6, 2024Updated 2 years ago
- ☆31Aug 27, 2024Updated last year
- Python Speech Recognition, Voice Recognition, Text-to-Speech and Voice Command Engine☆23Jan 21, 2023Updated 3 years ago
- [ECCV24] MixDQ: Memory-Efficient Few-Step Text-to-Image Diffusion Models with Metric-Decoupled Mixed Precision Quantization☆14Nov 27, 2024Updated last year
- Making AI & LLM APPs components reusable, replaceable, portable, and flexible.☆23Apr 28, 2024Updated 2 years ago
- ☆12Dec 27, 2024Updated last year
- [EMNLP 2024] Quantize LLM to extremely low-bit, and finetune the quantized LLMs☆15Jul 18, 2024Updated 2 years ago
- [ECCV24] The official code repository for paper "Training-Free Model Merging for Multi-target Domain Adaptation".☆18Sep 27, 2024Updated last year
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- ☆19Nov 10, 2024Updated last year
- Efficient encoder-decoder architecture for small language models (≤1B parameters) with cross-architecture knowledge distillation and visi…☆32Feb 7, 2025Updated last year
- Usable implementation of Mogrifier, a circuit for enhancing LSTMs and potentially other networks, from Deepmind☆22Jun 9, 2024Updated 2 years ago
- A Linux SVG-based theme engine for Qt and KDE☆14Dec 26, 2024Updated last year
- [AAAI 2024] DenoSent: A Denoising Objective for Self-Supervised Sentence Representation Learning☆15Apr 29, 2024Updated 2 years ago
- The official implementation of "Well Begun is Half Done: Low-resource Preference Alignment by Weak-to-Strong Decoding"☆22Jun 26, 2025Updated last year
- ☆15Mar 10, 2024Updated 2 years ago