State-of-the-art Parameter-Efficient MoE Fine-tuning Method
☆208Aug 22, 2024Updated 2 years ago
Alternatives and similar repositories for MixLoRA
Users that are interested in MixLoRA are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- An Efficient LLM Fine-Tuning Factory Optimized for MoE PEFT☆144Mar 11, 2025Updated last year
- The codes for 'Progressive cross-primitive consistency for open-world compositional zero-shot learning'☆34Mar 21, 2024Updated 2 years ago
- ☆50Mar 21, 2024Updated 2 years ago
- ☆30Oct 13, 2022Updated 3 years ago
- LoRAMoE: Revolutionizing Mixture of Experts for Maintaining World Knowledge in Language Model Alignment☆404Apr 29, 2024Updated 2 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- [SIGIR'24] The official implementation code of MOELoRA.☆197Jul 22, 2024Updated 2 years ago
- ☆179Jul 22, 2024Updated 2 years ago
- [EMNLP'24] Code and data for paper "Med-MoE: Mixture of Domain-Specific Experts for Lightweight Medical Vision-Language Models"☆157Jul 7, 2025Updated last year
- X-LoRA: Mixture of LoRA Experts☆285Aug 4, 2024Updated 2 years ago
- [CVPR 2023] Diversity-Aware Meta Visual Prompting☆84Nov 30, 2023Updated 2 years ago
- Mixture of Lora Experts☆11Apr 7, 2024Updated 2 years ago
- ☆74Dec 2, 2024Updated last year
- ☆17May 2, 2024Updated 2 years ago
- [COLM 2024] LoraHub: Efficient Cross-Task Generalization via Dynamic LoRA Composition☆670Jul 22, 2024Updated 2 years ago
- End-to-end encrypted cloud storage - Proton Drive • AdSpecial offer: 40% Off Yearly / 80% Off First Month. Protect your most important files, photos, and documents from prying eyes.
- [NeurIPS'24 Oral] HydraLoRA: An Asymmetric LoRA Architecture for Efficient Fine-Tuning☆243Dec 3, 2024Updated last year
- Load multiple LoRA modules simultaneously and automatically switch the appropriate combination of LoRA modules to generate the best answe…☆163Feb 9, 2024Updated 2 years ago
- ☆279Oct 31, 2023Updated 2 years ago
- [CVPR'24] HallusionBench: You See What You Think? Or You Think What You See? An Image-Context Reasoning Benchmark Challenging for GPT-4V(…☆343Oct 14, 2025Updated 11 months ago
- ☆15Jun 6, 2023Updated 3 years ago
- Source code of paper: A Stronger Mixture of Low-Rank Experts for Fine-Tuning Foundation Models. (ICML 2025)☆39Apr 2, 2025Updated last year
- Experiments and data for the paper "When and why vision-language models behave like bags-of-words, and what to do about it?" Oral @ ICLR …☆295Jun 7, 2023Updated 3 years ago
- Papers about Hallucination in Multi-Modal Large Language Models (MLLMs)☆103Nov 21, 2024Updated last year
- [SIGIR'24] The official implementation code of MOELoRA.☆37Aug 3, 2024Updated 2 years ago
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- [ECCV 2024] FairDomain: Achieving Fairness in Cross-Domain Medical Image Segmentation and Classification☆39Apr 15, 2026Updated 5 months ago
- [ ICLR 2025 ] Making LLMs More Effective with Hierarchical Mixture of LoRA Experts☆32Oct 9, 2025Updated 11 months ago
- Official repository for MATES: Model-Aware Data Selection for Efficient Pretraining with Data Influence Models [NeurIPS 2024]☆79Nov 14, 2024Updated last year
- ☆126Jul 6, 2024Updated 2 years ago
- ☆10Apr 16, 2024Updated 2 years ago
- Implementation for MTL-LoRA: Low-Rank Adaptation for Multi-Task Learning☆23Apr 1, 2025Updated last year
- Parameter-Efficient Sparsity Crafting From Dense to Mixture-of-Experts for Instruction Tuning on General Tasks (EMNLP'24)☆143Sep 20, 2024Updated 2 years ago
- SlowFast-LLaVA: A Strong Training-Free Baseline for Video Large Language Models☆294Sep 11, 2026Updated last week
- [ICML 2025] SparseLoRA: Accelerating LLM Fine-Tuning with Contextual Sparsity☆79Mar 10, 2026Updated 6 months ago
- Bare Metal GPUs on DigitalOcean Gradient AI • AdPurpose-built for serious AI teams training foundational models, running large-scale inference, and pushing the boundaries of what's possible.
- Code for http://proceedings.mlr.press/v80/dvurechensky18a.html☆17Jul 31, 2018Updated 8 years ago
- Analyzing and Reducing Catastrophic Forgetting in Parameter Efficient Tuning☆39Nov 17, 2024Updated last year
- Repo - Paper "Capturing Semantics for Imputation with Pre-trained Language Models." [ICDE 2021]☆10Mar 13, 2022Updated 4 years ago
- [ACL 2023] Official resources of "HAHE: Hierarchical Attention for Hyper-Relational Knowledge Graphs in Global and Local Level".☆28Aug 18, 2025Updated last year
- ☆11May 11, 2022Updated 4 years ago
- DiWA: Diverse Weight Averaging for Out-of-Distribution Generalization☆31Jan 31, 2023Updated 3 years ago
- AdaMoLE: Adaptive Mixture of LoRA Experts☆38Oct 11, 2024Updated last year