Delta-CoMe can achieve near loss-less 1-bit compressin which has been accepted by NeurIPS 2024
☆59Nov 16, 2024Updated last year
Alternatives and similar repositories for Delta-CoMe
Users that are interested in Delta-CoMe are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- ACL 2024: LoRA-Flow Dynamic LoRA Fusion for Large Language Models in Generative Tasks☆25Oct 9, 2024Updated last year
- [ICLR 2026] Official Implementation of ProxyThinker: Test-Time Guidance through Small Visual Reasoners.☆22Sep 24, 2025Updated 11 months ago
- An Open-Source Knowledge-Enhanced Multilingual Supervised Fine-tuning Dataset☆27Jan 19, 2025Updated last year
- ☆16Jul 23, 2024Updated 2 years ago
- [NeurIPS 2024 Spotlight] EMR-Merging: Tuning-Free High-Performance Model Merging☆82Mar 1, 2025Updated last year
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- OMGEval😮: An Open Multilingual Generative Evaluation Benchmark for Foundation Models☆36Jul 19, 2024Updated 2 years ago
- [arXiv] "Linear Dynamics in the RLVR Training of Large Language Models"☆19May 25, 2026Updated 3 months ago
- ☆21Jul 3, 2025Updated last year
- ☆15Sep 10, 2019Updated 6 years ago
- ☆18Aug 19, 2024Updated 2 years ago
- Model merging is a highly efficient approach for long-to-short reasoning.☆103Oct 15, 2025Updated 10 months ago
- official code repo for paper "Merging Models on the Fly Without Retraining: A Sequential Approach to Scalable Continual Model Merging"☆25Oct 11, 2025Updated 10 months ago
- [ICML 2026] Hybrid Policy Distillation (HPD) is a practical distillation framework for reasoning-oriented language models. This repositor…☆24Apr 24, 2026Updated 4 months ago
- Codes for Merging Large Language Models☆37Aug 7, 2024Updated 2 years ago
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- ☆23Oct 22, 2024Updated last year
- Official code for "Breaking the Ceiling in On-Policy Distillation via Multi-Agent Debate" (arXiv:2605.01347).☆37Jul 31, 2026Updated last month
- Adapter-X: A Novel General Parameter-Efficient Fine-Tuning Framework for Vision☆11Jul 22, 2024Updated 2 years ago
- Code used for analysis and visualiation of ocean model data during my postdoc☆12Mar 1, 2023Updated 3 years ago
- ☆30Jan 22, 2025Updated last year
- A Framework for Decoupling and Assessing the Capabilities of VLMs☆44Jun 28, 2024Updated 2 years ago
- [ICLR 2025] Official implementation of paper "Dynamic Low-Rank Sparse Adaptation for Large Language Models".☆25Mar 16, 2025Updated last year
- MIXQ: Taming Dynamic Outliers in Mixed-Precision Quantization by Online Prediction☆94Oct 29, 2024Updated last year
- [NAACL 2025] Representing Rule-based Chatbots with Transformers☆23Feb 9, 2025Updated last year
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- ☆15Apr 11, 2024Updated 2 years ago
- ☆47Nov 25, 2024Updated last year
- 对llava官方代码的一些学习笔记☆28Oct 11, 2024Updated last year
- ☆13Mar 5, 2025Updated last year
- [ACL 2025] An official pytorch implement of the paper: Condor: Enhance LLM Alignment with Knowledge-Driven Data Synthesis and Refinement☆40May 28, 2025Updated last year
- [AAAI 2025] PAT: Pruning-Aware Tuning for Large Language Models☆38Feb 1, 2025Updated last year
- Llama3开源模型中文版-全方位测评,基于SuperCLUE基准 | Llama3 Chinese Evaluation with SuperCLUE☆16Apr 21, 2024Updated 2 years ago
- This repository includes the introduction to uncertain label in Chest X-Ray diagnosis.☆10Oct 20, 2024Updated last year
- ☆19Jun 29, 2025Updated last year
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- Efficient and Effective Weight-Ensembling Mixture of Experts for Multi-Task Model Merging. Arxiv, 2024.☆16Oct 28, 2024Updated last year
- Spitzers Architecture School Urban Lab for Unit 26. This repository explores designing and codifying urban systems from the bottom up in …☆14Mar 29, 2022Updated 4 years ago
- Source Code for our ICLR'26 paper☆17Feb 22, 2026Updated 6 months ago
- ☆12Jan 8, 2025Updated last year
- Dateset Reset Policy Optimization☆30Apr 12, 2024Updated 2 years ago
- [ACL 2026 Oral] SPPO: Sequence-Level PPO for Long-Horizon Reasoning Tasks official repos.☆28May 18, 2026Updated 3 months ago
- Accepted By The 39th Annual Conference on Neural Information Processing Systems Datasets and Benchmarks Track☆25Nov 17, 2025Updated 9 months ago