Official repository for ICML 2024 paper "MoRe Fine-Tuning with 10x Fewer Parameters"
☆22Oct 14, 2025Updated 11 months ago
Alternatives and similar repositories for sparse_matrix_fine_tuning
Users that are interested in sparse_matrix_fine_tuning are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Mini Model Daemon☆13Nov 9, 2024Updated last year
- Official code for the CVPR 2024 Paper "Can Biases in ImageNet Models Explain Generalization?".☆13Jun 24, 2024Updated 2 years ago
- Direct Preference Optimization for RWKV, aiming for RWKV-5 and 6.☆11Mar 1, 2024Updated 2 years ago
- Official Chinese documentation for RWKV | RWKV官方中文文档☆15Sep 4, 2026Updated 2 weeks ago
- ☆10Jan 26, 2025Updated last year
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- RWKV v5,v6 LoRA Trainer on Cuda and Rocm Platform. RWKV is a RNN with transformer-level LLM performance. It can be directly trained like …☆13Mar 24, 2024Updated 2 years ago
- ☆15Apr 11, 2024Updated 2 years ago
- Generate an FPGA design for a TWN☆11Nov 4, 2019Updated 6 years ago
- ☆12Jul 25, 2023Updated 3 years ago
- Synchronizing Claude Code conversations across machines☆17Aug 31, 2026Updated 3 weeks ago
- A program that allows you to chat on VRChat using ChatGPT.☆15Mar 22, 2023Updated 3 years ago
- ☆17Jan 1, 2025Updated last year
- A new DRAM substrate that mitigates the excessive energy consumption from both (i) transmitting unused data on the memory channel and (i…☆14Aug 23, 2024Updated 2 years ago
- Implementation for paper "BATMANN: A Binarized-All-Through Memory-Augmented Neural Network for Efficient In-Memory Computing"☆12Jan 12, 2022Updated 4 years ago
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- ☆21Sep 5, 2023Updated 3 years ago
- Compare how fine-tuned AI video models interpret the same prompts☆14Jan 29, 2025Updated last year
- Simulator or Non-Uniform Cache Architectures☆10Aug 27, 2018Updated 8 years ago
- ☆14Jan 21, 2024Updated 2 years ago
- OmniByteFormer is a generalized Transformer model that can process any type of data by converting it into byte sequences, bypassing tradi…☆17Updated this week
- ☆15Jan 31, 2021Updated 5 years ago
- ☆17Oct 25, 2022Updated 3 years ago
- A self-hosted version of WaterCrawl, a powerful web crawling and data extraction platform.☆13Jul 27, 2025Updated last year
- Memory Simulator and Optimizer☆22Oct 23, 2019Updated 6 years ago
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- ☆32May 26, 2024Updated 2 years ago
- Automata Benchmark Suite☆23Oct 23, 2023Updated 2 years ago
- ☆16Dec 11, 2024Updated last year
- a method for efficient large integer arithmetic in cryptography☆16Sep 16, 2025Updated last year
- Running massive simulations using RNNs on CPUs for building bots and all kinds of things.☆12Jun 13, 2021Updated 5 years ago
- RWKV centralised docs for the community☆35Jan 17, 2026Updated 8 months ago
- Converting Mixtral-8x7B to Mixtral-[1~7]x7B☆22Mar 4, 2024Updated 2 years ago
- ☆54Jul 18, 2024Updated 2 years ago
- Computer vision pipeline☆28Jul 25, 2024Updated 2 years ago
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- Papers and codes about Quantized Networks for easier survey and reference.☆20Dec 3, 2021Updated 4 years ago
- ☆41Apr 30, 2025Updated last year
- A framework and CLI toolkit for orchestrating teams of loosely-coupled AI agents.☆19Aug 9, 2026Updated last month
- Source code & scripts for experimental characterization and demonstration of 1) simultaneous many-row activation, 2) up to nine-input maj…☆18May 17, 2024Updated 2 years ago
- ☆23Sep 29, 2024Updated last year
- Deep Neural Network Optimization Platform with Gradient-based, Gradient-Free Algorithms☆11Jan 13, 2020Updated 6 years ago
- [NeurIPS 2025] Official PyTorch implementation of paper "Learning Grouped Lattice Vector Quantizers for Low-Bit LLM Compression".☆16Oct 24, 2025Updated 10 months ago