[ICLR 2024] Dynamic Sparse Training with Structured Sparsity
☆26Apr 12, 2024Updated 2 years ago
Alternatives and similar repositories for condensed-sparsity
Users that are interested in condensed-sparsity are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- A compiler for decaf programming language written in python using Lark. Target language is MIPS.☆12Aug 10, 2020Updated 5 years ago
- Invariant representation learning from EEG within discriminative CNN architectures via adversarial censoring.☆14Feb 5, 2020Updated 6 years ago
- Gumbel-Softmax post-training quantization for LLMs (1–3 bit scalar, INT/GGUF-compatible).☆16Jul 11, 2026Updated 2 weeks ago
- Reproducing RigL (ICML 2020) as a part of ML Reproducibility Challenge 2020☆29Jan 6, 2022Updated 4 years ago
- Source code of paper ''KVSharer: Efficient Inference via Layer-Wise Dissimilar KV Cache Sharing''☆31Oct 24, 2024Updated last year
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- code for EMNLP 2024 paper: Interpreting Arithmetic Mechanism in Large Language Models through Comparative Neuron Analysis☆12Nov 17, 2024Updated last year
- [ICLR 2025] RaSA: Rank-Sharing Low-Rank Adaptation☆10May 19, 2025Updated last year
- HALO: Hadamard-Assisted Low-Precision Optimization and Training method for finetuning LLMs. 🚀 The official implementation of https://arx…☆31Feb 17, 2025Updated last year
- Official PyTorch implementation of CD-MOE☆12Mar 18, 2026Updated 4 months ago
- An extention of pytorch for low precision training / inference☆10Aug 28, 2023Updated 2 years ago
- (AAAI 2026) First-Order Error Matters: Accurate Compensation for Quantized Large Language Models☆16Apr 16, 2026Updated 3 months ago
- ZYN: Zero-Shot Reward Models with Yes-No Questions☆34Aug 15, 2023Updated 2 years ago
- Code to reproduce the experiments of the ICLR24-paper: "Sparse Model Soups: A Recipe for Improved Pruning via Model Averaging"☆12Oct 14, 2025Updated 9 months ago
- hardware (ASIC) DEFLATE designed for low-latency page-granularity memory compression and implemented in Chisel☆16Nov 15, 2024Updated last year
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- ☆16Jan 20, 2021Updated 5 years ago
- Using Test Vector Leakage Assessment methodology to analyze side-channel attacks on hardware implementations of AES-128☆13Oct 20, 2021Updated 4 years ago
- Exploration of automated dataset selection approaches at large scales.☆55Mar 4, 2025Updated last year
- ☆17Feb 3, 2022Updated 4 years ago
- ☆44Jan 25, 2024Updated 2 years ago
- The implementation for MLSys 2023 paper: "Cuttlefish: Low-rank Model Training without All The Tuning"☆44May 10, 2023Updated 3 years ago
- Side Channel script☆25Mar 1, 2023Updated 3 years ago
- Lightweight torch implementation of rigl, a sparse-to-sparse optimizer.☆60Jul 6, 2026Updated 2 weeks ago
- ☆16Mar 26, 2025Updated last year
- Deploy open-source AI quickly and easily - Special Bonus Offer • AdRunpod Hub is built for open source. One-click deployment and autoscaling endpoints without provisioning your own infrastructure.
- A block oriented training approach for inference time optimization.☆34Aug 19, 2024Updated last year
- A dock created with React, TypeScript, Electron and Flask to support SWE's with ADHD☆20Mar 24, 2026Updated 4 months ago
- Code for the paper "A Light Recipe to Train Robust Vision Transformers" [SaTML 2023]☆54Feb 6, 2023Updated 3 years ago
- Differentiable Logic Networks in PyTorch☆36Jul 15, 2026Updated last week
- PyTorch Implementation of the Sequential Multiagent Rollout algorithm☆11Jun 28, 2024Updated 2 years ago
- The perfectest coding lang☆28Dec 24, 2023Updated 2 years ago
- Include a dataset of the accepted papers in ACM FSE and IEEE/ACM ICSE during the last 5 years (2019-2023)☆16Nov 30, 2023Updated 2 years ago
- Code for "Routing Manifold Alignment Improves Generalization of Mixture-of-Experts LLMs"☆19Nov 6, 2025Updated 8 months ago
- Jax wrapper for autograd-differentiable functions☆13Oct 20, 2025Updated 9 months ago
- Bare Metal GPUs on DigitalOcean Gradient AI • AdPurpose-built for serious AI teams training foundational models, running large-scale inference, and pushing the boundaries of what's possible.
- Research code for "What to Pre-Train on? Efficient Intermediate Task Selection", EMNLP 2021☆37Dec 21, 2021Updated 4 years ago
- ☆22Jul 5, 2024Updated 2 years ago
- Factorized Neural Layers☆31Jul 11, 2023Updated 3 years ago
- Logic Reinforcement Learning☆21Oct 20, 2025Updated 9 months ago
- [ICML 2021] "Do We Actually Need Dense Over-Parameterization? In-Time Over-Parameterization in Sparse Training" by Shiwei Liu, Lu Yin, De…☆46Nov 11, 2023Updated 2 years ago
- ☆21Jun 22, 2022Updated 4 years ago
- The code for paper "MemGym: a Long-Horizon Memory Environment for LLM Agents".☆18Jun 2, 2026Updated last month