This repository contains code for the MicroAdam paper.
β21Dec 14, 2024Updated last year
Alternatives and similar repositories for MicroAdam
Users that are interested in MicroAdam are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Github Repo for OATS: Outlier-Aware Pruning through Sparse and Low Rank Decompositionβ20Apr 16, 2025Updated last year
- HALO: Hadamard-Assisted Low-Precision Optimization and Training method for finetuning LLMs. π The official implementation of https://arxβ¦β31Feb 17, 2025Updated last year
- Resources regarding evML (edge verified machine learning)β24Jan 4, 2025Updated last year
- SGLang Kernel Wheel Indexβ24Updated this week
- β27Aug 25, 2023Updated 2 years ago
- Managed Database hosting by DigitalOcean β’ AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- Code for "RSQ: Learning from Important Tokens Leads to Better Quantized LLMs"β23Mar 25, 2026Updated 3 months ago
- An extention to the GaLore paper, to perform Natural Gradient Descent in low rank subspaceβ19Oct 21, 2024Updated last year
- Privacy-focused metasearch engine with one-click setup, beautiful admin panel, and AI integration. Fork of SearXNG.β20Sep 7, 2025Updated 10 months ago
- [ICML2024 Spotlight] Fine-Tuning Pre-trained Large Language Models Sparselyβ24Jun 26, 2024Updated 2 years ago
- Repository for Sparse Finetuning of LLMs via modified version of the MosaicML llmfoundryβ43Jan 15, 2024Updated 2 years ago
- Official Code of The Combinatorial Brain Surgeon: Pruning Weights That Cancel One Another in Neural Networks[ICML2022]β16Sep 20, 2022Updated 3 years ago
- β14Nov 3, 2025Updated 8 months ago
- Code for NeurIPS 2024 Spotlight: "Scaling Laws and Compute-Optimal Training Beyond Fixed Training Durations"β93Oct 30, 2024Updated last year
- [ICDCS 2023] Evaluation and Optimization of Gradient Compression for Distributed Deep Learningβ10Apr 28, 2023Updated 3 years ago
- Virtual machines for every use case on DigitalOcean β’ AdGet dependable uptime with 99.99% SLA, simple security tools, and predictable monthly pricing with DigitalOcean's virtual machines, called Droplets.
- Physics Master is a model fine-tuned from llama3-8B-Instruct. It can answer your physics question!β16Aug 24, 2024Updated last year
- [ICML 2024] Official Implementation of SLEB: Streamlining LLMs through Redundancy Verification and Elimination of Transformer Blocksβ41Feb 4, 2025Updated last year
- Pytorch distributed backend extension with compression supportβ17Mar 24, 2025Updated last year
- Training with Block Minifloat number representationβ18May 2, 2021Updated 5 years ago
- β10Jun 19, 2023Updated 3 years ago
- β16Updated this week
- β33Nov 11, 2024Updated last year
- Boosting 4-bit inference kernels with 2:4 Sparsityβ96Sep 4, 2024Updated last year
- An implementation of the DISP-LLM method from the NeurIPS 2024 paper: Dimension-Independent Structural Pruning for Large Language Models.β24Aug 6, 2025Updated 11 months ago
- Deploy to Railway using AI coding agents - Free Credits Offer β’ AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- CLI tool designed to manage MCI (Model Context Interface) schemas and dynamically run MCP servers using defined MCI toolsetsβ16Nov 12, 2025Updated 8 months ago
- β60Jun 10, 2024Updated 2 years ago
- Landing repository for the paper "Softpick: No Attention Sink, No Massive Activations with Rectified Softmax"β91Sep 12, 2025Updated 10 months ago
- β17Apr 7, 2025Updated last year
- Artifact evaluation for HPCA'24 paper Lightening-Transformer: A Dynamically-operated Optically-interconnected Photonic Transformer Acceleβ¦β11Mar 3, 2024Updated 2 years ago
- Repository for AI model benchmarking on TT-Budaβ16Feb 9, 2026Updated 5 months ago
- QLoRA: Efficient Finetuning of Quantized LLMsβ11Jul 22, 2023Updated 2 years ago
- Code for ICML 2022 paper "SPDY: Accurate Pruning with Speedup Guarantees"β20May 3, 2023Updated 3 years ago
- Implementation of RankE: End-to-End Discrete Text-to-Image Post-Training via Rank-Consistent Alignmentβ20May 27, 2026Updated last month
- Deploy to Railway using AI coding agents - Free Credits Offer β’ AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- International Address formatter which considers the standard formatting rules of the countryβ14Nov 21, 2024Updated last year
- [ICLR'25] R-Sparse: Rank-Aware Activation Sparsity for Efficient LLM Inferenceβ21Apr 28, 2025Updated last year
- [ICML24] Official Implementation of "ETHER: Efficient Finetuning of Large-Scale Models with Hyperplane Reflections"β16May 31, 2024Updated 2 years ago
- new optimizerβ20Aug 4, 2024Updated last year
- Testing KAN-based text generation GPT modelsβ19May 6, 2024Updated 2 years ago
- Official github repo for "Fractional Reasoning via Latent Steering Vectors Improves Inference Time Compute"β17Jun 30, 2025Updated last year
- β15Sep 24, 2023Updated 2 years ago