Implementation of CoLA: Compute-Efficient Pre-Training of LLMs via Low-Rank Activation
☆26Feb 18, 2025Updated last year
Alternatives and similar repositories for CoLA
Users that are interested in CoLA are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- ☆20Feb 2, 2026Updated 5 months ago
- The official implementation of TinyTrain [ICML '24]☆27Jul 19, 2024Updated 2 years ago
- [NeurIPS 2021] Code for Unsupervised Learning of Compositional Energy Concepts☆60Sep 21, 2022Updated 3 years ago
- Apply CP, Tucker, TT/TR, HT to compress neural networks. Train from scratch.☆17Nov 26, 2020Updated 5 years ago
- Official repo for paper: Universal Reasoner: A Single, Composable Plug-and-Play Reasoner for Frozen LLMs☆20Nov 26, 2025Updated 7 months ago
- GPUs on demand by Runpod - Special Offer Available • AdRun AI, ML, and HPC workloads on powerful cloud GPUs—without limits or wasted spend. Deploy GPUs in under a minute and pay by the second.
- ☆11Feb 19, 2023Updated 3 years ago
- [EMNLP 2024] Quantize LLM to extremely low-bit, and finetune the quantized LLMs☆15Jul 18, 2024Updated 2 years ago
- [EMNLP 25] An effective and interpretable weight-editing method for mitigating overly short reasoning in LLMs, and a mechanistic study un…☆19Dec 17, 2025Updated 7 months ago
- ☆29May 13, 2025Updated last year
- My tests and experiments with some popular dl frameworks.☆17Sep 11, 2025Updated 10 months ago
- ☆19Nov 6, 2023Updated 2 years ago
- So, I trained a Llama a 130M architecture I coded from ground up to build a small instruct model from scratch. Trained on FineWeb dataset…☆18Mar 26, 2025Updated last year
- Codebase for ICML submission "DOGE: Domain Reweighting with Generalization Estimation"☆21Feb 29, 2024Updated 2 years ago
- Implementation of LaViC (KDD 2025)☆13Jun 1, 2025Updated last year
- Simple, predictable pricing with DigitalOcean hosting • AdAlways know what you'll pay with monthly caps and flat pricing. Enterprise-grade infrastructure trusted by 600k+ customers.
- [NAACL 24 Oral] LoRETTA: Low-Rank Economic Tensor-Train Adaptation for Ultra-Low-Parameter Fine-Tuning of Large Language Models☆39Jan 9, 2025Updated last year
- ☆14May 4, 2024Updated 2 years ago
- This is anonymous repository for submitting our work to a conference☆14Dec 17, 2024Updated last year
- Codebase for adaptive continual memory☆15Aug 15, 2023Updated 2 years ago
- ☆22Dec 23, 2024Updated last year
- ☆20Oct 13, 2024Updated last year
- VQE on photonic quantum devices☆19Updated this week
- Fast and memory-efficient exact attention☆21Updated this week
- [ICML 2024] SPP: Sparsity-Preserved Parameter-Efficient Fine-Tuning for Large Language Models☆22May 28, 2024Updated 2 years ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- Implementation of NAACL'25 "Empowering Retrieval-based Conversational Recommendation with Contrasting User Preferences"☆14Sep 9, 2025Updated 10 months ago
- The Stream-51 dataset for streaming classification and novelty detection from videos.☆17Feb 22, 2022Updated 4 years ago
- [ICML 2024] Sparse Model Inversion: Efficient Inversion of Vision Transformers with Less Hallucination☆14Apr 29, 2025Updated last year
- [CVPR 2020] A generative model with latent factors that are independent and localized.☆12Mar 27, 2025Updated last year
- pytorch implementation of multitask training of RNN☆17Jul 2, 2019Updated 7 years ago
- ☆12Nov 1, 2024Updated last year
- ☆19Jan 3, 2025Updated last year
- Source code for PECRS (EACL 2024)☆12Feb 3, 2024Updated 2 years ago
- ☆10Feb 22, 2023Updated 3 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- Official Implementation of ConceptLM.☆23Mar 18, 2026Updated 4 months ago
- The official source code for "Vision Language Model is NOT All You Need: Augmentation Strategies for Molecule Language Model".☆14Jul 23, 2024Updated last year
- MDM-Prime-v2: Binary Encoding and Index Shuffling Enable Scaling of Diffusion Language Models☆27May 23, 2026Updated last month
- The official source code for "Subgraph Federated Learning for Local Generalization (FedLoG)" at ICLR 2025 (Oral).☆17May 6, 2025Updated last year
- ☆15Jul 25, 2024Updated last year
- ☆11Apr 5, 2023Updated 3 years ago
- ☆27Jun 7, 2026Updated last month