The official code for Dropping Backward Propagation (DropBP)
☆32Oct 29, 2024Updated last year
Alternatives and similar repositories for dropbp
Users that are interested in dropbp are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- LISA: Layerwise Importance Sampling for Memory-Efficient Large Language Model Fine-Tuning☆39Apr 4, 2024Updated 2 years ago
- Official implementation of ECCV24 paper: POA☆24Aug 8, 2024Updated 2 years ago
- ☆22Dec 23, 2024Updated last year
- ☆22Dec 30, 2022Updated 3 years ago
- ☆11May 24, 2024Updated 2 years ago
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- Finetuning LLaMA with DeepSpeed☆10Apr 14, 2023Updated 3 years ago
- [EMNLP 2024] RoLoRA: Fine-tuning Rotated Outlier-free LLMs for Effective Weight-Activation Quantization☆40Sep 24, 2024Updated 2 years ago
- An awesome list that curates the best Flet tools, tutorials, blogs and more.☆10Jan 8, 2023Updated 3 years ago
- ☆13Nov 12, 2021Updated 4 years ago
- [EMNLP 2024] Source code for the paper "Learning Planning-based Reasoning with Trajectory Collection and Process Rewards Synthesizing".☆84Jan 14, 2025Updated last year
- [ICLR 2025] RaSA: Rank-Sharing Low-Rank Adaptation☆10May 19, 2025Updated last year
- [ACL 2024 Findings] Light-PEFT: Lightening Parameter-Efficient Fine-Tuning via Early Pruning☆13Sep 2, 2024Updated 2 years ago
- ☆18Apr 8, 2025Updated last year
- 🧮 Algebraic Positional Encodings.☆21Jun 5, 2026Updated 3 months ago
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- // clone this repo with --depth=1 to save disk size // toolchain compatible with Ubuntu 20.04+ //☆15Apr 28, 2022Updated 4 years ago
- Code for the paper "Knowledge-Aware Federated Active Learning with Non-IID Data", ICCV2023☆10Sep 8, 2023Updated 3 years ago
- Penn CIS 5650 (GPU Programming and Architecture) Final Project☆46Dec 11, 2023Updated 2 years ago
- Positional Skip-wise Training for Efficient Context Window Extension of LLMs to Extremely Length (ICLR 2024)☆208May 20, 2024Updated 2 years ago
- ☆16Apr 26, 2023Updated 3 years ago
- Unofficial implementation of the Ask-LLM paper 'How to Train Data-Efficient LLMs', arXiv:2402.09668.☆12Jun 19, 2024Updated 2 years ago
- ☆20Oct 25, 2022Updated 3 years ago
- Rookie's guide☆14Aug 10, 2024Updated 2 years ago
- My solution code to parallel architecture and programming Spring 2016☆12Aug 15, 2016Updated 10 years ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- EasyTTS是一个便捷的工具,旨在方便地使用第三方API服务来调用OpenAI的文本转语音(TTS)功能。 EasyTTS允许用户输入文本,并选择不同的模型、音色、格式来生成音频文件。☆10Nov 26, 2023Updated 2 years ago
- The ICDAR2015-TextSR dataset. This dataset was originally presented for the ICDAR2015 Competition on Text Image Super-Resolution. The rel…☆15Oct 8, 2019Updated 6 years ago
- ☆19Jan 3, 2025Updated last year
- Tensorflow implementation of TrialAttack (Triple Adversarial Learning for Influence based Poisoning Attack in Recommender Systems. KDD 20…☆11Sep 2, 2021Updated 5 years ago
- [ISCA'25] LIA: A Single-GPU LLM Inference Acceleration with Cooperative AMX-Enabled CPU-GPU Computation and CXL Offloading☆12Jun 28, 2025Updated last year
- ☆69Dec 3, 2024Updated last year
- Source code for the paper "Do Deep Neural Network Solutions form a Star Domain?"☆11May 26, 2024Updated 2 years ago
- nanoGPT using Equinox☆15Mar 3, 2023Updated 3 years ago
- Official implementation for the paper "Controlled Sparsity via Constrained Optimization"☆12Aug 10, 2022Updated 4 years ago
- Virtual machines for every use case on DigitalOcean • AdGet dependable uptime with 99.99% SLA, simple security tools, and predictable monthly pricing with DigitalOcean's virtual machines, called Droplets.
- A light-weight data management system for large-scale pretraining☆21May 17, 2025Updated last year
- Source code of FedAttack.☆11Feb 9, 2022Updated 4 years ago
- ☆14Jul 17, 2025Updated last year
- ☆16Feb 17, 2019Updated 7 years ago
- JAX Scalify: end-to-end scaled arithmetics☆18Oct 30, 2024Updated last year
- itertree python package - full featured tree data structure☆15Sep 8, 2025Updated last year
- Hierarchical Decomposition of Prompt-Based Continual Learning: Rethinking Obscured Sub-optimality (NeurIPS 2023, Spotlight)☆92Nov 15, 2024Updated last year