Fast GPU error-bounded lossy compressor for floating-point data.
☆68Jun 10, 2026Updated 2 months ago
Alternatives and similar repositories for cuSZp
Users that are interested in cuSZp are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- FZ-GPU: A Fast and High-Ratio Lossy Compressor for Scientific Data on GPUs☆15Jun 21, 2026Updated last month
- A GPU accelerated error-bounded lossy compression for scientific data.☆100Jul 2, 2026Updated last month
- ☆18Jun 12, 2026Updated last month
- DeepSZ: A Novel Framework to Compress Deep Neural Networks by Using Error-Bounded Lossy Compression☆12Oct 7, 2020Updated 5 years ago
- MGARD: MultiGrid Adaptive Reduction of Data☆49Jul 26, 2026Updated 2 weeks ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- Herald: Accelerating Neural Recommendation Training with Embedding Scheduling (NSDI 2024)☆23May 9, 2024Updated 2 years ago
- A GPU-based LZSS compression algorithm, highly tuned for NVIDIA GPGPUs and for streaming data, leveraging the respective strengths of CPU…☆38Dec 10, 2015Updated 10 years ago
- Optimized communication collectives for the Cerebras waferscale engine☆17Jun 5, 2024Updated 2 years ago
- HDF5 Cache VOL connector for caching data on fast storage layers and moving data asynchronously to the parallel file system to hide I/O o…☆22Feb 10, 2026Updated 6 months ago
- PIN-based Fault-Injector is a fault injector based on the Intel PIN tool. For more information, please refer to the following paper:☆18Jul 6, 2018Updated 8 years ago
- High Performance Sorting Based Distributed memory K-mer counter☆15Dec 8, 2025Updated 8 months ago
- First open-source KVTC implementation (NVIDIA, ICLR 2026) -- 8-32x KV cache compression via PCA + adaptive quantization + entropy coding☆21Apr 17, 2026Updated 3 months ago
- ☆37Apr 10, 2024Updated 2 years ago
- Python Package for Tensor Completion Algorithms☆35Nov 6, 2019Updated 6 years ago
- Open source password manager - Proton Pass • AdSecurely store, share, and autofill your credentials with Proton Pass, the end-to-end encrypted password manager trusted by millions.
- MLCommons Science benchmarking working group☆14Apr 17, 2026Updated 3 months ago
- [ICML‘2024] "LoCoCo: Dropping In Convolutions for Long Context Compression", Ruisi Cai, Yuandong Tian, Zhangyang Wang, Beidi Chen☆17Sep 7, 2024Updated last year
- This repo models rowhammer in gem5.☆10Updated this week
- Ultra fast MSD radix sorter☆10Jun 23, 2020Updated 6 years ago
- Artifacts of VLDB'22 paper "COMET: A Novel Memory-Efficient Deep Learning TrainingFramework by Using Error-Bounded Lossy Compression"☆10Aug 2, 2022Updated 4 years ago
- [ICML 2021] "Do We Actually Need Dense Over-Parameterization? In-Time Over-Parameterization in Sparse Training" by Shiwei Liu, Lu Yin, De…☆46Nov 11, 2023Updated 2 years ago
- ☆13Jul 27, 2026Updated 2 weeks ago
- ☆48Apr 27, 2026Updated 3 months ago
- The Atlas multi-GPU quantum circuit simulator.☆15Aug 17, 2024Updated last year
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- Oversubscription of GPU Memory through Transparent Swapping☆15Mar 27, 2015Updated 11 years ago
- docker:dind with NVIDIA GPU support via NVIDIA container toolkit☆14Aug 3, 2026Updated last week
- Ok-Topk is a scheme for distributed training with sparse gradients. Ok-Topk integrates a novel sparse allreduce algorithm (less than 6k c…☆27Dec 10, 2022Updated 3 years ago
- Minimal HTTP 1.1 Proxy & Load-balancer☆10Sep 23, 2017Updated 8 years ago
- ☆12Mar 26, 2024Updated 2 years ago
- Official implementation of "MaxK-GNN: Extremely Fast GPU Kernel Design for Accelerating Graph Neural Networks Training"☆42Mar 4, 2024Updated 2 years ago
- Enabling on-the-fly manipulations with LLVM IR code of CUDA sources☆124Apr 18, 2025Updated last year
- A 32-bit 5-stage RISC-V pipeline processor core with traps, S privilege mode, virtual memory, cache, branch prediction and TLB. Powered b…☆15Feb 14, 2024Updated 2 years ago
- ☆13Aug 25, 2021Updated 4 years ago
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- WIPE implementation☆12Nov 26, 2023Updated 2 years ago
- ☆15Jun 4, 2024Updated 2 years ago
- ☆14Nov 7, 2024Updated last year
- A branch predictor simulator in C++ that tests 6 different types of branch predictors.☆13Apr 26, 2018Updated 8 years ago
- ☆11Oct 11, 2023Updated 2 years ago
- OBsan: An Out-Of-Bound Sanitizer to Harden DNN Executables☆17Feb 28, 2023Updated 3 years ago
- Reproducing RigL (ICML 2020) as a part of ML Reproducibility Challenge 2020☆29Jan 6, 2022Updated 4 years ago