COCCL: Compression and precision co-aware collective communication library
☆39Sep 2, 2026Updated this week
Alternatives and similar repositories for COCCL
Users that are interested in COCCL are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- FZ-GPU: A Fast and High-Ratio Lossy Compressor for Scientific Data on GPUs☆15Jun 21, 2026Updated 2 months ago
- HDF5 Cache VOL connector for caching data on fast storage layers and moving data asynchronously to the parallel file system to hide I/O o…☆22Feb 10, 2026Updated 6 months ago
- GPULZ: Optimizing LZSS Lossless Compression for Multi-byte Data on Modern GPUs☆16Apr 18, 2025Updated last year
- Service-aware KV-cache compression for bandwidth-efficient disaggregated LLM serving.☆24Updated this week
- ☆18Jun 12, 2026Updated 2 months ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- Heterogeneous Accelerator Memory Resource☆14Nov 2, 2023Updated 2 years ago
- DeepSZ: A Novel Framework to Compress Deep Neural Networks by Using Error-Bounded Lossy Compression☆12Oct 7, 2020Updated 5 years ago
- official implementation of paper SDP4Bit: Toward 4-bit Communication Quantization in Sharded Data Parallelism for LLM Training☆44Dec 11, 2024Updated last year
- TACOS: [T]opology-[A]ware [Co]llective Algorithm [S]ynthesizer for Distributed Machine Learning☆37Jun 13, 2025Updated last year
- [NSDI25] AutoCCL: Automated Collective Communication Tuning for Accelerating Distributed and Parallel DNN Training☆35May 2, 2025Updated last year
- A GPU accelerated error-bounded lossy compression for scientific data.☆100Aug 26, 2026Updated last week
- a library to characterize the data and check the compression results of lossy compressors☆20Aug 31, 2025Updated last year
- Tutorials for Timemory☆21Aug 1, 2024Updated 2 years ago
- ☆17Nov 26, 2023Updated 2 years ago
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- Online Anomaly Detection for HPC Performance Data☆11Jun 25, 2018Updated 8 years ago
- GPU-accelerated LLM Training Simulator☆22Jun 26, 2025Updated last year
- SParse AcceleRation on Tensor Architecture☆18Apr 15, 2026Updated 4 months ago
- High Performance Linpack for Next-Generation AMD HPC Accelerators☆73Apr 21, 2026Updated 4 months ago
- High Performance Sorting Based Distributed memory K-mer counter☆15Dec 8, 2025Updated 8 months ago
- Material for the SC22 Deep Learning at Scale Tutorial☆41Jul 14, 2023Updated 3 years ago
- NCCL Profiling Kit☆156Jul 1, 2024Updated 2 years ago
- NCCL Examples from Official NVIDIA NCCL Developer Guide.☆21May 29, 2018Updated 8 years ago
- ☆56Aug 27, 2024Updated 2 years ago
- Open source password manager - Proton Pass • AdSecurely store, share, and autofill your credentials with Proton Pass, the end-to-end encrypted password manager trusted by millions.
- Logger for MPI communication☆28Jul 12, 2023Updated 3 years ago
- An Automated Performance Optimization Framework for P4-Programmable SmartNICs☆28Nov 18, 2023Updated 2 years ago
- A searchable Python interface to the SuiteSparse Matrix Collection☆60Apr 6, 2022Updated 4 years ago
- [CF ’20] Verified Instruction-Level Energy Consumption Measurement for NVIDIA GPUs☆15Dec 11, 2020Updated 5 years ago
- A library to abstract between different lossless and lossy compressors☆42Aug 7, 2026Updated 3 weeks ago
- Drishti provides I/O insights to help you improve your application's I/O performance.☆26Mar 3, 2026Updated 5 months ago
- LaunchMON is a software infrastructure that enables HPC run-time tools to co-locate tool daemons with a parallel job. Its API allows a to…☆13Aug 17, 2026Updated 2 weeks ago
- Sparse Boolean linear algebra for Nvidia Cuda, OpenCL and CPU computations☆16Aug 19, 2022Updated 4 years ago
- Directed Acyclic Graph Execution Engine (DAGEE) is a C++ library that enables programmers to express computation and data movement, as ta…☆49Oct 12, 2021Updated 4 years ago
- Managed Database hosting by DigitalOcean • AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- [DEPRECATED] Moved to ROCm/rocm-systems repo☆420Updated this week
- Third version of larcv. This is a complete replacement for larcv2.☆11Jun 24, 2024Updated 2 years ago
- A Symbolic Emulator for Shuffle Synthesis on the NVIDIA PTX Code☆16Mar 19, 2023Updated 3 years ago
- Library for exact linear algebra, a C++ template-library based originally on LinBox intended for F4-like implementations☆18Dec 15, 2012Updated 13 years ago
- Benchmark implementation of CosmoFlow in TensorFlow Keras☆22Feb 7, 2024Updated 2 years ago
- Artifacts of VLDB'22 paper "COMET: A Novel Memory-Efficient Deep Learning TrainingFramework by Using Error-Bounded Lossy Compression"☆10Aug 2, 2022Updated 4 years ago
- ☆47Dec 13, 2024Updated last year