A lightweight parameter server interface
☆90Jan 13, 2023Updated 3 years ago
Alternatives and similar repositories for ps-lite
Users that are interested in ps-lite are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- A high performance and generic framework for distributed DNN training☆3,718Oct 3, 2023Updated 2 years ago
- A lightweight parameter server interface☆1,561Mar 2, 2026Updated 5 months ago
- ☆12Jun 25, 2021Updated 5 years ago
- This is an official GitHub repository for the paper, "Towards timeout-less transport in commodity datacenter networks.".☆17Oct 12, 2021Updated 4 years ago
- Examples of usage for Mellanox HW offloads☆17Jan 18, 2022Updated 4 years ago
- Managed Database hosting by DigitalOcean • AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- verbs profiling library☆22Sep 22, 2023Updated 2 years ago
- eRPC library for Rust☆14Jan 16, 2020Updated 6 years ago
- ☆78Sep 13, 2016Updated 9 years ago
- ☆26May 19, 2021Updated 5 years ago
- The prototype for NSDI paper "NetHint: White-Box Networking for Multi-Tenant Data Centers"☆26Feb 2, 2024Updated 2 years ago
- THC: Accelerating Distributed Deep Learning Using Tensor Homomorphic Compression☆20Jul 30, 2024Updated 2 years ago
- REDM是一套基于商业化标准的开源directui界面框架,不仅能提供完善的项目管理方案、详细的文档框架,也可轻松协助完成可视化界面设计,其核心库的稳定性已在内部多个大型商化项目中通过验证。 http://hgy413.com/3426.html☆14Aug 19, 2018Updated 7 years ago
- libvirttime provides transparent time virtualization, all in userspace.☆15Nov 18, 2021Updated 4 years ago
- Reading seminar in Harvard Cloud Networking and Systems Group☆16Aug 29, 2022Updated 3 years ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- gossip: Efficient Communication Primitives for Multi-GPU Systems☆62Jul 1, 2022Updated 4 years ago
- RDMA and SHARP plugins for nccl library☆234Apr 3, 2026Updated 4 months ago
- ☆22Apr 2, 2023Updated 3 years ago
- Code for reproducing experiments performed for Accoridon☆13Jun 11, 2021Updated 5 years ago
- MPI Library Memory Consumption Utilities☆20Jun 10, 2026Updated 2 months ago
- Tutorials to GPU programming. Reading notes.☆19Apr 27, 2023Updated 3 years ago
- Unified Communication X (mailing list - https://elist.ornl.gov/mailman/listinfo/ucx-group)☆1,690Updated this week
- This is the Group-Meeting collections of HKUST System NetworkING (SING) Research Group.☆27Oct 3, 2019Updated 6 years ago
- Enhanced networking support for TensorFlow. Maintained by SIG-networking.☆99Nov 19, 2021Updated 4 years ago
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- High performance NCCL plugin for Bagua.☆15Sep 15, 2021Updated 4 years ago
- Implementation based on OSDI paper☆20Feb 11, 2018Updated 8 years ago
- NVIDIA NVSHMEM is a parallel programming interface for NVIDIA GPUs based on OpenSHMEM. NVSHMEM can significantly reduce multi-process com…☆571Updated this week
- ☆10Updated this week
- [ICML 2026] Less Is More: Training-Free Sparse Attention with Global Locality for Efficient Reasoning☆35Sep 12, 2025Updated 10 months ago
- An experimental communicating attention kernel based on DeepEP.☆34Jul 29, 2025Updated last year
- Collective communications library with various primitives for multi-machine training.☆1,441Jul 31, 2026Updated last week
- NUMA-Aware Reader-Writer Locks☆19Jun 12, 2014Updated 12 years ago
- MQ-ECN NS2 Simulation☆11Feb 26, 2016Updated 10 years ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- sparse word2vec☆108Jul 7, 2022Updated 4 years ago
- My Paper Reading Lists and Notes.☆26May 8, 2026Updated 3 months ago
- ☆11Apr 3, 2023Updated 3 years ago
- A Tool for Automatic Parallelization of Deep Learning Training in Distributed Multi-GPU Environments.☆130Feb 21, 2022Updated 4 years ago
- Boost hardware utilization for ML training workloads via Inter-model Horizontal Fusion☆32May 15, 2024Updated 2 years ago
- Tiresias is a GPU cluster manager for distributed deep learning training.☆166May 7, 2020Updated 6 years ago
- ☆17May 10, 2024Updated 2 years ago