Tensorflow is a computational library using data flow graphs for scalable machine learning, and Tensorflow-RDMA is the implementation over RDMA, which can get about 4.5x speedup on two nodes comparing with TCP/IP.
☆59Nov 27, 2022Updated 3 years ago
Alternatives and similar repositories for Tensorflow-RDMA
Users that are interested in Tensorflow-RDMA are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- LITE Kernel RDMA Support for Datacenter Applications. SOSP 2017.☆112Jul 9, 2020Updated 6 years ago
- Distributed Shared Persistent Memory. SoCC 2017☆70Jul 20, 2020Updated 6 years ago
- An RDMA-enabled Distributed Persistent Memory File System☆163Oct 14, 2017Updated 8 years ago
- Frog is Asynchronous Graph Processing on GPU with Hybrid Coloring Model. The fundamental idea is based on Pareto principle (or 80-20 rule…☆36May 29, 2021Updated 5 years ago
- [Archived] A Fast Multi-tiered Distributed Storage System based on User-Level I/O☆75Mar 2, 2018Updated 8 years ago
- Deploy open-source AI quickly and easily - Special Bonus Offer • AdRunpod Hub is built for open source. One-click deployment and autoscaling endpoints without provisioning your own infrastructure.
- Source code for our OSDI 2016 paper☆109Nov 11, 2018Updated 7 years ago
- [FAST 2022] FORD: Fast One-sided RDMA-based Distributed Transactions for Disaggregated Persistent Memory☆62Jun 22, 2024Updated 2 years ago
- This is the Group-Meeting collections of HKUST System NetworkING (SING) Research Group.☆27Oct 3, 2019Updated 6 years ago
- A tutorial on RDMA based programming using code examples☆643Jan 3, 2020Updated 6 years ago
- ☆21Nov 29, 2022Updated 3 years ago
- This is the implementation repository of our SOSP'24 paper: Aceso: Achieving Efficient Fault Tolerance in Memory-Disaggregated Key-Value …☆24Oct 20, 2024Updated last year
- Reading seminar in Harvard Cloud Networking and Systems Group☆16Aug 29, 2022Updated 4 years ago
- A P4 implementation of a 5G UPF for BMv2☆15Oct 11, 2021Updated 4 years ago
- A framework to understand RDMA☆413Oct 12, 2023Updated 2 years ago
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- Mirror of Apache crail (Incubating)☆152Jul 3, 2022Updated 4 years ago
- ☆33Mar 31, 2021Updated 5 years ago
- Rcmp: Reconstructing RDMA-based Memory Disaggregation via CXL☆64Dec 26, 2023Updated 2 years ago
- ☆21Jan 2, 2023Updated 3 years ago
- Quartz: A DRAM-based performance emulator for NVM☆162Jul 28, 2019Updated 7 years ago
- RDMA core userspace libraries and daemons☆2,355Updated this week
- ☆70May 1, 2017Updated 9 years ago
- A scalable RAID system to aggregate performance and capacity of the next-generation storage.☆17Jan 3, 2024Updated 2 years ago
- A Write-friendly and Cache-optimized Hashing Scheme for Non-volatile Memory Systems (MSST 2017, TPDS 2018)☆30Apr 11, 2018Updated 8 years ago
- Managed Database hosting by DigitalOcean • AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- Gengar, a distributed shared hybrid memory pool with RDMA support. Gengar allows applications to access remote DRAM/NVM in a large and gl…☆24May 24, 2022Updated 4 years ago
- SmartNIC☆14Dec 13, 2018Updated 7 years ago
- A Progam-Behavior-Guided Far Memory System☆36Oct 26, 2023Updated 2 years ago
- Sample code from thegeekinthecorner.com☆280Sep 13, 2020Updated 5 years ago
- Simulator of a memory controller to connect DRAMSim and FlashDIMMSim into one unified memory☆18Apr 4, 2024Updated 2 years ago
- OFI Programmer's Guide☆52Dec 29, 2022Updated 3 years ago
- THC: Accelerating Distributed Deep Learning Using Tensor Homomorphic Compression☆20Jul 30, 2024Updated 2 years ago
- A new version for Pytheas (formally DDN), a control platform for enabling data-driven control for network applications☆14Nov 28, 2016Updated 9 years ago
- ICNP'25-THEMIS: Addressing Congestion-Induced Unfairness in Long-Haul RDMA Networks☆16Aug 22, 2026Updated last week
- Serverless GPU API endpoints on Runpod - Get Bonus Credits • AdSkip the infrastructure headaches. Auto-scaling, pay-as-you-go, no-ops approach lets you focus on innovating your application.
- ddl-benchmarks: Benchmarks for Distributed Deep Learning☆36May 29, 2020Updated 6 years ago
- GPU-accelerated LLM Training Simulator☆22Jun 26, 2025Updated last year
- CoRM: Compactable Remote Memory over RDMA☆20Jun 18, 2021Updated 5 years ago
- ☆25Aug 1, 2016Updated 10 years ago
- MQ-ECN NS2 Simulation☆11Feb 26, 2016Updated 10 years ago
- Mallacc: Accelerating Memory Allocation☆13Jan 2, 2018Updated 8 years ago
- Primo: Practical Learning-Augmented Systems with Interpretable Models☆19Dec 26, 2023Updated 2 years ago