GPUDirect example
☆63Oct 19, 2021Updated 4 years ago
Alternatives and similar repositories for gpudma
Users that are interested in gpudma are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Minimal HW-based demo of GPUDirect RDMA on NVIDIA Jetson AGX Xavier running L4T☆227Jul 15, 2024Updated 2 years ago
- GPUDirect Async support for IB Verbs☆138Nov 10, 2022Updated 3 years ago
- PCIe library for the Xilinx 7 series FPGAs in the Bluespec language☆83Mar 22, 2022Updated 4 years ago
- NVIDIA GPU direct RDMA using SISCI API☆18Apr 8, 2018Updated 8 years ago
- ☆400Apr 23, 2024Updated 2 years ago
- Deploy open-source AI quickly and easily - Special Bonus Offer • AdRunpod Hub is built for open source. One-click deployment and autoscaling endpoints without provisioning your own infrastructure.
- A fast GPU memory copy library based on NVIDIA GPUDirect RDMA technology☆1,408Jul 14, 2026Updated last month
- A user-space test platform for testing the p2pdma Linux kernel framework with NVMe CMBs and other PCIe BAR memory.☆64May 16, 2023Updated 3 years ago
- BlueField RSHIM driver☆21Aug 27, 2020Updated 5 years ago
- FROZEN: the master branch has merged with the libfabric git repo☆31Oct 3, 2018Updated 7 years ago
- GPUDirect Async suite☆16Dec 5, 2018Updated 7 years ago
- ☆239Mar 28, 2026Updated 4 months ago
- Magnum IO community repo☆121Jun 22, 2026Updated last month
- Linux extra (out of tree) kernel modules for ntrdma.☆25May 23, 2025Updated last year
- A (hacky) Linux kernel driver for PCI end points that implement p2pmem on the device.☆27Mar 18, 2022Updated 4 years ago
- GPUs on demand by Runpod - Special Offer Available • AdRun AI, ML, and HPC workloads on powerful cloud GPUs—without limits or wasted spend. Deploy GPUs in under a minute and pay by the second.
- Build userspace NVMe drivers and storage applications with CUDA support☆444Dec 18, 2023Updated 2 years ago
- NVIDIA GPUDirect Storage Driver☆377Jun 1, 2026Updated 2 months ago
- RDMA and SHARP plugins for nccl library☆234Updated this week
- This tutorial demonstrates how to use CUDA-Aware MPI☆41May 16, 2023Updated 3 years ago
- Donard: A PCIe Peer-2-Peer kernel patch and library that builds on top of NVM. Express. Also see https://github.com/sbates130272/linux-do…☆32Nov 17, 2016Updated 9 years ago
- Infiniband Verbs Performance Tests☆1,007Updated this week
- these are custom recipes of nvidia nsight system post collection analysis.☆16Nov 7, 2025Updated 9 months ago
- HLS project modeling various sparse accelerators.☆12Jan 11, 2022Updated 4 years ago
- GPU Affinity is a package to automatically set the CPU process affinity to match the hardware architecture on a given platform☆29Dec 8, 2023Updated 2 years ago
- Managed Database hosting by DigitalOcean • AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- LineFS: Efficient SmartNIC Offload of a Distributed File System with Pipeline Parallelism☆90Dec 24, 2021Updated 4 years ago
- This is a tool for managing GPU partitions for NVIDIA Fabric Manager’s Shared NVSwitch.☆18Jul 2, 2026Updated last month
- CUDAAdvisor: a GPU profiling tool☆53Aug 24, 2018Updated 7 years ago
- The Task-Aware MPI (TAMPI) library extends the functionality of standard MPI libraries by providing new mechanisms for improving the inte…☆27Jun 15, 2026Updated 2 months ago
- ☆20Oct 24, 2024Updated last year
- Unifies OS page cache for heterogeneous systems☆13Jul 26, 2019Updated 7 years ago
- Infiniband verbs performance tests (fork of git://git.openfabrics.org/~grockah/perftest.git)☆21Jan 14, 2016Updated 10 years ago
- The test of different distributed-training methods on High-Flyer AIHPC☆27Oct 18, 2022Updated 3 years ago
- minimal program to monitor and print cpu usage of a container☆12May 19, 2020Updated 6 years ago
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- ☆23Jun 21, 2023Updated 3 years ago
- ☆43Jun 13, 2025Updated last year
- Cross-platform Rust wrappers for the PCI ID Repository☆12Oct 3, 2025Updated 10 months ago
- Source code of "Accelerating Truss Decomposition on Heterogeneous Processors", accepted by VLDB'20 - By Yulin Che, Zhuohang Lai, Shixuan …☆16May 25, 2020Updated 6 years ago
- C++ Asio language bindings for OFI libfabric☆14Jul 16, 2021Updated 5 years ago
- PyTorch Memory Efficient Sparse Sparse Matrix Multiplication☆12Aug 12, 2024Updated 2 years ago
- NCCL Fast Socket is a transport layer plugin to improve NCCL collective communication performance on Google Cloud.☆125Nov 15, 2023Updated 2 years ago