Prototype of OpenSHMEM for NVIDIA GPUs, developed as part of DoE Design Forward
☆24Apr 26, 2018Updated 8 years ago
Alternatives and similar repositories for df-nvshmem-prototype
Users that are interested in df-nvshmem-prototype are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Aries Network Performance Counters Monitoring Library☆11Nov 19, 2020Updated 5 years ago
- Guides and examples to help achieve optimal performance on a NVIDIA Grace CPU☆18Aug 9, 2024Updated 2 years ago
- GPUDirect Async implementation of HPGMG-FV CUDA☆11May 11, 2018Updated 8 years ago
- ☆15Jun 30, 2026Updated 2 months ago
- High Performance C++ Turbulent flow Lattice Boltzmann code☆17Sep 19, 2019Updated 6 years ago
- Proton VPN Special Offer - Get 70% off • AdSpecial partner offer. Trusted by over 100 million users worldwide. Tested, Approved and Recommended by Experts.
- Open Fabric Interfaces☆16Jul 16, 2020Updated 6 years ago
- Material point method proxy application based on Cabana.☆12Feb 19, 2025Updated last year
- A Monte Carlo transport mini-app for studying new parallel algorithms☆19Jul 28, 2026Updated last month
- CPE change log and release notes☆26Sep 3, 2024Updated last year
- ☆20Jan 17, 2024Updated 2 years ago
- MPI accelerator-integrated communication extensions☆39Apr 4, 2023Updated 3 years ago
- Please visit http://nrel.github.io/OpenWARP/ for more information.☆24Nov 18, 2016Updated 9 years ago
- ☆17Sep 15, 2021Updated 4 years ago
- Fortran 2003 wrappers for POSIX threads☆12Oct 13, 2017Updated 8 years ago
- Managed Database hosting by DigitalOcean • AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- A CUDA-based multi-GPU vertex-centric graph processing framework based on Warp Segmentation and Vertex Refinement techniques.☆12Mar 20, 2017Updated 9 years ago
- NCCL Fast Socket is a transport layer plugin to improve NCCL collective communication performance on Google Cloud.☆125Nov 15, 2023Updated 2 years ago
- An Open-Source Community Supported Fortran layer for AMD HIP☆10May 20, 2020Updated 6 years ago
- GPU implementation of classical molecular dynamics proxy application.☆31Jan 30, 2017Updated 9 years ago
- Pragmatic, Productive, and Portable Affinity for HPC☆52Updated this week
- Comb is a communication performance benchmarking tool.☆25Feb 27, 2023Updated 3 years ago
- https://eth-cscs.github.io/uenv/☆13Nov 18, 2024Updated last year
- OpenGL based 3D engine with a viewer API for pointcloud, surface meshes☆14May 10, 2023Updated 3 years ago
- Flexible and accessible Python framework for cardiac electrophysiology simulations using finite-difference methods.☆13Aug 13, 2026Updated 2 weeks ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- A simple pseudo-spectral solver for the Direct Numerical Simulation (DNS) of the 3D Taylor-Green Vortex in the Julia programming language☆11Jun 6, 2022Updated 4 years ago
- A Benchmark for Surface Reconstruction☆12Oct 5, 2015Updated 10 years ago
- General interest repository for CSCS users☆52Feb 20, 2025Updated last year
- C library containing high resolution timer implementation for several platforms.☆10Oct 20, 2020Updated 5 years ago
- MiniAMR Adaptive Mesh Refinement (AMR) Mini-App☆39Nov 12, 2024Updated last year
- ☆10Feb 17, 2026Updated 6 months ago
- ☆15Apr 6, 2016Updated 10 years ago
- MoSAIC: Modular system for Acceleration Integration MoSAIC☆10Jul 16, 2026Updated last month
- 3D WENO 5th Order Code in Fortran☆10May 26, 2018Updated 8 years ago
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- Backprop with Low-Precision Activations☆11Oct 28, 2019Updated 6 years ago
- RDMA and SHARP plugins for nccl library☆231Aug 12, 2026Updated 2 weeks ago
- CUDA Finite Difference Library☆16Aug 21, 2020Updated 6 years ago
- ☆14Apr 24, 2024Updated 2 years ago
- Header-only C++20 wrapper for MPI 4.0.☆16Oct 20, 2023Updated 2 years ago
- A Distributed Multi-GPU System for Fast Graph Processing☆65Oct 25, 2018Updated 7 years ago
- A simple Fortran program of Discontinuous Galerkin method(no Limiter now) solving 2D Euler Equation with the Isentropic Vortex initial va…☆12Jul 2, 2022Updated 4 years ago