Prototype of OpenSHMEM for NVIDIA GPUs, developed as part of DoE Design Forward
☆24Apr 26, 2018Updated 8 years ago
Alternatives and similar repositories for df-nvshmem-prototype
Users that are interested in df-nvshmem-prototype are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Aries Network Performance Counters Monitoring Library☆11Nov 19, 2020Updated 5 years ago
- Guides and examples to help achieve optimal performance on a NVIDIA Grace CPU☆17Aug 9, 2024Updated 2 years ago
- GPUDirect Async implementation of HPGMG-FV CUDA☆11May 11, 2018Updated 8 years ago
- ☆14Jun 30, 2026Updated last month
- High Performance C++ Turbulent flow Lattice Boltzmann code☆17Sep 19, 2019Updated 6 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- Open Fabric Interfaces☆16Jul 16, 2020Updated 6 years ago
- Material point method proxy application based on Cabana.☆12Feb 19, 2025Updated last year
- CPE change log and release notes☆26Sep 3, 2024Updated last year
- ☆20Jan 17, 2024Updated 2 years ago
- MPI accelerator-integrated communication extensions☆39Apr 4, 2023Updated 3 years ago
- Please visit http://nrel.github.io/OpenWARP/ for more information.☆24Nov 18, 2016Updated 9 years ago
- ☆17Sep 15, 2021Updated 4 years ago
- Fortran 2003 wrappers for POSIX threads☆12Oct 13, 2017Updated 8 years ago
- A CUDA-based multi-GPU vertex-centric graph processing framework based on Warp Segmentation and Vertex Refinement techniques.☆12Mar 20, 2017Updated 9 years ago
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- This repository contains an implementation for Portals4. Portals4 is a Network Programming Interface which allows high-performance networ…☆14Sep 3, 2024Updated last year
- NCCL Fast Socket is a transport layer plugin to improve NCCL collective communication performance on Google Cloud.☆125Nov 15, 2023Updated 2 years ago
- An Open-Source Community Supported Fortran layer for AMD HIP☆10May 20, 2020Updated 6 years ago
- GPU implementation of classical molecular dynamics proxy application.☆31Jan 30, 2017Updated 9 years ago
- Pragmatic, Productive, and Portable Affinity for HPC☆52Updated this week
- Comb is a communication performance benchmarking tool.☆25Feb 27, 2023Updated 3 years ago
- Space-Time Variable Code Generator and Solver☆17Nov 2, 2023Updated 2 years ago
- https://eth-cscs.github.io/uenv/☆12Nov 18, 2024Updated last year
- A matlab toolkit to calculate numerical differentiation using WENO5 scheme. Mainly for level set simulation.☆10Jul 11, 2016Updated 10 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- Artifact of paper "Exploiting Recent SIMD Architectural Advances for Irregular Applications"☆11Jun 23, 2016Updated 10 years ago
- A simple pseudo-spectral solver for the Direct Numerical Simulation (DNS) of the 3D Taylor-Green Vortex in the Julia programming language☆11Jun 6, 2022Updated 4 years ago
- A Benchmark for Surface Reconstruction☆12Oct 5, 2015Updated 10 years ago
- C library containing high resolution timer implementation for several platforms.☆10Oct 20, 2020Updated 5 years ago
- MiniAMR Adaptive Mesh Refinement (AMR) Mini-App☆39Nov 12, 2024Updated last year
- Effective transpose on Hopper GPU☆29Sep 6, 2025Updated 11 months ago
- ☆10Feb 17, 2026Updated 5 months ago
- ☆15Apr 6, 2016Updated 10 years ago
- Variable-density incompressible Navier-Stokes code in F90☆13Nov 16, 2021Updated 4 years ago
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- 2-D inviscid flow and adjoint solver☆14Mar 8, 2014Updated 12 years ago
- MoSAIC: Modular system for Acceleration Integration MoSAIC☆10Jul 16, 2026Updated 3 weeks ago
- Backprop with Low-Precision Activations☆11Oct 28, 2019Updated 6 years ago
- RDMA and SHARP plugins for nccl library☆234Apr 3, 2026Updated 4 months ago
- CUDA Finite Difference Library☆16Aug 21, 2020Updated 5 years ago
- ☆14Apr 24, 2024Updated 2 years ago
- Header-only C++20 wrapper for MPI 4.0.☆16Oct 20, 2023Updated 2 years ago