This repo includes everything you need to know about deploying GPU nodes on OCI
☆55Jul 24, 2026Updated this week
Alternatives and similar repositories for oci-hpc-oke
Users that are interested in oci-hpc-oke are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Terraform examples for deploying HPC clusters on OCI☆68May 26, 2026Updated 2 months ago
- GPU & cluster health and performance monitoring solution for OCI☆15Mar 12, 2026Updated 4 months ago
- NVIDIA NCCL Tests for Distributed Training☆150Updated this week
- Oracle Sharded database deployment automation and tools for use in client applications.☆38May 27, 2026Updated last month
- htop-like TUI for real-time RDMA network monitoring.☆77Updated this week
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- ☆11Sep 1, 2025Updated 10 months ago
- Clusterscope is a CLI and python library to extract information from HPC Clusters and Jobs.☆24Jul 14, 2026Updated last week
- Deploy, manage and monitor Gen AI workloads in minutes in your own tenancy and GPU infrastructure resources.☆55Apr 23, 2026Updated 3 months ago
- A toolkit for discovering cluster network topology.☆148Updated this week
- InfiniBand fabric monitoring daemon written in Go☆32May 22, 2025Updated last year
- ☆51Updated this week
- The Terraform OKE Module Installer for Oracle Cloud Infrastructure provides a Terraform module that provisions the necessary resources fo…☆202Jun 26, 2026Updated last month
- Hands on lab to learn WebLogic Operator (WebLogic on Kubernetes)☆19Oct 27, 2021Updated 4 years ago
- Python library for simple and complex indels.☆12Jan 22, 2024Updated 2 years ago
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- Where GPUs get cooked 👩🍳🔥☆402May 26, 2026Updated 2 months ago
- A command line utility to manage the configuration of a system's high performance network interfaces for RoCE deployments☆36Jul 25, 2023Updated 3 years ago
- slack bot written in go that processes DM's / app mentions and replies with chat-gpt's response☆14May 2, 2023Updated 3 years ago
- This repository contains the results and code for the MLPerf™ Training v4.0 benchmark.☆13Jun 11, 2024Updated 2 years ago
- ocifs provides a POSIX-compatible API wrapping Oracle Cloud Infrastructure's (OCI) Object Storage. ocifs is a python library that relies …☆23Nov 6, 2025Updated 8 months ago
- Guides and examples to help achieve optimal performance on a NVIDIA Grace CPU☆17Aug 9, 2024Updated last year
- RPerf: Accurate Latency Measurement Framework for RDMA☆15Apr 14, 2026Updated 3 months ago
- ☆14Mar 8, 2017Updated 9 years ago
- Process SMRT sequencing kinetic summary to predict regional methylation on large genome☆13Dec 4, 2018Updated 7 years ago
- End-to-end encrypted cloud storage - Proton Drive • AdSpecial offer: 40% Off Yearly / 80% Off First Month. Protect your most important files, photos, and documents from prying eyes.
- cotainr - a user space Apptainer/Singularity container builder.☆31Jul 9, 2026Updated 2 weeks ago
- A small POC using Caddy as a TLS-terminating MQTT proxy☆12Aug 31, 2022Updated 3 years ago
- The container images of PXE/HTTPBoot Server☆13Dec 26, 2023Updated 2 years ago
- A toolkit to design standard primers, multiplexed primers, and primers around SV's☆13Oct 29, 2022Updated 3 years ago
- monitor InfiniBand usage by job or host☆26Jan 11, 2012Updated 14 years ago
- pytorch code examples for measuring the performance of collective communication calls in AI workloads☆21Sep 18, 2025Updated 10 months ago
- A service-aware RoCE network monitoring system based on end- to-end probing.☆30Updated this week
- NVIDIA Network Operator☆357Updated this week
- Build a Slurm Cluster using SaltStack in virtual machines☆12Nov 26, 2018Updated 7 years ago
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- Run Slurm in Kubernetes☆404Updated this week
- HPC Monitoring Tool☆41Jun 19, 2026Updated last month
- Docker XPRA HTML5 Image with OpenGL support for NVIDIA cards☆13Oct 28, 2020Updated 5 years ago
- NVIDIA Fleet Intelligence Agent - Host agent for GPU telemetry collection and attestation☆44Updated this week
- Terraform module for production-ready HA Wireguard VPN server in AWS Amazon☆14Jul 19, 2021Updated 5 years ago
- SAML Authentication Plugin for Caddy v2☆16Sep 25, 2020Updated 5 years ago
- NVSentinel is a cross-platform fault remediation service designed to rapidly remediate runtime node-level issues in GPU-accelerated compu…☆350Updated this week