Recipes for reproducing training and serving benchmarks for large machine learning models using GPUs on Google Cloud.
☆140Aug 25, 2026Updated this week
Alternatives and similar repositories for gpu-recipes
Users that are interested in gpu-recipes are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- ☆121Aug 22, 2026Updated last week
- ☆23Updated this week
- xpk (Accelerated Processing Kit, pronounced x-p-k,) is a software tool to help Cloud developers to orchestrate training jobs on accelerat…☆193Updated this week
- ☆37Oct 31, 2025Updated 9 months ago
- Cluster Toolkit is an open-source software offered by Google Cloud which makes it easy for customers to deploy AI/ML and HPC environments…☆360Updated this week
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- Package of Pathways-on-Cloud utilities☆33Updated this week
- This repository is a collection of accelerated platform best practices, reference architectures, example use cases, reference implementat…☆102Updated this week
- ☆49May 5, 2026Updated 3 months ago
- A simplified and automated orchestration workflow to perform ML end-to-end (E2E) model tests and benchmarking on Cloud VMs across differe…☆64Updated this week
- ☆16Mar 13, 2025Updated last year
- ☆383Updated this week
- ☆60Updated this week
- ☆22Mar 11, 2026Updated 5 months ago
- JetStream is a throughput and memory optimized engine for LLM inference on XLA devices, starting with TPUs (and GPUs in future -- PRs wel…☆456Jan 5, 2026Updated 7 months ago
- Managed Database hosting by DigitalOcean • AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- ☆40Aug 17, 2026Updated last week
- TPU inference for vLLM, with unified JAX and PyTorch support.☆418Updated this week
- Empowering LLM Agents for Real-World Computer System Optimization☆17Sep 10, 2025Updated 11 months ago
- NVIDIA Inference Benchmarks provide recipes in ready-to-use templates for evaluating platform speed. Validate your platform across speci…☆71Updated this week
- AI on GKE is a collection of examples, best-practices, and prebuilt solutions to help build, deploy, and scale AI Platforms on Google Kub…☆330Jun 23, 2025Updated last year
- Exemplar Performance provides recipes in ready-to-use templates for evaluating performance of specific AI use cases across hardware and s…☆106Updated this week
- PyTorch/XLA integration with JetStream (https://github.com/google/JetStream) for LLM inference"☆85Dec 18, 2025Updated 8 months ago
- This repository compiles code samples and notebooks demonstrating how to use Generative AI on Google Cloud Vertex AI.☆848Jan 6, 2026Updated 7 months ago
- Tokamax: A GPU and TPU kernel library.☆271Updated this week
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- ☆13Dec 6, 2022Updated 3 years ago
- ☆15May 11, 2025Updated last year
- ☆17Jan 23, 2026Updated 7 months ago
- ☆18Apr 8, 2026Updated 4 months ago
- Training NVIDIA NeMo Megatron Large Language Model (LLM) using NeMo Framework on Google Kubernetes Engine☆16Apr 28, 2025Updated last year
- ☆19Feb 18, 2026Updated 6 months ago
- This repository compiles prescriptive guidance and code samples demonstrating how to operationalize Google Research T5X framework on Goog…☆56Jan 21, 2026Updated 7 months ago
- GCP PCI-DSS 3.2.1 InSpec Profile☆18May 26, 2021Updated 5 years ago
- Benchmarking guide for the Azure AI Infrastructure.☆41Jun 9, 2026Updated 2 months ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- NVIDIA Resiliency Extension is a python package for framework developers and users to implement fault-tolerant features. It improves the …☆324Updated this week
- ☆13Jul 28, 2026Updated last month
- JAX backend for SGL☆346Updated this week
- A simple, performant, and scalable Jax LLM!☆2,407Updated this week
- CUDA Template Functions☆20Dec 16, 2025Updated 8 months ago
- ☆67Updated this week
- Orbax provides common checkpointing and persistence utilities for JAX users☆533Updated this week