PyTorch/XLA integration with JetStream (https://github.com/google/JetStream) for LLM inference"
☆85Dec 18, 2025Updated 9 months ago
Alternatives and similar repositories for jetstream-pytorch
Users that are interested in jetstream-pytorch are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- JetStream is a throughput and memory optimized engine for LLM inference on XLA devices, starting with TPUs (and GPUs in future -- PRs wel…☆458Jan 5, 2026Updated 8 months ago
- torchprime is a reference model implementation for PyTorch on TPU.☆49Mar 3, 2026Updated 6 months ago
- Google TPU optimizations for transformers models☆136Jan 23, 2026Updated 7 months ago
- ☆19Feb 18, 2026Updated 7 months ago
- ☆34Jun 3, 2024Updated 2 years ago
- GPUs on demand by Runpod - Special Offer Available • AdRun AI, ML, and HPC workloads on powerful cloud GPUs—without limits or wasted spend. Deploy GPUs in under a minute and pay by the second.
- ☆127Sep 5, 2026Updated 2 weeks ago
- ☆41Sep 2, 2026Updated 2 weeks ago
- ☆22Apr 27, 2026Updated 4 months ago
- ☆15May 11, 2025Updated last year
- ☆17Jan 23, 2026Updated 7 months ago
- ☆22Mar 11, 2026Updated 6 months ago
- MLIR-based partitioning system☆210Updated this week
- ☆388Updated this week
- Go Wrappers for OpenXLA PJRT☆40Dec 12, 2025Updated 9 months ago
- GPUs on demand by Runpod - Special Offer Available • AdRun AI, ML, and HPC workloads on powerful cloud GPUs—without limits or wasted spend. Deploy GPUs in under a minute and pay by the second.
- xpk (Accelerated Processing Kit, pronounced x-p-k,) is a software tool to help Cloud developers to orchestrate training jobs on accelerat…☆193Aug 27, 2026Updated 3 weeks ago
- TPU inference for vLLM, with unified JAX and PyTorch support.☆435Updated this week
- A simplified and automated orchestration workflow to perform ML end-to-end (E2E) model tests and benchmarking on Cloud VMs across differe…☆64Updated this week
- torchax is a PyTorch frontend for JAX. It gives JAX the ability to author JAX programs using familiar PyTorch syntax. It also provides JA…☆245Updated this week
- A simple, performant, and scalable Jax LLM!☆2,421Updated this week
- Recipes for reproducing training and serving benchmarks for large machine learning models using GPUs on Google Cloud.☆140Sep 8, 2026Updated last week
- Latent Large Language Models☆19Aug 24, 2024Updated 2 years ago
- ☆55Apr 23, 2024Updated 2 years ago
- ☆16Mar 13, 2025Updated last year
- Managed Database hosting by DigitalOcean • AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- ☆156Updated this week
- ☆16Apr 10, 2022Updated 4 years ago
- ☆30Oct 26, 2024Updated last year
- Orbax provides common checkpointing and persistence utilities for JAX users☆533Updated this week
- JaxPP is a library for JAX that enables flexible MPMD pipeline parallelism for large-scale LLM training☆83Updated this week
- A collection of reusable, high-performance, well-documented, thorough-tested layers and models in Jax☆23Jun 8, 2025Updated last year
- ☆81Updated this week
- Fast and easy distributed model training examples.☆12Nov 26, 2024Updated last year
- Enabling PyTorch on XLA Devices (e.g. Google TPU)☆2,803May 27, 2026Updated 3 months ago
- Managed Database hosting by DigitalOcean • AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- PyTorch distributed training acceleration framework☆56Aug 13, 2025Updated last year
- JAX backend for SGL☆352Updated this week
- Automatic differentiation for Triton Kernels☆29Aug 12, 2025Updated last year
- Example of applying CUDA graphs to LLaMA-v2☆11Aug 25, 2023Updated 3 years ago
- Legible, Scalable, Reproducible Foundation Models with Named Tensors and Jax☆714Jan 26, 2026Updated 7 months ago
- Implementation of Direct Preference Optimization☆17Jul 17, 2023Updated 3 years ago
- Fork of Triton repository for OpenXLA uses of the Triton language and compiler☆17Feb 24, 2026Updated 6 months ago