DeepSpeed is a deep learning optimization library that makes distributed training and inference easy, efficient, and effective.
☆14Jan 8, 2026Updated 7 months ago
Alternatives and similar repositories for DeepSpeed
Users that are interested in DeepSpeed are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- ☆16Jun 4, 2026Updated 2 months ago
- Intel Gaudi's Megatron DeepSpeed Large Language Models for training☆18Dec 19, 2024Updated last year
- SynapseAI Core is a reference implementation of the SynapseAI API running on Habana Gaudi☆46Feb 3, 2025Updated last year
- Easy and lightning fast training of 🤗 Transformers on Habana Gaudi processor (HPU)☆212Updated this week
- A high-throughput and memory-efficient inference and serving engine for LLMs☆90Updated this week
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- Full End-to-End examples showing how to use First-gen Gaudi and Gaudi2 in common use cases☆13Dec 2, 2024Updated last year
- Reference models for Intel(R) Gaudi(R) AI Accelerator☆172Jan 8, 2026Updated 7 months ago
- Large Language Model Text Generation Inference on Habana Gaudi☆34Mar 20, 2025Updated last year
- ☆26Oct 9, 2025Updated 10 months ago
- Intel® Extension for DeepSpeed* is an extension to DeepSpeed that brings feature support with SYCL kernels on Intel GPU(XPU) device. Note…☆65May 27, 2026Updated 3 months ago
- Explainable AI Tooling (XAI). XAI is used to discover and explain a model's prediction in a way that is interpretable to the user. Releva…☆39Sep 22, 2025Updated 11 months ago
- ☆20Apr 9, 2019Updated 7 years ago
- ☆18Updated this week
- ☆106Updated this week
- Deploy open-source AI quickly and easily - Special Bonus Offer • AdRunpod Hub is built for open source. One-click deployment and autoscaling endpoints without provisioning your own infrastructure.
- ☆15Mar 3, 2025Updated last year
- Software kit for Qualcomm Cloud AI 100☆19Dec 15, 2025Updated 8 months ago
- ☆61Mar 6, 2026Updated 5 months ago
- Tutorials for running models on First-gen Gaudi and Gaudi2 for Training and Inference. The source files for the tutorials on https://dev…☆65Sep 18, 2025Updated 11 months ago
- Computation using data flow graphs for scalable machine learning☆67Updated this week
- ☆20Mar 27, 2023Updated 3 years ago
- ☆19Jul 26, 2024Updated 2 years ago
- Combining deep neural networks with PCA and k-NN classification for abdominal organ recognition in ultrasound images.☆28Oct 12, 2021Updated 4 years ago
- OpenVINO™ optimization for PointPillars*☆32May 5, 2025Updated last year
- Bare Metal GPUs on DigitalOcean Gradient AI • AdPurpose-built for serious AI teams training foundational models, running large-scale inference, and pushing the boundaries of what's possible.
- ☆61Dec 18, 2024Updated last year
- ☆19Aug 20, 2026Updated last week
- ☆37Dec 22, 2025Updated 8 months ago
- oneAPI Level Zero Specification Headers and Loader☆333Updated this week
- PArallelLOOPgEneratoR: Threaded Loops Code Generation Infrastructure targeting Tensor Contraction Applications such as GEMMs, Convolution…☆19Aug 4, 2026Updated 3 weeks ago
- Setup and Installation Instructions for Habana binaries, docker image creation☆28Updated this week
- A Strong FuxiCTR Baseline for News CTR Challenge at RecSys 2024☆19Jul 13, 2024Updated 2 years ago
- A Gradio Web UI for running local LLM on Intel GPU (e.g., local PC with iGPU, discrete GPU such as Arc, Flex and Max) using IPEX-LLM.☆17Updated this week
- Intel® End-to-End AI Optimization Kit☆30Jul 18, 2024Updated 2 years ago
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- TextEmbed is a REST API crafted for high-throughput and low-latency embedding inference. It accommodates a wide variety of embedding mode…☆28Sep 5, 2024Updated last year
- Official implementation of NanoNet: Real-time medical Image segmentation architecture (IEEE CBMS)☆32Oct 17, 2023Updated 2 years ago
- The Intel® Automated Self-Checkout Reference Package provides critical components required to build and deploy a self-checkout use case u…☆34Jul 14, 2026Updated last month
- Deep Learning for End-to-End Kidney Cancer Diagnosis on Multi-Phase Abdominal Computed Tomography☆23Dec 13, 2023Updated 2 years ago
- Intel® Tensor Processing Primitives extension for Pytorch*☆19Updated this week
- Fast and memory-efficient exact attention☆21Updated this week
- Make AI-generated UI actually look designed.☆27Aug 19, 2026Updated last week