DeepSpeed is a deep learning optimization library that makes distributed training and inference easy, efficient, and effective.
☆14Jan 8, 2026Updated 9 months ago
Alternatives and similar repositories for DeepSpeed
Users that are interested in DeepSpeed are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- ☆16Jun 4, 2026Updated 4 months ago
- Intel Gaudi's Megatron DeepSpeed Large Language Models for training☆17Dec 19, 2024Updated last year
- SynapseAI Core is a reference implementation of the SynapseAI API running on Habana Gaudi☆46Feb 3, 2025Updated last year
- Easy and lightning fast training of 🤗 Transformers on Habana Gaudi processor (HPU)☆213Updated this week
- A high-throughput and memory-efficient inference and serving engine for LLMs☆91Sep 22, 2026Updated 2 weeks ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- Full End-to-End examples showing how to use First-gen Gaudi and Gaudi2 in common use cases☆13Dec 2, 2024Updated last year
- Reference models for Intel(R) Gaudi(R) AI Accelerator☆173Jan 8, 2026Updated 9 months ago
- ☆27Oct 9, 2025Updated last year
- ☆15Mar 1, 2025Updated last year
- Explainable AI Tooling (XAI). XAI is used to discover and explain a model's prediction in a way that is interpretable to the user. Releva…☆39Sep 22, 2025Updated last year
- ☆20Apr 9, 2019Updated 7 years ago
- ☆115Updated this week
- Software kit for Qualcomm Cloud AI 100☆19Sep 3, 2026Updated last month
- LeetCode plugin code debuging template.☆14Apr 11, 2025Updated last year
- Proton VPN Special Offer - Get 70% off • AdSpecial partner offer. Trusted by over 100 million users worldwide. Tested, Approved and Recommended by Experts.
- ☆28Jan 7, 2023Updated 3 years ago
- A repository of Dockerfiles, scripts, yaml files, Helm Charts, etc. used to build and scale the sample AI workflows with python, kubernet…☆12Feb 22, 2024Updated 2 years ago
- ☆14May 25, 2022Updated 4 years ago
- ☆19Jul 26, 2024Updated 2 years ago
- A fork of the Linux kernel for p2pmem enabled devices like NVMe devices with CMBs, Microsemi NVRAM card (and other devices that can expos…☆29Aug 31, 2026Updated last month
- ☆61Dec 18, 2024Updated last year
- SGLang kernel library for Intel XPU☆36Updated this week
- ☆197Updated this week
- ☆19Updated this week
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- ☆39Dec 22, 2025Updated 9 months ago
- oneAPI Level Zero Specification Headers and Loader☆337Updated this week
- ☆16Jun 12, 2024Updated 2 years ago
- Data Files for "Deep diversification of an AAV capsid protein by machine learning"☆18Mar 9, 2021Updated 5 years ago
- PArallelLOOPgEneratoR: Threaded Loops Code Generation Infrastructure targeting Tensor Contraction Applications such as GEMMs, Convolution…☆19Aug 4, 2026Updated 2 months ago
- Setup and Installation Instructions for Habana binaries, docker image creation☆28Aug 26, 2026Updated last month
- A Strong FuxiCTR Baseline for News CTR Challenge at RecSys 2024☆19Jul 13, 2024Updated 2 years ago
- A Gradio Web UI for running local LLM on Intel GPU (e.g., local PC with iGPU, discrete GPU such as Arc, Flex and Max) using IPEX-LLM.☆17Updated this week
- ☆19Apr 23, 2025Updated last year
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- Intel® End-to-End AI Optimization Kit☆30Jul 18, 2024Updated 2 years ago
- TextEmbed is a REST API crafted for high-throughput and low-latency embedding inference. It accommodates a wide variety of embedding mode…☆28Sep 5, 2024Updated 2 years ago
- Intel® Extension for MLIR. A staging ground for MLIR dialects and tools for Intel devices using the MLIR toolchain.☆157Updated this week
- Official implementation of NanoNet: Real-time medical Image segmentation architecture (IEEE CBMS)☆32Oct 17, 2023Updated 2 years ago
- ☆16Oct 27, 2024Updated last year
- Intel® Tensor Processing Primitives extension for Pytorch*☆20Aug 24, 2026Updated last month
- Fast and memory-efficient exact attention☆22Sep 10, 2026Updated 3 weeks ago