DeepSpeed is a deep learning optimization library that makes distributed training and inference easy, efficient, and effective.
☆14Jan 8, 2026Updated 8 months ago
Alternatives and similar repositories for DeepSpeed
Users that are interested in DeepSpeed are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Intel Gaudi's Megatron DeepSpeed Large Language Models for training☆18Dec 19, 2024Updated last year
- SynapseAI Core is a reference implementation of the SynapseAI API running on Habana Gaudi☆46Feb 3, 2025Updated last year
- A high-throughput and memory-efficient inference and serving engine for LLMs☆91Sep 2, 2026Updated 2 weeks ago
- Full End-to-End examples showing how to use First-gen Gaudi and Gaudi2 in common use cases☆13Dec 2, 2024Updated last year
- Reference models for Intel(R) Gaudi(R) AI Accelerator☆172Jan 8, 2026Updated 8 months ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- Large Language Model Text Generation Inference on Habana Gaudi☆34Mar 20, 2025Updated last year
- ☆15Mar 1, 2025Updated last year
- Intel® Extension for DeepSpeed* is an extension to DeepSpeed that brings feature support with SYCL kernels on Intel GPU(XPU) device. Note…☆65May 27, 2026Updated 3 months ago
- Demo on iGPU for FFmpeg decode and scale, OpenVINO inference. this is zero-copy solution, which means No frame data copy from CPU to iGPU…☆17Jan 25, 2023Updated 3 years ago
- ☆11Jun 29, 2021Updated 5 years ago
- Software kit for Qualcomm Cloud AI 100☆19Sep 3, 2026Updated 2 weeks ago
- Velocity And Luminance Adaptive Rasterization☆16Mar 31, 2023Updated 3 years ago
- LeetCode plugin code debuging template.☆14Apr 11, 2025Updated last year
- ☆65Mar 6, 2026Updated 6 months ago
- Serverless GPU API endpoints on Runpod - Get Bonus Credits • AdSkip the infrastructure headaches. Auto-scaling, pay-as-you-go, no-ops approach lets you focus on innovating your application.
- Pretrain, finetune and serve LLMs on Intel platforms with Ray☆129Sep 23, 2025Updated 11 months ago
- A Prot paper related materials☆11Sep 5, 2022Updated 4 years ago
- Computation using data flow graphs for scalable machine learning☆67Updated this week
- A repository of Dockerfiles, scripts, yaml files, Helm Charts, etc. used to build and scale the sample AI workflows with python, kubernet…☆12Feb 22, 2024Updated 2 years ago
- ☆10Aug 5, 2022Updated 4 years ago
- ☆19Jul 26, 2024Updated 2 years ago
- Combining deep neural networks with PCA and k-NN classification for abdominal organ recognition in ultrasound images.☆28Oct 12, 2021Updated 4 years ago
- A fork of the Linux kernel for p2pmem enabled devices like NVMe devices with CMBs, Microsemi NVRAM card (and other devices that can expos…☆29Aug 31, 2026Updated 2 weeks ago
- A conda-smithy repository for ambertools.☆12Jul 15, 2026Updated 2 months ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- OpenVINO LLM Benchmark☆11Dec 7, 2023Updated 2 years ago
- ☆19Updated this week
- ☆39Dec 22, 2025Updated 8 months ago
- Data Files for "Deep diversification of an AAV capsid protein by machine learning"☆18Mar 9, 2021Updated 5 years ago
- PArallelLOOPgEneratoR: Threaded Loops Code Generation Infrastructure targeting Tensor Contraction Applications such as GEMMs, Convolution…☆19Aug 4, 2026Updated last month
- Assets for AnyLabeling app☆14May 5, 2023Updated 3 years ago
- 1st to MICCAI DigestPath2019 challenge (https://digestpath2019.grand-challenge.org/Home/) on colonoscopy tissue segmentation and classifi…☆17Mar 25, 2021Updated 5 years ago
- Train deepseek r1-like reasoning LLM with ease | 轻松训练1个deepseek r1类的推理LLM☆21Feb 15, 2025Updated last year
- A Strong FuxiCTR Baseline for News CTR Challenge at RecSys 2024☆19Jul 13, 2024Updated 2 years ago
- Serverless GPU API endpoints on Runpod - Get Bonus Credits • AdSkip the infrastructure headaches. Auto-scaling, pay-as-you-go, no-ops approach lets you focus on innovating your application.
- A Gradio Web UI for running local LLM on Intel GPU (e.g., local PC with iGPU, discrete GPU such as Arc, Flex and Max) using IPEX-LLM.☆17Updated this week
- ☆19Apr 23, 2025Updated last year
- Intel® End-to-End AI Optimization Kit☆30Jul 18, 2024Updated 2 years ago
- Official implementation of NanoNet: Real-time medical Image segmentation architecture (IEEE CBMS)☆32Oct 17, 2023Updated 2 years ago
- ☆16Oct 27, 2024Updated last year
- The Exocore is an easily navigable personal hypertext database for text and images— a personal wiki which, over time, serves as a faithfu…☆18Aug 23, 2024Updated 2 years ago
- Intel® Tensor Processing Primitives extension for Pytorch*☆19Aug 24, 2026Updated 3 weeks ago