Tune efficiently any LLM model from HuggingFace using distributed training (multiple GPU) and DeepSpeed. Uses Ray AIR to orchestrate the training on multiple AWS GPU instances
☆60Jun 20, 2023Updated 3 years ago
Alternatives and similar repositories for LLM-distributed-finetune
Users that are interested in LLM-distributed-finetune are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Tracking Ray Enhancement Proposals☆68Updated this week
- Ray - A curated list of resources: https://github.com/ray-project/ray☆83Jun 18, 2026Updated 2 months ago
- RayLLM - LLMs on Ray (Archived). Read README for more info.☆1,261Mar 13, 2025Updated last year
- Some microbenchmarks and design docs before commencement☆11Feb 1, 2021Updated 5 years ago
- Distributed XGBoost on Ray☆153Jun 25, 2024Updated 2 years ago
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- ☆12Mar 16, 2026Updated 5 months ago
- ☆11Apr 5, 2021Updated 5 years ago
- End-to-End Learning from Complex Multigraphs with Latent-Graph Convolutional Networks☆15Jul 25, 2024Updated 2 years ago
- Linear-chain LSTM-CRFs and Convolutional CRFs in PyTorch.☆22Aug 11, 2017Updated 9 years ago
- [ACL'21 Findings] Why Machine Reading Comprehension Models Learn Shortcuts?☆16Aug 8, 2023Updated 3 years ago
- [SIGMOD'25] Source code for the paper: Debunking the Myth of Join Ordering: Toward Robust SQL Analytics☆22Sep 4, 2025Updated 11 months ago
- ☆12Apr 30, 2024Updated 2 years ago
- Pygloo provides Python bindings for Gloo.☆22Jul 7, 2025Updated last year
- Example of applying CUDA graphs to LLaMA-v2☆11Aug 25, 2023Updated 2 years ago
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- Standalone Flash Attention v2 kernel without libtorch dependency☆113Sep 10, 2024Updated last year
- Fast GPU based tensor core reductions☆12Jan 13, 2023Updated 3 years ago
- ☆15Nov 24, 2018Updated 7 years ago
- SeqGAN implementation with Tensorflow☆18Jan 14, 2018Updated 8 years ago
- ☆11May 4, 2022Updated 4 years ago
- Flyte Backend Plugins contributed by the Flyte community.☆29Oct 9, 2023Updated 2 years ago
- socket program to send data with encryption☆13Jun 1, 2021Updated 5 years ago
- WIP. Veloce is a low-code Ray-based parallelization library that makes machine learning computation novel, efficient, and heterogeneous.☆17Aug 4, 2022Updated 4 years ago
- Federated Learning - PyTorch☆15Jun 27, 2021Updated 5 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- Solana Airdrop Faucet: A simple web application that allows users to receive free SOL tokens on the Solana Devnet. Built with Next.js, th…☆11Sep 22, 2024Updated last year
- CUDA and Triton implementations of Flash Attention with SoftmaxN.☆75May 26, 2024Updated 2 years ago
- crumbling large graphs into connected components☆12Jan 8, 2018Updated 8 years ago
- ☆17May 8, 2020Updated 6 years ago
- A Rust-based toolkit for building and packaging DuckDB extensions without Python dependencies.☆16Jul 21, 2026Updated 3 weeks ago
- ☆11Feb 9, 2024Updated 2 years ago
- Implementation of a simple client-server program which will facilitate a client to register itself to the server. It is implemented in py…☆17May 29, 2020Updated 6 years ago
- ☆10Nov 5, 2016Updated 9 years ago
- Distinguishing between anime and hentai☆16Jan 29, 2017Updated 9 years ago
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- ☆23Updated this week
- Piper based VoiceDock TTS implementation☆11Aug 12, 2023Updated 3 years ago
- To mitigate position bias in LLMs, especially in long-context scenarios, we scale only one dimension of LLMs, reducing position bias and …☆12Jun 18, 2024Updated 2 years ago
- 🧪 Experimental nodes for MLOPs☆12Aug 3, 2023Updated 3 years ago
- A distributed key-value store based on the implementation of Raft consensus algorithm following the paper, which applies multithreading t…☆19Nov 30, 2019Updated 6 years ago
- Deep Point Process by PyTorch☆22Nov 10, 2019Updated 6 years ago
- The open-source project for "Mandheling: Mixed-Precision On-Device DNN Training with DSP Offloading"[MobiCom'2022]☆20Aug 4, 2022Updated 4 years ago