Decentralized deep learning in PyTorch. Built to train models on thousands of volunteers across the world.
β2,509Jan 11, 2026Updated 6 months ago
Alternatives and similar repositories for hivemind
Users that are interested in hivemind are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Official code for "Distributed Deep Learning in Open Collaborations" (NeurIPS 2021)β119Jan 13, 2022Updated 4 years ago
- πΈ Run LLMs at home, BitTorrent-style. Fine-tuning and inference up to 10x faster than offloadingβ10,483Sep 7, 2024Updated last year
- β398Jan 31, 2026Updated 6 months ago
- "Towards Crowdsourced Training of Large Neural Networks using Decentralized Mixture-of-Experts" (NeurIPS 2020), original PyTorch implemenβ¦β56Nov 5, 2020Updated 5 years ago
- Official code for "SWARM Parallelism: Training Large Models Can Be Surprisingly Communication-Efficient"β150Dec 11, 2023Updated 2 years ago
- Managed Kubernetes at scale on DigitalOcean β’ AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- Training a model similar to OpenAI DALL-E with volunteers from all over the Internet using hivemind and dalle-pytorch (NeurIPS 2021 demo)β27May 29, 2023Updated 3 years ago
- Memory-efficient transformer. Work in progress.β19Sep 17, 2022Updated 3 years ago
- PyTorch extensions for high performance and large scale training.β3,409Apr 26, 2025Updated last year
- "Moshpit SGD: Communication-Efficient Decentralized Training on Heterogeneous Unreliable Devices", official implementationβ30Feb 4, 2025Updated last year
- An implementation of model parallel autoregressive transformers on GPUs, based on the Megatron and DeepSpeed librariesβ7,450Jun 11, 2026Updated last month
- Accessible large language models via k-bit quantization for PyTorch.β8,403Updated this week
- Efficient Deep Learning Systems course materialsβ1,021May 28, 2026Updated 2 months ago
- Pretrain, finetune ANY AI model of ANY size on 1 or 10,000+ GPUs with zero code changes.β31,282Updated this week
- prime is a framework for efficient, globally distributed training of AI models over the internet.β854Jul 17, 2026Updated 3 weeks ago
- Serverless GPU API endpoints on Runpod - Get Bonus Credits β’ AdSkip the infrastructure headaches. Auto-scaling, pay-as-you-go, no-ops approach lets you focus on innovating your application.
- Supercharge Your Model Trainingβ5,493Apr 29, 2026Updated 3 months ago
- OpenDiLoCo: An Open-Source Framework for Globally Distributed Low-Communication Trainingβ586Jul 17, 2026Updated 3 weeks ago
- π A simple way to launch, train, and use PyTorch models on almost any device and distributed configuration, automatic mixed precision (iβ¦β9,807Aug 3, 2026Updated last week
- Flexible and powerful tensor operations for readable and reliable code (for pytorch, jax, TF and others)β9,569Jul 5, 2026Updated last month
- A Python-level JIT compiler designed to make unmodified PyTorch programs faster.β1,078Apr 17, 2024Updated 2 years ago
- Running large language models on a single GPU for throughput-oriented scenarios.β9,358Oct 28, 2024Updated last year
- DeepSpeed is a deep learning optimization library that makes distributed training and inference easy, efficient, and effective.β42,888Updated this week
- A repo for distributed training of language models with Reinforcement Learning via Human Feedback (RLHF)β4,752Jan 8, 2024Updated 2 years ago
- The AI Compute Platform for frontier teams. SkyPilot turns fragmented AI compute into one AI supercomputer, so frontier AI teams build cuβ¦β10,465Updated this week
- Simple, predictable pricing with DigitalOcean hosting β’ AdAlways know what you'll pay with monthly caps and flat pricing. Enterprise-grade infrastructure trusted by 600k+ customers.
- RWKV (pronounced RwaKuv) is an RNN with great LLM performance, which can also be directly trained like a GPT transformer (parallelizable)β¦β14,655Jul 23, 2026Updated 2 weeks ago
- Development repository for the Triton language and compilerβ19,908Updated this week
- a libp2p-backed daemon wrapping the functionalities of go-libp2p for use in other languagesβ11Feb 9, 2025Updated last year
- Accelerated deep learning R&Dβ3,380Jul 8, 2026Updated last month
- Composable transformations of Python+NumPy programs: differentiate, vectorize, JIT to GPU/TPU, and moreβ36,125Updated this week
- Library for 8-bit optimizers and quantization routines.β779Aug 18, 2022Updated 3 years ago
- AITemplate is a Python framework which renders neural network into high performance CUDA/HIP C++ code. Specialized for FP16 TensorCore (Nβ¦β4,725Updated this week
- functorch is JAX-like composable function transforms for PyTorch.β1,434Aug 21, 2025Updated 11 months ago
- Fast and memory-efficient exact attentionβ24,658Updated this week
- Deploy to Railway using AI coding agents - Free Credits Offer β’ AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- A Data Streaming Library for Efficient Neural Network Trainingβ1,543Jun 25, 2026Updated last month
- Incentivized Training over Wide Web with 1000x model compression.β22Oct 30, 2024Updated last year
- A data augmentations library for audio, image, text, and video.β5,087Jul 16, 2026Updated 3 weeks ago
- A minimal PyTorch re-implementation of the OpenAI GPT (Generative Pretrained Transformer) trainingβ24,780Aug 15, 2024Updated last year
- Go ahead and axolotl questionsβ12,331Updated this week
- Type annotations and dynamic checking for a tensor's shape, dtype, names, etc.β1,484May 2, 2025Updated last year
- Facebook AI Research Sequence-to-Sequence Toolkit written in Python.β32,244Sep 30, 2025Updated 10 months ago