Decentralized deep learning in PyTorch. Built to train models on thousands of volunteers across the world.
β2,495Jan 11, 2026Updated 6 months ago
Alternatives and similar repositories for hivemind
Users that are interested in hivemind are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Official code for "Distributed Deep Learning in Open Collaborations" (NeurIPS 2021)β118Jan 13, 2022Updated 4 years ago
- πΈ Run LLMs at home, BitTorrent-style. Fine-tuning and inference up to 10x faster than offloadingβ10,300Sep 7, 2024Updated last year
- β398Jan 31, 2026Updated 5 months ago
- "Towards Crowdsourced Training of Large Neural Networks using Decentralized Mixture-of-Experts" (NeurIPS 2020), original PyTorch implemenβ¦β56Nov 5, 2020Updated 5 years ago
- Official code for "SWARM Parallelism: Training Large Models Can Be Surprisingly Communication-Efficient"β150Dec 11, 2023Updated 2 years ago
- Simple, predictable pricing with DigitalOcean hosting β’ AdAlways know what you'll pay with monthly caps and flat pricing. Enterprise-grade infrastructure trusted by 600k+ customers.
- Training a model similar to OpenAI DALL-E with volunteers from all over the Internet using hivemind and dalle-pytorch (NeurIPS 2021 demo)β27May 29, 2023Updated 3 years ago
- Memory-efficient transformer. Work in progress.β19Sep 17, 2022Updated 3 years ago
- PyTorch extensions for high performance and large scale training.β3,411Apr 26, 2025Updated last year
- "Moshpit SGD: Communication-Efficient Decentralized Training on Heterogeneous Unreliable Devices", official implementationβ30Feb 4, 2025Updated last year
- An implementation of model parallel autoregressive transformers on GPUs, based on the Megatron and DeepSpeed librariesβ7,444Jun 11, 2026Updated last month
- Accessible large language models via k-bit quantization for PyTorch.β8,333Updated this week
- Pretrain, finetune ANY AI model of ANY size on 1 or 10,000+ GPUs with zero code changes.β31,241Updated this week
- Efficient Deep Learning Systems course materials (HSE, YSDA)β1,015May 28, 2026Updated last month
- prime is a framework for efficient, globally distributed training of AI models over the internet.β851Updated this week
- Managed hosting for WordPress and PHP on Cloudways β’ AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- Supercharge Your Model Trainingβ5,487Apr 29, 2026Updated 2 months ago
- OpenDiLoCo: An Open-Source Framework for Globally Distributed Low-Communication Trainingβ582Updated this week
- A Smart, Automatic, Fast and Lightweight Web Scraper for Pythonβ7,657Jun 9, 2025Updated last year
- π A simple way to launch, train, and use PyTorch models on almost any device and distributed configuration, automatic mixed precision (iβ¦β9,785Updated this week
- Flexible and powerful tensor operations for readable and reliable code (for pytorch, jax, TF and others)β9,553Jul 5, 2026Updated 2 weeks ago
- A Python-level JIT compiler designed to make unmodified PyTorch programs faster.β1,078Apr 17, 2024Updated 2 years ago
- A repo for distributed training of language models with Reinforcement Learning via Human Feedback (RLHF)β4,753Jan 8, 2024Updated 2 years ago
- DeepSpeed is a deep learning optimization library that makes distributed training and inference easy, efficient, and effective.β42,752Updated this week
- Running large language models on a single GPU for throughput-oriented scenarios.β9,359Oct 28, 2024Updated last year
- Deploy to Railway using AI coding agents - Free Credits Offer β’ AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- Run, manage, and scale AI workloads on any AI infrastructure. Use one system to access & manage all AI compute (Kubernetes, Slurm, 20+ clβ¦β10,319Updated this week
- Development repository for the Triton language and compilerβ19,738Updated this week
- RWKV (pronounced RwaKuv) is an RNN with great LLM performance, which can also be directly trained like a GPT transformer (parallelizable)β¦β14,629Updated this week
- a libp2p-backed daemon wrapping the functionalities of go-libp2p for use in other languagesβ11Feb 9, 2025Updated last year
- Accelerated deep learning R&Dβ3,376Jul 8, 2026Updated last week
- Composable transformations of Python+NumPy programs: differentiate, vectorize, JIT to GPU/TPU, and moreβ36,025Updated this week
- Library for 8-bit optimizers and quantization routines.β779Aug 18, 2022Updated 3 years ago
- AITemplate is a Python framework which renders neural network into high performance CUDA/HIP C++ code. Specialized for FP16 TensorCore (Nβ¦β4,725Updated this week
- functorch is JAX-like composable function transforms for PyTorch.β1,434Aug 21, 2025Updated 11 months ago
- GPU virtual machines on DigitalOcean Gradient AI β’ AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- Fast and memory-efficient exact attentionβ24,497Updated this week
- A Data Streaming Library for Efficient Neural Network Trainingβ1,534Jun 25, 2026Updated 3 weeks ago
- A data augmentations library for audio, image, text, and video.β5,087Updated this week
- Incentivized Training over Wide Web with 1000x model compression.β22Oct 30, 2024Updated last year
- A minimal PyTorch re-implementation of the OpenAI GPT (Generative Pretrained Transformer) trainingβ24,721Aug 15, 2024Updated last year
- Type annotations and dynamic checking for a tensor's shape, dtype, names, etc.β1,484May 2, 2025Updated last year
- Go ahead and axolotl questionsβ12,219Updated this week