Decentralized deep learning in PyTorch. Built to train models on thousands of volunteers across the world.
β2,516Jan 11, 2026Updated 7 months ago
Alternatives and similar repositories for hivemind
Users that are interested in hivemind are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Official code for "Distributed Deep Learning in Open Collaborations" (NeurIPS 2021)β119Jan 13, 2022Updated 4 years ago
- πΈ Run LLMs at home, BitTorrent-style. Fine-tuning and inference up to 10x faster than offloadingβ10,527Sep 7, 2024Updated last year
- β398Jan 31, 2026Updated 6 months ago
- "Towards Crowdsourced Training of Large Neural Networks using Decentralized Mixture-of-Experts" (NeurIPS 2020), original PyTorch implemenβ¦β56Nov 5, 2020Updated 5 years ago
- Official code for "SWARM Parallelism: Training Large Models Can Be Surprisingly Communication-Efficient"β150Dec 11, 2023Updated 2 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer β’ AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- Training a model similar to OpenAI DALL-E with volunteers from all over the Internet using hivemind and dalle-pytorch (NeurIPS 2021 demo)β27May 29, 2023Updated 3 years ago
- Memory-efficient transformer. Work in progress.β19Sep 17, 2022Updated 3 years ago
- PyTorch extensions for high performance and large scale training.β3,407Apr 26, 2025Updated last year
- "Moshpit SGD: Communication-Efficient Decentralized Training on Heterogeneous Unreliable Devices", official implementationβ30Feb 4, 2025Updated last year
- An implementation of model parallel autoregressive transformers on GPUs, based on the Megatron and DeepSpeed librariesβ7,462Jun 11, 2026Updated 2 months ago
- Accessible large language models via k-bit quantization for PyTorch.β8,446Updated this week
- Efficient Deep Learning Systems course materialsβ1,029May 28, 2026Updated 3 months ago
- Pretrain, finetune ANY AI model of ANY size on 1 or 10,000+ GPUs with zero code changes.β31,315Updated this week
- prime is a framework for efficient, globally distributed training of AI models over the internet.β855Jul 17, 2026Updated last month
- Deploy on Railway without the complexity - Free Credits Offer β’ AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- Supercharge Your Model Trainingβ5,496Apr 29, 2026Updated 4 months ago
- OpenDiLoCo: An Open-Source Framework for Globally Distributed Low-Communication Trainingβ591Jul 17, 2026Updated last month
- A Smart, Automatic, Fast and Lightweight Web Scraper for Pythonβ7,913Jul 29, 2026Updated last month
- π A simple way to launch, train, and use PyTorch models on almost any device and distributed configuration, automatic mixed precision (iβ¦β9,841Updated this week
- Flexible and powerful tensor operations for readable and reliable code (for pytorch, jax, TF and others)β9,582Updated this week
- A Python-level JIT compiler designed to make unmodified PyTorch programs faster.β1,078Apr 17, 2024Updated 2 years ago
- Running large language models on a single GPU for throughput-oriented scenarios.β9,352Oct 28, 2024Updated last year
- DeepSpeed is a deep learning optimization library that makes distributed training and inference easy, efficient, and effective.β43,019Updated this week
- A repo for distributed training of language models with Reinforcement Learning via Human Feedback (RLHF)β4,755Jan 8, 2024Updated 2 years ago
- Deploy open-source AI quickly and easily - Special Bonus Offer β’ AdRunpod Hub is built for open source. One-click deployment and autoscaling endpoints without provisioning your own infrastructure.
- The AI Compute Platform for frontier teams. SkyPilot turns fragmented AI compute into one AI supercomputer, so frontier AI teams build cuβ¦β10,534Updated this week
- Development repository for the Triton language and compilerβ20,037Updated this week
- RWKV (pronounced RwaKuv) is an RNN with great LLM performance, which can also be directly trained like a GPT transformer (parallelizable)β¦β14,687Updated this week
- a libp2p-backed daemon wrapping the functionalities of go-libp2p for use in other languagesβ11Feb 9, 2025Updated last year
- Accelerated deep learning R&Dβ3,382Jul 8, 2026Updated last month
- Composable transformations of Python+NumPy programs: differentiate, vectorize, JIT to GPU/TPU, and moreβ36,225Updated this week
- Library for 8-bit optimizers and quantization routines.β778Aug 18, 2022Updated 4 years ago
- AITemplate is a Python framework which renders neural network into high performance CUDA/HIP C++ code. Specialized for FP16 TensorCore (Nβ¦β4,725Aug 7, 2026Updated 3 weeks ago
- functorch is JAX-like composable function transforms for PyTorch.β1,434Aug 21, 2025Updated last year
- Wordpress hosting with auto-scaling - Free Trial Offer β’ AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- Fast and memory-efficient exact attentionβ24,801Updated this week
- A Data Streaming Library for Efficient Neural Network Trainingβ1,552Jun 25, 2026Updated 2 months ago
- Incentivized Training over Wide Web with 1000x model compression.β22Oct 30, 2024Updated last year
- A data augmentations library for audio, image, text, and video.β5,090Updated this week
- A minimal PyTorch re-implementation of the OpenAI GPT (Generative Pretrained Transformer) trainingβ24,842Aug 15, 2024Updated 2 years ago
- Go ahead and axolotl questionsβ12,421Updated this week
- Type annotations and dynamic checking for a tensor's shape, dtype, names, etc.β1,482May 2, 2025Updated last year