GPU-accelerated LLM Training Simulator
☆22Jun 26, 2025Updated last year
Alternatives and similar repositories for multiverse
Users that are interested in multiverse are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- GPU-accelerated LLM Training Simulator☆54Jun 26, 2025Updated last year
- ☆25Oct 7, 2025Updated 10 months ago
- A Reliable and Resilient Collective Communication Library for NCCL and others☆18Aug 8, 2026Updated 3 weeks ago
- TACCL: Guiding Collective Algorithm Synthesis using Communication Sketches☆83Jul 25, 2023Updated 3 years ago
- TACOS: [T]opology-[A]ware [Co]llective Algorithm [S]ynthesizer for Distributed Machine Learning☆37Jun 13, 2025Updated last year
- Serverless GPU API endpoints on Runpod - Get Bonus Credits • AdSkip the infrastructure headaches. Auto-scaling, pay-as-you-go, no-ops approach lets you focus on innovating your application.
- ☆21Jun 29, 2022Updated 4 years ago
- ☆1,151Updated this week
- ☆47Dec 13, 2024Updated last year
- ☆22Nov 3, 2025Updated 9 months ago
- ☆24Sep 10, 2025Updated 11 months ago
- ICNP'25-THEMIS: Addressing Congestion-Induced Unfairness in Long-Haul RDMA Networks☆16Aug 22, 2026Updated last week
- Efficient GPU communication over multiple NICs.☆30Nov 20, 2025Updated 9 months ago
- [NSDI'22] Differential Network Analysis☆14Jun 2, 2022Updated 4 years ago
- Artifact of Chimera☆18May 6, 2025Updated last year
- Deploy open-source AI quickly and easily - Special Bonus Offer • AdRunpod Hub is built for open source. One-click deployment and autoscaling endpoints without provisioning your own infrastructure.
- An Automated Performance Optimization Framework for P4-Programmable SmartNICs☆28Nov 18, 2023Updated 2 years ago
- [NSDI 2023] TopoOpt: Optimizing the Network Topology for Distributed DNN Training☆43Sep 10, 2024Updated last year
- An external memory allocator example for PyTorch.☆16Aug 10, 2025Updated last year
- ☆16Mar 15, 2026Updated 5 months ago
- A fast and user-transparent parallel simulator implementation for ns-3☆110Nov 4, 2025Updated 9 months ago
- ☆16Mar 18, 2025Updated last year
- COCCL: Compression and precision co-aware collective communication library☆39Updated this week
- ☆56Aug 27, 2024Updated 2 years ago
- ☆25Jul 7, 2024Updated 2 years ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- GPGPU-Sim 中文注释版代码,包含 GPGPU-Sim 模拟器的最新版代码,经过中文注释,以帮助中文用户更好地理解和使用该模拟器。☆30Dec 18, 2024Updated last year
- A codebase for pretraining multi-billion-scale sparse GPTs.☆28Feb 9, 2026Updated 6 months ago
- ☆23Apr 2, 2023Updated 3 years ago
- blogs about Coimpiler & Virtual Machine☆12Jun 15, 2025Updated last year
- Venus Collective Communication Library, supported by SII and Infrawaves.☆151Jun 24, 2026Updated 2 months ago
- A Triton-only attention backend for vLLM☆28Jul 14, 2026Updated last month
- A lightweight design for computation-communication overlap.☆245Jan 20, 2026Updated 7 months ago
- The 3rd Iteration of the Berkeley RISC-V DMA Accelerator☆28Nov 21, 2019Updated 6 years ago
- Intel® SHMEM - Device initiated shared memory based communication library☆33Aug 5, 2026Updated 3 weeks ago
- Managed Database hosting by DigitalOcean • AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- 基于FPGA的PCIe 板卡,支持 离散量输入输出、ARINC429协议☆18Dec 1, 2022Updated 3 years ago
- ☆32Aug 21, 2021Updated 5 years ago
- An open platform for exploring scale-up network systems.☆22Mar 16, 2026Updated 5 months ago
- ☆24Apr 28, 2024Updated 2 years ago
- An experimental communicating attention kernel based on DeepEP.☆34Jul 29, 2025Updated last year
- LIBRA: Enabling Workload-aware Multi-dimensional Network Topology Optimization for Distributed Training of Large AI Models☆12May 7, 2024Updated 2 years ago
- ☆416Oct 4, 2023Updated 2 years ago