LiBai(李白): A Toolbox for Large-Scale Distributed Parallel Training
☆403Jul 31, 2025Updated 11 months ago
Alternatives and similar repositories for libai
Users that are interested in libai are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Models and examples built with OneFlow☆100Oct 16, 2024Updated last year
- Datasets, Transforms and Models specific to Computer Vision☆91Nov 17, 2023Updated 2 years ago
- A toolkit for developers to simplify the transformation of nn.Module instances. It's now corresponding to Pytorch.fx.☆13Apr 7, 2023Updated 3 years ago
- OneFlow->ONNX☆42Apr 19, 2023Updated 3 years ago
- Deep Learning ❤️ OneFlow☆19Aug 26, 2021Updated 4 years ago
- Managed Database hosting by DigitalOcean • AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- ☆24Apr 25, 2023Updated 3 years ago
- DeepLearning Framework Performance Profiling Toolkit☆292Mar 28, 2022Updated 4 years ago
- OneFlow is a deep learning framework designed to be user-friendly, scalable and efficient.☆9,410Dec 4, 2025Updated 7 months ago
- OneFlow models for benchmarking.☆103Aug 7, 2024Updated last year
- OneFlow Serving☆20Apr 10, 2025Updated last year
- A more efficient yolov5 with oneflow backend 🎉🎉🎉☆216Jul 10, 2025Updated last year
- ☆17Jan 1, 2024Updated 2 years ago
- oneflow documentation☆69Jun 26, 2024Updated 2 years ago
- Easy Parallel Library (EPL) is a general and efficient deep learning framework for distributed model training.☆272Mar 31, 2023Updated 3 years ago
- Proton VPN Special Offer - Get 70% off • AdSpecial partner offer. Trusted by over 100 million users worldwide. Tested, Approved and Recommended by Experts.
- ☆11Dec 26, 2025Updated 6 months ago
- Automatically Discovering Fast Parallelization Strategies for Distributed Deep Neural Network Training☆1,896Jul 1, 2026Updated 2 weeks ago
- Transformer related optimization, including BERT, GPT☆6,442Mar 27, 2024Updated 2 years ago
- LightSeq: A High Performance Library for Sequence Processing and Generation☆3,296May 16, 2023Updated 3 years ago
- Ongoing research training transformer models at scale☆17,125Updated this week
- OneDiff: An out-of-the-box acceleration library for diffusion models.☆1,964Dec 4, 2025Updated 7 months ago
- ☆12Mar 13, 2023Updated 3 years ago
- TVMScript kernel for deformable attention☆25Dec 15, 2021Updated 4 years ago
- ☆12Aug 10, 2022Updated 3 years ago
- Bare Metal GPUs on DigitalOcean Gradient AI • AdPurpose-built for serious AI teams training foundational models, running large-scale inference, and pushing the boundaries of what's possible.
- optimized BERT transformer inference on NVIDIA GPU. https://arxiv.org/abs/2210.03052☆479Mar 15, 2024Updated 2 years ago
- Running BERT without Padding☆479Mar 18, 2022Updated 4 years ago
- auto deploy neovim like chxuan/vimplus☆12Apr 22, 2025Updated last year
- A Python-level JIT compiler designed to make unmodified PyTorch programs faster.☆1,078Apr 17, 2024Updated 2 years ago
- Ongoing research training transformer language models at scale, including: BERT & GPT-2☆2,257Aug 14, 2025Updated 11 months ago
- A high-throughput and memory-efficient inference and serving engine for LLMs☆17Jun 3, 2024Updated 2 years ago
- A flexible and efficient deep neural network (DNN) compiler that generates high-performance executable from a DNN model description.☆1,002Sep 19, 2024Updated last year
- A library for accelerating Transformer models on NVIDIA GPUs, including using 8-bit and 4-bit floating point (FP8 and FP4) precision on H…☆3,435Updated this week
- ☆144Jan 30, 2025Updated last year
- Simple, predictable pricing with DigitalOcean hosting • AdAlways know what you'll pay with monthly caps and flat pricing. Enterprise-grade infrastructure trusted by 600k+ customers.
- ☆221Aug 17, 2023Updated 2 years ago
- Training and serving large-scale neural networks with auto parallelization.☆3,178Dec 9, 2023Updated 2 years ago
- A more efficient GLM implementation!☆54Feb 18, 2023Updated 3 years ago
- ☆79May 4, 2021Updated 5 years ago
- Efficient Training (including pre-training and fine-tuning) for Big Models☆624Jul 7, 2026Updated last week
- gossip: Efficient Communication Primitives for Multi-GPU Systems☆62Jul 1, 2022Updated 4 years ago
- [CVPR-2023] Towards Any Structural Pruning☆18Apr 27, 2023Updated 3 years ago