real Transformer TeraFLOPS on various GPUs
☆912Jan 9, 2024Updated 2 years ago
Alternatives and similar repositories for transformers-benchmarks
Users that are interested in transformers-benchmarks are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- 深度学习经典、新论文逐段精读☆33,839Mar 22, 2025Updated last year
- Ongoing research training transformer models at scale☆17,947Updated this week
- Ongoing research training transformer language models at scale, including: BERT & GPT-2☆2,264Aug 14, 2025Updated last year
- Fast and memory-efficient exact attention☆24,976Updated this week
- Making large AI models cheaper, faster and more accessible☆41,437Updated this week
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- Transformer related optimization, including BERT, GPT☆6,448Mar 27, 2024Updated 2 years ago
- A library for accelerating Transformer models on NVIDIA GPUs, including using 8-bit and 4-bit floating point (FP8 and FP4) precision on H…☆3,544Updated this week
- Best practice for training LLaMA models in Megatron-LM☆667Jan 2, 2024Updated 2 years ago
- DeepSpeed is a deep learning optimization library that makes distributed training and inference easy, efficient, and effective.☆43,142Updated this week
- Example models using DeepSpeed☆6,849Updated this week
- optimized BERT transformer inference on NVIDIA GPU. https://arxiv.org/abs/2210.03052☆479Mar 15, 2024Updated 2 years ago
- LightSeq: A High Performance Library for Sequence Processing and Generation☆3,294May 16, 2023Updated 3 years ago
- 🤗 PEFT: State-of-the-art Parameter-Efficient Fine-Tuning.☆21,703Updated this week
- GLM-130B: An Open Bilingual Pre-Trained Model (ICLR 2023)☆7,648Jul 25, 2023Updated 3 years ago
- Serverless GPU API endpoints on Runpod - Get Bonus Credits • AdSkip the infrastructure headaches. Auto-scaling, pay-as-you-go, no-ops approach lets you focus on innovating your application.
- 中文LLaMA&Alpaca大语言模型+本地CPU/GPU训练部署 (Chinese LLaMA & Alpaca LLMs)☆18,942Apr 19, 2026Updated 5 months ago
- This is an official implementation for "Swin Transformer: Hierarchical Vision Transformer using Shifted Windows".☆16,080Jul 24, 2024Updated 2 years ago
- An open platform for training, serving, and evaluating large language models. Release repo for Vicuna and Chatbot Arena.☆39,539May 1, 2026Updated 4 months ago
- Hackable and optimized Transformers building blocks, supporting a composable construction.☆10,548Updated this week
- 用文本编辑器剪视频☆7,807Oct 5, 2024Updated last year
- An open-source, tool-augmented conversational language model from Fudan University☆12,251Sep 6, 2026Updated last week
- The largest collection of PyTorch image encoders / backbones. Including train, eval, inference, export scripts, and pretrained weights --…☆37,159Updated this week
- High performance NCCL plugin for Bagua.☆15Sep 15, 2021Updated 5 years ago
- A large-scale 7B pretraining language model developed by BaiChuan-Inc.☆5,649Jul 18, 2024Updated 2 years ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- LMDeploy is a toolkit for compressing, deploying, and serving LLMs.☆8,082Updated this week
- 🚀 A simple way to launch, train, and use PyTorch models on almost any device and distributed configuration, automatic mixed precision (i…☆9,874Updated this week
- Inference code for Llama models☆59,619Jan 26, 2025Updated last year
- LightLLM is a Python-based LLM (Large Language Model) inference and serving framework, notable for its lightweight design, easy scalabili…☆4,295Updated this week
- Development repository for the Triton language and compiler☆20,197Updated this week
- ☆13Feb 22, 2023Updated 3 years ago
- PyTorch extensions for high performance and large scale training.☆3,406Apr 26, 2025Updated last year
- The official GitHub page for the survey paper "A Survey of Large Language Models".☆12,217Mar 11, 2025Updated last year
- Unified Efficient Fine-Tuning of 100+ LLMs & VLMs (ACL 2024)☆74,888Updated this week
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- Large-scale Self-supervised Pre-training Across Tasks, Languages, and Modalities☆22,220Updated this week
- A PyTorch Extension: Tools for easy mixed precision and distributed training in Pytorch☆8,997Updated this week
- MII makes low-latency and high-throughput inference possible, powered by DeepSpeed.☆2,113Jun 30, 2025Updated last year
- 《动手学深度学习》:面向中文读者、能运行、可讨论。中英文版被70多个国家的500多所大学用于教学。☆80,809Jul 30, 2024Updated 2 years ago
- Train transformer language models with reinforcement learning.☆19,342Updated this week
- 🤗 Transformers: the model-definition framework for state-of-the-art machine learning models in text, vision, audio, and multimodal model…☆166,391Updated this week
- TensorRT LLM provides users with an easy-to-use Python API to define Large Language Models (LLMs) and supports state-of-the-art optimizat…☆14,666Updated this week