Facebook AI Research Sequence-to-Sequence Toolkit written in Python.
β32,228Sep 30, 2025Updated 11 months ago
Alternatives and similar repositories for fairseq
Users that are interested in fairseq are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- π€ Transformers: the model-definition framework for state-of-the-art machine learning models in text, vision, audio, and multimodal modelβ¦β164,665Updated this week
- Large-scale Self-supervised Pre-training Across Tasks, Languages, and Modalitiesβ22,200Updated this week
- An open-source NLP research library, built on PyTorch.β11,886Nov 22, 2022Updated 3 years ago
- Open Source Neural Machine Translation and (Large) Language Models in PyTorchβ7,014Oct 14, 2025Updated 10 months ago
- TensorFlow code and pre-trained models for BERTβ40,047Jul 23, 2024Updated 2 years ago
- Deploy on Railway without the complexity - Free Credits Offer β’ AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- Unsupervised text tokenizer for Neural Network-based text generation.β12,052Updated this week
- End-to-End Speech Processing Toolkitβ9,947Updated this week
- Google Researchβ38,659Updated this week
- DeepSpeed is a deep learning optimization library that makes distributed training and inference easy, efficient, and effective.β43,036Updated this week
- Repository to track the progress in Natural Language Processing (NLP), including the datasets and the current state-of-the-art for the moβ¦β22,953Jul 28, 2024Updated 2 years ago
- PyTorch original implementation of Cross-lingual Language Model Pretraining.β2,920Feb 14, 2023Updated 3 years ago
- XLNet: Generalized Autoregressive Pretraining for Language Understandingβ6,189May 28, 2023Updated 3 years ago
- Pretrain, finetune ANY AI model of ANY size on 1 or 10,000+ GPUs with zero code changes.β31,313Updated this week
- A library for efficient similarity search and clustering of dense vectors.β40,832Updated this week
- Virtual machines for every use case on DigitalOcean β’ AdGet dependable uptime with 99.99% SLA, simple security tools, and predictable monthly pricing with DigitalOcean's virtual machines, called Droplets.
- State-of-the-Art Embeddings, Retrieval, and Rerankingβ19,051Updated this week
- Code for the paper "Exploring the Limits of Transfer Learning with a Unified Text-to-Text Transformer"β6,545Jul 8, 2026Updated last month
- A framework for training and evaluating AI models on a variety of openly available dialogue datasets.β10,620Jul 30, 2026Updated last month
- Library of deep learning models and datasets designed to make deep learning more accessible and accelerate ML research.β17,466Jun 2, 2023Updated 3 years ago
- A very simple framework for state-of-the-art Natural Language Processing (NLP)β14,383Oct 27, 2025Updated 10 months ago
- Ongoing research training transformer models at scaleβ17,687Updated this week
- A scalable generative AI framework built for researchers and developers working on Large Language Models, Multimodal, and Speech AI (Autoβ¦β18,368Updated this week
- A PyTorch-based Speech Toolkitβ11,798Updated this week
- Language-Agnostic SEntence Representationsβ3,660May 2, 2024Updated 2 years ago
- GPU virtual machines on DigitalOcean Gradient AI β’ AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- Library for fast text representation and classification.β26,534Mar 22, 2024Updated 2 years ago
- A PyTorch Extension: Tools for easy mixed precision and distributed training in Pytorchβ8,997Updated this week
- π Scalable embedding, reasoning, ranking for images and sentences with CLIPβ12,835Jan 23, 2024Updated 2 years ago
- π€ PEFT: State-of-the-art Parameter-Efficient Fine-Tuning.β21,615Updated this week
- An open platform for training, serving, and evaluating large language models. Release repo for Vicuna and Chatbot Arena.β39,527May 1, 2026Updated 4 months ago
- Code and documentation to train Stanford's Alpaca models, and generate the data.β30,246Jul 17, 2024Updated 2 years ago
- Fast and memory-efficient exact attentionβ24,813Updated this week
- Composable transformations of Python+NumPy programs: differentiate, vectorize, JIT to GPU/TPU, and moreβ36,230Updated this week
- π A simple way to launch, train, and use PyTorch models on almost any device and distributed configuration, automatic mixed precision (iβ¦β9,839Updated this week
- 1-Click AI Models by DigitalOcean Gradient β’ AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- A modular framework for vision & language multimodal research from Facebook AI Research (FAIR)β5,634Jul 7, 2026Updated last month
- Unsupervised Word Segmentation for Neural Machine Translation and Text Generationβ2,274Aug 7, 2024Updated 2 years ago
- CLIP (Contrastive Language-Image Pretraining), Predict the most relevant text snippet given an imageβ34,246Mar 25, 2026Updated 5 months ago
- Code for the paper "Language Models are Unsupervised Multitask Learners"β25,017Aug 14, 2024Updated 2 years ago
- Tensors and Dynamic neural networks in Python with strong GPU accelerationβ102,694Updated this week
- A natural language modeling framework based on PyTorchβ6,292Oct 17, 2022Updated 3 years ago
- Distributed training framework for TensorFlow, Keras, PyTorch, and Apache MXNet.β14,687Jul 29, 2026Updated last month