Facebook AI Research Sequence-to-Sequence Toolkit written in Python.
β32,243Sep 30, 2025Updated 10 months ago
Alternatives and similar repositories for fairseq
Users that are interested in fairseq are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- π€ Transformers: the model-definition framework for state-of-the-art machine learning models in text, vision, audio, and multimodal modelβ¦β163,802Updated this week
- Large-scale Self-supervised Pre-training Across Tasks, Languages, and Modalitiesβ22,185Jan 23, 2026Updated 6 months ago
- An open-source NLP research library, built on PyTorch.β11,888Nov 22, 2022Updated 3 years ago
- Open Source Neural Machine Translation and (Large) Language Models in PyTorchβ7,012Oct 14, 2025Updated 9 months ago
- TensorFlow code and pre-trained models for BERTβ40,057Jul 23, 2024Updated 2 years ago
- Managed Database hosting by DigitalOcean β’ AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- Unsupervised text tokenizer for Neural Network-based text generation.β12,018Updated this week
- End-to-End Speech Processing Toolkitβ9,917Updated this week
- Google Researchβ38,514Updated this week
- DeepSpeed is a deep learning optimization library that makes distributed training and inference easy, efficient, and effective.β42,902Updated this week
- Repository to track the progress in Natural Language Processing (NLP), including the datasets and the current state-of-the-art for the moβ¦β22,961Jul 28, 2024Updated 2 years ago
- PyTorch original implementation of Cross-lingual Language Model Pretraining.β2,921Feb 14, 2023Updated 3 years ago
- XLNet: Generalized Autoregressive Pretraining for Language Understandingβ6,186May 28, 2023Updated 3 years ago
- Pretrain, finetune ANY AI model of ANY size on 1 or 10,000+ GPUs with zero code changes.β31,284Updated this week
- A library for efficient similarity search and clustering of dense vectors.β40,714Updated this week
- GPU virtual machines on DigitalOcean Gradient AI β’ AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- State-of-the-Art Embeddings, Retrieval, and Rerankingβ18,987Updated this week
- Code for the paper "Exploring the Limits of Transfer Learning with a Unified Text-to-Text Transformer"β6,545Jul 8, 2026Updated last month
- A framework for training and evaluating AI models on a variety of openly available dialogue datasets.β10,625Jul 30, 2026Updated last week
- Library of deep learning models and datasets designed to make deep learning more accessible and accelerate ML research.β17,460Jun 2, 2023Updated 3 years ago
- A very simple framework for state-of-the-art Natural Language Processing (NLP)β14,381Oct 27, 2025Updated 9 months ago
- Ongoing research training transformer models at scaleβ17,393Updated this week
- A scalable generative AI framework built for researchers and developers working on Large Language Models, Multimodal, and Speech AI (Autoβ¦β18,095Updated this week
- A PyTorch-based Speech Toolkitβ11,748Jun 15, 2026Updated last month
- Language-Agnostic SEntence Representationsβ3,660May 2, 2024Updated 2 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer β’ AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- Library for fast text representation and classification.β26,544Mar 22, 2024Updated 2 years ago
- A PyTorch Extension: Tools for easy mixed precision and distributed training in Pytorchβ8,991Updated this week
- π Scalable embedding, reasoning, ranking for images and sentences with CLIPβ12,834Jan 23, 2024Updated 2 years ago
- Inference code for Llama modelsβ59,553Jan 26, 2025Updated last year
- π€ PEFT: State-of-the-art Parameter-Efficient Fine-Tuning.β21,528Updated this week
- An open platform for training, serving, and evaluating large language models. Release repo for Vicuna and Chatbot Arena.β39,519May 1, 2026Updated 3 months ago
- Code and documentation to train Stanford's Alpaca models, and generate the data.β30,244Jul 17, 2024Updated 2 years ago
- Fast and memory-efficient exact attentionβ24,681Updated this week
- Composable transformations of Python+NumPy programs: differentiate, vectorize, JIT to GPU/TPU, and moreβ36,140Updated this week
- Serverless GPU API endpoints on Runpod - Get Bonus Credits β’ AdSkip the infrastructure headaches. Auto-scaling, pay-as-you-go, no-ops approach lets you focus on innovating your application.
- π A simple way to launch, train, and use PyTorch models on almost any device and distributed configuration, automatic mixed precision (iβ¦β9,811Updated this week
- A modular framework for vision & language multimodal research from Facebook AI Research (FAIR)β5,635Jul 7, 2026Updated last month
- Unsupervised Word Segmentation for Neural Machine Translation and Text Generationβ2,272Aug 7, 2024Updated 2 years ago
- CLIP (Contrastive Language-Image Pretraining), Predict the most relevant text snippet given an imageβ34,155Mar 25, 2026Updated 4 months ago
- Code for the paper "Language Models are Unsupervised Multitask Learners"β25,030Aug 14, 2024Updated last year
- Tensors and Dynamic neural networks in Python with strong GPU accelerationβ102,312Updated this week
- A natural language modeling framework based on PyTorchβ6,293Oct 17, 2022Updated 3 years ago