☆548Feb 13, 2024Updated 2 years ago
Alternatives and similar repositories for byt5
Users that are interested in byt5 are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- ☆1,290Dec 15, 2022Updated 3 years ago
- ☆3,000Aug 3, 2026Updated 2 months ago
- A utility for storing and reading files for Korean LM training 💾☆35Jul 18, 2026Updated 2 months ago
- Binary Passage Retriever (BPR) - an efficient passage retriever for open-domain question answering☆176Jun 6, 2021Updated 5 years ago
- OSLO: Open Source framework for Large-scale model Optimization☆309Aug 25, 2022Updated 4 years ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- Library for Knowledge Intensive Language Tasks☆978Mar 31, 2022Updated 4 years ago
- M2D2: A Massively Multi-domain Language Modeling Dataset (EMNLP 2022) by Machel Reid, Victor Zhong, Suchin Gururangan, Luke Zettlemoyer☆55Nov 21, 2022Updated 3 years ago
- ☆183May 26, 2023Updated 3 years ago
- Code for the paper "Exploring the Limits of Transfer Learning with a Unified Text-to-Text Transformer"☆6,556Updated this week
- ☆1,575Jul 2, 2026Updated 3 months ago
- Pytorch Implementation of EncT5: Fine-tuning T5 Encoder for Non-autoregressive Tasks☆62Jan 22, 2022Updated 4 years ago
- QED: A Framework and Dataset for Explanations in Question Answering☆119Aug 3, 2021Updated 5 years ago
- Flexible components pairing 🤗 Transformers with Pytorch Lightning☆608Nov 21, 2022Updated 3 years ago
- NL-Augmenter 🦎 → 🐍 A Collaborative Repository of Natural Language Transformations☆785May 19, 2024Updated 2 years ago
- Bare Metal GPUs on DigitalOcean Gradient AI • AdPurpose-built for serious AI teams training foundational models, running large-scale inference, and pushing the boundaries of what's possible.
- Autoregressive Entity Retrieval☆800Jul 6, 2023Updated 3 years ago
- ☆74Jul 2, 2021Updated 5 years ago
- Shared repository for open-sourced projects from the Google AI Language team.☆1,814Jun 10, 2026Updated 3 months ago
- [ACL 2021] Learning Dense Representations of Phrases at Scale; EMNLP'2021: Phrase Retrieval Learns Passage Retrieval, Too https://arxiv.o…☆606Jun 15, 2022Updated 4 years ago
- Parallelformers: An Efficient Model Parallelization Toolkit for Deployment☆788Apr 24, 2023Updated 3 years ago
- Implementation of RETRO, Deepmind's Retrieval based Attention net, in Pytorch☆878Oct 30, 2023Updated 2 years ago
- GC4LM: A Colossal (Biased) language model for German☆13May 2, 2021Updated 5 years ago
- Task-based datasets, preprocessing, and evaluation for sequence models.☆596Updated this week
- SentAugment is a data augmentation technique for NLP that retrieves similar sentences from a large bank of sentences. It can be used in c…☆358Feb 22, 2022Updated 4 years ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- The implementation of DeBERTa☆2,255Sep 29, 2023Updated 3 years ago
- An efficient implementation of the popular sequence models for text generation, summarization, and translation tasks. https://arxiv.org/p…☆432Aug 17, 2022Updated 4 years ago
- ⚡ boost inference speed of T5 models by 5x & reduce the model size by 3x.☆586Apr 24, 2023Updated 3 years ago
- ELECTRA: Pre-training Text Encoders as Discriminators Rather Than Generators☆2,365Mar 23, 2024Updated 2 years ago
- Reproduce results and replicate training fo T0 (Multitask Prompted Training Enables Zero-Shot Task Generalization)☆463Nov 5, 2022Updated 3 years ago
- XtremeDistil framework for distilling/compressing massive multilingual neural network models to tiny and efficient models for AI at scale☆158Dec 20, 2023Updated 2 years ago
- 매주 목요일, 20:00 모임☆16Jul 24, 2020Updated 6 years ago
- This repository contains the code for "Exploiting Cloze Questions for Few-Shot Text Classification and Natural Language Inference"☆1,623Jun 12, 2023Updated 3 years ago
- Resources for the "CTRLsum: Towards Generic Controllable Text Summarization" paper☆149May 1, 2025Updated last year
- Simple, predictable pricing with DigitalOcean hosting • AdAlways know what you'll pay with monthly caps and flat pricing. Enterprise-grade infrastructure trusted by 600k+ customers.
- Implementation of the GBST block from the Charformer paper, in Pytorch☆118Jul 15, 2021Updated 5 years ago
- Repo for training MLMs, CLMs, or T5-type models on the OLM pretraining data, but it should work with any hugging face text dataset.☆98Feb 9, 2023Updated 3 years ago
- Code and data to support the paper "PAQ 65 Million Probably-Asked Questions andWhat You Can Do With Them"☆210Aug 31, 2021Updated 5 years ago
- ☆92Sep 29, 2021Updated 5 years ago
- ☆220Jun 8, 2020Updated 6 years ago
- [EMNLP 2021] LM-Critic: Language Models for Unsupervised Grammatical Error Correction☆118Sep 26, 2021Updated 5 years ago
- Code for the Shortformer model, from the ACL 2021 paper by Ofir Press, Noah A. Smith and Mike Lewis.☆145Jul 26, 2021Updated 5 years ago