Unofficial PyTorch implementation of Fastformer based on paper "Fastformer: Additive Attention Can Be All You Need"."
☆131Sep 6, 2021Updated 5 years ago
Alternatives and similar repositories for Fastformer-PyTorch
Users that are interested in Fastformer-PyTorch are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Implementation of Fast Transformer in Pytorch☆176Aug 26, 2021Updated 5 years ago
- ☆13Aug 13, 2020Updated 6 years ago
- FastFormers - highly efficient transformer models for NLU☆701Mar 21, 2025Updated last year
- Optimizing bit-level Jaccard Index and Population Counts for large-scale quantized Vector Search via Harley-Seal CSA and Lookup Tables☆22May 18, 2025Updated last year
- FairSeq repo with Apollo optimizer☆113Dec 20, 2023Updated 2 years ago
- Bare Metal GPUs on DigitalOcean Gradient AI • AdPurpose-built for serious AI teams training foundational models, running large-scale inference, and pushing the boundaries of what's possible.
- Trains Transformer model variants. Data isn't shuffled between batches.☆145Oct 5, 2022Updated 3 years ago
- transformers go brrr...☆148Feb 15, 2022Updated 4 years ago
- Another attempt at a long-context / efficient transformer by me☆38Apr 11, 2022Updated 4 years ago
- Interactive tree-maps with SBERT & Hierarchical Clustering (HAC)☆30Dec 31, 2024Updated last year
- Tensorflow Implementation of "Theory and Experiments on Vector Quantized Autoencoders"☆15Feb 27, 2019Updated 7 years ago
- ACL 2021: HiTransformer☆13May 29, 2021Updated 5 years ago
- [NAACL 2021] This is the code for our paper `Fine-Tuning Pre-trained Language Model with Weak Supervision: A Contrastive-Regularized Self…☆205Aug 17, 2022Updated 4 years ago
- An open-source AutoML Library based on PyTorch☆306Sep 3, 2026Updated 2 weeks ago
- Tagger for explicit cause-and-effect relationships in text☆11Jan 8, 2020Updated 6 years ago
- Managed Database hosting by DigitalOcean • AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- On Generating Extended Summaries of Long Documents☆78Jan 26, 2021Updated 5 years ago
- Code of the paper "Low-Latency Speech Separation Guided Diarization for Telephone Conversations"☆15Dec 22, 2022Updated 3 years ago
- WaveGlow vocoder with VQVAE☆61Jun 18, 2019Updated 7 years ago
- ☆31Jan 16, 2021Updated 5 years ago
- Implementation of Perceiver, General Perception with Iterative Attention, in Pytorch☆1,218Updated this week
- [NeurIPS 2021 Spotlight] Official code for "Focal Self-attention for Local-Global Interactions in Vision Transformers"☆559Mar 27, 2022Updated 4 years ago
- A series of BERT and Albert model checkpoints trained to reduce gendered correlations in pre-training☆11Oct 22, 2020Updated 5 years ago
- ☆11Jun 21, 2022Updated 4 years ago
- ☆74Jul 2, 2021Updated 5 years ago
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- speech-aligner,是一个从“人声语音”及其“语言文本”,产生音素级别时间对齐标注的工具。speech-aligner, is a tool that generate phoneme-level alignment between human speech an…☆15Dec 19, 2018Updated 7 years ago
- ☆17Sep 22, 2020Updated 6 years ago
- A comprehensive tool for linguistic analysis of communities☆49Oct 1, 2021Updated 4 years ago
- GPT-J 6B inference on TensorRT with INT-8 precision☆11Apr 5, 2023Updated 3 years ago
- (ACL-IJCNLP 2021) Convolutions and Self-Attention: Re-interpreting Relative Positions in Pre-trained Language Models.☆21Jul 13, 2022Updated 4 years ago
- Pytorch implementation of Compressive Transformers, from Deepmind☆165Oct 4, 2021Updated 4 years ago
- Implementation, trained models and result data for the paper "Aspect-based Document Similarity for Research Papers" #COLING2020☆64Apr 30, 2024Updated 2 years ago
- Extract statistics from Wikipedia Dump files.☆26Aug 2, 2021Updated 5 years ago
- Sessa: Selective State Space Attention☆18Apr 28, 2026Updated 4 months ago
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- Some demos using Nvidia RAPIDS for Cheminformatics☆14Aug 17, 2020Updated 6 years ago
- [ICLR 2022] Official implementation of cosformer-attention in cosFormer: Rethinking Softmax in Attention☆199Dec 2, 2022Updated 3 years ago
- ☆75Nov 19, 2022Updated 3 years ago
- An implementation of Performer, a linear attention-based transformer, in Pytorch☆1,182Feb 2, 2022Updated 4 years ago
- Demo for DART, Audio Imagination workshop submission in NeurIPS 2024☆16Apr 22, 2026Updated 5 months ago
- Unofficially Implements https://arxiv.org/abs/2112.05682 to get Linear Memory Cost on Attention for PyTorch☆12Jan 16, 2022Updated 4 years ago
- Neural Lexicon Reader: Reduce Pronunciation Errors in End-to-end TTS by Leveraging External Textual Knowledge☆21Jul 25, 2022Updated 4 years ago