Pervasive Attention: 2D Convolutional Networks for Sequence-to-Sequence Prediction
☆496May 8, 2021Updated 5 years ago
Alternatives and similar repositories for attn2d
Users that are interested in attn2d are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Code for the Eager Translation Model from the paper You May Not Need Attention☆293Dec 17, 2018Updated 7 years ago
- Phrase-Based & Neural Unsupervised Machine Translation☆1,499Sep 15, 2021Updated 4 years ago
- PyTorch Implementation of "Non-Autoregressive Neural Machine Translation"☆271Feb 12, 2022Updated 4 years ago
- PyTorch implementation of the Quasi-Recurrent Neural Network - up to 16 times faster than NVIDIA's cuDNN LSTM☆1,263Feb 12, 2022Updated 4 years ago
- Deep Learning Projects that Build Themselves☆360Jan 10, 2019Updated 7 years ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- Implementation of Adversarial Variational Optimization in PyTorch☆42Aug 7, 2018Updated 8 years ago
- Sequence-to-Sequence Framework in PyTorch☆392Jan 5, 2023Updated 3 years ago
- PyTorch implementation of PNASNet-5 on ImageNet☆321Aug 4, 2022Updated 4 years ago
- PyTorch original implementation of Cross-lingual Language Model Pretraining.☆2,920Feb 14, 2023Updated 3 years ago
- Latent Alignment and Variational Attention☆327Nov 5, 2018Updated 7 years ago
- Unsupervised Neural Machine Translation☆474Jul 8, 2020Updated 6 years ago
- ☆120Feb 20, 2019Updated 7 years ago
- Training RNNs as Fast as CNNs (https://arxiv.org/abs/1709.02755)☆2,107Jan 4, 2022Updated 4 years ago
- Translate - a PyTorch Language Library☆840Apr 27, 2023Updated 3 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- Fast, general, and tested differentiable structured prediction in PyTorch☆1,133Apr 20, 2022Updated 4 years ago
- Code and Data release for "Improving Multilingual Translation by Representation and Gradient Regularization" (Yang et al. EMNLP 2021), an…☆13Aug 12, 2024Updated 2 years ago
- Transformer training code for sequential tasks☆610Sep 14, 2021Updated 4 years ago
- Write PyTorch code at the level of individual examples, then run it efficiently on minibatches.☆486Feb 12, 2022Updated 4 years ago
- The Natural Language Decathlon: A Multitask Challenge for NLP☆2,338May 1, 2025Updated last year
- souce code for "Accelerating Neural Transformer via an Average Attention Network"☆78Jul 3, 2019Updated 7 years ago
- LSTM and QRNN Language Model Toolkit for PyTorch☆1,989Feb 12, 2022Updated 4 years ago
- Tools for PyTorch☆225Aug 10, 2022Updated 4 years ago
- nmtpy is a Python framework based on dl4mt-tutorial to experiment with Neural Machine Translation pipelines.☆126Mar 15, 2018Updated 8 years ago
- Bare Metal GPUs on DigitalOcean Gradient AI • AdPurpose-built for serious AI teams training foundational models, running large-scale inference, and pushing the boundaries of what's possible.
- Learning General Purpose Distributed Sentence Representations via Large Scale Multi-task Learning☆311Aug 18, 2020Updated 5 years ago
- Training Very Deep Neural Networks Without Skip-Connections☆590Jun 9, 2018Updated 8 years ago
- Sequence-to-Sequence learning using PyTorch☆518Nov 12, 2019Updated 6 years ago
- ☆24Oct 9, 2018Updated 7 years ago
- Sparse and structured neural attention mechanisms☆224Aug 31, 2020Updated 5 years ago
- Code and model for the paper "Improving Language Understanding by Generative Pre-Training"☆2,310Jan 25, 2019Updated 7 years ago
- Sequence modeling benchmarks and temporal convolutional networks☆4,544Mar 28, 2022Updated 4 years ago
- [ICLR'19] Trellis Networks for Sequence Modeling☆471Aug 20, 2019Updated 6 years ago
- 🐥A PyTorch implementation of OpenAI's finetuned transformer language model with a script to import the weights pre-trained by OpenAI☆1,523Aug 9, 2021Updated 5 years ago
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- Reference implementation of real-time autoregressive wavenet inference☆744Jan 19, 2021Updated 5 years ago
- Generative Query Network (GQN) in PyTorch as described in "Neural Scene Representation and Rendering"☆322Jun 24, 2019Updated 7 years ago
- Code for "Understanding and Improving Interpolation in Autoencoders via an Adversarial Regularizer"☆244Aug 6, 2018Updated 8 years ago
- Dilated RNNs in pytorch☆213Jun 24, 2019Updated 7 years ago
- Facebook AI Research Sequence-to-Sequence Toolkit☆3,725Sep 17, 2021Updated 4 years ago
- A short and easy implementation of Quantile Regression DQN | Distributional Reinforcement Learning☆97Sep 3, 2020Updated 5 years ago
- Sequence to Sequence Models with PyTorch☆740Mar 27, 2022Updated 4 years ago