Pytorch implementation of "Block Recurrent Transformers" (Hutchins & Schlag et al., 2022)
☆85May 14, 2022Updated 4 years ago
Alternatives and similar repositories for block-recurrent-transformer
Users that are interested in block-recurrent-transformer are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Implementation of Block Recurrent Transformer - Pytorch☆226Aug 20, 2024Updated 2 years ago
- dracut module using vdfuse to loop mount☆11Mar 21, 2021Updated 5 years ago
- ☆260Jun 6, 2025Updated last year
- ☆16Dec 9, 2023Updated 2 years ago
- Transformer-based Label Set Generation for Multi-modal Multi-label Emotion Detection☆14Dec 16, 2021Updated 4 years ago
- Bare Metal GPUs on DigitalOcean Gradient AI • AdPurpose-built for serious AI teams training foundational models, running large-scale inference, and pushing the boundaries of what's possible.
- Fine-Tuning Pre-trained Transformers into Decaying Fast Weights☆20Oct 9, 2022Updated 3 years ago
- Code for the ACL 2021 paper "Structural Guidance for Transformer Language Models"☆15Sep 17, 2025Updated 11 months ago
- ☆29Jul 12, 2022Updated 4 years ago
- CopulaGNN: Towards Integrating Representational and Correlational Roles of Graphs in Graph Neural Networks (ICLR 2021)☆14Dec 5, 2022Updated 3 years ago
- ☆16Mar 13, 2023Updated 3 years ago
- ☆14Aug 18, 2022Updated 4 years ago
- Creating DRL infrastructure for Dynamic Beta with Zipline and Keras☆14Dec 8, 2022Updated 3 years ago
- Control of 2D Rayleigh Benard Convection using Deep Reinforcement Learning with Tensorforce and Shenfun.☆23Jul 5, 2023Updated 3 years ago
- ☆69Aug 3, 2023Updated 3 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- ☆14Sep 11, 2022Updated 3 years ago
- Add NextJS to Moleculer! 🎉☆11Jun 21, 2018Updated 8 years ago
- Resources for the paper "PARE: A Simple and Strong Baseline for Monolingual and Multilingual Distantly Supervised Relation Extraction"☆13Jul 26, 2022Updated 4 years ago
- ☆10Feb 25, 2020Updated 6 years ago
- Source code of the paper "Synchronous Double-channel Recurrent Network for Aspect-Opinion Pair Extraction, ACL 2020."☆12Aug 10, 2020Updated 6 years ago
- Xfce Desktop container designed for direct access to the GPU with EGL using VirtualGL for GPUs. Does not require /tmp/.X11-unix host sock…☆10Jul 25, 2022Updated 4 years ago
- Top1 Solution on OGB Challenge (Graph Property Prediction on HIV dataset)☆10Oct 20, 2021Updated 4 years ago
- Codes for the paper "Learning Graph-Level Representations with Gated Recurrent Neural Networks"☆29Feb 11, 2019Updated 7 years ago
- A library for simplifying training with multi gpu setups in the HuggingFace / PyTorch ecosystem.☆16Jun 10, 2026Updated 2 months ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- Efficient Transformers with Dynamic Token Pooling☆68May 20, 2023Updated 3 years ago
- ☆22Sep 6, 2025Updated 11 months ago
- ☆13Jul 16, 2024Updated 2 years ago
- Code for the ICLR'23 paper "Temporal Dependencies in Feature Importance for Time Series Prediction"☆25Mar 31, 2023Updated 3 years ago
- Implementation of Gated State Spaces, from the paper "Long Range Language Modeling via Gated State Spaces", in Pytorch☆101Feb 25, 2023Updated 3 years ago
- ☆29May 4, 2024Updated 2 years ago
- Generalizing Natural Language Analysis through Span-relation Representations☆91Sep 22, 2025Updated 11 months ago
- Implementation of Convolutional Recurrent Neural Network (CRNN) to decode motor imagery EEG data.☆11Mar 25, 2023Updated 3 years ago
- Official implementation of the paper "ALADIN: Distilling Fine-grained Alignment Scores for Efficient Image-Text Matching and Retrieval"☆28Dec 6, 2023Updated 2 years ago
- Serverless GPU API endpoints on Runpod - Get Bonus Credits • AdSkip the infrastructure headaches. Auto-scaling, pay-as-you-go, no-ops approach lets you focus on innovating your application.
- A Dual-Stage Attention-Based Recurrent Neural Network for Time Series Prediction☆113Jul 4, 2024Updated 2 years ago
- Code for: "Cutting Down on Prompts and Parameters: Simple Few-Shot Learning with Language Models"☆19Feb 2, 2022Updated 4 years ago
- Code to related to my NIPS 2016 paper☆10Dec 4, 2016Updated 9 years ago
- ☆21May 16, 2024Updated 2 years ago
- ☆15Jul 14, 2022Updated 4 years ago
- Surface Electromyograph (SEMG) Control for a robotic hand☆15Mar 21, 2020Updated 6 years ago
- Code repository for the paper "Learning partial differential equations for biological transport models from noisy spatiotemporal data"☆11Jul 3, 2019Updated 7 years ago