Reference implementation of "An Algorithm for Routing Vectors in Sequences" (Heinsen, 2022) and "An Algorithm for Routing Capsules in All Domains" (Heinsen, 2019), for composing deep neural networks.
☆171Apr 13, 2023Updated 3 years ago
Alternatives and similar repositories for heinsen_routing
Users that are interested in heinsen_routing are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Reference implementation of "Softmax Attention with Constant Cost per Token" (Heinsen, 2024)☆25Jun 6, 2024Updated 2 years ago
- Implementation of deep implicit attention in PyTorch☆68Aug 2, 2021Updated 5 years ago
- Code implementing "Efficient Parallelization of a Ubiquitious Sequential Computation" (Heinsen, 2023)☆97Dec 5, 2024Updated last year
- Implementation of Token Shift GPT - An autoregressive model that solely relies on shifting the sequence space for mixing☆49Jan 27, 2022Updated 4 years ago
- ☆40Jan 5, 2024Updated 2 years ago
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- Hidden Engrams: Long Term Memory for Transformer Model Inference☆35Jun 26, 2021Updated 5 years ago
- A Building blocks for elixir CQRS segregated applications☆15Sep 25, 2019Updated 6 years ago
- A TensorFlow implementation of "Matrix Capsules with EM Routing" by Hinton et al. (2018).☆86Sep 17, 2025Updated last year
- A *tuned* minimal PyTorch re-implementation of the OpenAI GPT (Generative Pretrained Transformer) training☆118Aug 9, 2021Updated 5 years ago
- Mechanistic Interpretability for Transformer Models☆55Jun 1, 2022Updated 4 years ago
- A Haskell derived programming language for systems development.☆13Sep 18, 2018Updated 8 years ago
- Official code repository for the main conference paper in EMNLP 2022: SubeventWriter: Iterative Sub-event Sequence Generation with Cohere…☆11Oct 16, 2022Updated 3 years ago
- This demo showcase the use of onnxruntime-rs with a GPU on CUDA 11 to run Bert in a data pipeline with Rust.☆16Feb 7, 2022Updated 4 years ago
- A Skew Binomial Heap for Erlang.☆15Jun 22, 2011Updated 15 years ago
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- Noise-Contrastive Visualization☆53Nov 25, 2023Updated 2 years ago
- Embroid: Unsupervised Prediction Smoothing Can Improve Few-Shot Classification☆11Aug 12, 2023Updated 3 years ago
- Code for Neural Execution Engines: Learning to Execute Subroutines☆19Jan 11, 2021Updated 5 years ago
- A NIF module for Erlang to Mozilla's Spidermonkey Javascript runtime.☆14Updated this week
- JSON encode/decode library written in Erlang☆17Apr 12, 2025Updated last year
- Libhydrogen bindings for Erlang☆20Feb 10, 2019Updated 7 years ago
- MelGAN-VC: Voice Conversion and Audio Style Transfer on arbitrarily long samples using Spectrograms☆12Nov 25, 2021Updated 4 years ago
- Unofficially Implements https://arxiv.org/abs/2112.05682 to get Linear Memory Cost on Attention for PyTorch☆12Jan 16, 2022Updated 4 years ago
- A PyTorch Implementation of Matrix Capsules with EM Routing☆90Apr 1, 2020Updated 6 years ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- Statistical discontinuous constituent parsing☆11Feb 15, 2018Updated 8 years ago
- Identifying similar OCaml codes☆32Jul 30, 2024Updated 2 years ago
- Implementation of GateLoop Transformer in Pytorch and Jax☆93Jun 18, 2024Updated 2 years ago
- Implementation of Memorizing Transformers (ICLR 2022), attention net augmented with indexing and retrieval of memories using approximate …☆646Jul 17, 2023Updated 3 years ago
- Implementation of a Transformer using ReLA (Rectified Linear Attention) from https://arxiv.org/abs/2104.07012☆49Apr 6, 2022Updated 4 years ago
- GPU Accelerated, Distributed, Actor Model Language (WIP)☆30Jun 21, 2023Updated 3 years ago
- Official Pytorch code for (AAAI 2020) paper "Capsule Routing via Variational Bayes", https://arxiv.org/pdf/1905.11455.pdf☆102Jul 15, 2021Updated 5 years ago
- Explicit Alignment Objectives for Multilingual Bidirectional Encoders☆14Apr 14, 2021Updated 5 years ago
- ☆11Feb 18, 2022Updated 4 years ago
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- Variable-order CRFs with structure learning☆17Aug 1, 2024Updated 2 years ago
- Official Pytorch and JAX implementation of "Efficient-VDVAE: Less is more"☆199Aug 15, 2022Updated 4 years ago
- Fun with variational autoencoders.☆11Dec 15, 2017Updated 8 years ago
- The repository for the code of the UltraFastBERT paper☆518Mar 24, 2024Updated 2 years ago
- Implementation of the Remixer Block from the Remixer paper, in Pytorch☆36Sep 27, 2021Updated 4 years ago
- Memory-efficient transformer. Work in progress.☆19Sep 17, 2022Updated 4 years ago
- An Erlang/OTP parse transform to emulate Scheme's cut/cute syntax.☆18Jan 13, 2021Updated 5 years ago