A wrapper layer for stacking layers horizontally
☆228Jan 22, 2022Updated 4 years ago
Alternatives and similar repositories for keras-multi-head
Users that are interested in keras-multi-head are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Attention mechanism for processing sequential data that considers the context for each timestamp.☆657Jan 22, 2022Updated 4 years ago
- Transformer implemented in Keras☆369Jan 22, 2022Updated 4 years ago
- Layer normalization implemented in Keras☆60Jan 22, 2022Updated 4 years ago
- Keras library for building (Universal) Transformers, facilitating BERT and GPT models☆540May 30, 2020Updated 6 years ago
- Position embedding layers in Keras☆58Jan 22, 2022Updated 4 years ago
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- Keras Attention Layer (Luong and Bahdanau scores).☆2,809Mar 12, 2026Updated 4 months ago
- Adaptive embedding and softmax☆17Jan 22, 2022Updated 4 years ago
- Lookahead mechanism for optimizers in Keras.☆50Jun 24, 2021Updated 5 years ago
- Keras Layer implementation of Attention for Sequential models☆444Mar 25, 2023Updated 3 years ago
- AdaBound optimizer in Keras☆56Jul 11, 2020Updated 6 years ago
- Keras implementation of ABCNN by Yin & Schütze (WIP)☆23Jun 16, 2020Updated 6 years ago
- How to use ELMo embeddings in Keras with Tensorflow Hub☆259Dec 18, 2018Updated 7 years ago
- Collection of custom layers and utility functions for Keras which are missing in the main framework.☆62May 25, 2020Updated 6 years ago
- A Keras+TensorFlow Implementation of the Transformer: Attention Is All You Need☆720Sep 24, 2021Updated 4 years ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- Binary and Categorical Focal loss implementation in Keras.☆279Dec 20, 2024Updated last year
- SNAIL Attention Block for Keras.☆17Mar 30, 2020Updated 6 years ago
- An Attention Layer in Keras☆43Apr 23, 2019Updated 7 years ago
- Contains an implementation of the attention mechanism and a keras text classifier wrapper.☆29Sep 18, 2018Updated 7 years ago
- some attention implements☆1,449Nov 20, 2019Updated 6 years ago
- Keras implementation of BERT with pre-trained weights☆813Jul 26, 2019Updated 7 years ago
- 自注意力与文本分类☆119Nov 3, 2018Updated 7 years ago
- This is a drop-in Keras layer for ELMo embeddings.☆47Dec 29, 2018Updated 7 years ago
- SemEval 2019 Hyperpartisan News Detection - team Bertha von Suttner contribution☆23Aug 15, 2019Updated 6 years ago
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- keras implement of transformers for humans☆5,417Nov 11, 2024Updated last year
- A Keras TensorFlow 2.0 implementation of BERT, ALBERT and adapter-BERT.☆807Jan 13, 2023Updated 3 years ago
- This is a repository for AfriHate Project☆16Jun 9, 2025Updated last year
- Gradient accumulation for Keras☆35Jun 27, 2021Updated 5 years ago
- Implementation of the Transformer architecture described by Vaswani et al. in "Attention Is All You Need"☆28Apr 21, 2019Updated 7 years ago
- Implementations for a family of attention mechanisms, suitable for all kinds of natural language processing tasks and compatible with Ten…☆381Feb 6, 2024Updated 2 years ago
- RAdam implemented in Keras & TensorFlow☆324Jan 22, 2022Updated 4 years ago
- Graph convolutional layers☆62Jan 22, 2022Updated 4 years ago
- Implementation for QANet using Keras with Tensorflow backend☆12Dec 13, 2018Updated 7 years ago
- Open source password manager - Proton Pass • AdSecurely store, share, and autofill your credentials with Proton Pass, the end-to-end encrypted password manager trusted by millions.
- Keras community contributions☆1,586Oct 21, 2022Updated 3 years ago
- Keras/TF implementation of AdamW, SGDW, NadamW, Warm Restarts, and Learning Rate multipliers☆168Jan 6, 2022Updated 4 years ago
- Layer-wise Adaptive Moments optimizer for Batch training☆15Apr 3, 2019Updated 7 years ago
- Keras implementation of AdaBound☆130Nov 4, 2019Updated 6 years ago
- A Hyperparameter Tuning Library for Keras☆2,923Dec 1, 2025Updated 7 months ago
- Temporal Pattern Attention for Multivariate Time Series Forecasting☆735Nov 29, 2018Updated 7 years ago
- Neural Deconvolutions in Tensorflow☆12May 18, 2020Updated 6 years ago