CyberZHG / keras-adaptive-softmaxView external linksLinks
Adaptive embedding and softmax
☆17Jan 22, 2022Updated 4 years ago
Alternatives and similar repositories for keras-adaptive-softmax
Users that are interested in keras-adaptive-softmax are comparing it to the libraries listed below
Sorting:
- Transformer-XL with checkpoint loader☆68Jan 22, 2022Updated 4 years ago
- Gradient accumulation for Keras☆35Jun 27, 2021Updated 4 years ago
- Ordered Neurons LSTM☆30Jan 22, 2022Updated 4 years ago
- ☆11Sep 3, 2021Updated 4 years ago
- AdaBound optimizer in Keras☆56Jul 11, 2020Updated 5 years ago
- Tensorflow NCE loss in Keras☆34Oct 6, 2018Updated 7 years ago
- a language model with gated conv nets (implements https://arxiv.org/pdf/1612.08083v1.pdf)☆16Jan 5, 2021Updated 5 years ago
- Sampling Matters in Deep Embedding Learning (ICCV'17)☆16Oct 16, 2018Updated 7 years ago
- Transformer implemented in Keras☆369Jan 22, 2022Updated 4 years ago
- Load GPT-2 checkpoint and generate texts☆127Jan 22, 2022Updated 4 years ago
- Generative Adversarial Network with Weight Normalization + ResNet☆22Dec 18, 2017Updated 8 years ago
- Learning rate multiplier☆46Jun 22, 2021Updated 4 years ago
- Keras implementation of AdaBound☆130Nov 4, 2019Updated 6 years ago
- ‘京东杯’《用户对品类下的购买预测》-冠军团队‘AI你所想’解决方案☆24Dec 15, 2019Updated 6 years ago
- ☆12Mar 21, 2024Updated last year
- A library for minimizing the effects of confounding covariates☆15May 28, 2025Updated 8 months ago
- Code for the paper "Addressing Model Vulnerability to Distributional Shifts over Image Transformation Sets", ICCV 2019☆27Mar 17, 2020Updated 5 years ago
- Implementation in Keras of: Snapshot Ensembles: Train 1, get M for free (https://arxiv.org/abs/1704.00109)☆26Oct 19, 2018Updated 7 years ago
- lookahead optimizer for keras☆170Oct 14, 2019Updated 6 years ago
- Implementation of XLNet that can load pretrained checkpoints☆170Jan 22, 2022Updated 4 years ago
- Pytorch implementation of Dauphin et al. (2016) "Language Modeling with Gated Convolutional Networks"☆29Jan 10, 2023Updated 3 years ago
- RAdam optimizer for keras☆71Oct 14, 2019Updated 6 years ago
- Training RNNs as fast as CNNs. An unofficial tensorflow implementation.☆33Feb 23, 2018Updated 7 years ago
- Octave convolution☆34Jan 22, 2022Updated 4 years ago
- 90%+ with 40 labels. please see the readme for details.☆36Aug 18, 2020Updated 5 years ago
- Training with FP16 weights in PyTorch☆81Aug 7, 2019Updated 6 years ago
- Material of the European Summer School on Computational and Mathematical Modeling of Cognition☆11Jul 19, 2022Updated 3 years ago
- Cognitive Computational Neuroscience online Reading Club(CCN0RC)☆12Jul 30, 2021Updated 4 years ago
- ☆12Updated this week
- A Terraform module to create an Identity and Access Management (IAM) Role on Amazon Web Services (AWS). https://aws.amazon.com/iam☆10Apr 6, 2022Updated 3 years ago
- website of the educational workshop of the Montreal Artificial Intelligence and Neuroscience (MAIN) conference☆11Dec 16, 2025Updated last month
- A wrapper layer for stacking layers horizontally☆228Jan 22, 2022Updated 4 years ago
- Implementation of "Improving the Improved Training of Wasserstein GANs: A Consistency Term and Its Dual Effect" in pytorch☆37Jun 17, 2019Updated 6 years ago
- PyTorch volume toolkit. Efficient data loading, dataset conversions, visualization tools☆10Dec 7, 2022Updated 3 years ago
- Codebase accompanying the paper 'Widening the Representation Bottleneck in Neural Machine Translation with Lexical Shortcuts', (Emelin, D…☆11Feb 14, 2023Updated 3 years ago
- Layers, datasets and utilities for PyTorch☆10Nov 22, 2023Updated 2 years ago
- Listing my favorite research papers 📝 from different fields as I read them.☆10Oct 17, 2019Updated 6 years ago
- Graphics Engine Created From Apples Metal Shading Language☆11Nov 6, 2017Updated 8 years ago
- ☆10Apr 5, 2022Updated 3 years ago