Adaptive embedding and softmax
☆17Jan 22, 2022Updated 4 years ago
Alternatives and similar repositories for keras-adaptive-softmax
Users that are interested in keras-adaptive-softmax are comparing it to the libraries listed below
Sorting:
- Transformer-XL with checkpoint loader☆67Jan 22, 2022Updated 4 years ago
- Gradient accumulation for Keras☆35Jun 27, 2021Updated 4 years ago
- Ordered Neurons LSTM☆30Jan 22, 2022Updated 4 years ago
- ☆11Sep 3, 2021Updated 4 years ago
- Tensorflow NCE loss in Keras☆34Oct 6, 2018Updated 7 years ago
- Sampling Matters in Deep Embedding Learning (ICCV'17)☆16Oct 16, 2018Updated 7 years ago
- A package of Wide Residual Networks for image recognition in Keras.☆15Jan 22, 2022Updated 4 years ago
- Slides from various talks I gave☆18Oct 25, 2018Updated 7 years ago
- Transformer implemented in Keras☆369Jan 22, 2022Updated 4 years ago
- Load GPT-2 checkpoint and generate texts☆127Jan 22, 2022Updated 4 years ago
- Documentation for Chatstack: A Full Pipeline UI for building Chinese NLU System☆18Sep 7, 2019Updated 6 years ago
- Generative Adversarial Network with Weight Normalization + ResNet☆22Dec 18, 2017Updated 8 years ago
- Learning rate multiplier☆46Jun 22, 2021Updated 4 years ago
- Keras implementation of AdaBound☆130Nov 4, 2019Updated 6 years ago
- ‘京东杯’《用户对品类下的购买预测》-冠军团队‘AI你所想’解决方案☆24Dec 15, 2019Updated 6 years ago
- Code for the paper "Addressing Model Vulnerability to Distributional Shifts over Image Transformation Sets", ICCV 2019☆27Mar 17, 2020Updated 5 years ago
- lookahead optimizer for keras☆169Oct 14, 2019Updated 6 years ago
- Implementation of XLNet that can load pretrained checkpoints☆169Jan 22, 2022Updated 4 years ago
- Pytorch implementation of Dauphin et al. (2016) "Language Modeling with Gated Convolutional Networks"☆29Jan 10, 2023Updated 3 years ago
- RAdam optimizer for keras☆71Oct 14, 2019Updated 6 years ago
- Training RNNs as fast as CNNs. An unofficial tensorflow implementation.☆33Feb 23, 2018Updated 8 years ago
- Octave convolution☆34Jan 22, 2022Updated 4 years ago
- 90%+ with 40 labels. please see the readme for details.☆36Aug 18, 2020Updated 5 years ago
- Training with FP16 weights in PyTorch☆81Aug 7, 2019Updated 6 years ago
- website of the educational workshop of the Montreal Artificial Intelligence and Neuroscience (MAIN) conference☆11Dec 16, 2025Updated 2 months ago
- Material of the European Summer School on Computational and Mathematical Modeling of Cognition☆11Jul 19, 2022Updated 3 years ago
- A wrapper layer for stacking layers horizontally☆228Jan 22, 2022Updated 4 years ago
- Implementation of "Improving the Improved Training of Wasserstein GANs: A Consistency Term and Its Dual Effect" in pytorch☆37Jun 17, 2019Updated 6 years ago
- ☆10Apr 5, 2022Updated 3 years ago
- ☆12Nov 25, 2018Updated 7 years ago
- C++ PyTorch Examples☆10Aug 18, 2019Updated 6 years ago
- Pytorch code for the paper 'Attention-based Atrous Convolutional Neural Networks: Visualisation and Understanding Perspectives of Acousti…☆14Nov 12, 2020Updated 5 years ago
- Easier, dynamic mocking for Swift.☆10Aug 4, 2019Updated 6 years ago
- Pytorch implementation of ProtoAU for recommendation.☆10Dec 19, 2024Updated last year
- Python scripts to facilitate easy working☆11Jun 24, 2024Updated last year
- 提取出判决书中的金额项和金额数。☆11Apr 8, 2016Updated 9 years ago
- Materials for Graph Models and Graph Networks☆11Jul 6, 2018Updated 7 years ago
- This is an official implementation of our CVPR 2020 paper "Non-Local Neural Networks With Grouped Bilinear Attentional Transforms".☆12Jan 30, 2021Updated 5 years ago
- parallel corpora for any languages supported by glosbe.com☆10Feb 9, 2016Updated 10 years ago