Transparent Cudnn / Cublas / Eigen usage for the deep learning training using MNIST dataset.
☆18Sep 3, 2020Updated 5 years ago
Alternatives and similar repositories for DeepLearning-Training-Cuda
Users that are interested in DeepLearning-Training-Cuda are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- 利用Alexnet实现MNIST☆12Aug 27, 2018Updated 7 years ago
- CUDA for MNIST training/inference☆44Dec 30, 2023Updated 2 years ago
- ☆40Feb 28, 2020Updated 6 years ago
- Implementation of algorithms for memory optimized deep neural network training☆10Jul 23, 2020Updated 6 years ago
- Deep learning library written in C++ with CUDA support☆10May 23, 2021Updated 5 years ago
- Virtual machines for every use case on DigitalOcean • AdGet dependable uptime with 99.99% SLA, simple security tools, and predictable monthly pricing with DigitalOcean's virtual machines, called Droplets.
- Partial implementation of NVIDIA® cuDNN API for Coriander, OpenCL 1.2☆22Apr 21, 2025Updated last year
- A lightweight MATLAB deeplearning toolbox,based on gpuArray.☆53Sep 16, 2017Updated 8 years ago
- Forward and backward Attention DNN operators implementationed by LibTorch, cuDNN, and Eigen.☆31Jun 6, 2023Updated 3 years ago
- Analyze TensorFlow source code☆19Mar 13, 2017Updated 9 years ago
- Mirror of MAGMA - Next-generation linear algebra libraries for heterogeneous architectures. Please use the official repository, https://b…☆25Aug 29, 2017Updated 8 years ago
- Memory footprint reduction for transformer models☆11Jan 24, 2023Updated 3 years ago
- A lightweight deep learning framework made with ❤️☆33May 24, 2019Updated 7 years ago
- ☆16Aug 18, 2015Updated 10 years ago
- 自动上传markdown到知乎☆21Nov 29, 2018Updated 7 years ago
- Bare Metal GPUs on DigitalOcean Gradient AI • AdPurpose-built for serious AI teams training foundational models, running large-scale inference, and pushing the boundaries of what's possible.
- AutodiffEngine☆13Apr 1, 2019Updated 7 years ago
- Improving-Deep-Neural-Networks☆12Dec 31, 2017Updated 8 years ago
- antkillerfarm's crazy magic☆18Oct 3, 2024Updated last year
- ☆10Sep 23, 2023Updated 2 years ago
- Example code from Parallel Programming in C with MPI and OpenMP☆11Feb 24, 2021Updated 5 years ago
- 可以随机生成制定数量的车牌号,因为用到停车场的虚假数据生成,所以地区集中在一个地方。支持各类车辆的生成,只需在注释的地方修改即可。☆10May 30, 2021Updated 5 years ago
- ☆19Jan 4, 2024Updated 2 years ago
- DeepSparkHub selects hundreds of application algorithms and models, covering various fields of AI and general-purpose computing, to suppo…☆73Aug 3, 2026Updated last week
- Generic C-like Preprocessor for Rust☆14Jun 27, 2023Updated 3 years ago
- Deploy open-source AI quickly and easily - Special Bonus Offer • AdRunpod Hub is built for open source. One-click deployment and autoscaling endpoints without provisioning your own infrastructure.
- Fast CUDA Kernels for ResNet Inference.☆183May 26, 2019Updated 7 years ago
- Permuterm index system for effective document retrieval of wild card queries.☆14Mar 25, 2015Updated 11 years ago
- Environment control for benchmarks☆14Feb 10, 2025Updated last year
- Public repository made for Automated Feature Engineering workshop (Summer Data Conf, Odessa, 2018-07-21)☆19Jul 16, 2018Updated 8 years ago
- ☆11Apr 10, 2022Updated 4 years ago
- PyTorch-UVM on super-large language models.☆17Dec 21, 2020Updated 5 years ago
- Notes of paper reading☆20Dec 12, 2021Updated 4 years ago
- ☆17Aug 16, 2025Updated last year
- Optimized Parallel Tiled Approach to perform 2D Convolution by taking advantage of the lower latency, higher bandwidth shared memory as w…☆15Oct 17, 2017Updated 8 years ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- ☆10Mar 20, 2019Updated 7 years ago
- BytePS examples (Vision, NLP, GAN, etc)☆19Nov 24, 2022Updated 3 years ago
- [ICML 2025] LaCache: Ladder-Shaped KV Caching for Efficient Long-Context Modeling of Large Language Models☆16Nov 4, 2025Updated 9 months ago
- Social Disatancing Monitor using yolov3 and DPU HW acceleration for Xilinx adaptive computing challenge 2020☆12Feb 17, 2023Updated 3 years ago
- ☆20Nov 7, 2019Updated 6 years ago
- My blog☆11Jul 6, 2025Updated last year
- Scalable radix top-k selection on GPUs.☆23Jan 27, 2025Updated last year