parallel algorithm based on cuda
☆60Nov 27, 2017Updated 8 years ago
Alternatives and similar repositories for algorithms-cuda
Users that are interested in algorithms-cuda are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Shared memory overlap-and-save method for NVIDIA GPUs using CUDA☆18Aug 21, 2025Updated 11 months ago
- General Industry Camera Driver☆13Dec 22, 2020Updated 5 years ago
- Programming Test☆13Aug 17, 2015Updated 10 years ago
- ICML2017 MEC: Memory-efficient Convolution for Deep Neural Network C++实现(非官方)☆17Apr 9, 2019Updated 7 years ago
- Implementation of 3d non-separable convolution using CUDA & FFT Convolution☆20Jan 15, 2019Updated 7 years ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- darknet with CMake both linux and windows☆17Dec 19, 2018Updated 7 years ago
- 我的LeetCode做题代码分方向汇总集合☆37Feb 4, 2019Updated 7 years ago
- A simple implement for A Novel Approach For Finger Vein Verification Based on Self-Taught Learning https://arxiv.org/pdf/1508.03710.pdf☆11May 10, 2018Updated 8 years ago
- ☆270Jan 14, 2018Updated 8 years ago
- A pattern-based algorithmic autotuner for graph processing on GPUs.☆33Jun 25, 2025Updated last year
- CUDA implementation of complex single precision float FIR filter☆20Feb 20, 2019Updated 7 years ago
- Documentations for RELION☆15Mar 13, 2026Updated 4 months ago
- Code for the blog post on few-shot classification via task representation and communication.☆18May 24, 2017Updated 9 years ago
- Sim-to-Real via Sim-to-Sim using fast.ai's U-net☆10Nov 25, 2019Updated 6 years ago
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- Some source code about matrix multiplication implementation on CUDA☆34Sep 12, 2018Updated 7 years ago
- Mixed-Radix DIT FFT in C++11☆13Aug 27, 2018Updated 7 years ago
- GPU-enhanced parallel implementation of single particle cryo-EM image processing☆12Oct 2, 2017Updated 8 years ago
- This code generates the filter weights for polyphase filter banks with arbitrary numbers of channels, and with configurable windows.☆31Jul 10, 2024Updated 2 years ago
- Implementing the Wasserstein Loss Layer☆12Jan 19, 2016Updated 10 years ago
- A Rust style C++ library.☆19Sep 3, 2022Updated 3 years ago
- TensorRT Acceleration for PyTorch Native Eager Mode Quantization Models☆17Jul 22, 2024Updated 2 years ago
- A Simple and Efficient FFT Implementation in C☆11Aug 18, 2018Updated 7 years ago
- ☆10Jan 13, 2021Updated 5 years ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- Semantic point cloud segmentation with graph convolutional network☆10Dec 10, 2017Updated 8 years ago
- SC'25 UltraAttn: Efficiently Parallelizing Attention through Hierarchical Context-Tiling☆16Aug 14, 2025Updated 11 months ago
- GHive: Accelerating Analytical Query Processing in Apache Hive via CPU-GPU Heterogeneous Computing.☆14Nov 8, 2023Updated 2 years ago
- A Caffe implementation of PSROI-Align☆55Jan 1, 2018Updated 8 years ago
- 使用sklearn.ensemble.RandomForestRegressor和GridSearchCV进行成人死亡率预测☆12Sep 26, 2022Updated 3 years ago
- Binary image skeletonization algorithms for CPUs and Nvidia GPUs☆12Jun 15, 2017Updated 9 years ago
- Tensorflow implementation of deformable conv and pooling operations.☆10Jul 17, 2017Updated 9 years ago
- PyTorch code for full quantization of DNN using BCGD☆14Jul 24, 2019Updated 6 years ago
- IN2118 Databases Implementation on Modern CPU Architectures, SS 2020, TUM☆20Oct 10, 2020Updated 5 years ago
- Simple, predictable pricing with DigitalOcean hosting • AdAlways know what you'll pay with monthly caps and flat pricing. Enterprise-grade infrastructure trusted by 600k+ customers.
- Convolutional Neural Network of vgg19 model using Cuda to accelerate☆12Jun 11, 2018Updated 8 years ago
- MPI and MPI - CUDA accelerated Huffman encoding☆10Jul 26, 2017Updated 8 years ago
- StarPU Runtime system☆16Sep 22, 2010Updated 15 years ago
- Pytorch implementation of Blazingly Fast Video Object Segmentation with Pixel-Wise Metric Learning (Chen et al)☆27Jun 23, 2018Updated 8 years ago
- CrypTool 1 (CT1) is a free Windows program for cryptography and cryptanalysis☆18Jun 13, 2024Updated 2 years ago
- Large matrix multiplication in CUDA☆17Oct 20, 2023Updated 2 years ago
- Standalone SDR experiment using multicore MCU☆11Apr 12, 2018Updated 8 years ago