Use CUDA intrinsics with user-defined types
☆48Aug 14, 2014Updated 11 years ago
Alternatives and similar repositories for generics
Users that are interested in generics are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Full-speed Array of Structures access☆177Apr 25, 2023Updated 3 years ago
- Generalized Histograms for CUDA-capable GPUs☆43Sep 1, 2015Updated 10 years ago
- A Sound and Complete Verification Tool for Warp-Specialized GPU Kernels☆19Jun 17, 2015Updated 11 years ago
- An open-source attempt at a GPU-accelerated Adaptive Support Weight stereo matching algorithm☆13May 31, 2016Updated 10 years ago
- ☆18May 24, 2016Updated 10 years ago
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- 基于MxDraw+Dotnetbar开发的cad二次开发demo,实现了基本框架的搭建☆19Mar 18, 2018Updated 8 years ago
- An experimental game-engine☆13Jul 9, 2023Updated 3 years ago
- ☆16Jul 28, 2021Updated 5 years ago
- A Simple Autoencoder Implementation in C++☆15Jan 10, 2017Updated 9 years ago
- ☆18Aug 22, 2019Updated 6 years ago
- A fast and highly scalable GPU dynamic memory allocator☆111Mar 11, 2015Updated 11 years ago
- A 128 bit unsigned integer class for CUDA☆46Jan 3, 2025Updated last year
- [ARCHIVED] Cooperative primitives for CUDA C++. See https://github.com/NVIDIA/cccl☆1,841Oct 9, 2023Updated 2 years ago
- This repository mirrors the principal Gitlab repository of the Chebyshev Accelerated Subspace iteration Eigensolver. If you want to contr…☆21Jul 8, 2026Updated last month
- Serverless GPU API endpoints on Runpod - Get Bonus Credits • AdSkip the infrastructure headaches. Auto-scaling, pay-as-you-go, no-ops approach lets you focus on innovating your application.
- High speed client-server☆11May 27, 2017Updated 9 years ago
- GPU Optimization and Memory Abstraction Framework☆33Oct 31, 2019Updated 6 years ago
- SAP API Wrapper with C#☆20Apr 1, 2018Updated 8 years ago
- Facebook's CUDA extensions.☆284Mar 27, 2019Updated 7 years ago
- ☆10Dec 9, 2015Updated 10 years ago
- CUDA Flux is a profiler for GPU applications which reports the basic block executions frequencies of compute kernels☆33Mar 15, 2021Updated 5 years ago
- A C++ memory context☆11Jul 28, 2021Updated 5 years ago
- An example of Typescript Decorators usage☆10Aug 10, 2015Updated 11 years ago
- A vertex-centric CUDA/C++ API for large graph analytics on GPUs using the Gather-Apply-Scatter abstraction☆24May 4, 2014Updated 12 years ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- maxas Scott Grey's maxas assembler sgemm explaining the (for me) missing parts https://github.com/NervanaSystems/maxas☆17Dec 22, 2018Updated 7 years ago
- Exercises and sample code for a C# SIMD (vectorization) workshop☆38Oct 11, 2017Updated 8 years ago
- Google Calendar API v3. Haskell implementation☆12Sep 27, 2014Updated 11 years ago
- Lua wrapper for ArrayFire☆10Feb 27, 2017Updated 9 years ago
- Binary Stereo Matching ICPR 2012☆14Jul 15, 2014Updated 12 years ago
- Smoothing while preserving edges and coherent structures in images with anisotropic diffusion using lattice basis reduction☆10Apr 25, 2026Updated 3 months ago
- Find file by zoxide☆11Oct 19, 2024Updated last year
- Simulated Video Feedback on the GPU☆10Mar 28, 2019Updated 7 years ago
- OBD BT-ELM327-BMW☆34Jan 23, 2014Updated 12 years ago
- Open source password manager - Proton Pass • AdSecurely store, share, and autofill your credentials with Proton Pass, the end-to-end encrypted password manager trusted by millions.
- ☆12Sep 29, 2021Updated 4 years ago
- Automated PS3 integration tests for emulators☆17Mar 23, 2016Updated 10 years ago
- Portably Performant Physical Algebra☆12Jan 25, 2023Updated 3 years ago
- Radix sort analyses in parallel and serial ways.☆11Jan 21, 2016Updated 10 years ago
- A Nim Wrapper for ArrayFire☆28Nov 8, 2024Updated last year
- neon - nGraph integration☆13Jan 3, 2023Updated 3 years ago
- A simple profiler to count Nvidia PTX assembly instructions of OpenCL/SYCL/CUDA kernels for roofline model analysis.☆59Mar 20, 2025Updated last year