Fast integer division with divisor not known at compile time. To be used primarily in CUDA kernels.
☆75Nov 4, 2015Updated 10 years ago
Alternatives and similar repositories for int_fastdiv
Users that are interested in int_fastdiv are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Generate simple index ranges in C++ and CUDA C++☆39Jun 14, 2023Updated 3 years ago
- ☆16Jul 28, 2021Updated 4 years ago
- My very own vxsort re-implemented with "modern" C++ by a complete idiot (in C++)☆33Jun 12, 2026Updated last month
- OpenMP offload playground☆10Nov 16, 2024Updated last year
- Computes the Henry coefficient of methane in IRMOF-1☆10Oct 5, 2021Updated 4 years ago
- Deploy open-source AI quickly and easily - Special Bonus Offer • AdRunpod Hub is built for open source. One-click deployment and autoscaling endpoints without provisioning your own infrastructure.
- Finite State Coder☆15Apr 17, 2015Updated 11 years ago
- Repository with examples for the C++20 Coroutines video and article.☆29Jul 10, 2024Updated 2 years ago
- A library to benchmark CUDA code, similar to google benchmark.☆30Apr 18, 2021Updated 5 years ago
- An implemention of parallel marching cubes algorithm by CUDA☆10Sep 23, 2021Updated 4 years ago
- Common code library☆15Feb 3, 2018Updated 8 years ago
- Artifact for 'Register Optimizations for Stencils on GPUs'☆10Sep 18, 2018Updated 7 years ago
- Massively Parallel ANS Decoding on GPUs☆30Jul 26, 2019Updated 6 years ago
- "Guided Visibility Sampling++", an aggressive from-region visibility algorithm. The implementation is based on the Vulkan graphics API.☆22Apr 18, 2021Updated 5 years ago
- A single-header C++ library for simplifying the use of CUDA Runtime Compilation (NVRTC).☆573Sep 15, 2025Updated 10 months ago
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- ☆47Updated this week
- Materials for workshop on GPU computation for statistics, data science, machine learning applications.☆14Sep 8, 2016Updated 9 years ago
- Gray-Scott reaction-diffusion system in 3D using CUDA☆12Jun 8, 2019Updated 7 years ago
- ☆10Oct 1, 2024Updated last year
- Cahn Hilliard CUDA (Phase-Field Simulation of Spinodal Decomposition)☆13Jul 4, 2019Updated 7 years ago
- LOGAN: High-Performance Multi-GPU X-Drop Long-Read Alignment.☆30Sep 23, 2022Updated 3 years ago
- C implementation of the Landau-Vishkin algorithm☆36Apr 8, 2022Updated 4 years ago
- ☆16Dec 24, 2024Updated last year
- Sweep and Tiniest Queue & Tight-Inclusion GPU CCD☆21Jul 16, 2026Updated last week
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- Network based loader and flasher for Pano G2 devices☆15Jul 8, 2023Updated 3 years ago
- The source code of the paper An Eigenanalysis of Angle-Based Deformation Energies☆18Sep 12, 2023Updated 2 years ago
- CUDA Kernel Benchmarking Library☆905Updated this week
- RISC-V System on Chip Builder☆12Sep 27, 2020Updated 5 years ago
- ☆15Nov 30, 2023Updated 2 years ago
- THIS REPOSITORY HAS MOVED TO github.com/nvidia/cub, WHICH IS AUTOMATICALLY MIRRORED HERE.☆87Feb 21, 2024Updated 2 years ago
- Homemade Pixel Art Tool (WIP)☆17Oct 18, 2024Updated last year
- Scalable Integer Sort application for co-design in the exascale era☆19Apr 12, 2021Updated 5 years ago
- Taichi Implementation of "The Power Particle-in-Cell Method"☆21Aug 21, 2022Updated 3 years ago
- Open source password manager - Proton Pass • AdSecurely store, share, and autofill your credentials with Proton Pass, the end-to-end encrypted password manager trusted by millions.
- An Attention Superoptimizer☆22Jan 20, 2025Updated last year
- This is an enhanced GPU MPM framework with explicit solver☆19Mar 20, 2026Updated 4 months ago
- A simple example showing how to implement a DDA based screen-space ray marcher in Unity☆13Apr 18, 2017Updated 9 years ago
- Starlight: A Kernel Optimizer for GPU Processing☆16Jan 10, 2024Updated 2 years ago
- ☆32Mar 26, 2021Updated 5 years ago
- Zero-Overhead bare-metal GPGPU library for C++ on Windows.☆15Jan 29, 2017Updated 9 years ago
- Odyssey: a public, GPU-based GRRT (general relativistic radiative transfer) code☆22Aug 19, 2022Updated 3 years ago