Unit benchmarks of CUDA event APIs.
☆17Apr 23, 2024Updated 2 years ago
Alternatives and similar repositories for cuda_event_benchmark
Users that are interested in cuda_event_benchmark are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- outline and links for PLDI 2022 tutorial☆17Jun 13, 2022Updated 4 years ago
- Collection of CUDA benchmarks, with a focus on unified vs. explicit memory management.☆21Oct 15, 2019Updated 6 years ago
- ☆43Nov 15, 2025Updated 8 months ago
- Light SIMD library for modern C++☆13Mar 4, 2023Updated 3 years ago
- CUDA tool set for non-C++ languages that provides similar functionality like Thrust, with NVRTC at its core.☆59Aug 13, 2022Updated 4 years ago
- Deploy open-source AI quickly and easily - Special Bonus Offer • AdRunpod Hub is built for open source. One-click deployment and autoscaling endpoints without provisioning your own infrastructure.
- An HPL-AI implementation for Fugaku☆24Jun 29, 2021Updated 5 years ago
- Microbenchmarks showing relative performance of different Python functions/patterns.☆13Oct 3, 2025Updated 10 months ago
- ☆25Jun 24, 2022Updated 4 years ago
- A parser for PTX 6.5☆13Jun 19, 2023Updated 3 years ago
- HPC Game Platform☆11Apr 20, 2023Updated 3 years ago
- Generate simple index ranges in C++ and CUDA C++☆39Jun 14, 2023Updated 3 years ago
- A Rust interface to the x2apic interrupt architecture.☆17Feb 14, 2025Updated last year
- ☆11Jun 9, 2023Updated 3 years ago
- tensorflow fork with Salus integration☆12Jan 7, 2022Updated 4 years ago
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- Adaptive consistency replication with reinforcement learning for large scale globally distributed storage.☆13Sep 29, 2025Updated 10 months ago
- A library for storing interpolatable vector fields on co-processors☆17May 26, 2026Updated 2 months ago
- ☆72Jun 23, 2020Updated 6 years ago
- constant-size associative container backed by a simple array☆20Aug 6, 2023Updated 3 years ago
- MLIR dialect for libgccjit☆24Dec 3, 2024Updated last year
- Benchmark for popular fft libaries - fftw | cufftw | cufft☆18Dec 8, 2018Updated 7 years ago
- ☆41Nov 28, 2022Updated 3 years ago
- SIMD aligned data structures to work with `std::simd`.☆11Dec 14, 2024Updated last year
- A C++ smart pointer with copy-on-write semantics☆15Apr 4, 2016Updated 10 years ago
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- High-Performance Linpack Benchmark adopted version for GPU backend☆12Sep 12, 2022Updated 3 years ago
- PETSc Interface for Octave and MATLAB (Deprecated)☆10Nov 10, 2022Updated 3 years ago
- A pointer type for heap-allocated objects which heap storage can be re-used☆15Sep 8, 2024Updated last year
- Simple starter CMake project that uses NVBench.☆15May 6, 2025Updated last year
- Proposed fixit commands for cargo☆65Updated this week
- HTML/JS port of CUDA Occupancy Calculator☆17Nov 23, 2021Updated 4 years ago
- ☆77May 29, 2019Updated 7 years ago
- Nanoscale logging library in C++11☆14Dec 15, 2021Updated 4 years ago
- 2022 ECS CloudBuild Distributed Cache Contest - Final Round https://tianchi.aliyun.com/competition/entrance/531982/introduction☆17Dec 8, 2022Updated 3 years ago
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- Data structures for ASTs☆14Dec 6, 2022Updated 3 years ago
- nd-tree data structures and algorithms☆15Sep 29, 2015Updated 10 years ago
- Efficient-Tensor-Management-on-HM-for-Deep-Learning☆11Nov 15, 2021Updated 4 years ago
- Odds and ends — collection miscellania. Extra functionality for slices, strings and other things☆22Apr 11, 2020Updated 6 years ago
- Parallel Graph Input Output☆19Jul 18, 2023Updated 3 years ago
- Provides truly zero-cost alternatives to Iterator::step_by for both incrementing and decrementing any type that satisfies RangeBounds<T: …☆13Jan 5, 2022Updated 4 years ago
- ☆10Aug 4, 2022Updated 4 years ago