A GPU benchmark suite for autotuners
☆19Feb 20, 2024Updated 2 years ago
Alternatives and similar repositories for BAT
Users that are interested in BAT are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Parse data and generate plotting scripts based on plotly.☆12Jun 4, 2026Updated 2 months ago
- Instructions and templates for SC authors☆18Aug 22, 2021Updated 5 years ago
- ☆17Dec 8, 2023Updated 2 years ago
- ☆19Nov 4, 2020Updated 5 years ago
- CUDA/HIP header-only library for low-precision (16 bit, 8 bit) and vectorized GPU kernel development☆24Updated this week
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- OCCA Python API: JIT Compilation for Multiple Architectures☆11Dec 20, 2019Updated 6 years ago
- A little library giving you a live monitoring of MPI programs.☆25Oct 23, 2022Updated 3 years ago
- Record GPU memory accesses of a CUDA program and visualize the access pattern in a browser☆13Nov 17, 2020Updated 5 years ago
- A Symbolic Emulator for Shuffle Synthesis on the NVIDIA PTX Code☆16Mar 19, 2023Updated 3 years ago
- Virtual programming language☆10Dec 5, 2022Updated 3 years ago
- Distributed machine learning platform☆13Aug 20, 2015Updated 11 years ago
- Kernel Tuning Toolkit☆72Updated this week
- Base container for developing C++ and Fortran HPC applications☆18Jun 14, 2022Updated 4 years ago
- simple port of hpl-2.0 to use NVIDIA GPU accelation with CUBLAS☆29May 13, 2013Updated 13 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- SIMULATeQCD is a multi-GPU Lattice QCD framework that makes it easy for physicists to implement lattice QCD formulas while still providin…☆40Updated this week
- PIRA - Automatic Instrumentation Refinement☆18Mar 28, 2024Updated 2 years ago
- This repository contains the figures, tables and source code in the ICS'24 paper: "Accelerated Auto-Tuning of GPU Kernels for Tensor Comp…☆10Dec 5, 2024Updated last year
- Dark channel Haze removal algorithm with CUDA acceleration (typically 10x or more speedup using a Nvidia GPU)☆14Dec 7, 2017Updated 8 years ago
- A quick way of spawning many batch jobs☆14Oct 24, 2022Updated 3 years ago
- RISC-V vector extension ISA simulation☆18Jun 11, 2019Updated 7 years ago
- Ansible config for Cluster in the Cloud☆11Apr 25, 2024Updated 2 years ago
- nVidia's CUDA accelerated Spin Transformations of Discrete Surfaces, based on the original code and paper by Keenan Crane, Ulrich Pinkall…☆17Mar 14, 2018Updated 8 years ago
- TUI for browsing, canceling, and inspecting SLURM jobs☆13Nov 13, 2023Updated 2 years ago
- Open source password manager - Proton Pass • AdSecurely store, share, and autofill your credentials with Proton Pass, the end-to-end encrypted password manager trusted by millions.
- Runs a single CUDA/OpenCL kernel, taking its source from a file and arguments from the command-line☆26Jun 10, 2026Updated 2 months ago
- cuPC: CUDA-based Parallel PC Algorithm for Causal Structure Learning on GPU☆17Mar 19, 2021Updated 5 years ago
- A fast alternative to the standard C/C++ pow() function. With adjustable accuracy-space tradeoff.☆14Jul 12, 2013Updated 13 years ago
- ☆12May 18, 2024Updated 2 years ago
- Scripts for running various benchmarks on Isambard and other systems.☆29May 13, 2021Updated 5 years ago
- Enhancing the convergence speed by 2x and improving the training success of Physics-Informed Neural Networks (PINNs).☆13Oct 14, 2024Updated last year
- Escoin: Efficient Sparse Convolutional Neural Network Inference on GPUs☆16Feb 28, 2019Updated 7 years ago
- WIP · CUDA compatibility for Blaze · https://bitbucket.org/blaze-lib/blaze☆21Nov 18, 2019Updated 6 years ago
- Practice algorithms and data structure on different languages☆14Aug 21, 2022Updated 4 years ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- PyDTNN - Python Distributed Training of Neural Networks☆14Updated this week
- ☆13Sep 19, 2024Updated last year
- ☆20Sep 28, 2024Updated last year
- Tool to detect and report leaked MPI objects like MPI_Requests and MPI_Datatypes☆13Sep 17, 2014Updated 11 years ago
- A cross-platform visualization prototyping framework☆56Jul 16, 2026Updated last month
- Ansible role for managing Dell PowerConnect switches☆20Oct 31, 2023Updated 2 years ago
- pyCUDA implementation of forward propagation for Convolutional Neural Networks☆18Jan 4, 2019Updated 7 years ago