Record GPU memory accesses of a CUDA program and visualize the access pattern in a browser
☆13Nov 17, 2020Updated 5 years ago
Alternatives and similar repositories for cuda-memory-access-recorder
Users that are interested in cuda-memory-access-recorder are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Experimental iOS app embedding Fabrice Bellard's TinyEMU☆14May 26, 2020Updated 6 years ago
- The Haiku operating system. (Pull requests will be ignored; patches may be sent to https://review.haiku-os.org).☆12May 5, 2026Updated 3 months ago
- Levarage shared memory buffers between QEMU VMs in go☆12Aug 15, 2024Updated 2 years ago
- Play-with-compiler sandbox based on PWD☆10Oct 22, 2020Updated 5 years ago
- Goal: a website to automatically train and certify compiler researchers and developers☆10Nov 24, 2019Updated 6 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- Some simple examples of using SIMD CPU instructions.☆15Dec 16, 2018Updated 7 years ago
- ☆10May 20, 2022Updated 4 years ago
- We try to put source files of llvm tutorials here☆18Oct 6, 2020Updated 5 years ago
- Accelerate database with GPU☆13Dec 30, 2013Updated 12 years ago
- BinDex: A Two-Layered Index for Fast and Robust Scans (SIGMOD2020)☆10Jun 5, 2020Updated 6 years ago
- OCCA Python API: JIT Compilation for Multiple Architectures☆11Dec 20, 2019Updated 6 years ago
- Using TensorFlow for physics-informed neural networks for scientific machine learning (SciML)☆16Nov 30, 2020Updated 5 years ago
- Simple WebSocket server for QuickJS☆18Jul 11, 2021Updated 5 years ago
- Accelerating CNN's convolution operation on GPUs by using memory-efficient data access patterns.☆14Dec 8, 2017Updated 8 years ago
- Serverless GPU API endpoints on Runpod - Get Bonus Credits • AdSkip the infrastructure headaches. Auto-scaling, pay-as-you-go, no-ops approach lets you focus on innovating your application.
- Sample database with LLVM execution engine.☆11Jul 13, 2016Updated 10 years ago
- Linux dotfiles☆10Aug 5, 2026Updated last week
- A Minecraft demo in Haskell ;)☆12Mar 11, 2016Updated 10 years ago
- A CLI tool for getting screenshots of URLs using headless chrome☆26Apr 23, 2023Updated 3 years ago
- ☆13Feb 5, 2022Updated 4 years ago
- A library for extending `Command` in order to execute programs more easily.☆20Aug 1, 2026Updated 2 weeks ago
- 🔮 High-performance kaleidoscope effects for real-time applications☆15Aug 1, 2026Updated 2 weeks ago
- A graph coloring register allocator for LLVM.☆11Jan 23, 2017Updated 9 years ago
- Speech recognition with federated learning☆11Jan 9, 2020Updated 6 years ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- An llvm pass for counting global uncoalesced acceses for cuda code via dynamic analysis.☆14Nov 17, 2018Updated 7 years ago
- Google Earth Pro image extractor and alignment☆13Feb 9, 2018Updated 8 years ago
- Multi-GPU (CUDA-MPI) baseline implementation of Heat Equation and the inviscid Burgers' equation☆12Oct 17, 2017Updated 8 years ago
- A reference implementation of GPT-2 in Python, for teaching ML compilers☆23Jul 27, 2026Updated 3 weeks ago
- Apache NuttX RTOS in the Web Browser: TinyEMU with VirtIO☆27Mar 3, 2024Updated 2 years ago
- Minimal implementation of Contrastive Predictive Coding for audio.☆18Nov 17, 2019Updated 6 years ago
- ☆12May 30, 2021Updated 5 years ago
- Dynamic Control Flow Recovery☆25Apr 15, 2018Updated 8 years ago
- ☆13Feb 5, 2018Updated 8 years ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- Host software for running SSITH processors on AWS F1 FPGAs☆20Jul 20, 2021Updated 5 years ago
- Base container for developing C++ and Fortran HPC applications☆18Jun 14, 2022Updated 4 years ago
- Protecting Real-Time GPU Kernels on Integrated CPU-GPU SoC Platforms☆12Apr 9, 2018Updated 8 years ago
- A simple way to send realtime low latency high quality video over Internet☆13Jul 3, 2020Updated 6 years ago
- Analysis and investigating the confounding effect of accents in end-to-end Automatic Speech Recognition models.☆15Jun 27, 2020Updated 6 years ago
- Implementation of marching cubes in unity with c# job system☆11Oct 29, 2019Updated 6 years ago
- Presenting Collection of Pretrained Models. Links to pretrained models in NLP and voice.☆23Dec 27, 2019Updated 6 years ago