☆47Jul 16, 2025Updated last year
Alternatives and similar repositories for cuda-rt-hook
Users that are interested in cuda-rt-hook are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Hyperparameter: The High-Performance Configuration Library for AI Systems☆22Dec 14, 2025Updated 8 months ago
- TiledLower is a Dataflow Analysis and Codegen Framework written in Rust.☆13Nov 23, 2024Updated last year
- FA4-based Relative Attention Kernel developed by TML and Colfax☆18Jul 17, 2026Updated last month
- Study materials collected while studying☆51Apr 16, 2022Updated 4 years ago
- Handwritten GEMM using Intel AMX (Advanced Matrix Extension)☆17Jan 11, 2025Updated last year
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- High-performance GPU kernels for Ads and Recsys model training, independently implemented and optimized for real-world workloads and mode…☆41Aug 25, 2026Updated last week
- ☆20Jul 3, 2026Updated 2 months ago
- ☆12Sep 11, 2020Updated 5 years ago
- Muon in Int8 Precision Made Possible☆20Jun 18, 2026Updated 2 months ago
- Tmux sidebar for vibe coding. Manage sessions and monitor agents at a glance☆15Aug 25, 2026Updated last week
- Fast OS-level support for GPU checkpoint and restore☆290Sep 28, 2025Updated 11 months ago
- ☆37Aug 7, 2025Updated last year
- Hooked CUDA-related dynamic libraries by using automated code generation tools.☆173Dec 12, 2023Updated 2 years ago
- A practical way of learning Swizzle☆45Feb 3, 2025Updated last year
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- Synchronizing Claude Code conversations across machines☆17Updated this week
- Artifacts for ATC '22 paper "Faster Software Packet Processing on FPGA NICs with eBPF Program Warping"☆17May 20, 2022Updated 4 years ago
- ☆47Dec 13, 2024Updated last year
- WeChat official account crawler 微信公众号爬虫☆13Apr 13, 2024Updated 2 years ago
- Artifact evaluation repo for EuroSys'24.☆30Nov 7, 2023Updated 2 years ago
- Debug print operator for cudagraph debugging☆18Aug 2, 2024Updated 2 years ago
- A survey of manufacturer-provided DRAM operating parameters and timings as specified by DRAM chip datasheets from between 1970 and 2021. …☆11May 4, 2022Updated 4 years ago
- Paper list of federated learning: About system design☆13Apr 13, 2022Updated 4 years ago
- ☆11Apr 3, 2023Updated 3 years ago
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- ☆22Jul 16, 2022Updated 4 years ago
- ☆34Nov 7, 2022Updated 3 years ago
- ☆32Jul 2, 2025Updated last year
- My love.☆28Apr 1, 2026Updated 5 months ago
- Automatic virtualization of (general) accelerators.☆47Nov 28, 2022Updated 3 years ago
- TileFusion is an experimental C++ macro kernel template library that elevates the abstraction level in CUDA C for tile processing.☆117Aug 4, 2026Updated last month
- ☆39Dec 14, 2025Updated 8 months ago
- The source code for paper LeCo: Lightweight Compression via Learning Serial Correlations (SIGMOD'24).☆17Mar 26, 2024Updated 2 years ago
- An In-kernel Transparent Monitoring System for Microservice Systems with eBPF☆22Sep 11, 2022Updated 3 years ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- An external memory allocator example for PyTorch.☆16Aug 10, 2025Updated last year
- Repo for OSDI 2023 paper: "Ship your Critical Section Not Your Data: Enabling Transparent Delegation with TCLocks"☆21Nov 6, 2024Updated last year
- DINT: Fast In-Kernel Distributed Transactions with eBPF☆52Jul 6, 2024Updated 2 years ago
- ☆11May 13, 2025Updated last year
- Dynamic Memory Management for Serving LLMs without PagedAttention☆519Aug 24, 2026Updated last week
- 通过系统编程学习Rust☆11Mar 8, 2022Updated 4 years ago
- Benchmark tests supporting the TiledCUDA library.☆19Nov 19, 2024Updated last year