Artifact from "Hardware Compute Partitioning on NVIDIA GPUs". THIS IS A FORK OF BAKITAS REPO. I AM NOT ONE OF THE AUTHORS OF THE PAPER.
☆68Nov 24, 2025Updated 9 months ago
Alternatives and similar repositories for libsmctrl
Users that are interested in libsmctrl are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- ☆34Jul 13, 2026Updated 2 months ago
- ☆27Aug 19, 2022Updated 4 years ago
- Tutorials for NVIDIA CUPTI samples☆73Jul 22, 2026Updated 2 months ago
- ☆23Sep 6, 2026Updated 2 weeks ago
- An interference-aware scheduler for fine-grained GPU sharing☆164Nov 26, 2025Updated 9 months ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- a simple API to use CUPTI☆10Aug 19, 2025Updated last year
- Research prototype of PRISM — a cost-efficient multi-LLM serving system with flexible time- and space-based GPU sharing.☆77Mar 17, 2026Updated 6 months ago
- A GPU-accelerated DNN inference serving system that supports instant kernel preemption and biased concurrent execution in GPU scheduling.☆43May 29, 2022Updated 4 years ago
- [NeurIPS 2025] ClusterFusion: Expanding Operator Fusion Scope for LLM Inference via Cluster-Level Collective Primitive☆77Aug 8, 2026Updated last month
- ☆12Nov 5, 2024Updated last year
- ☆271Dec 25, 2025Updated 8 months ago
- ☆34Sep 9, 2020Updated 6 years ago
- A highly-flexible GPU simulator for AMD GPUs.☆267Sep 11, 2026Updated last week
- Code for paper "Modality Plug-and-Play: Elastic Modality Adaptation in Multimodal LLMs for Embodied AI"☆13Jan 19, 2024Updated 2 years ago
- GPUs on demand by Runpod - Special Offer Available • AdRun AI, ML, and HPC workloads on powerful cloud GPUs—without limits or wasted spend. Deploy GPUs in under a minute and pay by the second.
- ☆12Aug 17, 2022Updated 4 years ago
- eBPF for GPU UVM offloading and scheduling in Linux kernel☆69Sep 10, 2026Updated last week
- Virtualized Elastic KV Cache for Dynamic GPU Sharing and Beyond☆1,477Updated this week
- ☆25May 18, 2025Updated last year
- ☆79May 4, 2021Updated 5 years ago
- Automatic Parallelism Using LLVM☆10Aug 2, 2014Updated 12 years ago
- Medusa: Accelerating Serverless LLM Inference with Materialization [ASPLOS'25]☆49May 13, 2025Updated last year
- collection of benchmarks to measure basic GPU capabilities☆539Oct 24, 2025Updated 10 months ago
- [ACM EuroSys 2023] Fast and Efficient Model Serving Using Multi-GPUs with Direct-Host-Access☆56Aug 6, 2025Updated last year
- Deploy open-source AI quickly and easily - Special Bonus Offer • AdRunpod Hub is built for open source. One-click deployment and autoscaling endpoints without provisioning your own infrastructure.
- Efficient and easy multi-instance LLM serving☆562Mar 12, 2026Updated 6 months ago
- Experiments evaluating preemption on the NVIDIA Pascal architecture☆16Nov 10, 2016Updated 9 years ago
- Allow torch tensor memory to be released and resumed later☆275Updated this week
- Dynamic Memory Management for Serving LLMs without PagedAttention☆524Aug 24, 2026Updated 3 weeks ago
- GeminiFS: A Companion File System for GPUs☆85Aug 11, 2026Updated last month
- Hooked CUDA-related dynamic libraries by using automated code generation tools.☆173Dec 12, 2023Updated 2 years ago
- An efficient storage system for concurrent graph processing☆10Feb 1, 2021Updated 5 years ago
- [ASPLOS' 26] TetriServe: Efficiently Serving Mixed DiT Workloads☆19Mar 12, 2026Updated 6 months ago
- ☆89Apr 18, 2025Updated last year
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- Tacker: Tensor-CUDA Core Kernel Fusion for Improving the GPU Utilization while Ensuring QoS☆33Feb 10, 2025Updated last year
- Unofficial description of the CUDA assembly (SASS) instruction sets.☆250Jul 18, 2025Updated last year
- Artifacts for our NSDI'23 paper TGS☆98Jun 10, 2024Updated 2 years ago
- ☆18Apr 21, 2024Updated 2 years ago
- CUDA checkpoint and restore utility☆491Jul 6, 2026Updated 2 months ago
- Evaluation utilities based on SymPy.☆25Dec 12, 2024Updated last year
- Assembler and Decompiler for NVIDIA (Maxwell Pascal Volta Turing Ampere) GPUs.☆98Feb 23, 2023Updated 3 years ago