☆25Feb 14, 2026Updated 6 months ago
Alternatives and similar repositories for KernelMem
Users that are interested in KernelMem are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Official Repo of CudaForge☆88Dec 2, 2025Updated 9 months ago
- ICCAD'23 Best Paper Award candidate: Robust GNN-based Representation Learning for HLS☆28May 23, 2024Updated 2 years ago
- A reference implementation of the Mind Mappings Framework.☆30Dec 2, 2021Updated 4 years ago
- ☆19Updated this week
- LLM4Kernel: A Survey of Large Language Models for GPU Kernel Development☆82Mar 31, 2026Updated 5 months ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- ☆13Jan 7, 2025Updated last year
- ☆12Aug 4, 2022Updated 4 years ago
- ☆18Apr 27, 2026Updated 4 months ago
- Review automated kernel generation in the era of LLMs☆309Jun 25, 2026Updated 2 months ago
- PyTorch code for our paper "AdaSVD: Adaptive Singular Value Decomposition for Large Language Models"☆16Mar 9, 2025Updated last year
- A CUDA kernel optimization toolkit for validation, benchmarking, Nsight Compute profiling, bottleneck analysis, and iterative tuning. It …☆204Sep 5, 2026Updated last week
- UniQL official repository (ICLR 2026)☆19Jan 27, 2026Updated 7 months ago
- [ICML‘25] Official code for paper "Occult: Optimizing Collaborative Communication across Experts for Accelerated Parallel MoE Training an…☆14Apr 17, 2025Updated last year
- [NeurIPS'25] KVCOMM: Online Cross-context KV-cache Communication for Efficient LLM-based Multi-agent Systems☆18Nov 1, 2025Updated 10 months ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- KernelBenchX: A Comprehensive Benchmark for Evaluating LLM-Generated GPU Kernels☆51Jun 1, 2026Updated 3 months ago
- 从 幻觉翻译 获取基于 LaTex 源码翻译的arXiv文章☆22Jul 20, 2026Updated last month
- ☆19Jul 1, 2024Updated 2 years ago
- Optimize tensor program fast with Felix, a gradient descent autotuner.☆33Mar 5, 2026Updated 6 months ago
- KernelBench: Can LLMs Write GPU Kernels? - Benchmark + Toolkit with Torch -> CUDA (+ more DSLs)☆1,239Mar 24, 2026Updated 5 months ago
- ☆10Apr 24, 2023Updated 3 years ago
- IntelliKit is a collection of intelligent tools designed to make GPU kernel development, profiling, and validation accessible to LLMs and…☆33Updated this week
- The official repository for the experiments included in the paper titled "Patch-level Routing in Mixture-of-Experts is Provably Sample-ef…☆14Feb 12, 2026Updated 7 months ago
- SJTU 中文简约 LaTeX 报告模板☆10Jun 7, 2021Updated 5 years ago
- Proton VPN Special Offer - Get 70% off • AdSpecial partner offer. Trusted by over 100 million users worldwide. Tested, Approved and Recommended by Experts.
- Autonomous GPU Kernel Generation & Optimization via Deep Agents☆545Updated this week
- [AAAI 2026] Official implementation of "FlashSVD: Memory-Efficient Inference with Streaming for Low-Rank Models". If you find this reposi…☆19Aug 26, 2026Updated 2 weeks ago
- HLS-Eval: A Benchmark and Framework for Evaluating LLMs on High-Levels Synthesis Design Tasks☆22Updated this week
- This repo implements an interface to GTAV for SCENIC language.☆11Dec 7, 2019Updated 6 years ago
- DATE'24 paper: "Hierarchical Source-to-Post-Route QoR Prediction in High-Level Synthesis with GNNs"☆21Dec 10, 2024Updated last year
- [ACM MM 2026 Oral]⚡ZEUS accelerates your diffuser. Any modality. Any model. Any scheduler. https://yixiao-wang-stats.github.io/zeus/☆22Jun 2, 2026Updated 3 months ago
- Official code for the paper "HEXA-MoE: Efficient and Heterogeneous-Aware MoE Acceleration with Zero Computation Redundancy"☆15Mar 6, 2025Updated last year
- Fast GPU based tensor core reductions☆12Jan 13, 2023Updated 3 years ago
- Keyformer proposes KV Cache reduction through key tokens identification and without the need for fine-tuning☆59Mar 26, 2024Updated 2 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- Source code of the IPDPS '21 paper: "TileSpMV: A Tiled Algorithm for Sparse Matrix-Vector Multiplication on GPUs" by Yuyao Niu, Zhengyang…☆13Aug 12, 2022Updated 4 years ago
- ☆19Oct 17, 2025Updated 10 months ago
- Artifact of Chimera☆18May 6, 2025Updated last year
- ☆16Apr 11, 2025Updated last year
- The model for edge classification by transforming edges to nodes.☆15Dec 22, 2020Updated 5 years ago
- [CVPRW2024, Official Code] for paper "Exploring AIGC Video Quality: A Focus on Visual Harmony, Video-Text Consistency and Domain Distribu…☆13Jun 14, 2024Updated 2 years ago
- This project is on how to Develop 1D Convolutional Neural Network Models for Human Activity Recognition Below is an example video of a s…☆12May 11, 2020Updated 6 years ago