☆26Feb 14, 2026Updated 7 months ago
Alternatives and similar repositories for KernelMem
Users that are interested in KernelMem are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Official Repo of CudaForge☆90Dec 2, 2025Updated 10 months ago
- ICCAD'23 Best Paper Award candidate: Robust GNN-based Representation Learning for HLS☆28May 23, 2024Updated 2 years ago
- A reference implementation of the Mind Mappings Framework.☆30Dec 2, 2021Updated 4 years ago
- ☆19Sep 24, 2026Updated last week
- LLM4Kernel: A Survey of Large Language Models for GPU Kernel Development☆82Mar 31, 2026Updated 6 months ago
- Managed Database hosting by DigitalOcean • AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- ☆13Jan 7, 2025Updated last year
- ☆12Aug 4, 2022Updated 4 years ago
- ☆13Apr 15, 2025Updated last year
- ☆18Apr 27, 2026Updated 5 months ago
- Review automated kernel generation in the era of LLMs☆313Jun 25, 2026Updated 3 months ago
- [DAC2024] Explainable Fuzzy Neural Network with Multi-Fidelity Reinforcement Learning for Micro-Architecture Design Space Exploration☆10Jul 21, 2026Updated 2 months ago
- PyTorch code for our paper "AdaSVD: Adaptive Singular Value Decomposition for Large Language Models"☆16Mar 9, 2025Updated last year
- ☆16Jan 24, 2024Updated 2 years ago
- A blog for LLVM(v11.0.0) beginner, step by step, with detailed documents and comments. Record the way I learn LLVM.☆13Jun 17, 2022Updated 4 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- Heron: Automatically Constrained High-Performance Library Generation for Deep Learning Accelerators☆25Jan 30, 2024Updated 2 years ago
- A CUDA kernel optimization toolkit for validation, benchmarking, Nsight Compute profiling, bottleneck analysis, and iterative tuning. It …☆211Sep 5, 2026Updated 3 weeks ago
- UniQL official repository (ICLR 2026)☆19Jan 27, 2026Updated 8 months ago
- ☆16Oct 28, 2025Updated 11 months ago
- [ICML‘25] Official code for paper "Occult: Optimizing Collaborative Communication across Experts for Accelerated Parallel MoE Training an…☆14Apr 17, 2025Updated last year
- [NeurIPS'25] KVCOMM: Online Cross-context KV-cache Communication for Efficient LLM-based Multi-agent Systems☆18Nov 1, 2025Updated 11 months ago
- KernelBenchX: A Comprehensive Benchmark for Evaluating LLM-Generated GPU Kernels☆52Jun 1, 2026Updated 4 months ago
- ☆21Jul 1, 2024Updated 2 years ago
- Optimize tensor program fast with Felix, a gradient descent autotuner.☆33Mar 5, 2026Updated 6 months ago
- Deploy open-source AI quickly and easily - Special Bonus Offer • AdRunpod Hub is built for open source. One-click deployment and autoscaling endpoints without provisioning your own infrastructure.
- KernelBench: Can LLMs Write GPU Kernels? - Benchmark + Toolkit with Torch -> CUDA (+ more DSLs)☆1,273Mar 24, 2026Updated 6 months ago
- ☆10Apr 24, 2023Updated 3 years ago
- The official repository for the experiments included in the paper titled "Patch-level Routing in Mixture-of-Experts is Provably Sample-ef…☆15Feb 12, 2026Updated 7 months ago
- This is Official implementation for T2S-Bench & Structure-of-Thought: Benchmarking and Prompting Comprehensive Text-to-Structure Reasonin…☆24Mar 5, 2026Updated 6 months ago
- ☆104Nov 22, 2025Updated 10 months ago
- Autonomous GPU Kernel Generation & Optimization via Deep Agents☆573Sep 8, 2026Updated 3 weeks ago
- SimPET is a framework intended to setup and launch PET imaging Monte Carlo simulations on a simple way. It uses popular tools such as Sim…☆13Aug 22, 2026Updated last month
- HLS-Eval: A Benchmark and Framework for Evaluating LLMs on High-Levels Synthesis Design Tasks☆26Sep 24, 2026Updated last week
- This repo implements an interface to GTAV for SCENIC language.☆11Dec 7, 2019Updated 6 years ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- [ACM MM 2026 Oral]⚡ZEUS accelerates your diffuser. Any modality. Any model. Any scheduler. https://yixiao-wang-stats.github.io/zeus/☆22Jun 2, 2026Updated 4 months ago
- Official code for the paper "HEXA-MoE: Efficient and Heterogeneous-Aware MoE Acceleration with Zero Computation Redundancy"☆15Mar 6, 2025Updated last year
- Fast GPU based tensor core reductions☆12Jan 13, 2023Updated 3 years ago
- Keyformer proposes KV Cache reduction through key tokens identification and without the need for fine-tuning☆59Mar 26, 2024Updated 2 years ago
- A library that assists in traversing and downloading from Wikimedia Data Dumps and their mirrors.☆12May 12, 2026Updated 4 months ago
- ☆19Oct 17, 2025Updated 11 months ago
- Artifact of Chimera☆18May 6, 2025Updated last year