π· Build compute kernels
β213Apr 6, 2026Updated 4 months ago
Alternatives and similar repositories for kernel-builder
Users that are interested in kernel-builder are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Build compute kernels and load them from the Hub.β728Updated this week
- Kernel sources for https://huggingface.co/kernels-communityβ141Updated this week
- β19Dec 31, 2025Updated 7 months ago
- Hugging Face on Microsoft Azure (documentation, examples and more)β16Aug 19, 2026Updated last week
- ANE accelerated embedding models!β20Dec 11, 2024Updated last year
- GPUs on demand by Runpod - Special Offer Available β’ AdRun AI, ML, and HPC workloads on powerful cloud GPUsβwithout limits or wasted spend. Deploy GPUs in under a minute and pay by the second.
- β29May 26, 2026Updated 3 months ago
- Hugging Face Jobsβ20Jul 11, 2025Updated last year
- β12Jan 4, 2024Updated 2 years ago
- (ECCV 2026): Official code for Flash-BoN: Instant Drafts for Inference-Time Scaling in Diffusion Modelsβ21Jul 9, 2026Updated last month
- Google TPU optimizations for transformers modelsβ136Jan 23, 2026Updated 7 months ago
- β15Dec 4, 2024Updated last year
- A Python wrapper around HuggingFace's TGI (text-generation-inference) and TEI (text-embedding-inference) servers.β32Sep 19, 2025Updated 11 months ago
- Automatically derive Python dunder methods for your Rust codeβ25May 26, 2026Updated 3 months ago
- β24May 26, 2026Updated 3 months ago
- End-to-end encrypted cloud storage - Proton Drive β’ AdSpecial offer: 40% Off Yearly / 80% Off First Month. Protect your most important files, photos, and documents from prying eyes.
- Making Flux go brrr on GPUs.β172Jan 5, 2026Updated 7 months ago
- Experimental compiler for deep learning modelsβ75Sep 18, 2025Updated 11 months ago
- ModernBERT model optimized for Apple Neural Engine.β39Jan 10, 2025Updated last year
- Tile primitives for speedy kernelsβ3,658Updated this week
- Minimal implementation of a Byte Pair Encoding (BPE) tokenizer in Zigβ15Apr 7, 2025Updated last year
- Graph model execution API for Candleβ18Jul 27, 2025Updated last year
- A pytorch quantization backend for optimumβ1,054Updated this week
- A high-performance constrained decoding engine based on context free grammar in Rustβ59May 22, 2025Updated last year
- Minimalistic large language model 3D-parallelism trainingβ2,804May 26, 2026Updated 3 months ago
- End-to-end encrypted cloud storage - Proton Drive β’ AdSpecial offer: 40% Off Yearly / 80% Off First Month. Protect your most important files, photos, and documents from prying eyes.
- Rust client for the huggingface hub aiming for minimal subset of features over `huggingface-hub` python packageβ327Aug 17, 2026Updated last week
- Testbed for LLM inference with cutile-rs.β72Jul 1, 2026Updated last month
- Supporting code for "LLMs for your iPhone: Whole-Tensor 4 Bit Quantization"β11Mar 31, 2024Updated 2 years ago
- Implementation of the HuggingFace Xet Protocol.β28Aug 13, 2026Updated 2 weeks ago
- β25Sep 4, 2025Updated 11 months ago
- A repo of resource for the GPU Mode talk on OpenEnv.β16Jan 14, 2026Updated 7 months ago
- Mount Hugging Face Buckets and repos as local filesystems. No download, no copy, no waiting.β789Aug 16, 2026Updated 2 weeks ago
- xet client tech, used in huggingface_hubβ564Updated this week
- Efficient implementation of DeepSeek Ops (Blockwise FP8 GEMM, MoE, and MLA) for AMD Instinct MI300Xβ80Feb 11, 2026Updated 6 months ago
- Managed hosting for WordPress and PHP on Cloudways β’ AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- π€ HuggingFace Inference Toolkit for Google Cloud Vertex AI (similar to SageMaker's Inference Toolkit, but for Vertex AI and unofficial)β17Mar 20, 2024Updated 2 years ago
- PyTorch native quantization for training and inferenceβ2,959Updated this week
- FlashInfer: Kernel Library for LLM Servingβ6,281Updated this week
- Kernels, of the mega variety :)β814May 26, 2026Updated 3 months ago
- Tritonbench is a collection of PyTorch custom operators with example inputs to measure their performance.β366Updated this week
- Helpful tools and examples for working with flex-attentionβ1,229Updated this week
- π€ Tokenizers.js: A pure JS/TS implementation of today's most used tokenizersβ55Jul 28, 2026Updated last month