A comprehensive knowledge base for Huawei Ascend NPU development, structured as distributed Agent Skills. https://ascend-ai-coding.github.io/awesome-ascend-skills/
☆174Oct 9, 2026Updated this week
Alternatives and similar repositories for awesome-ascend-skills
Users that are interested in awesome-ascend-skills are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Mojo Opset is a collection of different high-performance kernel implementations for LLM and multimodal.☆57Updated this week
- ☆23Sep 29, 2026Updated last week
- ☆31Sep 29, 2026Updated last week
- Community maintained hardware plugin for vLLM on Huawei Ascend☆2,934Updated this week
- See vLLM official support: https://github.com/vllm-project/vllm-ascend☆11Feb 5, 2025Updated last year
- Bare Metal GPUs on DigitalOcean Gradient AI • AdPurpose-built for serious AI teams training foundational models, running large-scale inference, and pushing the boundaries of what's possible.
- Ascend TileLang adapter☆403Updated this week
- Triton language and compiler for Ascend NPU☆182Updated this week
- LMCache-Ascend is a plugin for running LMCache on the Ascend NPU.☆90Sep 29, 2026Updated last week
- ☆17Oct 3, 2026Updated last week
- A high-throughput and memory-efficient inference and serving engine for LLMs☆16Jul 19, 2026Updated 2 months ago
- SGLang kernel library for NPU☆191Updated this week
- MetaAttention: A Unified and Performant Attention Framework Across Hardware Backends(PPoPP'26)☆18Aug 6, 2026Updated 2 months ago
- Triton adapter for Ascend. Mirror of https://gitcode.com/ascend/triton-ascend☆129May 18, 2026Updated 4 months ago
- ☆71Sep 24, 2026Updated 2 weeks ago
- Bare Metal GPUs on DigitalOcean Gradient AI • AdPurpose-built for serious AI teams training foundational models, running large-scale inference, and pushing the boundaries of what's possible.
- ☆74Sep 22, 2026Updated 2 weeks ago
- ☆48Updated this week
- 🌈 Solutions of LeetGPU☆109Jun 11, 2026Updated 3 months ago
- ☆932Updated this week
- A demonstrative example of running SGLang Diffusion with DP router☆26Mar 15, 2026Updated 6 months ago
- ☆48May 7, 2026Updated 5 months ago
- ☆29Mar 30, 2026Updated 6 months ago
- High-performance Rust benchmark client for vLLM serving endpoints.☆51Aug 3, 2026Updated 2 months ago
- Anderson points-to analysis implementation based on LLVM☆12Jan 3, 2021Updated 5 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- MultiArchKernelBench: A Multi-Platform Benchmark for Kernel Generation☆70Jul 8, 2026Updated 3 months ago
- code for the paper titled "Adaptive Cross-Layer Attention for Image Restoration"☆14Nov 6, 2025Updated 11 months ago
- torchcomms: a modern PyTorch communications API☆400Updated this week
- A lightweight library for xPU kernel JIT compilation☆374Updated this week
- Shadowsocks/ShadowsocksR 账号在线监控☆12Nov 25, 2018Updated 7 years ago
- Vue3版网络拓扑图(技术栈Vue3+TS+Antv/G6+Axios+Mockjs)☆15Jun 20, 2024Updated 2 years ago
- Omni_Infer is a suite of inference accelerators designed for the Ascend NPU platform, offering native support and an expanding feature se…☆131Sep 20, 2026Updated 2 weeks ago
- Evaluating Lossy Compression Rates of Deep Generative Models☆15Jan 20, 2021Updated 5 years ago
- interview question for AI infra☆18Mar 22, 2026Updated 6 months ago
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- ☆14Feb 7, 2020Updated 6 years ago
- The public blockchain vulnerability dataset released in our FSE'22 paper☆10Aug 22, 2022Updated 4 years ago
- SGLang is a high-performance serving framework for large language models and multimodal models.☆23Updated this week
- DomainPlus: Cross-Transform Domain Learning towards High Dynamic Range Imaging☆12Oct 11, 2022Updated 3 years ago
- An LLM post-training framework with vLLM for RL Scaling☆484Updated this week
- 分享AI Infra知识&代码练习:PyTorch、vLLM/SGLang、slime/vime框架入门⚡️、性能加速🚀、大模型基础🧠、AI软硬件🔧等☆4,132Sep 20, 2026Updated 2 weeks ago
- 基于YOLOv8的小目标检测算法,部署在华为晟腾开发板(Atlas 200I DK A2)。☆16Apr 27, 2025Updated last year