A comprehensive knowledge base for Huawei Ascend NPU development, structured as distributed Agent Skills. https://ascend-ai-coding.github.io/awesome-ascend-skills/
☆143Jul 24, 2026Updated this week
Alternatives and similar repositories for awesome-ascend-skills
Users that are interested in awesome-ascend-skills are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- ☆22Jun 29, 2026Updated 3 weeks ago
- Community maintained hardware plugin for vLLM on Ascend☆2,472Updated this week
- Triton language and compiler for Ascend NPU☆115Updated this week
- Ascend TileLang adapter☆338Updated this week
- LMCache on Ascend☆82Updated this week
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- ☆17Mar 21, 2026Updated 4 months ago
- A high-throughput and memory-efficient inference and serving engine for LLMs☆15Updated this week
- 基于Ascend(昇腾910B)纯国产显卡复刻MiniMind,🚀🚀 「大模型」2小时完全从0训练26M的小参数GPT!🌏 Train a 26M-parameter GPT from scratch in just 2h!☆25Mar 2, 2026Updated 4 months ago
- SGLang kernel library for NPU☆170Updated this week
- Provide performance insight capabilities for RL frameworks.☆47Updated this week
- MetaAttention: A Unified and Performant Attention Framework Across Hardware Backends(PPoPP'26)☆16Dec 31, 2025Updated 6 months ago
- Triton adapter for Ascend. Mirror of https://gitcode.com/ascend/triton-ascend☆127May 18, 2026Updated 2 months ago
- 基于 MindSpore 框架 MS-Serving 服务适配的 Langchain-Chatchat(原Langchain-ChatGLM)☆19Mar 21, 2024Updated 2 years ago
- AISBench Benchmark is a model evaluation tool built on OpenCompass, compatible with OpenCompass’s configuration system, dataset structure…☆208Updated this week
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- ☆67Updated this week
- ☆74Updated this week
- A high-performance inference engine for LLM, VLM, DiT and REC models, optimized for diverse AI accelerators. It is hosted in OpenAtom Fou…☆1,489Updated this week
- ☆18Oct 6, 2023Updated 2 years ago
- A framework and CLI toolkit for orchestrating teams of loosely-coupled AI agents.☆18Updated this week
- 用于在昇腾设备上高性能推理PaddleOCR模型☆52Aug 1, 2025Updated 11 months ago
- 🌈 Solutions of LeetGPU☆94Jun 11, 2026Updated last month
- ☆696Jul 14, 2026Updated last week
- ☆14Nov 7, 2024Updated last year
- Serverless GPU API endpoints on Runpod - Get Bonus Credits • AdSkip the infrastructure headaches. Auto-scaling, pay-as-you-go, no-ops approach lets you focus on innovating your application.
- A demonstrative example of running SGLang Diffusion with DP router☆17Mar 15, 2026Updated 4 months ago
- [CVPRW 2021] DUVE network for NTIRE 2021 Quality enhancement of heavily compressed videos - Track 3 Fixed bit-rate☆10Oct 17, 2024Updated last year
- Partial Redundancy Elimination Pass in LLVM☆15May 20, 2019Updated 7 years ago
- ☆29Mar 30, 2026Updated 3 months ago
- An experimental communicating attention kernel based on DeepEP.☆34Jul 29, 2025Updated 11 months ago
- MultiArchKernelBench: A Multi-Platform Benchmark for Kernel Generation☆64Jul 8, 2026Updated 2 weeks ago
- torchcomms: a modern PyTorch communications API☆380Updated this week
- Examples of MolScore implementations☆12May 30, 2024Updated 2 years ago
- [ECCV 2024] "Prediction Exposes Your Face: Black-box Model Inversion via Prediction Alignment"☆15Mar 12, 2025Updated last year
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- 16-fold memory access reduction with nearly no loss☆107Mar 26, 2025Updated last year
- Omni_Infer is a suite of inference accelerators designed for the Ascend NPU platform, offering native support and an expanding feature se…☆127Updated this week
- Vue3版网络拓扑图(技术栈Vue3+TS+Antv/G6+Axios+Mockjs)☆13Jun 20, 2024Updated 2 years ago
- interview question for AI infra☆18Mar 22, 2026Updated 4 months ago
- TurboServe: Serving Streaming Video Generation Efficiently and Economically☆37Jul 12, 2026Updated last week
- ☆14Feb 7, 2020Updated 6 years ago
- SGLang is a high-performance serving framework for large language models and multimodal models.☆18Updated this week