A comprehensive knowledge base for Huawei Ascend NPU development, structured as distributed Agent Skills. https://ascend-ai-coding.github.io/awesome-ascend-skills/
☆167Aug 29, 2026Updated this week
Alternatives and similar repositories for awesome-ascend-skills
Users that are interested in awesome-ascend-skills are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Mojo Opset is a collection of different high-performance kernel implementations for LLM and multimodal.☆52Updated this week
- ☆22Jun 29, 2026Updated 2 months ago
- ☆30Jul 3, 2026Updated last month
- Community maintained hardware plugin for vLLM on Ascend☆2,730Updated this week
- Learning and Debugging for FSDP/FSDP2 Training☆17Feb 7, 2026Updated 6 months ago
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- Ascend TileLang adapter☆358Updated this week
- LMCache-Ascend is a plugin for running LMCache on the Ascend NPU.☆87Updated this week
- ☆17Mar 21, 2026Updated 5 months ago
- 基于Ascend(昇腾910B)纯国产显卡复刻MiniMind,🚀🚀 「大模型」2小时完全从0训练26M的小参数GPT!🌏 Train a 26M-parameter GPT from scratch in just 2h!☆27Mar 2, 2026Updated 5 months ago
- SGLang kernel library for NPU☆174Updated this week
- [ICML 2025] This is the official PyTorch implementation of "OmniBal: Towards Fast Instruction-Tuning for Vision-Language Models via Omniv…☆27Jun 16, 2025Updated last year
- MetaAttention: A Unified and Performant Attention Framework Across Hardware Backends(PPoPP'26)☆17Aug 6, 2026Updated 3 weeks ago
- Triton adapter for Ascend. Mirror of https://gitcode.com/ascend/triton-ascend☆127May 18, 2026Updated 3 months ago
- RWKV6 in native pytorch and triton:)☆11Aug 4, 2024Updated 2 years ago
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- 基于 MindSpore 框架 MS-Serving 服务适配的 Langchain-Chatchat(原Langchain-ChatGLM)☆19Mar 21, 2024Updated 2 years ago
- ☆31Updated this week
- ☆74Updated this week
- ☆18Oct 6, 2023Updated 2 years ago
- A high-performance inference engine for LLM, VLM, DiT and REC models, optimized for diverse AI accelerators. It is hosted in OpenAtom Fou…☆1,541Updated this week
- A framework and CLI toolkit for orchestrating teams of loosely-coupled AI agents.☆18Aug 9, 2026Updated 3 weeks ago
- 🌈 Solutions of LeetGPU☆102Jun 11, 2026Updated 2 months ago
- ☆774Aug 23, 2026Updated last week
- ☆37May 7, 2026Updated 3 months ago
- Virtual machines for every use case on DigitalOcean • AdGet dependable uptime with 99.99% SLA, simple security tools, and predictable monthly pricing with DigitalOcean's virtual machines, called Droplets.
- A demonstrative example of running SGLang Diffusion with DP router☆19Mar 15, 2026Updated 5 months ago
- Partial Redundancy Elimination Pass in LLVM☆15May 20, 2019Updated 7 years ago
- High-performance Rust benchmark client for vLLM serving endpoints.☆52Aug 3, 2026Updated 3 weeks ago
- An experimental communicating attention kernel based on DeepEP.☆34Jul 29, 2025Updated last year
- MultiArchKernelBench: A Multi-Platform Benchmark for Kernel Generation☆67Jul 8, 2026Updated last month
- Anderson points-to analysis implementation based on LLVM☆12Jan 3, 2021Updated 5 years ago
- code for the paper titled "Adaptive Cross-Layer Attention for Image Restoration"☆14Nov 6, 2025Updated 9 months ago
- torchcomms: a modern PyTorch communications API☆390Updated this week
- 个人学习编译原理、理解创造一个编译器主体流程的小项目☆10Oct 7, 2020Updated 5 years ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- Official repository for "SSA: Sparse Sparse Attention by Aligning Full and Sparse Attention Outputs in Feature Space"☆29May 7, 2026Updated 3 months ago
- Shadowsocks/ShadowsocksR 账号在线监控☆12Nov 25, 2018Updated 7 years ago
- 16-fold memory access reduction with nearly no loss☆107Mar 26, 2025Updated last year
- [AAAI 2024] Official repository for "Rethinking Dimensional Rationale in Graph Contrastive Learning from Causal Perspective"☆14Dec 8, 2024Updated last year
- Omni_Infer is a suite of inference accelerators designed for the Ascend NPU platform, offering native support and an expanding feature se…☆129Updated this week
- Evaluating Lossy Compression Rates of Deep Generative Models☆15Jan 20, 2021Updated 5 years ago
- TurboServe: Serving Streaming Video Generation Efficiently and Economically☆44Jul 12, 2026Updated last month