Mojo Opset is a collection of different high-performance kernel implementations for LLM and multimodal.
☆52Aug 25, 2026Updated this week
Alternatives and similar repositories for mojo_opset
Users that are interested in mojo_opset are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- A torch compile backend for multi-targets☆51May 27, 2026Updated 3 months ago
- Triton language and compiler for Ascend NPU☆147Updated this week
- ☆22Jun 29, 2026Updated 2 months ago
- 面向多平台编译优化的深度学习中间表示☆10Oct 28, 2024Updated last year
- A comprehensive knowledge base for Huawei Ascend NPU development, structured as distributed Agent Skills. https://ascend-ai-coding.github…☆167Updated this week
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- [HPCA 2026] AI Accelerator Benchmark focuses on evaluating AI Accelerators from a practical production perspective, including the ease of…☆378Apr 22, 2026Updated 4 months ago
- Learning and Debugging for FSDP/FSDP2 Training☆17Feb 7, 2026Updated 6 months ago
- Ascend operator generation☆37Jun 17, 2026Updated 2 months ago
- Provide performance insight capabilities for RL frameworks.☆75Updated this week
- ☆11Jul 28, 2026Updated last month
- Guide to build and use Tensorflow XLA/AOT on Windows☆13Dec 26, 2018Updated 7 years ago
- See vLLM official support: https://github.com/vllm-project/vllm-ascend☆11Feb 5, 2025Updated last year
- My tests and experiments with some popular dl frameworks.☆17Sep 11, 2025Updated 11 months ago
- 模型量化工程 Base pretrained models and datasets in pytorch (MNIST, SVHN, CIFAR10, CIFAR100, STL10, AlexNet, VGG16, VGG19, ResNet, Inception,…☆12Aug 3, 2018Updated 8 years ago
- Bare Metal GPUs on DigitalOcean Gradient AI • AdPurpose-built for serious AI teams training foundational models, running large-scale inference, and pushing the boundaries of what's possible.
- ☆57Mar 15, 2025Updated last year
- Chorus: Heterogeneous GPU+CPU Multiple Protein Sequences Alignment Search for Large Database☆17Apr 7, 2025Updated last year
- The public blockchain vulnerability dataset released in our FSE'22 paper☆10Aug 22, 2022Updated 4 years ago
- SGLang is a high-performance serving framework for large language models and multimodal models.☆18Updated this week
- Triton adapter for Ascend. Mirror of https://gitcode.com/ascend/triton-ascend☆127May 18, 2026Updated 3 months ago
- ☆12Oct 19, 2014Updated 11 years ago
- compare the theory attention gradient with PyTorch attention gradient☆16Apr 1, 2024Updated 2 years ago
- ☆11Nov 13, 2020Updated 5 years ago
- gups mirror☆12Oct 25, 2015Updated 10 years ago
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- [NeurIPS 2023] Token-Scaled Logit Distillation for Ternary Weight Generative Language Models☆18Dec 6, 2023Updated 2 years ago
- Labs of 2019 Web Information Processing and Application in USTC.☆11Jan 15, 2020Updated 6 years ago
- 用以批量下载 ClassIn 上的课程录像。仅在中国科学技术大学(USTC)的 Blackboard + ClassIn 平台下测试可用。☆10Dec 8, 2022Updated 3 years ago
- Code for paper "FuSeConv Fully Separable Convolutions for Fast Inference on Systolic Arrays" published at DATE 2021☆18Aug 23, 2021Updated 5 years ago
- A Triton JIT runtime and ffi provider in C++☆40Aug 7, 2026Updated 3 weeks ago
- A simple and experimental c/c++ package manager☆12Jul 10, 2026Updated last month
- (AAAI 2026) OSVBench, a new benchmark for evaluating Large Language Models (LLMs) in generating complete specification code pertaining to…☆16May 13, 2025Updated last year
- ☆28Oct 21, 2020Updated 5 years ago
- ☆14May 28, 2023Updated 3 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- Public skills collected from well-known open-source projects focused on LLM infrastructure, GPU kernels, compiler/operator development☆31May 7, 2026Updated 3 months ago
- An easy to use and efficient memory pool allocator written in C++☆11Jun 9, 2017Updated 9 years ago
- ☆12Oct 9, 2020Updated 5 years ago
- A Unified Framework for Training, Mapping and Simulation of ReRAM-Based Convolutional Neural Network Acceleration☆37May 19, 2022Updated 4 years ago
- MultiArchKernelBench: A Multi-Platform Benchmark for Kernel Generation☆67Jul 8, 2026Updated last month
- A fast seed-embed-extend based sequence mapper and aligner☆23Aug 28, 2024Updated 2 years ago
- 利用图神经网络进行CTR预估☆15Nov 22, 2019Updated 6 years ago