Mojo Opset is a collection of different high-performance kernel implementations for LLM and multimodal.
☆50Aug 6, 2026Updated this week
Alternatives and similar repositories for mojo_opset
Users that are interested in mojo_opset are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- A torch compile backend for multi-targets☆51May 27, 2026Updated 2 months ago
- Triton language and compiler for Ascend NPU☆136Updated this week
- ☆22Jun 29, 2026Updated last month
- 面向多平台编译优化的深度学习中间表示☆10Oct 28, 2024Updated last year
- A comprehensive knowledge base for Huawei Ascend NPU development, structured as distributed Agent Skills. https://ascend-ai-coding.github…☆157Updated this week
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- [HPCA 2026] AI Accelerator Benchmark focuses on evaluating AI Accelerators from a practical production perspective, including the ease of…☆372Apr 22, 2026Updated 3 months ago
- Learning and Debugging for FSDP/FSDP2 Training☆17Feb 7, 2026Updated 6 months ago
- Ascend operator generation☆33Jun 17, 2026Updated last month
- Provide performance insight capabilities for RL frameworks.☆53Updated this week
- ☆11Jul 28, 2026Updated 2 weeks ago
- Guide to build and use Tensorflow XLA/AOT on Windows☆13Dec 26, 2018Updated 7 years ago
- See vLLM official support: https://github.com/vllm-project/vllm-ascend☆11Feb 5, 2025Updated last year
- My tests and experiments with some popular dl frameworks.☆17Sep 11, 2025Updated 11 months ago
- A model compilation solution for various hardware☆476Aug 20, 2025Updated 11 months ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- 模型量化工程 Base pretrained models and datasets in pytorch (MNIST, SVHN, CIFAR10, CIFAR100, STL10, AlexNet, VGG16, VGG19, ResNet, Inception,…☆12Aug 3, 2018Updated 8 years ago
- ☆57Mar 15, 2025Updated last year
- A LogGOPS (LogP, LogGP, LogGPS) Simulator and Simulation Framework☆16Aug 20, 2024Updated last year
- Chorus: Heterogeneous GPU+CPU Multiple Protein Sequences Alignment Search for Large Database☆17Apr 7, 2025Updated last year
- ☆14Feb 7, 2020Updated 6 years ago
- The public blockchain vulnerability dataset released in our FSE'22 paper☆10Aug 22, 2022Updated 3 years ago
- SGLang is a high-performance serving framework for large language models and multimodal models.☆17Updated this week
- ☆12Oct 19, 2014Updated 11 years ago
- ☆22Mar 28, 2022Updated 4 years ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- compare the theory attention gradient with PyTorch attention gradient☆16Apr 1, 2024Updated 2 years ago
- Evaluating Program Reasoning of LLMs via Formal Specification Inference (ACL 2025)☆15Sep 21, 2025Updated 10 months ago
- ☆11Nov 13, 2020Updated 5 years ago
- gups mirror☆12Oct 25, 2015Updated 10 years ago
- [NeurIPS 2023] Token-Scaled Logit Distillation for Ternary Weight Generative Language Models☆18Dec 6, 2023Updated 2 years ago
- Bayesian optimization based on Gaussian processes☆13Dec 2, 2022Updated 3 years ago
- Labs of 2019 Web Information Processing and Application in USTC.☆11Jan 15, 2020Updated 6 years ago
- 用以批量下载 ClassIn 上的课程录像。仅在中国科学技术大学(USTC)的 Blackboard + ClassIn 平台下测试可用。☆10Dec 8, 2022Updated 3 years ago
- Code for paper "FuSeConv Fully Separable Convolutions for Fast Inference on Systolic Arrays" published at DATE 2021☆18Aug 23, 2021Updated 4 years ago
- Simple, predictable pricing with DigitalOcean hosting • AdAlways know what you'll pay with monthly caps and flat pricing. Enterprise-grade infrastructure trusted by 600k+ customers.
- A "standard library" of Triton kernels.☆26Oct 2, 2025Updated 10 months ago
- [TMLR] Official PyTorch implementation of paper "Quantization Variation: A New Perspective on Training Transformers with Low-Bit Precisio…☆50Sep 27, 2024Updated last year
- A Triton JIT runtime and ffi provider in C++☆38Updated this week
- ICM-Assistant: Instruction-tuning Multimodal Large Language Models for Rule-based Explainable Image Content Moderation. AAAI, 2025☆16Aug 25, 2025Updated 11 months ago
- DLBlas: clean and efficient kernels☆46Jul 28, 2026Updated 2 weeks ago
- ☆14May 28, 2023Updated 3 years ago
- The source code of my website.☆14Dec 8, 2021Updated 4 years ago