☆164Jul 26, 2026Updated 2 weeks ago
Alternatives and similar repositories for aios
Users that are interested in aios are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Based on Nano-vLLM, a simple replication of vLLM with self-contained paged attention and flash attention implementation☆964Jul 22, 2026Updated 2 weeks ago
- CUDA 13.1 Tutorial Series for RTX 5090 (Blackwell) - Chinese teaching materials☆29Jan 18, 2026Updated 6 months ago
- Implementation of the paper - Fast Training of Convolutional Networks through FFTs (CUDA for parallelization)☆10May 8, 2020Updated 6 years ago
- learn TensorRT from scratch🥰☆18Sep 29, 2024Updated last year
- GPU implementation of Winograd convolution☆10Oct 23, 2017Updated 8 years ago
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- Nano vLLM☆26Aug 11, 2025Updated 11 months ago
- paper-read-notes☆13Sep 26, 2024Updated last year
- This repository organizes materials, recordings, and schedules related to AI-infra learning meetings.☆556Mar 1, 2026Updated 5 months ago
- ☆14May 28, 2019Updated 7 years ago
- 宝藏问题:汇总了大模型的重要资源,整理了宝藏问题手稿,宝藏论文参考文献☆226Updated this week
- Mirror of http://gitlab.hpcrl.cse.ohio-state.edu/chong/ppopp19_ae, refactoring for understanding☆17Oct 20, 2021Updated 4 years ago
- Examples illustrating usage of the rocBLAS library☆17Aug 12, 2024Updated last year
- Implementation of 3d non-separable convolution using CUDA & FFT Convolution☆20Jan 15, 2019Updated 7 years ago
- 分享AI Infra知识&代码练习:PyTorch、vLLM/SGLang、slime/vime框架入门⚡️、性能加速🚀、大模型基础🧠、AI软硬件🔧等☆3,423Updated this week
- Managed Database hosting by DigitalOcean • AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- 3D mesh generator based on the Octree with mixed-elements at the end.☆11Feb 12, 2024Updated 2 years ago
- ☆12Apr 26, 2021Updated 5 years ago
- ☆22Nov 6, 2025Updated 9 months ago
- is a tiny 3D mesher library for C++.☆13Sep 19, 2025Updated 10 months ago
- ☆51Apr 16, 2026Updated 3 months ago
- 基于pytorch的不平衡数据的文本分类☆12Dec 26, 2021Updated 4 years ago
- udp并发实现代码,含udp server,udp client请求建立测试代码☆16Oct 26, 2024Updated last year
- Linux Traffic Control (TC) based implementation of Kubernetes NPWG MultiNetworkPolicy API☆12Jul 24, 2026Updated 2 weeks ago
- ☆23Nov 30, 2020Updated 5 years ago
- Managed Database hosting by DigitalOcean • AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- A Winograd Minimal Filter Implementation in CUDA☆31Aug 25, 2021Updated 4 years ago
- ☆161Mar 5, 2026Updated 5 months ago
- CUDA project for uni subject☆26Oct 26, 2020Updated 5 years ago
- Code for "Boosting Semi-supervised Image Segmentation with Global and Local Mutual Information Regularization"☆13Jul 14, 2021Updated 5 years ago
- the frontend of lmcache☆18Apr 22, 2026Updated 3 months ago
- CNI plugin to override routes☆16May 23, 2026Updated 2 months ago
- "FastSAM_Awsome_Openvino" 项目展示了如何通过 OpenVINO 框架高效部署 FastSAM 模型,实现了令人瞩目的实例分割功能。该项目提供了 C++ 版本和 Python 版本两种实现,为开发者提供了在不同语言环境下使用 FastSAM 模型的选…☆37Dec 13, 2023Updated 2 years ago
- METIS partitioner for mesh and graph☆19Aug 5, 2021Updated 5 years ago
- ToyLLM: Learning LLM from Scratch☆25Updated this week
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- 收录SC小组在学习高性能计算、分布式架构、数据挖掘与人工智能方向的笔记和材料☆15Oct 29, 2021Updated 4 years ago
- [IEEE TPAMI] LRANet++: Low-Rank Approximation Network for Accurate and Efficient Text Spotting☆31Dec 2, 2025Updated 8 months ago
- 基于自由度(熵)、凝固度 新词发现算法实现☆12Oct 7, 2018Updated 7 years ago
- A book about Training and Inferring Technology Stack of LLM☆29May 29, 2026Updated 2 months ago
- A Row Decomposition-based Approach for Sparse Matrix Multiplication on GPUs☆30Nov 29, 2023Updated 2 years ago
- ☆17Jun 1, 2022Updated 4 years ago
- ☆13Feb 23, 2023Updated 3 years ago