☆36Apr 2, 2026Updated 5 months ago
Alternatives and similar repositories for OmniInfer-LLM
Users that are interested in OmniInfer-LLM are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- ☆34Feb 10, 2026Updated 7 months ago
- ☆33Jun 20, 2026Updated 2 months ago
- Your on-phone / mobile AI Agent / Claw, capable of operating terminals and performing a wide range of tasks in the Android world || 你的手机 …☆1,981Updated this week
- 为visinger SVS系统写的展示系统~本质仍然是个音乐播放器☆11Apr 18, 2023Updated 3 years ago
- MLSys competition for the best MOE NKI kernels☆48May 29, 2026Updated 3 months ago
- Bare Metal GPUs on DigitalOcean Gradient AI • AdPurpose-built for serious AI teams training foundational models, running large-scale inference, and pushing the boundaries of what's possible.
- ☆20May 28, 2025Updated last year
- This code is a version of implement of the essay named Deep Inception Networks: A General End-to-End Framework for Multi-asset Quantitati…☆14Mar 15, 2024Updated 2 years ago
- 这是我个人的本科毕业设计项目,总体上不是特别成熟。但对MOT和边缘计算的发展可能有非常微小的启发意义,欢迎交流!☆10Aug 16, 2022Updated 4 years ago
- 同济大学 数据库课设 TJU Database Curriculum Project☆10May 30, 2020Updated 6 years ago
- This is an official PyTorch implementation of ASDA (accepted by ACMMM 2024).☆26Oct 22, 2024Updated last year
- ☆21Updated this week
- ☆28Oct 6, 2021Updated 4 years ago
- A Library for intra-GPU/Inter-SM parallelsim☆12Aug 7, 2026Updated last month
- hexagon tutorial☆66Mar 29, 2026Updated 5 months ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- SC'25 UltraAttn: Efficiently Parallelizing Attention through Hierarchical Context-Tiling☆16Aug 14, 2025Updated last year
- 该项目实现了对以DLA34作为骨干网络的FairMOT的TensorRT迁移加速,同时这也是一个NVIDIA-阿里云的Hackathon2021竞赛项目。☆16May 22, 2021Updated 5 years ago
- ☆12Apr 30, 2024Updated 2 years ago
- Seq2act: Mapping Natural Language Instructions to Mobile UI Action Sequences from Google research☆15Jul 13, 2020Updated 6 years ago
- Fast GPU based tensor core reductions☆12Jan 13, 2023Updated 3 years ago
- High-speed and easy-use LLM serving framework for local deployment☆166Aug 7, 2025Updated last year
- text to speech☆10Mar 19, 2024Updated 2 years ago
- ☆16Jul 25, 2023Updated 3 years ago
- Appling the asynchronous tensor swapping to PyTorch framework.☆31Jun 20, 2023Updated 3 years ago
- Managed Database hosting by DigitalOcean • AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- ☆17Mar 22, 2024Updated 2 years ago
- ☆185Jun 21, 2026Updated 2 months ago
- Source code for CSWAP-CLUSTER'21 and CSWAP+-TPDS'22☆24Mar 2, 2022Updated 4 years ago
- unofficial pytorch implementation of HiFi-GAN with fast MISR.☆15Mar 21, 2023Updated 3 years ago
- setup pytorch on android☆12Mar 2, 2020Updated 6 years ago
- code examples in Python / Theano / Blocks / Foxhound to go along with my blog post "Neural Image Captioning for Mortals"☆22Nov 30, 2017Updated 8 years ago
- Federated Learning - PyTorch☆15Jun 27, 2021Updated 5 years ago
- ☆25Mar 15, 2023Updated 3 years ago
- ☆10Jan 14, 2025Updated last year
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- AQUATOPE: QoS-and-Uncertainty-Aware Resource Management for Multi-Stage Serverless Workflows (ASPLOS'23)☆25Mar 13, 2024Updated 2 years ago
- An image process system.☆16Oct 27, 2023Updated 2 years ago
- Source code for ChunkGraph-ATC'24☆27Jul 13, 2024Updated 2 years ago
- ☆11Dec 6, 2020Updated 5 years ago
- ☆29Feb 12, 2025Updated last year
- MiniWear - Miniature & Wearable Electronics Modules☆12Sep 28, 2016Updated 9 years ago
- For advanced physics-driven combined with neural network enhancement force field.☆19Mar 9, 2026Updated 6 months ago