This repository serves as an example of deploying the YOLO models on Triton Server for performance and testing purposes
☆71Oct 20, 2025Updated 9 months ago
Alternatives and similar repositories for triton-server-yolo
Users that are interested in triton-server-yolo are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- This repository utilizes the Triton Inference Server Client, which streamlines the complexity of model deployment.☆21Sep 1, 2024Updated last year
- Implementation of Nvidia DeepStream 7 with YOLOv9 Models.☆15Jun 22, 2024Updated 2 years ago
- Provides an ensemble model to deploy a YoloV8 ONNX model to Triton☆42Oct 19, 2023Updated 2 years ago
- Implementation of YOLOv9 QAT optimized for deployment on TensorRT platforms.☆139Apr 24, 2025Updated last year
- The Purpose of this repository is to create a DeepStream/Triton-Server sample application that utilizes yolov7, yolov7-qat, yolov9 models…☆20Apr 1, 2024Updated 2 years ago
- Managed Database hosting by DigitalOcean • AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- Implementation of End-to-End YOLO Models for DeepStream☆75Feb 26, 2026Updated 5 months ago
- C++ application to perform computer vision tasks using Nvidia Triton Server for model inference☆30Jul 21, 2026Updated last week
- This repository implements the YOLOv9 model on Jetson Orin Nano☆19Aug 28, 2024Updated last year
- Provides an ensemble model to deploy a YOLOv8 TensorRT model to Triton☆13Mar 28, 2024Updated 2 years ago
- ☆18Mar 28, 2024Updated 2 years ago
- ffmpeg+cuvid+tensorrt+multicamera☆12Dec 31, 2024Updated last year
- 搜藏的希望的代码片段☆13Jun 6, 2023Updated 3 years ago
- YOLOV7 Face Detection☆21Dec 15, 2022Updated 3 years ago
- YOLOv12 TensorRT 端到端模型加速推理和INT8量化实现☆14Mar 5, 2025Updated last year
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- This is a repo with a Triton Server deployment template☆24Aug 4, 2024Updated last year
- ☆12Aug 10, 2022Updated 3 years ago
- NVIDIA DeepStream SDK 8.0 / 7.1 / 7.0 / 6.4 / 6.3 / 6.2 / 6.1.1 / 6.1 / 6.0.1 / 6.0 application for YOLO-Face models☆80Oct 13, 2025Updated 9 months ago
- DETR tensor去除推理过程无用辅助头+fp16部署再次加速+解决转tensorrt 输出全为0问题的新方法。☆12Jan 9, 2024Updated 2 years ago
- A project demonstrating how to make DeepStream docker images.☆95Apr 20, 2026Updated 3 months ago
- RT-DETRv2 tensorrt C++ 部署☆26Oct 29, 2024Updated last year
- Native GStreamer plugins that integrate SAHI (Slicing Aided Hyper Inference) into NVIDIA DeepStream for real-time small object detection …☆30Jun 8, 2026Updated last month
- FastSAM 部署版本,便于移植不同平,部署简单、运行速度快。☆25May 30, 2024Updated 2 years ago
- FastSAM 部署rknn C++ 代码☆13May 30, 2024Updated 2 years ago
- Proton VPN Special Offer - Get 70% off • AdSpecial partner offer. Trusted by over 100 million users worldwide. Tested, Approved and Recommended by Experts.
- Try to export the ONNX QDQ model that conforms to the AXERA NPU quantization specification. Currently, only w8a8 is supported.☆11Sep 10, 2024Updated last year
- This is a repository to practice multi-thread programming in C++☆31Feb 21, 2024Updated 2 years ago
- Cpp and python implementation of YOLOv9 using TensorRT API☆123Sep 30, 2024Updated last year
- 大模型API性能指标比较 - 深入分析TTFT、TPS等关键指标☆20Sep 12, 2024Updated last year
- ☆15Jan 10, 2023Updated 3 years ago
- A shared library of on-demand DeepStream Pipeline Services for Python and C/C++☆345Mar 17, 2025Updated last year
- Accelerating SAHI-based inference on YOLO models using TensorRT.☆103Jan 6, 2026Updated 6 months ago
- This project introduces how to implement high-performance deployment of YOLOv8-SAHI with Int8 Engine on embedded devices such as Jetson. …☆19Sep 2, 2025Updated 10 months ago
- ☆17Oct 16, 2023Updated 2 years ago
- Bare Metal GPUs on DigitalOcean Gradient AI • AdPurpose-built for serious AI teams training foundational models, running large-scale inference, and pushing the boundaries of what's possible.
- The repository is created to showcase software concepts used to maintain and scale software apps☆12Dec 10, 2022Updated 3 years ago
- HunyuanDiT with TensorRT and libtorch☆18May 22, 2024Updated 2 years ago
- ☆13Sep 25, 2020Updated 5 years ago
- ☆25Oct 6, 2022Updated 3 years ago
- a ai infra framework for edge device base on nndeploy☆18Nov 27, 2025Updated 8 months ago
- tensorrt for yolo series (YOLOv11,YOLOv10,YOLOv9,YOLOv8,YOLOv7,YOLOv6,YOLOX,YOLOv5), nms plugin support☆1,161Oct 15, 2025Updated 9 months ago
- 使用mnn-llm对GOT-OCR2.0进行推理☆14Oct 2, 2024Updated last year