Deep insight tensorrt, including but not limited to qat, ptq, plugin, triton_inference, cuda
☆24Jul 28, 2026Updated 2 weeks ago
Alternatives and similar repositories for tensorrt-insight
Users that are interested in tensorrt-insight are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Geek Car, An Autonomous Application Based On Cyber RT http://for-geeks.com☆13Feb 5, 2026Updated 6 months ago
- BEV-RoadSeg for Freespace Detection in PyTorch, including Python onnx and tensorRT API versions.☆12Sep 16, 2021Updated 4 years ago
- ☆27Aug 5, 2022Updated 4 years ago
- The ComponentContainer and Executor that assign a dedicated thread for each callback group.☆10Jun 20, 2025Updated last year
- Quick and Self-Contained TensorRT Custom Plugin Implementation and Integration☆87May 26, 2025Updated last year
- End-to-end encrypted email - Proton Mail • AdSpecial offer: 40% Off Yearly / 80% Off First Month. All Proton services are open source and independently audited for security.
- ☆21Jul 6, 2026Updated last month
- 📦 eCAL C/C++ runtime core☆12Oct 21, 2024Updated last year
- Provides a demo of micro-ROS based on ST Disco L475 IOT01 board.☆13Jul 14, 2021Updated 5 years ago
- PoC SOME/IP to ROS2 topics☆26Apr 24, 2026Updated 3 months ago
- ROS2 image transport plugin using libav(ffmpeg) for generating foxglove CompressedVideo messages☆14Jun 29, 2026Updated last month
- learning materials of driveos from nvidia drive sdk.☆13Jun 10, 2026Updated 2 months ago
- 🔴 Accelerated GStreamer utilities for NVIDIA Jetson Nano.☆10May 8, 2021Updated 5 years ago
- PointPillars TensorRT version pretrained on MMDetection3d with WaymoOpenDataset☆23Aug 11, 2022Updated 4 years ago
- A lite version based on cyber RT☆14Feb 5, 2026Updated 6 months ago
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- ☆17Dec 19, 2022Updated 3 years ago
- Experimental GPU language with meta-programming☆32Sep 6, 2024Updated last year
- Implement Flash Attention using Cute.☆111Dec 17, 2024Updated last year
- deepstream_tools will serve as a parent repo to hold various tools to be released for DeepStream SDK.☆35Jul 15, 2026Updated last month
- ☆22Jul 31, 2026Updated 2 weeks ago
- ROS 2 Type adapters for common C++ frameworks☆18Mar 26, 2025Updated last year
- Acceleration-friendly component architecture framework☆20May 1, 2026Updated 3 months ago
- Some common CUDA kernel implementations (Not the fastest).☆30Jun 24, 2026Updated last month
- Examples of using Isaac ROS GEMs together☆19Jul 7, 2026Updated last month
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- ROS 2 C++ executor bringing low CPU usage, low latency and deterministic ordering☆51Jul 18, 2024Updated 2 years ago
- ☆29May 6, 2020Updated 6 years ago
- fake CUTLASS to get peformance☆26Apr 28, 2026Updated 3 months ago
- Flash Attention in ~100 lines of CUDA (forward pass only)☆12Jun 10, 2024Updated 2 years ago
- A method to automatically calibrate lidar and camera☆21Jun 11, 2024Updated 2 years ago
- ☆37Jul 30, 2023Updated 3 years ago
- ros interface for cupoch☆23Mar 10, 2022Updated 4 years ago
- Optimizing Tensor Computation Graphs with Equality Saturation and Monte Carlo Tree Search☆15Aug 9, 2024Updated 2 years ago
- Quantize yolov7 using pytorch_quantization.🚀🚀🚀☆12Oct 20, 2023Updated 2 years ago
- GPUs on demand by Runpod - Special Offer Available • AdRun AI, ML, and HPC workloads on powerful cloud GPUs—without limits or wasted spend. Deploy GPUs in under a minute and pay by the second.
- ROS2 package that allows recording without interprocess communication☆22Apr 16, 2026Updated 3 months ago
- TensorRT deploy and PTQ/QAT tools development for FastBEV, total time only need 6.9ms!!!☆313Dec 8, 2023Updated 2 years ago
- 喋りすぎているヤツを邪魔するブラウザ拡張☆11Dec 16, 2021Updated 4 years ago
- ☆324May 11, 2022Updated 4 years ago
- Optimizing Deep Convolutional Neural Network with Ternarized Weights and High Accuracy☆16Jan 27, 2019Updated 7 years ago
- ☆26Nov 7, 2024Updated last year
- Performance of the C++ interface of flash attention and flash attention v2 in large language model (LLM) inference scenarios.☆15Aug 31, 2023Updated 2 years ago