☆13Oct 8, 2024Updated last year
Alternatives and similar repositories for YOLO-MultiModal
Users that are interested in YOLO-MultiModal are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- This project uses three types of images as inputs RGB, Depth, and thermal images to perform object detection with YOLOv8.☆31Jul 23, 2024Updated last year
- ☆10Jun 6, 2024Updated 2 years ago
- Cross-Modality Attentive Feature Fusion for Object Detection in Multispectral Remote Sensing Imagery☆16Oct 7, 2022Updated 3 years ago
- This repo contains the Pytorch implementation of the AAAI'18 paper - Deep Reinforcement Learning for Unsupervised Video Summarization wit…☆11Jun 5, 2023Updated 3 years ago
- EACL 2023 paper "MLASK: Multimodal Summarization of Video-based News Articles"☆11Nov 7, 2023Updated 2 years ago
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- [ICIP 2020]"Multispectral Fusion for Object Detection with Cyclic Fuse-and-Refine Blocks"☆14Oct 6, 2020Updated 5 years ago
- An innovative object detection system for visually impaired individuals. Using YOLO V3 algorithm and the extensive COCO dataset, our syst…☆12Jul 13, 2024Updated 2 years ago
- [MICCAI'22] Unsupervised Contrastive Learning on Gall Bladder Ultrasound Videos☆11May 28, 2023Updated 3 years ago
- ☆11Jun 5, 2021Updated 5 years ago
- Vehicle counting system with YOLOv8 and DeepSORT☆10Aug 23, 2023Updated 2 years ago
- ☆63Nov 26, 2024Updated last year
- I fine-tuned (p-tuning) Tsinghua’s open-source large language model, ChatGLM2-6B, using several years of my WeChat chat history. Inspired…☆12Mar 6, 2024Updated 2 years ago
- ☆13Jun 17, 2023Updated 3 years ago
- Official implementation of “CAT-ViL: Co-Attention Gated Vision-Language Embedding for Visual Question Localized-Answering in Robotic Surg…☆18Jul 7, 2024Updated 2 years ago
- Virtual machines for every use case on DigitalOcean • AdGet dependable uptime with 99.99% SLA, simple security tools, and predictable monthly pricing with DigitalOcean's virtual machines, called Droplets.
- SpeedVision is an AI-powered tool that detects and calculates vehicle speed from video footage using YOLO-based object detection and fram…☆14Sep 22, 2024Updated last year
- ☆10May 1, 2021Updated 5 years ago
- This project used Yolov8/AnimeGAN and Flask to accomplish the task of background segmentation , background remove and background replacem…☆12Apr 12, 2024Updated 2 years ago
- ☆15Jun 27, 2023Updated 3 years ago
- Deformable Convolutional Networks v2 with Pytorch☆33Dec 2, 2020Updated 5 years ago
- ☆21Sep 9, 2022Updated 3 years ago
- Implementation of LTC-SUM: Lightweight Client-driven Personalized Video Summarization Framework Using 2D CNN☆22Jul 11, 2023Updated 3 years ago
- 3D scene mapping system that using PyTorch's MiDaS model to estimate scene point cloud☆13Jan 10, 2025Updated last year
- 基于optitrack定位的无人机目标跟踪(target tracking)☆13Oct 16, 2024Updated last year
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- Data Science Exercises based on real-world scenarios with explanatory comments and prettified output.☆15May 8, 2023Updated 3 years ago
- ☆12May 24, 2023Updated 3 years ago
- AST-GCN: Attribute-Augmented Spatiotemporal Graph Convolutional Network for Traffic Forecasting. This is my implementation of this model …☆11Aug 31, 2023Updated 2 years ago
- 深度学习500问,以问答形式对常用的概率知识、线性代数、机器学习、深度学习、计算机视觉等热点问题进行阐述,以帮助自己及有需要的读者。 全书分为17个章节,20多万字。由于水平有限,书中不妥之处恳请广大读者批评指正。 未完待续............ 如有意合作,联系sc…☆19Nov 12, 2018Updated 7 years ago
- ☆17Jul 18, 2023Updated 3 years ago
- Simple video summarisation Python package.☆25Jan 29, 2024Updated 2 years ago
- Pytorch code for paper Contrastive Losses Are Natural Criteria for Unsupervised Video Summarization☆22Jan 7, 2023Updated 3 years ago
- Video Summarization With Spatiotemporal Vision Transformer☆23Jul 5, 2023Updated 3 years ago
- A Chat with AI☆11May 11, 2026Updated 2 months ago
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- XiaoLoong Dev-C++, improved fork of Orwell Dev-C++☆21Jan 24, 2026Updated 5 months ago
- 基于yolov5,在woodscape数据集上实现旋转框目标检测+语义分割☆13Mar 4, 2024Updated 2 years ago
- Example of YOLOv8 pose detection (estimation) on browser. It shows implementations powered by ONNX and TFJS served through JavaScript wit…☆15Jun 9, 2024Updated 2 years ago
- Multimodal summarization of user-generated videos from wearable cameras☆23Jun 22, 2025Updated last year
- 关于Pytorch-Geometric的学习,包括官方文档的基本内容和部分API的使用方式,以及官方源码中的示例代码和Pytorch-Geometric的部分源码实现☆21Dec 2, 2020Updated 5 years ago
- 3D LiDAR Object Detection using YOLOv8-obb (oriented bounding box).☆16Sep 6, 2024Updated last year
- EchoGLAD: Hierarchical Graph Neural Networks for Left Ventricle Landmark Detection on Echocardiograms☆21Apr 17, 2023Updated 3 years ago