本项目结合了YOLO的目标检测与分割能力,以及CLIP微调分类能力,YOLO初步框选物体,CLIP对框选物体进行细致化分类,能够在复杂场景下实现物体的精准定位与属性提取。通过对目标物体的检测、分割和细粒度分类,项目特别适用于商品分类、智能货柜管理等任务
☆31Nov 26, 2024Updated last year
Alternatives and similar repositories for YOLO_CLIP_targetDetection
Users that are interested in YOLO_CLIP_targetDetection are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- 用open-cv检测物体的大小,是实时的☆17Jan 7, 2022Updated 4 years ago
- 用Paddle复现论文ChineseBERT: Chinese Pretraining Enhanced by Glyph and Pinyin Information(ACL2021)☆10Nov 15, 2021Updated 4 years ago
- 基于MCP协议和LangChain框架实现的企业级AI多Agent多模态系统,包含RAG技术增强的知识检索能力。☆44Mar 10, 2025Updated last year
- Implementation of "Personalized Federated Fine-Tuning for LLMs via Data-Driven Heterogeneous Model Architectures"☆16Jan 21, 2026Updated 7 months ago
- ☆25Dec 2, 2024Updated last year
- Simple, predictable pricing with DigitalOcean hosting • AdAlways know what you'll pay with monthly caps and flat pricing. Enterprise-grade infrastructure trusted by 600k+ customers.
- ☆15Aug 31, 2025Updated last year
- code for "A Full-Scale Hierarchical Encoder-Decoder Network with Cascading Edge-prior for Infrared and Visible Image Fusion".☆15Jan 3, 2024Updated 2 years ago
- Official PyTorch implementation of “FusionGCN: Multi-focus image fusion using superpixel features generation GCN and pixel-level feature …☆13Jul 21, 2026Updated last month
- 桥梁病害检测分割系统后端项目☆18Oct 13, 2025Updated 10 months ago
- ☆23Updated this week
- This project is for vehicle Tracking and detecting using yolo v11 (latest )☆15Sep 30, 2024Updated last year
- Classify Traffic Signs.☆10Jan 31, 2017Updated 9 years ago
- 🎮 东方红魔乡AI智能体 - 基于深度强化学习的弹幕游戏自动化项目 一个集成了计算机视觉、深度强化学习和游戏AI技术的综合性项目,专门为《东方红魔乡》弹幕游戏开发的智能代理。 🚀 核心特性: • 🧠 PPO强化学习算法 - 智能决策和策略优化 • 👁️ YOLO…☆17Jul 1, 2025Updated last year
- ☆20Jul 18, 2024Updated 2 years ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- Semantic-Geometric-Physical-Driven Robot Manipulation Skill Transfer via Skill Library and Tactile Representation☆16Mar 31, 2026Updated 5 months ago
- Device Authentication for Wi-Fi Based on Deep Learning and Radio Frequency Fingerprint☆15Aug 26, 2024Updated 2 years ago
- The official repository for the paper "Statler: State-Maintaining Language Models for Embodied Reasoning"☆13Jun 10, 2024Updated 2 years ago
- ☆18Oct 13, 2024Updated last year
- ReSemAct: Advancing Fine-Grained Robotic Manipulation via Semantic Structuring and Affordance Refinement☆17Jan 5, 2026Updated 8 months ago
- code for A Multi-scale Information Integration Framework for Infrared and Visible Image Fusion☆19Jul 15, 2024Updated 2 years ago
- Implementation of Dat2Vec2.0 for vision☆18Feb 6, 2023Updated 3 years ago
- Simple voice activity detection (VAD) algorithm in Python☆15Aug 10, 2023Updated 3 years ago
- Transform natural language into SQL queries using Azure OpenAI. Visualize database results with interactive charts and explore data effor…☆34Feb 9, 2026Updated 6 months ago
- Simple, predictable pricing with DigitalOcean hosting • AdAlways know what you'll pay with monthly caps and flat pricing. Enterprise-grade infrastructure trusted by 600k+ customers.
- This sample shows how to use the oneAPI Video Processing Library (oneVPL) to perform a single and multi-source video decode and preproces…☆15Jun 15, 2023Updated 3 years ago
- 记录自己用的BILSTM-CRF、ELMo、BERT等来做NER任务的代码。☆26Feb 6, 2020Updated 6 years ago
- ☆22Jan 8, 2024Updated 2 years ago
- Manipulating semantic data within Python☆19Jan 14, 2025Updated last year
- 面向6G边缘智能的大语言模型驱动轻量化射频指纹识别☆16Jul 3, 2025Updated last year
- A basic tutorial (theory and practicals) for Visual Place Recognition.☆17Mar 4, 2024Updated 2 years ago
- [NeurIPS 2021] "Hyperparameter Tuning is All You Need for LISTA" by Xiaohan Chen, Jialin Liu, Zhangyang Wang and Wotao Yin☆23Dec 30, 2021Updated 4 years ago
- NeoNav: Improving the Generalization of Visual Navigation via Generating Next Expected Observations☆17Jul 26, 2020Updated 6 years ago
- 信号调制方式识别与参数估计装置(D 题)2023年全国大学生电子设计竞赛试题☆17Nov 7, 2025Updated 10 months ago
- Serverless GPU API endpoints on Runpod - Get Bonus Credits • AdSkip the infrastructure headaches. Auto-scaling, pay-as-you-go, no-ops approach lets you focus on innovating your application.
- 3D_lut generate for surround view☆13Jul 31, 2019Updated 7 years ago
- This dataset contains time-series signals of radar active jamming, which are used for research related to radar active jammingrecognition…☆19Mar 15, 2024Updated 2 years ago
- [SRIBD Project 2024] A Realtime Wi-Fi Sensing System Demo☆16Jan 14, 2026Updated 7 months ago
- [Paper][CCKS2023] CausE: Towards Causal Knowledge Graph Embedding☆17Jul 30, 2023Updated 3 years ago
- Code for REACT: Real-time Efficient Attribute Clustering and Transfer for Updatable 3D Scene Graph☆17Feb 12, 2026Updated 6 months ago
- 优化wav2lip的执行步骤,将头脸分离、嘴型替换、回补背景三个步骤分离,添加gfpgan强化面部功能,实现提前解帧,流式循环处理,对接obs☆82Dec 16, 2024Updated last year
- Zeroshot Active VIsual Search☆15Jun 18, 2023Updated 3 years ago