本项目结合了YOLO的目标检测与分割能力,以及CLIP微调分类能力,YOLO初步框选物体,CLIP对框选物体进行细致化分类,能够在复杂场景下实现物体的精准定位与属性提取。通过对目标物体的检测、分割和细粒度分类,项目特别适用于商品分类、智能货柜管理等任务
☆29Nov 26, 2024Updated last year
Alternatives and similar repositories for YOLO_CLIP_targetDetection
Users that are interested in YOLO_CLIP_targetDetection are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- A survey of occlusion handling in object detection☆26Apr 15, 2021Updated 5 years ago
- 基于MCP协议和LangChain框架实现的企业级AI多Agent多模态系统,包含RAG技术增强的知识检索能力。☆43Mar 10, 2025Updated last year
- 智能医疗问诊系统 - LangGraph Agent + RAG + React☆17Jan 27, 2026Updated 6 months ago
- ☆26Dec 2, 2024Updated last year
- Polyp-SAM++ is the first text-guided polyp-segmentation method using segment anything model (SAM).☆13Aug 23, 2023Updated 2 years ago
- GPUs on demand by Runpod - Special Offer Available • AdRun AI, ML, and HPC workloads on powerful cloud GPUs—without limits or wasted spend. Deploy GPUs in under a minute and pay by the second.
- ☆15Aug 31, 2025Updated 10 months ago
- code for "A Full-Scale Hierarchical Encoder-Decoder Network with Cascading Edge-prior for Infrared and Visible Image Fusion".☆15Jan 3, 2024Updated 2 years ago
- Official PyTorch implementation of “FusionGCN: Multi-focus image fusion using superpixel features generation GCN and pixel-level feature …☆13Jul 21, 2026Updated last week
- 一个面向研究场景的 multi-agent workflow,重点解决检索、整理、交接、续跑、防遗忘和多源 fallback。主要面向OpenClaw, 兼容ClaudeCode/Codex☆16Apr 26, 2026Updated 3 months ago
- Code for RA-L paper "One-shot Learning for Task-oriented Grasping"☆12May 9, 2024Updated 2 years ago
- ☆15Aug 3, 2025Updated 11 months ago
- ☆23Jan 19, 2023Updated 3 years ago
- Classify Traffic Signs.☆10Jan 31, 2017Updated 9 years ago
- 🎮 东方红魔乡AI智能体 - 基于深度强化学习的弹幕游戏自动化项目 一个集成了计算机视觉、深度强化学习和游戏AI技术的综合性项目,专门为《东方红魔乡》弹幕游戏开发的智能代理。 🚀 核心特性: • 🧠 PPO强化学习算法 - 智能决策和策略优化 • 👁️ YOLO…☆16Jul 1, 2025Updated last year
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- An Open Dataset for Wireless Cellular Spectrum Monitoring and Anomaly Detection☆16Mar 16, 2026Updated 4 months ago
- Semantic-Geometric-Physical-Driven Robot Manipulation Skill Transfer via Skill Library and Tactile Representation☆16Mar 31, 2026Updated 3 months ago
- The official repository for the paper "Statler: State-Maintaining Language Models for Embodied Reasoning"☆13Jun 10, 2024Updated 2 years ago
- 开源自定义唤醒词☆17Dec 24, 2025Updated 7 months ago
- ReSemAct: Advancing Fine-Grained Robotic Manipulation via Semantic Structuring and Affordance Refinement☆17Jan 5, 2026Updated 6 months ago
- ☆17Oct 8, 2024Updated last year
- The vocal folds dataset.☆25Jan 5, 2019Updated 7 years ago
- Implementation of Dat2Vec2.0 for vision☆18Feb 6, 2023Updated 3 years ago
- [ECCV2024] "SAM-COD: SAM-guided Unified Framework for Weakly-Supervised Camouflaged Object Detection"☆20Jul 3, 2025Updated last year
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- Simple voice activity detection (VAD) algorithm in Python☆15Aug 10, 2023Updated 2 years ago
- This sample shows how to use the oneAPI Video Processing Library (oneVPL) to perform a single and multi-source video decode and preproces…☆15Jun 15, 2023Updated 3 years ago
- 记录自己用的BILSTM-CRF、ELMo、BERT等来做NER任务的代码。☆26Feb 6, 2020Updated 6 years ago
- 课程代码☆24Apr 30, 2026Updated 2 months ago
- ☆22Jan 8, 2024Updated 2 years ago
- Manipulating semantic data within Python☆19Jan 14, 2025Updated last year
- ☆22Jan 14, 2026Updated 6 months ago
- [NeurIPS 2021] "Hyperparameter Tuning is All You Need for LISTA" by Xiaohan Chen, Jialin Liu, Zhangyang Wang and Wotao Yin☆24Dec 30, 2021Updated 4 years ago
- ☆26Aug 5, 2024Updated last year
- Virtual machines for every use case on DigitalOcean • AdGet dependable uptime with 99.99% SLA, simple security tools, and predictable monthly pricing with DigitalOcean's virtual machines, called Droplets.
- NeoNav: Improving the Generalization of Visual Navigation via Generating Next Expected Observations☆17Jul 26, 2020Updated 6 years ago
- 信号调制方式识别与参数估计装置(D 题)2023年全国大学生电子设计竞赛试题☆17Nov 7, 2025Updated 8 months ago
- ☆30Apr 8, 2025Updated last year
- Unbiased Directed Object Attention Graph for Object Navigation☆15Nov 28, 2022Updated 3 years ago
- 3D_lut generate for surround view☆13Jul 31, 2019Updated 6 years ago
- ☆35May 27, 2025Updated last year
- Instance segmentation for robotics using Mask-RCNN☆18Mar 12, 2020Updated 6 years ago