[ACM MM23] CLIP-Count: Towards Text-Guided Zero-Shot Object Counting
☆124Mar 20, 2024Updated 2 years ago
Alternatives and similar repositories for CLIP-Count
Users that are interested in CLIP-Count are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Includes FSC-147-D and the code for training and testing the CounTX model from the paper Open-world Text-specified Object Counting.☆42Sep 27, 2024Updated last year
- [AAAI 2024] VLCounter: Text-aware Visual Representation for Zero-Shot Object Counting☆44Nov 19, 2024Updated last year
- CounTR: Transformer-based Generalised Visual Counting☆127Jul 11, 2024Updated 2 years ago
- [CVPR 2023] CrowdCLIP: Unsupervised Crowd Counting via Vision-Language Model☆92Jul 28, 2023Updated 3 years ago
- Learning to Count without Annotations☆24May 24, 2024Updated 2 years ago
- Managed Database hosting by DigitalOcean • AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- CVPR2023 Zero-shot Counting☆60Mar 23, 2025Updated last year
- This is the official implementation of: Learning to Count Anything: Reference-less Class-agnostic Counting with Weak Supervision Michael …☆39Jul 12, 2024Updated 2 years ago
- ☆55Dec 14, 2023Updated 2 years ago
- LOCA - A Low-Shot Object Counting Network With Iterative Prototype Adaptation (ICCV 2023)☆65Jul 3, 2024Updated 2 years ago
- The codes for ACM Multimedia 2023 paper 'DAOT: Domain-Agnostically Aligned Optimal Transport for Domain-Adaptive Crowd Counting. '☆14Jan 12, 2024Updated 2 years ago
- Official PyTorch implementation of FusionCount: Efficient Crowd Counting via Multiscale Feature Fusion☆13Oct 25, 2022Updated 3 years ago
- [ECCV 2022] An End-to-End Transformer Model for Crowd Localization☆116Mar 20, 2023Updated 3 years ago
- The official implementation of the crowd counting model CLIP-EBC.☆99Jul 17, 2024Updated 2 years ago
- Spatio-channel Attention Blocks for Cross-modal Crowd Counting -- Official Pytorch Implementation (ACCV'22, Oral)☆28Dec 4, 2023Updated 2 years ago
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- Official implement of CVPR2025 paper: "T2ICount: Enhancing Cross-modal Understanding for zero-shot Counting"☆27Apr 9, 2025Updated last year
- Official Implement of CVPR 2022 paper 'Boosting Crowd Counting via Multifaceted Attention'☆125May 10, 2025Updated last year
- [WACV 2023] Few-shot Object Counting with Similarity-Aware Feature Enhancement☆142Oct 10, 2023Updated 2 years ago
- ☆28Feb 21, 2025Updated last year
- ☆30Jun 10, 2024Updated 2 years ago
- GeckoNum Benchmark for T2I Model Eval.☆15Dec 5, 2024Updated last year
- Code and datasets for "Text encoders are performance bottlenecks in contrastive vision-language models". Coming soon!☆11May 24, 2023Updated 3 years ago
- PyTorch implementations of the paper: "DR.VIC: Decomposition and Reasoning for Video Individual Counting, CVPR, 2022"☆61Jun 12, 2023Updated 3 years ago
- an empirical study on few-shot counting using segment anything (SAM)☆95Apr 25, 2023Updated 3 years ago
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- [CVPR 2024] Contrasting Intra-Modal and Ranking Cross-Modal Hard Negatives to Enhance Visio-Linguistic Fine-grained Understanding☆55Apr 7, 2025Updated last year
- (CVPR 2024) Point, Segment and Count: A Generalized Framework for Object Counting☆127Nov 12, 2024Updated last year
- PSGCNet: A Pyramidal Scale and Global Context Guided Network for Dense Object Counting in Remote-Sensing Images☆20Jun 13, 2022Updated 4 years ago
- Official repo for CVPR2024 paper "Single Domain Generalization for Crowd Counting"☆99Mar 31, 2025Updated last year
- This method uses Segment Anything and CLIP to ground and count any object that matches a custom text prompt, without requiring any point …☆183Apr 22, 2023Updated 3 years ago
- Includes the code for training and testing the CountGD++ model from the paper CountGD++: Generalized Prompting for Open-World Counting.☆68Aug 22, 2026Updated last month
- Official Implement of ECCV 2024 paper "Multi-modal Crowd Counting via a Broker Modality"☆18Mar 19, 2026Updated 6 months ago
- ☆17May 19, 2026Updated 4 months ago
- Disentangled Graph Variational Auto-Encoder for Multimodal Recommendation with Interpretability, IEEE TMM☆16Jun 3, 2025Updated last year
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- SVL-Adapter: Self-Supervised Adapter for Vision-Language Pretrained Models☆21Jan 11, 2024Updated 2 years ago
- MAtch, eXpand and Improve: Unsupervised Finetuning for Zero-Shot Action Recognition with Language Knowledge (ICCV 2023)☆31Sep 5, 2023Updated 3 years ago
- Examples of Verbalized Machine Learning (VML)☆16Mar 16, 2025Updated last year
- ☆15Feb 24, 2023Updated 3 years ago
- CVPR2024: Dual Memory Networks: A Versatile Adaptation Approach for Vision-Language Models☆97Jul 4, 2024Updated 2 years ago
- ☆146Apr 2, 2024Updated 2 years ago
- [TIP 2023] Redesigning Multi-Scale Neural Network for Crowd Counting☆24Jul 2, 2024Updated 2 years ago