AIVIETNAM-Hub / Hybrid-Unified-and-Iterative-A-Novel-Framework-for-Text-based-Person-Anomaly-RetrievalView on GitHub
☆15Jul 23, 2025Updated last year
Alternatives and similar repositories for Hybrid-Unified-and-Iterative-A-Novel-Framework-for-Text-based-Person-Anomaly-Retrieval
Users that are interested in Hybrid-Unified-and-Iterative-A-Novel-Framework-for-Text-based-Person-Anomaly-Retrieval are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- ☆19Mar 5, 2024Updated 2 years ago
- A codebase for flexible and efficient Image Text Representation Alignment☆24Jun 20, 2023Updated 3 years ago
- [NeurIPS 2025] VADTree: Explainable Training-Free Video Anomaly Detection via Hierarchical Granularity-Aware Tree☆19Jun 9, 2026Updated 2 months ago
- The official code of "Beyond Walking: A Large-Scale Image-Text Benchmark for Text-based Person Anomaly Search"☆37Jul 25, 2026Updated last month
- ☆36Dec 31, 2024Updated last year
- Bare Metal GPUs on DigitalOcean Gradient AI • AdPurpose-built for serious AI teams training foundational models, running large-scale inference, and pushing the boundaries of what's possible.
- This is a repo for the paper "Networking Systems for Video Anomaly Detection: A Tutorial and Survey". Paper: https://arxiv.org/abs/2405.1…☆33Mar 26, 2025Updated last year
- Source code of our AAAI 2024 paper "Cross-Modal and Uni-Modal Soft-Label Alignment for Image-Text Retrieval"☆55Mar 28, 2024Updated 2 years ago
- Code for Retrieval-Augmented Perception (ICML 2025)☆75Apr 22, 2026Updated 4 months ago
- [EMNLP-2025 Oral] ZoomEye: Enhancing Multimodal LLMs with Human-Like Zooming Capabilities through Tree-Based Image Exploration☆91Nov 20, 2025Updated 9 months ago
- [CVPR 2025 Highlight] Your Large Vision-Language Model Only Needs A Few Attention Heads For Visual Grounding☆84Aug 31, 2025Updated last year
- [ICCV 2025] Rethinking Multi-modal Object Detection from the Perspective of Mono-Modality Feature Learning☆87Mar 24, 2026Updated 5 months ago
- Implement of the paper "Time Travelling Pixels: Bitemporal Features Integration with Foundation Model for Remote Sensing Image Change Det…☆95Apr 25, 2024Updated 2 years ago
- [ICLR'25] Official code for the paper 'MLLMs Know Where to Look: Training-free Perception of Small Visual Details with Multimodal LLMs'☆387Apr 20, 2025Updated last year
- Chat with RS-ChatGPT and get the remote sensing interpretation results and the response!☆242Mar 27, 2024Updated 2 years ago
- Virtual machines for every use case on DigitalOcean • AdGet dependable uptime with 99.99% SLA, simple security tools, and predictable monthly pricing with DigitalOcean's virtual machines, called Droplets.
- Hong, D., Zhang, B., Li, X., Li, Y., Li, C., Yao, J., Yokoya, N., Li, H., Ghamisi, P., Jia, X., Plaza, A. and Gamba, P., Benediktsson, J.…☆281Dec 28, 2024Updated last year
- ICAFusion: Iterative Cross-Attention Guided Feature Fusion for Multispectral Object Detection, Pattern Recognition☆273Jun 3, 2026Updated 3 months ago
- RGB-T Fusion, RGB-T SOD, RGB-T Vehicle Detection, RGB-T Crowd Counting, RGB-T Pedestrian Detection, RGB-T Semantic Segmeantaion, RGB-T Tr…☆237Aug 9, 2026Updated 3 weeks ago
- VadCLIP official Pytorch implementation☆238Mar 10, 2024Updated 2 years ago
- PyTorch Implementation of "V* : Guided Visual Search as a Core Mechanism in Multimodal LLMs"☆713Jan 7, 2024Updated 2 years ago
- [CVPR 2023] Revisiting Weak-to-Strong Consistency in Semi-Supervised Semantic Segmentation☆582Oct 15, 2024Updated last year
- [NeurIPS 2021] LoveDA: A Remote Sensing Land-Cover Dataset for Domain Adaptive Semantic Segmentation☆561Aug 29, 2026Updated last week
- 📖 This is a repository for organizing papers, codes and other resources related to unified multimodal models.☆831Oct 10, 2025Updated 10 months ago
- 🛰️ Official repository of paper "RemoteCLIP: A Vision Language Foundation Model for Remote Sensing" (IEEE TGRS)☆591Jun 27, 2024Updated 2 years ago
- End-to-end encrypted cloud storage - Proton Drive • AdSpecial offer: 40% Off Yearly / 80% Off First Month. Protect your most important files, photos, and documents from prying eyes.
- YOLOv11-RGBT: Towards a Comprehensive Single-Stage Multispectral Object Detection Framework(Supports RGBT detection for all YOLO series f…☆719Dec 15, 2025Updated 8 months ago
- A collection of deep learning based RGB-T-Fusion methods, codes, and datasets. The main directions involved are Multispectral Pedestrian …☆755Apr 17, 2026Updated 4 months ago
- Easily compute clip embeddings and build a clip retrieval system with them☆2,795Mar 28, 2026Updated 5 months ago
- Collection of AWESOME vision-language models for vision tasks☆3,126Oct 14, 2025Updated 10 months ago
- NanoDet-Plus⚡Super fast and lightweight anchor-free object detection model. 🔥Only 980 KB(int8) / 1.8MB (fp16) and run 97FPS on cellphone…☆6,254Aug 8, 2024Updated 2 years ago
- An open source implementation of CLIP.☆14,119Updated this week
- A vector index built on TurboQuant, written in Rust with Python bindings☆16,676Aug 21, 2026Updated 2 weeks ago
- Advanced AI Explainability for computer vision. Support for CNNs, Vision Transformers, Classification, Object detection, Segmentation, I…☆12,965Aug 13, 2026Updated 3 weeks ago
- OpenMMLab Semantic Segmentation Toolbox and Benchmark.☆9,941Aug 13, 2024Updated 2 years ago
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- 21 Lessons, Get Started Building with Generative AI☆119,265Updated this week
- Tongyi Deep Research, the Leading Open-source Deep Research Agent☆19,914Feb 27, 2026Updated 6 months ago
- The official Meta Llama 3 GitHub site☆29,232Jan 26, 2025Updated last year
- Stealth Chromium that passes every bot detection test. Drop-in Playwright replacement with source-level fingerprint patches. 30/30 tests …☆31,227Updated this week
- 100+ AI Agents, Agent Skills and RAG Apps - Free and Open Source.☆136,362Updated this week
- 🧑🏫 60+ Implementations/tutorials of deep learning papers with side-by-side notes 📝; including transformers (original, xl, switch, fee…☆67,394Jan 22, 2026Updated 7 months ago
- A latent text-to-image diffusion model☆73,385Jun 18, 2024Updated 2 years ago