☆20Apr 21, 2026Updated 3 months ago
Alternatives and similar repositories for CLEAR
Users that are interested in CLEAR are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- [ACMMM'25] Referring Expression Instance Retrieval and A Strong End-to-End Baseline☆19Apr 7, 2026Updated 3 months ago
- [ACMMM 2026] PLUME: Latent Reasoning Based Universal Multimodal Embedding☆25Apr 29, 2026Updated 3 months ago
- A composed retrieval project☆17Apr 9, 2026Updated 3 months ago
- [Findings of ACL 2024]Optimal Transport Guided Correlation Assignment for Multimodal Entity Linking☆17Jun 12, 2025Updated last year
- ☆19Mar 25, 2025Updated last year
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- [CVPR 2026] WISER: Wider Search, Deeper Thinking, and Adaptive Fusion for Training-Free Zero-Shot Composed Image Retrieval☆21Jun 17, 2026Updated last month
- Official implementation of ICLR 2026: Let LLMs Speak Embedding Languages: Generative Text Embeddings via Iterative Contrastive Refinement☆15May 24, 2026Updated 2 months ago
- [ACM MM 2025] ViTCoT: Video-Text Interleaved Chain-of-Thought for Boosting Video Understanding in Large Language Models☆18Jul 15, 2025Updated last year
- 机器学习乐园:主要包括机器学习基础,深度学习实践,工业应用。☆15Nov 14, 2022Updated 3 years ago
- [ICML 2025 Oral] This is the official repository of the paper "What Limits Virtual Agent Application? OmniBench: A Scalable Multi-Dimensi…☆23Jun 12, 2025Updated last year
- ☆12Mar 31, 2025Updated last year
- ☆10Jun 14, 2024Updated 2 years ago
- "Unsupervised Cumulative Domain Adaptation for Foggy Scene Optical Flow", accepted by CVPR 2023☆15Jan 6, 2026Updated 6 months ago
- Unsupervised Variational Translator for Bridging Image Restoration and High-Level Vision Tasks☆14May 12, 2025Updated last year
- GPUs on demand by Runpod - Special Offer Available • AdRun AI, ML, and HPC workloads on powerful cloud GPUs—without limits or wasted spend. Deploy GPUs in under a minute and pay by the second.
- [CVPR 2026] LongVideo-R1: Smart Navigation for Low-cost Long Video Understanding☆51Jul 7, 2026Updated 3 weeks ago
- [ICLR 2026] Official repo for "FrameThinker: Learning to Think with Long Videos via Multi-Turn Frame Spotlighting"☆50Oct 9, 2025Updated 9 months ago
- ☆30Jul 23, 2025Updated last year
- This is the official code for the paper "Reconstruct before Query: Continual Missing Modality Learning with Decomposed Prompt Collaborati…☆12Aug 13, 2024Updated last year
- Some basic visulization functions☆13Jul 27, 2023Updated 3 years ago
- The official code of "CaLa: Complementary Association Learning for Augmenting Composed Image Retrieval"☆15Sep 19, 2024Updated last year
- SciAssess is a comprehensive benchmark for evaluating Large Language Models' proficiency in scientific literature analysis across various…☆89May 21, 2025Updated last year
- [EMNLP'2023 Findings] MoqaGPT, for zero-shot multimodal question answering with LLMs☆13Dec 28, 2024Updated last year
- ☆15Nov 14, 2022Updated 3 years ago
- Managed Database hosting by DigitalOcean • AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- This is the official repository of Daily-Omni: Towards Audio-Visual Reasoning with Temporal Alignment across Modalities☆46Jul 26, 2026Updated last week
- The official repository of "HoliSDiP: Image Super-Resolution via Holistic Semantics and Diffusion Prior"☆10Jan 24, 2025Updated last year
- The official implementation of the paper "Are We Making Progress in Multimodal Domain Generalization? A Comprehensive Benchmark Study"☆15May 8, 2026Updated 2 months ago
- ParaVT: Taming the Tool Prior Paradox for Parallel Tool Use in Agentic Video Reinforcement Learning☆56Jun 2, 2026Updated 2 months ago
- 包含作业代码及代码分析☆10Aug 13, 2021Updated 4 years ago
- ☆12Jun 12, 2024Updated 2 years ago
- Implementation of paper "Vision Language Model for Interpretable and Fine-grained Detection of Safety Compliance in Diverse Workplaces"☆15Jan 17, 2025Updated last year
- This is an official implementation of "Self-supervised Image Denoising with Downsampled Invariance Loss and Conditional Blind-Spot Networ…☆18Sep 4, 2023Updated 2 years ago
- Extending context length of visual language models☆12Dec 18, 2024Updated last year
- Bare Metal GPUs on DigitalOcean Gradient AI • AdPurpose-built for serious AI teams training foundational models, running large-scale inference, and pushing the boundaries of what's possible.
- [CVPR 2026] VideoSeek: Long-Horizon Video Agent with Tool-Guided Seeking☆64Mar 23, 2026Updated 4 months ago
- Code of the paper "FBSDiff: Plug-and-Play Frequency Band Substitution of Diffusion Features for Highly Controllable Text-Driven Image Tra…☆22Apr 28, 2025Updated last year
- Official Code of "Random Parameter Pruning Attack (Accepeted by CVPR26)"☆16Feb 26, 2026Updated 5 months ago
- Visual Speech Recongnition☆21Dec 24, 2024Updated last year
- [ICLR 2026] "VideoReasonBench: Can MLLMs Perform Vision-Centric Complex Video Reasoning?", Yuanxin Liu, Kun Ouyang, Haoning Wu, Yi Liu, L…☆41Jan 30, 2026Updated 6 months ago
- [ICML 2026] Scripting Multi-Scene Videos with Time-Aware and Structural Audio-Visual Captions☆52Jun 29, 2026Updated last month
- [EMNLP 2025] The official implementation of "Zero-shot Multimodal Document Retrieval via Cross-Modal Question Generation"☆15Aug 26, 2025Updated 11 months ago