☆20Apr 21, 2026Updated 4 months ago
Alternatives and similar repositories for CLEAR
Users that are interested in CLEAR are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- [ACMMM 2026] PLUME: Latent Reasoning Based Universal Multimodal Embedding☆25Apr 29, 2026Updated 3 months ago
- [Findings of ACL 2024]Optimal Transport Guided Correlation Assignment for Multimodal Entity Linking☆17Jun 12, 2025Updated last year
- ☆19Mar 25, 2025Updated last year
- The official implementation of COOPER: A Unified Model for Cooperative Perception and Reasoning in Spatial Intelligence.☆38Jul 1, 2026Updated last month
- [ACM MM 2025] ViTCoT: Video-Text Interleaved Chain-of-Thought for Boosting Video Understanding in Large Language Models☆18Jul 15, 2025Updated last year
- Deploy open-source AI quickly and easily - Special Bonus Offer • AdRunpod Hub is built for open source. One-click deployment and autoscaling endpoints without provisioning your own infrastructure.
- Latest Advances on Modality Priors in Multimodal Large Language Models☆31Dec 10, 2025Updated 8 months ago
- 机器学习乐园:主要包括机器学习基础,深度学习实践,工业应用。☆15Nov 14, 2022Updated 3 years ago
- ☆12Mar 31, 2025Updated last year
- ☆12Feb 25, 2024Updated 2 years ago
- "Unsupervised Cumulative Domain Adaptation for Foggy Scene Optical Flow", accepted by CVPR 2023☆15Jan 6, 2026Updated 7 months ago
- ☆27Feb 3, 2026Updated 6 months ago
- Unsupervised Variational Translator for Bridging Image Restoration and High-Level Vision Tasks☆14May 12, 2025Updated last year
- [CVPR 2026] LongVideo-R1: Smart Navigation for Low-cost Long Video Understanding☆52Jul 7, 2026Updated last month
- [ICLR 2026] Official repo for "FrameThinker: Learning to Think with Long Videos via Multi-Turn Frame Spotlighting"☆56Oct 9, 2025Updated 10 months ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- This is the official code for the paper "Reconstruct before Query: Continual Missing Modality Learning with Decomposed Prompt Collaborati…☆12Aug 13, 2024Updated 2 years ago
- Codes and data for ACL 2023 Findings paper "Click: Controllable Text Generation with Sequence Likelihood Contrastive Learning"☆18Feb 26, 2024Updated 2 years ago
- SciAssess is a comprehensive benchmark for evaluating Large Language Models' proficiency in scientific literature analysis across various…☆89May 21, 2025Updated last year
- [EMNLP'2023 Findings] MoqaGPT, for zero-shot multimodal question answering with LLMs☆13Dec 28, 2024Updated last year
- Fast and memory-efficient exact attention☆21Aug 17, 2026Updated last week
- This is the official repository of Daily-Omni: Towards Audio-Visual Reasoning with Temporal Alignment across Modalities☆47Jul 26, 2026Updated 3 weeks ago
- The official repository of "HoliSDiP: Image Super-Resolution via Holistic Semantics and Diffusion Prior"☆10Jan 24, 2025Updated last year
- ☆16Dec 23, 2025Updated 8 months ago
- ParaVT: Taming the Tool Prior Paradox for Parallel Tool Use in Agentic Video Reinforcement Learning☆55Jun 2, 2026Updated 2 months ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- 包含作业代码及代码分析☆10Aug 13, 2021Updated 5 years ago
- ☆12Jun 12, 2024Updated 2 years ago
- ☆27Jul 8, 2026Updated last month
- This is an official implementation of "Self-supervised Image Denoising with Downsampled Invariance Loss and Conditional Blind-Spot Networ…☆18Sep 4, 2023Updated 2 years ago
- The official code of paper "OMS-DPM: Optimizing Model Schedule for Diffusion Probabilistic Model" accepted by ICML 2023☆24Oct 11, 2023Updated 2 years ago
- Extending context length of visual language models☆12Dec 18, 2024Updated last year
- [CVPR 2026] VideoSeek: Long-Horizon Video Agent with Tool-Guided Seeking☆68Mar 23, 2026Updated 5 months ago
- Official Code of "Random Parameter Pruning Attack (Accepeted by CVPR26)"☆16Feb 26, 2026Updated 5 months ago
- [ICLR 2026] "VideoReasonBench: Can MLLMs Perform Vision-Centric Complex Video Reasoning?", Yuanxin Liu, Kun Ouyang, Haoning Wu, Yi Liu, L…☆41Jan 30, 2026Updated 6 months ago
- GPUs on demand by Runpod - Special Offer Available • AdRun AI, ML, and HPC workloads on powerful cloud GPUs—without limits or wasted spend. Deploy GPUs in under a minute and pay by the second.
- ☆20May 15, 2026Updated 3 months ago
- [ICML 2026] Scripting Multi-Scene Videos with Time-Aware and Structural Audio-Visual Captions☆54Jun 29, 2026Updated last month
- [EMNLP 2025] The official implementation of "Zero-shot Multimodal Document Retrieval via Cross-Modal Question Generation"☆15Aug 26, 2025Updated 11 months ago
- ☆25Feb 21, 2025Updated last year
- Can VLMs understand students' hand-drawn math work?☆19Jan 20, 2026Updated 7 months ago
- code of cvpr26 paper Symphony☆17Apr 7, 2026Updated 4 months ago
- The code implementation for UME-R1: Exploring Reasoning-Driven Generative Multimodal Embeddings (ICLR 2026).☆71Feb 25, 2026Updated 5 months ago