☆16Aug 18, 2025Updated 11 months ago
Alternatives and similar repositories for DeCAP
Users that are interested in DeCAP are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- TAG: A Simple Yet Effective Temporal-Aware Approach for Zero-Shot Video Temporal Grounding☆24Nov 18, 2025Updated 8 months ago
- 푸시알림 커스텀 서비스, Knocknock☆14Jan 29, 2023Updated 3 years ago
- CIFAR10 ResNets implemented in JAX+Flax☆12Apr 6, 2022Updated 4 years ago
- This is a repository contains the implementation of our NeurIPS'24 paper "Temporal Sentence Grounding with Relevance Feedback in Videos"☆13Aug 22, 2025Updated 10 months ago
- 2019.04.13부터 시작한 스터디 모임의 repository입니다. 주교재: https://www.gitbook.com/book/ddanggle/interpy-kr☆19Aug 4, 2021Updated 4 years ago
- Deploy open-source AI quickly and easily - Special Bonus Offer • AdRunpod Hub is built for open source. One-click deployment and autoscaling endpoints without provisioning your own infrastructure.
- Python Package reimplementation of Holistically-Nested Edge Detection in PyTorch☆12Jan 5, 2021Updated 5 years ago
- Codebase for VidHal: Benchmarking Hallucinations in Vision LLMs☆14Apr 23, 2026Updated 2 months ago
- [EMNLP 2025 Findings] MEXA: Towards General Multimodal Reasoning with Dynamic Multi-Expert Aggregation☆15Aug 22, 2025Updated 10 months ago
- [CVPR'2025] Synthetic Data is an Elegant GIFT for Continual Vision-Language Models☆25Jun 29, 2025Updated last year
- [ECCV2024] Mitigating Background Shift in Class-Incremental Semantic Segmentation☆35Aug 16, 2024Updated last year
- Code for the paper "Finetuning CLIP to Reason about Pairwise Differences"☆21Oct 1, 2024Updated last year
- [ECCV 2024] BeyondScene: Higher-Resolution Human-Centric Scene Generation With Pretrained Diffusion☆21Jul 2, 2024Updated 2 years ago
- [ACL 2026] G2RPO-A: Guided Group Relative Policy Optimization with Adaptive Guidance☆16May 20, 2026Updated 2 months ago
- Code for "Skill-based Chain-of-Thoughts for Domain-Adaptive Video Reasoning [EMNLP 2025 Findings]"☆18Aug 27, 2025Updated 10 months ago
- Deploy open-source AI quickly and easily - Special Bonus Offer • AdRunpod Hub is built for open source. One-click deployment and autoscaling endpoints without provisioning your own infrastructure.
- CVPR2021: Out-of-Distribution Detection Using Union of 1-Dimensional Subspaces☆22Jul 28, 2021Updated 4 years ago
- Paper-Study☆26Nov 9, 2022Updated 3 years ago
- A batched implementation for efficient Qwen2.5-VL inference.☆25Jul 16, 2025Updated last year
- ☆20Jul 21, 2025Updated last year
- ProactiveBench: A Comprehensive Benchmark for VideoLLM Proactive Interaction Evaluation☆18Jan 8, 2026Updated 6 months ago
- Contrast is All You Need For High-Fidelity Text-to-Image Diffusion Models [CVPR 2024]☆27Oct 7, 2024Updated last year
- Training code for CLIP-FlanT5☆31Jul 29, 2024Updated last year
- Parallel Continuous Chain-of-Thought with Jacobi Iteration. Accepted to EMNLP 2025.☆23Mar 29, 2026Updated 3 months ago
- VisualGPTScore for visio-linguistic reasoning☆27Oct 7, 2023Updated 2 years ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- ☆58Dec 18, 2018Updated 7 years ago
- ☆51Nov 7, 2024Updated last year
- Implementation of Em_Garde: a proposal-retrieval framework for streaming video understanding☆26Jun 24, 2026Updated 3 weeks ago
- Evaluation and dataset construction code for the CVPR 2025 paper "Vision-Language Models Do Not Understand Negation"☆47Feb 26, 2026Updated 4 months ago
- Papers of Implicit Reasoning in LLMs.☆25Mar 13, 2025Updated last year
- 파이토치 딥러닝 프로젝트 모음집☆37Dec 18, 2022Updated 3 years ago
- Code for "VideoRepair: Improving Text-to-Video Generation via Misalignment Evaluation and Localized Refinement [ACL 2026 Findings]"☆52Apr 7, 2026Updated 3 months ago
- [NeurIPS2025] The official PyTorch implementation of the "Eyes Wide Open: Ego Proactive Video-LLM for Streaming Video".☆34Dec 25, 2025Updated 6 months ago
- Awesome Vision-Language Compositionality, a comprehensive curation of research papers in literature.☆40Feb 13, 2025Updated last year
- GPUs on demand by Runpod - Special Offer Available • AdRun AI, ML, and HPC workloads on powerful cloud GPUs—without limits or wasted spend. Deploy GPUs in under a minute and pay by the second.
- AdaMoLE: Adaptive Mixture of LoRA Experts☆38Oct 11, 2024Updated last year
- [ICLR 2026] MMDuet2: Enhancing Proactive Interaction of Video MLLMs with Multi-Turn Reinforcement Learning☆40Jan 14, 2026Updated 6 months ago
- A Pokemon card grading system using Deep Learning☆50Dec 23, 2021Updated 4 years ago
- Chain-of-Frames [CVPR 2026]☆40Jul 2, 2025Updated last year
- ☆40Nov 5, 2025Updated 8 months ago
- (ICCV2025) Official repository of paper "ViSpeak: Visual Instruction Feedback in Streaming Videos"☆52Jul 1, 2025Updated last year
- ☆40Jun 18, 2026Updated last month