πΎ E.T. Bench: Towards Open-Ended Event-Level Video-Language Understanding (NeurIPS 2024)
β73Jan 20, 2025Updated last year
Alternatives and similar repositories for ETBench
Users that are interested in ETBench are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- [ECCV 2024] Learning Video Context as Interleaved Multimodal Sequencesβ47Mar 11, 2025Updated last year
- [NeurIPS 2024] Mitigating Object Hallucination via Concentric Causal Attentionβ69Aug 30, 2025Updated last year
- β32Jul 29, 2024Updated 2 years ago
- β81Nov 24, 2024Updated last year
- [CVPR 2025 Oral] VideoEspresso: A Large-Scale Chain-of-Thought Dataset for Fine-Grained Video Reasoning via Core Frame Selectionβ143Jul 28, 2025Updated last year
- 1-Click AI Models by DigitalOcean Gradient β’ AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- β18Jul 10, 2024Updated 2 years ago
- Data release for Step Differences in Instructional Video (CVPR24)β15Jun 19, 2024Updated 2 years ago
- [CVPR'2024 Highlight] Official PyTorch implementation of the paper "VTimeLLM: Empower LLM to Grasp Video Moments".β295Jun 13, 2024Updated 2 years ago
- [NeurIPS 2023] Rewrite Caption Semantics: Bridging Semantic Gaps for Language-Supervised Semantic Segmentationβ21Jan 3, 2024Updated 2 years ago
- β13Apr 13, 2026Updated 5 months ago
- UniMD: Towards Unifying Moment retrieval and temporal action Detectionβ56Jul 5, 2024Updated 2 years ago
- β29Nov 17, 2024Updated last year
- [ACL 2024 Findings] "TempCompass: Do Video LLMs Really Understand Videos?", Yuanxin Liu, Shicheng Li, Yi Liu, Yuxiang Wang, Shuhuai Ren, β¦β132Apr 4, 2025Updated last year
- A Versatile Video-LLM for Long and Short Video Understanding with Superior Temporal Localization Abilityβ109Nov 28, 2024Updated last year
- Managed hosting for WordPress and PHP on Cloudways β’ AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- [CVPR 2025] LLaVA-ST: A Multimodal Large Language Model for Fine-Grained Spatial-Temporal Understandingβ84Jul 4, 2025Updated last year
- [CVPR 2026] TimeLens: Rethinking Video Temporal Grounding with Multimodal LLMsβ184Jul 21, 2026Updated 2 months ago
- π§ VideoMind: A Chain-of-LoRA Agent for Temporal-Grounded Video Reasoning (ICLR 2026)β360Feb 8, 2026Updated 8 months ago
- [ICCV 2025] LVBench: An Extreme Long Video Understanding Benchmarkβ155Jul 9, 2025Updated last year
- β29Apr 8, 2025Updated last year
- β18Jan 26, 2026Updated 8 months ago
- [EMNLP 2025 Main] The official repo of MMLU-ProX benchmark.β29Aug 26, 2025Updated last year
- [CVPR2026] VideoAuto-R1: Video Auto Reasoning via Thinking Once, Answering Twiceβ90Feb 27, 2026Updated 7 months ago
- Official Implementation of "Chrono: A Simple Blueprint for Representing Time in MLLMs"β96Mar 9, 2025Updated last year
- Deploy on Railway without the complexity - Free Credits Offer β’ AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- [EMNLP 2025 Findings] Grounded-VideoLLM: Sharpening Fine-grained Temporal Grounding in Video Large Language Modelsβ149Aug 21, 2025Updated last year
- β15Oct 30, 2023Updated 2 years ago
- The official code of Towards Balanced Alignment: Modal-Enhanced Semantic Modeling for Video Moment Retrieval (AAAI2024)β32Mar 29, 2024Updated 2 years ago
- Official Repository of paper VideoGPT+: Integrating Image and Video Encoders for Enhanced Video Understandingβ294Sep 5, 2026Updated last month
- Official PyTorch code of GroundVQA (CVPR'24)β63Sep 13, 2024Updated 2 years ago
- [ICCV 2025] Dynamic-VLMβ28Dec 16, 2024Updated last year
- TEMPURA enables video-language models to reason about causal event relationships and generate fine-grained, timestamped descriptions of uβ¦β29Sep 26, 2026Updated last week
- [Neurips 24' D&B] Official Dataloader and Evaluation Scripts for LongVideoBench.β139Jul 27, 2024Updated 2 years ago
- [AAAI 2025] VTG-LLM: Integrating Timestamp Knowledge into Video LLMs for Enhanced Video Temporal Groundingβ131Dec 10, 2024Updated last year
- Managed hosting for WordPress and PHP on Cloudways β’ AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- F-16 is a powerful video large language model (LLM) that perceives high-frame-rate videos, which is developed by the Department of Electrβ¦β43Jul 3, 2025Updated last year
- [ICML 2025] Official PyTorch implementation of LongVUβ435May 8, 2025Updated last year
- Code for the paper: "Sentence Specified Dynamic Video Thumbnail Generation"β34Aug 8, 2019Updated 7 years ago
- β10Jul 5, 2024Updated 2 years ago
- A Comprehensive Survey on Evaluating Reasoning Capabilities in Multimodal Large Language Models.β76Mar 18, 2025Updated last year
- R1-like Video-LLM for Temporal Groundingβ137Jun 20, 2025Updated last year
- Official pytorch implementation of "Explore-And-Match: Bridging Proposal-Based and Proposal-Free With Transformer for Sentence Grounding β¦β42Aug 5, 2022Updated 4 years ago