llyx97 / video_reason_benchView external linksLinks
[ICLR 2026] "VideoReasonBench: Can MLLMs Perform Vision-Centric Complex Video Reasoning?", Yuanxin Liu, Kun Ouyang, Haoning Wu, Yi Liu, Lin Sui, Xinhao Li, Yan Zhong, Y. Charles, Xinyu Zhou, Xu Sun
☆37Jan 30, 2026Updated 2 weeks ago
Alternatives and similar repositories for video_reason_bench
Users that are interested in video_reason_bench are comparing it to the libraries listed below
Sorting:
- Code for our EMNLP-2022 paper: "Language Prior Is Not the Only Shortcut: A Benchmark for Shortcut Learning in VQA"☆40Nov 1, 2022Updated 3 years ago
- 机器学习乐园:主要包括机器学习基础,深度学习实践,工业应用。☆15Nov 14, 2022Updated 3 years ago
- [ICCV 2025] Official Implementation of "Shot-by-Shot: Film-Grammar-Aware Training-Free Audio Description Generation". Junyu Xie, Tengda H…☆20Jul 26, 2025Updated 6 months ago
- ☆27Jul 23, 2025Updated 6 months ago
- Official repository of the video reasoning benchmark MMR-V. Can Your MLLMs "Think with Video"?☆38Jun 23, 2025Updated 7 months ago
- [CVPR 2025] GPS as a Control Signal for Image Generation☆25Mar 18, 2025Updated 10 months ago
- [ACMMM2025] Official released code for VQA² series models☆61Oct 19, 2025Updated 3 months ago
- VideoAuto-R1: Video Auto Reasoning via Thinking Once, Answering Twice☆61Jan 9, 2026Updated last month
- ☆32Jul 29, 2024Updated last year
- ☆37Nov 8, 2024Updated last year
- This repository collects awesome representative papers and resources for "From Pre-training to Post-training: A Survey on Time Series Fou…☆30Feb 1, 2026Updated 2 weeks ago
- [ICCV 2025] "Player-Centric Multimodal Prompt Generation for Large Language Model Based Identity-Aware Basketball Video Captioning".☆15Dec 11, 2025Updated 2 months ago
- Embodied-Planner-R1: Unleashing Embodied Task Planning Ability in LLMs via Reinforcement Learning☆24Jan 5, 2026Updated last month
- ☆13Nov 5, 2024Updated last year
- [NeurIPS'2024] Invertible Consistency Distillation for Text-Guided Image Editing in Around 7 Steps☆101Jul 4, 2024Updated last year
- [ICLR 2026] Official Implementation of ProxyThinker: Test-Time Guidance through Small Visual Reasoners.☆19Sep 24, 2025Updated 4 months ago
- ☆10Aug 3, 2022Updated 3 years ago
- [NeurIPS2022] Perceptual Attacks of No-Reference Image Quality Models with Human-in-the-Loop☆14Apr 13, 2023Updated 2 years ago
- F-16 is a powerful video large language model (LLM) that perceives high-frame-rate videos, which is developed by the Department of Electr…☆34Jul 3, 2025Updated 7 months ago
- [NeurIPS 2025] Code for Low-Rank Head Avatar Personalization with Registers☆17Dec 9, 2025Updated 2 months ago
- The PyTorch implementation of DSM (EMNLP 2022).☆10Mar 26, 2024Updated last year
- Official repository of paper "LOVE-R1: Advancing Long Video Understanding with Adaptive Zoom-in Mechanism via Multi-Step Reasoning"☆20Nov 1, 2025Updated 3 months ago
- Implementation of the algorithms in the research paper iNeRF☆10Oct 1, 2021Updated 4 years ago
- Generating Summaries with Controllable Readability Levels (EMNLP 2023)☆14Aug 6, 2025Updated 6 months ago
- ☆11Jan 27, 2020Updated 6 years ago
- Self Evolving Large Multimodal Models with Continuous Rewards☆19Nov 21, 2025Updated 2 months ago
- ☆21Feb 3, 2026Updated last week
- [AAAI 2026] Official Code for VQAThinker: Exploring Generalizable and Explainable Video Quality Assessment via Reinforcement Learning☆19Nov 28, 2025Updated 2 months ago
- Official InfiniBench: A Benchmark for Large Multi-Modal Models in Long-Form Movies and TV Shows☆19Nov 4, 2025Updated 3 months ago
- ☆12May 28, 2020Updated 5 years ago
- [ECCV 2024] Official code repository of paper titled "Efficient 3D-Aware Facial Image Editing Via Attribute-Specific Prompt Learning"☆10Aug 2, 2024Updated last year
- ☆15Feb 11, 2025Updated last year
- This is a repository contains materials for future survey submission☆12Jan 17, 2024Updated 2 years ago
- CVPR 2025 Accepted Papers☆23Dec 20, 2025Updated last month
- ☆18Apr 10, 2025Updated 10 months ago
- This is the official repository of Daily-Omni: Towards Audio-Visual Reasoning with Temporal Alignment across Modalities☆36Jul 4, 2025Updated 7 months ago
- ☆64Feb 4, 2026Updated last week
- Official Implementation for "SiLVR : A Simple Language-based Video Reasoning Framework"☆19Jan 18, 2026Updated 3 weeks ago
- [MICCAI'25 Oral] Endo3R: Unified Online Reconstruction from Dynamic Monocular Endoscopic Video☆31Sep 25, 2025Updated 4 months ago