Official code for MotionBench (CVPR 2025)
☆81Mar 3, 2025Updated last year
Alternatives and similar repositories for MotionBench
Users that are interested in MotionBench are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Accepted By The 39th Annual Conference on Neural Information Processing Systems Datasets and Benchmarks Track☆25Nov 17, 2025Updated 10 months ago
- [ICLR 2026] "VideoReasonBench: Can MLLMs Perform Vision-Centric Complex Video Reasoning?", Yuanxin Liu, Kun Ouyang, Haoning Wu, Yi Liu, L…☆41Jan 30, 2026Updated 7 months ago
- ☆13Apr 13, 2026Updated 5 months ago
- [ICLR 2026] MotionSight's official code implementation.☆48Apr 24, 2026Updated 4 months ago
- F-16 is a powerful video large language model (LLM) that perceives high-frame-rate videos, which is developed by the Department of Electr…☆41Jul 3, 2025Updated last year
- Proton VPN Special Offer - Get 70% off • AdSpecial partner offer. Trusted by over 100 million users worldwide. Tested, Approved and Recommended by Experts.
- [ICCV 2025] LVBench: An Extreme Long Video Understanding Benchmark☆150Jul 9, 2025Updated last year
- Evaluating Knowledge Acquisition from Multi-Discipline Professional Videos☆72Sep 5, 2025Updated last year
- Pixels, Patterns, but no Poetry: To See the World like Humans☆18Aug 11, 2025Updated last year
- [ICLR 2026] Official implementation of the paper "Map the Flow: Revealing Hidden Pathways of Information in VideoLLMs"☆26Mar 3, 2026Updated 6 months ago
- ☆11Aug 4, 2024Updated 2 years ago
- A Shortcut-aware Video-QA Benchmark for Physical Understanding via Minimal Video Pairs☆39Sep 22, 2025Updated 11 months ago
- Extending context length of visual language models☆12Dec 18, 2024Updated last year
- ☆44Nov 8, 2024Updated last year
- This is the official repository of Daily-Omni: Towards Audio-Visual Reasoning with Temporal Alignment across Modalities☆48Jul 26, 2026Updated last month
- Managed Database hosting by DigitalOcean • AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- CoMA: Compositional Human Motion Generation with Multi-modal Agents☆17Sep 9, 2026Updated last week
- https://huggingface.co/datasets/multimodal-reasoning-lab/Zebra-CoT☆140Jan 30, 2026Updated 7 months ago
- ☆18Apr 9, 2026Updated 5 months ago
- [ICCV 2025] Implementation of the paper "Q-Frame: Query-aware Frame Selection and Multi-Resolution Adaptation for Video-LLMs"☆82Oct 25, 2025Updated 10 months ago
- [CVPR 2025] PVC: Progressive Visual Token Compression for Unified Image and Video Processing in Large Vision-Language Models☆54Jun 12, 2025Updated last year
- T* keyframe search for long-form video understanding (CVPR 2025) + LV-Haystack temporal search benchmark code☆97Aug 23, 2026Updated 3 weeks ago
- [ICML 2026] Scripting Multi-Scene Videos with Time-Aware and Structural Audio-Visual Captions☆58Jun 29, 2026Updated 2 months ago
- [ACM MM 2025] ViTCoT: Video-Text Interleaved Chain-of-Thought for Boosting Video Understanding in Large Language Models☆18Jul 15, 2025Updated last year
- ☆29Aug 9, 2025Updated last year
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- [CVPR 2025 Oral] VideoEspresso: A Large-Scale Chain-of-Thought Dataset for Fine-Grained Video Reasoning via Core Frame Selection☆142Jul 28, 2025Updated last year
- VCapsBench: A Large-scale Fine-grained Benchmark for Video Caption Quality Evaluation☆20Jun 2, 2025Updated last year
- Soft-QMIX: Integrating Maximum Entropy For Monotonic Value Function Factorization☆16Jul 3, 2024Updated 2 years ago
- ACDiT: Interpolating Autoregressive Conditional Modeling and Diffusion Transformer☆42Jan 29, 2026Updated 7 months ago
- (3DV 2026) Pytorch implementation of “InterPose: Learning to Generate Human-Object Interactions from Large-Scale Web Videos”☆28Mar 16, 2026Updated 6 months ago
- ☆30Sep 4, 2025Updated last year
- (ICCV2025) Official repository of paper "ViSpeak: Visual Instruction Feedback in Streaming Videos"☆54Jul 1, 2025Updated last year
- TVBench: Redesigning Video-Language Evaluation☆16Jun 9, 2025Updated last year
- Scaling Motion Generation Model with Million-Level Human Motions (ICML 2025)☆68May 14, 2025Updated last year
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- ☆13May 17, 2025Updated last year
- [ICLR 2025] Ready-to-React: Online Reaction Policy for Two-Character Interaction Generation