MVU-Eval @NeurIPS DB 2025
☆16Nov 11, 2025Updated 10 months ago
Alternatives and similar repositories for MVU-Eval
Users that are interested in MVU-Eval are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- The Source Code for ViDiC-1K☆15Mar 13, 2026Updated 6 months ago
- The Source Code for MT-Video-Bench @ ACL Findings 2026☆21Jan 20, 2026Updated 7 months ago
- The Source Code for OmniVideoBench @ICLR 2026☆78Feb 12, 2026Updated 7 months ago
- [AAAI 2026] CrossVid: A Comprehensive Benchmark for Evaluating Cross-Video Reasoning in Multimodal Large Language Models☆23Jul 9, 2026Updated 2 months ago
- The Source Code for DR3-Eval☆40Aug 12, 2026Updated last month
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- The Source Code for T2AV-Compass @ ICML 2026☆22Jun 21, 2026Updated 2 months ago
- The Source Code for WebCompass☆25May 2, 2026Updated 4 months ago
- ☆19Aug 28, 2025Updated last year
- [ICML2026] OmniSIFT: Modality-Asymmetric Token Compression for Efficient Omni-modal Large Language Models☆30May 21, 2026Updated 3 months ago
- The official repository for paper "FlexSelect: Flexible Token Selection for Efficient Long Video Understanding".☆31Sep 19, 2025Updated 11 months ago
- ☆40Jun 21, 2026Updated 2 months ago
- ACM MM 2022 paper_AVQA: A Dataset for Audio-Visual Question Answering on Videos☆15Aug 17, 2023Updated 3 years ago
- Code for the paper "Controllable Video Captioning with an Exemplar Sentence"☆12Apr 14, 2021Updated 5 years ago
- Starter Code for VALUE benchmark☆79Aug 23, 2022Updated 4 years ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- [ICLR 2026] Mixing Importance with Diversity: Joint Optimization for KV Cache Compression in Large Vision-Language Models☆31Mar 21, 2026Updated 5 months ago
- ~ ZeroShot Learning for ZJU AI Competition (GAN Approach)☆12Nov 3, 2025Updated 10 months ago
- VideoEval-Pro: Robust and Realistic Long Video Understanding Evaluation [TMLR26]☆15Jun 1, 2026Updated 3 months ago
- [ICLR 2023] Soft Neighbors are Positive Supporters in Contrastive Visual Representation Learning☆15Aug 2, 2023Updated 3 years ago
- MMhops-R1: Multimodal Multi-hop Reasoning☆17Aug 17, 2026Updated last month
- PyTorch implementation of Graph Convolutional Networks in Feature Space for Image Deblurring and Super-resolution, IJCNN 2021.☆12Nov 14, 2021Updated 4 years ago
- F-16 is a powerful video large language model (LLM) that perceives high-frame-rate videos, which is developed by the Department of Electr…☆41Jul 3, 2025Updated last year
- Controllable mage captioning model with unsupervised modes☆21Apr 14, 2023Updated 3 years ago
- [ICLR-2026] Official Implementation of our paper "THOR: Tool-Integrated Hierarchical Optimization via RL for Mathematical Reasoning".☆32Aug 4, 2026Updated last month
- Virtual machines for every use case on DigitalOcean • AdGet dependable uptime with 99.99% SLA, simple security tools, and predictable monthly pricing with DigitalOcean's virtual machines, called Droplets.
- Pipelined MIPS architecture created in Verilog. Includes data forwarding and hazard detection.☆16Apr 1, 2018Updated 8 years ago
- A minimal, educational HEVC (H.265) encoder written in Python.☆53Feb 23, 2026Updated 6 months ago
- ☆13Jul 3, 2024Updated 2 years ago
- build vgg16 with pytorch 0.4.0 for classification of CIFAR datasets☆10Mar 31, 2019Updated 7 years ago
- [CVPR 2026] Variation-aware Vision Token Dropping for Faster Large Vision-Language Models☆35May 27, 2026Updated 3 months ago
- ☐ ☐ A simple, out-of-the-box and cross-platform bbox annotation tool by Python. Try it by `pip install easybox`☆10May 28, 2021Updated 5 years ago
- Seeing is Not Reasoning: MVPBench for Graph-based Evaluation of Multi-path Visual Physical CoT☆14Jul 30, 2025Updated last year
- UMB: Understanding Model Behavior for Open-World object Detection (NeurIPS 2024)☆12May 26, 2024Updated 2 years ago
- ☆12May 15, 2025Updated last year
- Simple, predictable pricing with DigitalOcean hosting • AdAlways know what you'll pay with monthly caps and flat pricing. Enterprise-grade infrastructure trusted by 600k+ customers.
- The official source code of our AAAI25 paper "D&M: Enriching E-commerce Videos with Sound Effects by Key Moment Detection and SFX Matchin…☆10Feb 9, 2025Updated last year
- TAG: A Simple Yet Effective Temporal-Aware Approach for Zero-Shot Video Temporal Grounding☆25Nov 18, 2025Updated 9 months ago
- [NeurIPS 2025] VideoRFT: Incentivizing Video Reasoning Capability in MLLMs via Reinforced Fine-Tuning☆66Jan 6, 2026Updated 8 months ago
- codes for ICML2021 paper iDARTS: Differentiable Architecture Search with Stochastic Implicit Gradients☆10May 27, 2021Updated 5 years ago
- a survey on deep research☆48Sep 9, 2025Updated last year
- ☆10Nov 27, 2024Updated last year
- 🫧 Code for Holistic Reasoning with Long-Context LMs: A Benchmark for Database Operations on Massive Textual Data (Maekawa*, Iso* et al.…☆13Feb 25, 2025Updated last year