Learning to cut end-to-end pretrained modules
☆39Apr 17, 2025Updated last year
Alternatives and similar repositories for MovieCuts
Users that are interested in MovieCuts are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- This is the official repository for our ECCV 2022 paper titled, "The Anatomy of Video Editing: A Dataset and Benchmark Suite for AI-Assis…☆56Nov 28, 2022Updated 3 years ago
- [ICCV 2025] Official Implementation of "Shot-by-Shot: Film-Grammar-Aware Training-Free Audio Description Generation". Junyu Xie, Tengda H…☆28May 16, 2026Updated 4 months ago
- [ECCV 2022] AutoTransition: Learning to Recommend Video Transition Effects☆69Mar 6, 2025Updated last year
- MAD: A Scalable Dataset for Language Grounding in Videos from Movie Audio Descriptions☆177Oct 22, 2023Updated 2 years ago
- [CVPR'23 Highlight] AutoAD: Movie Description in Context.☆105Nov 6, 2024Updated last year
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- Tools for movie and video research☆315Jun 20, 2022Updated 4 years ago
- Implementation of Pix2Seq in PyTorch☆10Feb 3, 2022Updated 4 years ago
- [CVPR 2025] VidComposition: Can MLLMs Analyze Compositions in Compiled Videos?☆31May 10, 2025Updated last year
- ☆21Aug 26, 2025Updated last year
- ☆11Nov 22, 2019Updated 6 years ago
- This repository contains the dataset, codebase, and benchmarks for our paper: <CNVid-3.5M: Build, Filter, and Pre-train the Large-scale P…☆26Nov 28, 2023Updated 2 years ago
- A dataset for Audio-Visual Sound Event Detection in Movies☆26Jan 23, 2023Updated 3 years ago
- Official Codebase of "A Unified Audio-Visual Learning Framework for Localization, Separation, and Recognition" (ICML 2023)☆12Jun 1, 2023Updated 3 years ago
- ☆30Mar 3, 2025Updated last year
- Virtual machines for every use case on DigitalOcean • AdGet dependable uptime with 99.99% SLA, simple security tools, and predictable monthly pricing with DigitalOcean's virtual machines, called Droplets.
- ☆13May 17, 2025Updated last year
- Codebase for CVPR2020 A Local-to-Global Approach to Multi-modal Movie Scene Segmentation☆241May 20, 2024Updated 2 years ago
- Code for the CVPR 2020 paper "Learning Instance Occlusion for Panoptic Segmentation"☆13Jun 17, 2020Updated 6 years ago
- Narrative movie understanding benchmark☆76Jun 11, 2025Updated last year
- ☆13Nov 15, 2022Updated 3 years ago
- Classification of the video file into one of the 5 classes (Static, Pan, Tilt, Zoom, Motion-Still) based on the camera/object motion in t…☆16Dec 27, 2018Updated 7 years ago
- ☆16Sep 6, 2024Updated 2 years ago
- ☆11Oct 2, 2024Updated 2 years ago
- Database of cinematographic data of real films through film annotations.☆15Aug 4, 2020Updated 6 years ago
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- Code and datasets for "Text encoders are performance bottlenecks in contrastive vision-language models". Coming soon!☆11May 24, 2023Updated 3 years ago
- Consistent Human Image and Video Generation with Spatially Conditioned Diffusion☆17Sep 1, 2025Updated last year
- [COLM 2025] Official code for "When To Solve, When To Verify: Compute-Optimal Problem Solving and Generative Verification for LLM Reasoni…☆15Oct 31, 2025Updated 11 months ago
- Official implementation of "Interpreting and Controlling Vision Foundation Models via Text Explanations"☆14May 29, 2024Updated 2 years ago
- ☆163Jan 16, 2025Updated last year
- [ACCV 2024] Official Implementation of "AutoAD-Zero: A Training-Free Framework for Zero-Shot Audio Description". Junyu Xie, Tengda Han, M…☆33May 16, 2026Updated 4 months ago
- ☆39Oct 19, 2024Updated last year
- ☆60Jun 4, 2025Updated last year
- A new multi-shot video understanding benchmark Shot2Story with comprehensive video summaries and detailed shot-level captions.☆181Jan 30, 2025Updated last year
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- ShotBench: Expert-Level Cinematic Understanding in Vision-Language Models☆107Sep 12, 2025Updated last year
- ☆144Jan 3, 2024Updated 2 years ago
- NAACL 2022 paper on Analyzing Modality Robustness in Multimodal Sentiment Analysis☆31Jan 21, 2023Updated 3 years ago
- GeckoNum Benchmark for T2I Model Eval.☆15Dec 5, 2024Updated last year
- [CVPR21] Visual Semantic Role Labeling for Video Understanding (https://arxiv.org/abs/2104.00990)☆61Aug 17, 2021Updated 5 years ago
- ☆15Feb 24, 2023Updated 3 years ago
- Danmuku dataset☆12Jul 7, 2023Updated 3 years ago