[Arxiv 2024] Official code for MMTrail: A Multimodal Trailer Video Dataset with Language and Music Descriptions
☆34Feb 6, 2025Updated last year
Alternatives and similar repositories for MMTrail
Users that are interested in MMTrail are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- [Arxiv2022] Interpreting Class Conditional GANs with Channel Awareness☆17Apr 4, 2022Updated 4 years ago
- ☆11Sep 28, 2023Updated 2 years ago
- Curriculum Vitae of Quan Wang☆15Aug 27, 2026Updated 3 weeks ago
- [ICLR 2024] ViDA: Homeostatic Visual Domain Adapter for Continual Test Time Adaptation☆78Apr 25, 2024Updated 2 years ago
- Audio Generation model working with GPT-2 and VQVAE compressed representation of MelSpectrograms☆18Oct 8, 2023Updated 2 years ago
- Deploy open-source AI quickly and easily - Special Bonus Offer • AdRunpod Hub is built for open source. One-click deployment and autoscaling endpoints without provisioning your own infrastructure.
- ☆25Jan 14, 2021Updated 5 years ago
- Unofficial pytorch implementation of the paper "Learnable Fourier Features for Multi-Dimensional Spatial Positional Encoding", NeurIPS 20…☆13Apr 24, 2024Updated 2 years ago
- [CVPR 2023] Dataset and Code Pytorch Implementation of "Diverse 3D Hand Gesture Prediction from Body Dynamics by Bilateral Hand Disentan…☆23Oct 3, 2023Updated 2 years ago
- FuseLIP: Multimodal Embeddings via Early Fusion of Discrete Tokens☆17Sep 8, 2025Updated last year
- Code for the AAAI 2024 paper: "AGS: Affordable and Generalizable Substitute Training for Transferable Adversarial Attack" (accepted).☆13Mar 28, 2024Updated 2 years ago
- [IROS 2021] Official code for "Stereo Waterdrop Removal with Row-wise Dilated Attention"☆35Aug 21, 2021Updated 5 years ago
- A dataset for Audio-Visual Sound Event Detection in Movies☆26Jan 23, 2023Updated 3 years ago
- On Path to Multimodal Generalist: General-Level and General-Bench☆22Jul 11, 2025Updated last year
- [ECCV 2022 Oral] 3D-Aware Indoor Scene Synthesis with Depth Priors☆71Nov 24, 2022Updated 3 years ago
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- [ICCV 2023] Video Background Music Generation: Dataset, Method and Evaluation☆79Mar 29, 2024Updated 2 years ago
- ☆14Oct 16, 2023Updated 2 years ago
- official code for CVPR'24 paper Diff-BGM☆71Oct 12, 2024Updated last year
- Transfering images python from server to client UI python socket☆13Sep 30, 2020Updated 5 years ago
- ☆12Dec 10, 2018Updated 7 years ago
- Implementation of "Reconstruction-based Anomaly Detection with Completely Random Forest," SIAM International Conference on Data Mining (S…☆10Feb 16, 2021Updated 5 years ago
- ☆20Aug 11, 2025Updated last year
- PyTorch tool for training with bigger batch size on the GPU☆11Feb 26, 2021Updated 5 years ago
- [TPAMI 2026] Learning Long-form Movie Prior via Large Language Models☆34Aug 25, 2026Updated 3 weeks ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- ☆31Dec 16, 2024Updated last year
- ☆26Nov 26, 2024Updated last year
- VL-GPT: A Generative Pre-trained Transformer for Vision and Language Understanding and Generation☆86Sep 12, 2024Updated 2 years ago
- The repo host the code and model of MAViL.☆45Jul 24, 2023Updated 3 years ago
- ☆14Oct 7, 2021Updated 4 years ago
- This repo contains docs for FDU NISL servers. It will be maintained by server adminstrators.☆17Jul 1, 2024Updated 2 years ago
- Memory Oriented Transfer Learning for Semi-Supervised Image Deraining☆28Nov 23, 2023Updated 2 years ago
- A Framework for Symbolic MUsic Graph Explanations☆11Jul 30, 2025Updated last year
- ☆12Jun 1, 2024Updated 2 years ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- ☆85Dec 4, 2022Updated 3 years ago
- Code Release for the paper "TriBERT: Full-body Human-centric Audio-visual Representation Learning for Visual Sound Separation" in NeurIPS…☆13Dec 9, 2021Updated 4 years ago
- Fast Image Restoration with Multi-bin Trainable Linear Units.☆11Dec 23, 2019Updated 6 years ago
- ☆12Feb 2, 2024Updated 2 years ago
- Repository for "Training Audio Captioning Models without Audio"☆10Sep 26, 2023Updated 2 years ago
- Music production for silent film clips.☆34Apr 30, 2025Updated last year
- ☆10Sep 25, 2024Updated last year