Video dataset dedicated to portrait-mode video recognition.
☆57Oct 13, 2025Updated 11 months ago
Alternatives and similar repositories for Portrait-Mode-Video
Users that are interested in Portrait-Mode-Video are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Official implementation of MTM☆21Aug 30, 2023Updated 3 years ago
- ☆12Sep 11, 2021Updated 5 years ago
- ☆14Nov 22, 2022Updated 3 years ago
- Repo for "Human-Centric Foundation Models: Perception, Generation and Agentic Modeling" (https://arxiv.org/abs/2502.08556)☆60Feb 15, 2025Updated last year
- Video Diffusion State Space Models☆19Mar 27, 2024Updated 2 years ago
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- Offical Code for Paper "Exploring Inter-Channel Correlation for Diversity-preserved Knowledge Distillation"☆17Jan 19, 2022Updated 4 years ago
- An efficient multi-modal instruction-following data synthesis tool and the official implementation of Oasis https://arxiv.org/abs/2503.08…☆40Jun 4, 2025Updated last year
- Official repository for "Attend to Not Attended: Structure-then-Detail Token Merging for Post-training DiT Acceleration", which has been …☆17Sep 29, 2025Updated 11 months ago
- Code for paper: "Executing Arithmetic: Fine-Tuning Large Language Models as Turing Machines"☆10Oct 11, 2024Updated last year
- [CVPR 2025 Oral] VideoEspresso: A Large-Scale Chain-of-Thought Dataset for Fine-Grained Video Reasoning via Core Frame Selection☆142Jul 28, 2025Updated last year
- Code and demo for paper: Zhao et al., "Q&A: Query-Based Representation Learning for Multi-Track Symbolic Music re-Arrangement," IJCAI 202…☆21May 2, 2024Updated 2 years ago
- (CVPR 2022) Automated Progressive Learning for Efficient Training of Vision Transformers☆25Feb 26, 2025Updated last year
- Official implementation of Next Block Prediction: Video Generation via Semi-Autoregressive Modeling☆42Feb 12, 2025Updated last year
- [ICLR 2023] Towards Smooth Video Composition☆84Jun 12, 2023Updated 3 years ago
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- The official implementation for the paper 'mmSampler: Efficient Frame Sampler for Multimodal Video Retrieval'.☆11Aug 23, 2022Updated 4 years ago
- [ICLR 2026] M2-Miner: Multi-Agent Enhanced MCTS for Mobile GUI Agent Data Mining☆55Apr 22, 2026Updated 5 months ago
- NSAS code for CVPR review☆27Jun 2, 2021Updated 5 years ago
- [CVPR 2024 Highlight] ImageNet-D☆47Jul 24, 2026Updated 2 months ago
- ICML2024-ReconBoost: Boosting Can Achieve Modality Reconcilement☆30May 2, 2025Updated last year
- ☆11Jul 26, 2024Updated 2 years ago
- ☆16Apr 7, 2024Updated 2 years ago
- ☆11Sep 30, 2024Updated last year
- ☆13Feb 2, 2025Updated last year
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- [CVPR 2024] Adapting Short-Term Transformers for Action Detection in Untrimmed Videos☆11Jun 11, 2024Updated 2 years ago
- [ICCV 2023] One-Shot Generative Domain Adaptation☆57Dec 23, 2021Updated 4 years ago
- Edge-Aware Mirror Network for Camouflaged Object Detection (EAMNet, IEEE ICME 2023).☆13Jul 8, 2023Updated 3 years ago
- [AAAI'25 Oral] NightReID: A Large-Scale Nighttime Person Re-Identification Benchmark☆11Jun 10, 2025Updated last year
- [ICLR 2024] Code for FreeNoise based on AnimateDiff☆112Jan 22, 2024Updated 2 years ago
- Synthesizing Efficient Data with Diffusion Models for Person Re-Identification Pre-Training☆11Jan 23, 2024Updated 2 years ago
- [CVPR 2026 Main] MultiBanana: A Challenging Benchmark for Multi-Reference Text-to-Image Generation☆29Aug 4, 2026Updated last month
- Official InfiniBench: A Benchmark for Large Multi-Modal Models in Long-Form Movies and TV Shows☆20Nov 4, 2025Updated 10 months ago
- ECCV 2024 STMA & CVPR 2024 1st MOSE & 1st VOT Challenge & 1st LSVOS v6☆12Oct 16, 2024Updated last year
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- [NAACL 2025] Representing Rule-based Chatbots with Transformers☆23Feb 9, 2025Updated last year
- ☆16Jul 29, 2025Updated last year
- ☆27Feb 11, 2025Updated last year
- A ShuffleBatchNorm layer to shuffle BatchNorm statistics across multiple GPUs☆57Mar 17, 2022Updated 4 years ago
- ☆18Jun 25, 2026Updated 3 months ago
- This repository contains the video files (download links) and corresponding annotations used in the paper "Long-Term Face Tracking for Cr…☆14Dec 18, 2020Updated 5 years ago
- [CVPR 2024] Panda-70M: Captioning 70M Videos with Multiple Cross-Modality Teachers☆707Oct 25, 2024Updated last year