Video dataset dedicated to portrait-mode video recognition.
☆57Oct 13, 2025Updated 9 months ago
Alternatives and similar repositories for Portrait-Mode-Video
Users that are interested in Portrait-Mode-Video are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Official implementation of MTM☆21Aug 30, 2023Updated 2 years ago
- ☆12Sep 11, 2021Updated 4 years ago
- [ECCV 2024] 3DPE: Real-time 3D-aware Portrait Editing from a Single Image☆22Sep 15, 2025Updated 10 months ago
- ☆14Nov 22, 2022Updated 3 years ago
- Repo for "Human-Centric Foundation Models: Perception, Generation and Agentic Modeling" (https://arxiv.org/abs/2502.08556)☆58Feb 15, 2025Updated last year
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- Video Diffusion State Space Models☆19Mar 27, 2024Updated 2 years ago
- An efficient multi-modal instruction-following data synthesis tool and the official implementation of Oasis https://arxiv.org/abs/2503.08…☆40Jun 4, 2025Updated last year
- Official repository for "Attend to Not Attended: Structure-then-Detail Token Merging for Post-training DiT Acceleration", which has been …☆17Sep 29, 2025Updated 10 months ago
- Code for paper: "Executing Arithmetic: Fine-Tuning Large Language Models as Turing Machines"☆11Oct 11, 2024Updated last year
- The official repo for "VisualWebInstruct: Scaling up Multimodal Instruction Data through Web Search" [EMNLP25]☆39Feb 1, 2026Updated 5 months ago
- [CVPR 2025 Oral] VideoEspresso: A Large-Scale Chain-of-Thought Dataset for Fine-Grained Video Reasoning via Core Frame Selection☆140Jul 28, 2025Updated last year
- Code and demo for paper: Zhao et al., "Q&A: Query-Based Representation Learning for Multi-Track Symbolic Music re-Arrangement," IJCAI 202…☆21May 2, 2024Updated 2 years ago
- (CVPR 2022) Automated Progressive Learning for Efficient Training of Vision Transformers☆25Feb 26, 2025Updated last year
- Official implementation of Next Block Prediction: Video Generation via Semi-Autoregressive Modeling☆42Feb 12, 2025Updated last year
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- [ICLR 2023] Towards Smooth Video Composition☆85Jun 12, 2023Updated 3 years ago
- code of the paper "Vision-Language Navigation with Multi-granularity Observation and Auxiliary Reasoning Tasks"☆23Mar 23, 2021Updated 5 years ago
- [IJCAI 2024] Official implementation of the paper "Integrating View Conditions for Image Synthesis"☆25Aug 27, 2024Updated last year
- [ICLR 2026] M2-Miner: Multi-Agent Enhanced MCTS for Mobile GUI Agent Data Mining☆55Apr 22, 2026Updated 3 months ago
- NSAS code for CVPR review☆27Jun 2, 2021Updated 5 years ago
- Accepted by AAAI2022☆21Apr 10, 2022Updated 4 years ago
- retouching ptoto,remove moles/buffing/face-lift☆16Aug 12, 2021Updated 4 years ago
- Official code of *Towards Event-oriented Long Video Understanding*☆12Jul 26, 2024Updated 2 years ago
- [CVPR 2024 Highlight] ImageNet-D☆47Updated this week
- End-to-end encrypted email - Proton Mail • AdSpecial offer: 40% Off Yearly / 80% Off First Month. All Proton services are open source and independently audited for security.
- 🔥🔥MLVU: Multi-task Long Video Understanding Benchmark☆264Apr 13, 2026Updated 3 months ago
- ICML2024-ReconBoost: Boosting Can Achieve Modality Reconcilement☆29May 2, 2025Updated last year
- ☆11Jul 26, 2024Updated 2 years ago
- ☆11Sep 30, 2024Updated last year
- ☆16Apr 7, 2024Updated 2 years ago
- ☆13Feb 2, 2025Updated last year
- Official repo for paper "MiraData: A Large-Scale Video Dataset with Long Durations and Structured Captions"☆527Sep 2, 2024Updated last year
- [ICCV 2023] One-Shot Generative Domain Adaptation☆57Dec 23, 2021Updated 4 years ago
- Edge-Aware Mirror Network for Camouflaged Object Detection (EAMNet, IEEE ICME 2023).☆13Jul 8, 2023Updated 3 years ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- [AAAI'25 Oral] NightReID: A Large-Scale Nighttime Person Re-Identification Benchmark☆11Jun 10, 2025Updated last year
- [ICLR 2024] Code for FreeNoise based on AnimateDiff☆112Jan 22, 2024Updated 2 years ago
- [CVPR 2026 Main] MultiBanana: A Challenging Benchmark for Multi-Reference Text-to-Image Generation☆29Jul 6, 2026Updated 3 weeks ago
- Official InfiniBench: A Benchmark for Large Multi-Modal Models in Long-Form Movies and TV Shows☆20Nov 4, 2025Updated 8 months ago
- NeurIPS 2023 - Real-world super-resolution as multi-task learning☆25Mar 15, 2024Updated 2 years ago
- [NAACL 2025] Representing Rule-based Chatbots with Transformers☆23Feb 9, 2025Updated last year
- ☆16Jul 29, 2025Updated last year