Educational repository for applying the main video data curation techniques presented in the Stable Video Diffusion paper.
☆81Dec 30, 2023Updated 2 years ago
Alternatives and similar repositories for single-video-curation-svd
Users that are interested in single-video-curation-svd are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- This repository hosts code for converting the original MLP Mixer models (JAX) to TensorFlow.☆15Sep 29, 2021Updated 4 years ago
- Implementation of CaiT models in TensorFlow and ImageNet-1k checkpoints. Includes code for inference and fine-tuning.☆12Jun 9, 2023Updated 3 years ago
- ☆13Nov 10, 2021Updated 4 years ago
- EILeV: Eliciting In-Context Learning in Vision-Language Models for Videos Through Curated Data Distributional Properties☆133Nov 10, 2024Updated last year
- Dataset splits and evaluation code for the paper "Benchmark for Compositional Text-to-Image Synthesis" (NeurIPS 2021)☆45May 3, 2022Updated 4 years ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- [SIGGRAPH 2025] MotionCanvas: Cinematic Shot Design with Controllable Image-to-Video Generation☆36Aug 5, 2025Updated 11 months ago
- [TBench 2024] Official implementation of "AIGCBench: Comprehensive Evaluation of Image-to-Video Content Generated by AI"☆48Jan 30, 2024Updated 2 years ago
- Faster generation with text-to-image diffusion models.☆234Jun 28, 2025Updated last year
- [ICLR 2024] Code for FreeNoise based on VideoCrafter☆429Aug 25, 2025Updated 10 months ago
- ☆17Jan 10, 2024Updated 2 years ago
- ☆17Dec 28, 2023Updated 2 years ago
- Implementation of P+: Extended Textual Conditioning in Text-to-Image Generation☆49Mar 26, 2023Updated 3 years ago
- [IJCV 2024] LaVie: High-Quality Video Generation with Cascaded Latent Diffusion Models☆952Nov 13, 2024Updated last year
- [ICLR 2024] Code for FreeNoise based on AnimateDiff☆112Jan 22, 2024Updated 2 years ago
- Bare Metal GPUs on DigitalOcean Gradient AI • AdPurpose-built for serious AI teams training foundational models, running large-scale inference, and pushing the boundaries of what's possible.
- ☆25Dec 22, 2023Updated 2 years ago
- ☆66Jun 4, 2024Updated 2 years ago
- ☆26Mar 18, 2024Updated 2 years ago
- This repository implements the idea of "caption upsampling" from DALL-E 3 with Zephyr-7B and gathers results with SDXL.☆158Oct 25, 2023Updated 2 years ago
- [ICCV 2023] Label-Efficient Online Continual Object Detection in Streaming Video☆23Jan 8, 2024Updated 2 years ago
- Code of StyleCrafter on SDXL☆20Jun 25, 2024Updated 2 years ago
- ☆16Mar 25, 2024Updated 2 years ago
- The official repository of paper "ScaleLong: Towards More Stable Training of Diffusion Model via Scaling Network Long Skip Connection" (N…☆50Oct 23, 2023Updated 2 years ago
- maskrcnn with Latent Graph Neural Network, experiments of "LatentGNN"(ICML2019)☆14Sep 2, 2019Updated 6 years ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- Official implementation of Würstchen: Efficient Pretraining of Text-to-Image Models☆555Apr 6, 2024Updated 2 years ago
- Easily create large video dataset from video urls☆661Jul 30, 2024Updated last year
- Stable Video Diffusion Training Code and Extensions.☆733Jul 25, 2024Updated 2 years ago
- A repository for hacking Generative Fill with Open Source Tools☆37Mar 15, 2024Updated 2 years ago
- Code repo for "SketchODE: Learning neural sketch representation in continuous time" published in ICLR 2022☆11Apr 19, 2022Updated 4 years ago
- [ECCV 2024 Oral] MotionDirector: Motion Customization of Text-to-Video Diffusion Models.☆1,049Aug 21, 2024Updated last year
- This is a port of Mistral-7B model in JAX☆34Jul 1, 2024Updated 2 years ago
- [ECCV 2024] HiFi-123: Towards High-fidelity One Image to 3D Content Generation☆67Jul 12, 2024Updated 2 years ago
- [CVPR 2024] Panda-70M: Captioning 70M Videos with Multiple Cross-Modality Teachers☆700Oct 25, 2024Updated last year
- Managed Database hosting by DigitalOcean • AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- 📺 An End-to-End Solution for High-Resolution and Long Video Generation Based on Transformer Diffusion☆2,268Mar 6, 2025Updated last year
- Description and applications of OpenAI's paper about DALL-E (2021) and implementation of other (CLIP-guided) zero-shot text-to-image gene…☆33Aug 11, 2022Updated 3 years ago
- ☆24Dec 10, 2023Updated 2 years ago
- ☆13Oct 12, 2023Updated 2 years ago
- ☆17Jan 2, 2024Updated 2 years ago
- PyTorch implementation of CLIP Maximum Mean Discrepancy (CMMD) for evaluating image generation models.☆167Apr 5, 2024Updated 2 years ago
- [ICLR 2024] Code for FreeNoise based on LaVie☆34Jan 28, 2024Updated 2 years ago