Kandinsky 5.0: A family of diffusion models for Video & Image generation
☆827Aug 7, 2026Updated last month
Alternatives and similar repositories for kandinsky-5
Users that are interested in kandinsky-5 are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- KVAE tokenizers☆73Aug 14, 2026Updated last month
- ☆22Nov 25, 2025Updated 9 months ago
- Pusa: Thousands Timesteps Video Diffusion Model☆686Feb 13, 2026Updated 7 months ago
- [ICLR 2026] UniVideo: Unified Understanding, Generation, and Editing for Videos☆558Jul 3, 2026Updated 2 months ago
- rCM & Causal-rCM: Leading and Unified Algorithms/Infrastructures for Bidirectional/Autoregressive Video Diffusion Distillation at Scale☆809Jun 25, 2026Updated 2 months ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- [NeurIPS 2025 Oral]Infinity⭐️: Unified Spacetime AutoRegressive Modeling for Visual Generation☆786Apr 16, 2026Updated 5 months ago
- Ovis-Image is a 7B text-to-image model specifically optimized for high-quality text rendering, designed to operate efficiently under stri…☆319May 15, 2026Updated 4 months ago
- Wan: Open and Advanced Large-Scale Video Generative Models☆31Jul 28, 2025Updated last year
- A unified inference and post-training framework for accelerated video generation.☆4,476Updated this week
- [CVPR'26 Highlight] Ditto: Scaling Instruction-Based Video Editing with a High-Quality Synthetic Dataset☆629Jun 1, 2026Updated 3 months ago
- [ICLR 2026] ChronoEdit: Towards Temporal Reasoning for Image Editing and World Simulation☆706Nov 20, 2025Updated 10 months ago
- ComfyUI nodes for Kandinsky 5 DiT. T2V model generates high-quality video with advanced text conditioning.☆18Oct 17, 2025Updated 11 months ago
- 📹 A more flexible framework that can generate videos at any resolution and creates videos from images.☆2,252Updated this week
- ☆1,760Nov 15, 2025Updated 10 months ago
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- A Unified Visual Generator with Interleaved OmniModal Context☆235Mar 5, 2026Updated 6 months ago
- Wan2.2-Lightning: Speed up wan2.2 model with distillation☆313Nov 7, 2025Updated 10 months ago
- [CVPR2026]We present FlashPortrait, an end-to-end video diffusion transformer capable of synthesizing ID-preserving, infinite-length vide…☆481Feb 21, 2026Updated 6 months ago
- Tiny AutoEncoder for Hunyuan Video (and other video models)☆536Aug 31, 2026Updated 2 weeks ago
- HunyuanImage-3.0: A Powerful Native Multimodal Model for Image Generation☆3,266Jun 23, 2026Updated 2 months ago
- TurboDiffusion: 100–200× Acceleration for Video Diffusion Models☆3,732Aug 27, 2026Updated 3 weeks ago
- Industry-level video foundation model for unified Text-to-Video (T2V) and Image-to-Video (I2V) generation.☆954Aug 27, 2025Updated last year
- (CVPR 2025) From Slow Bidirectional to Fast Autoregressive Video Diffusion Models☆1,440Aug 7, 2025Updated last year
- Helios: Real Real-Time Long Video Generation Model☆2,150Aug 24, 2026Updated 3 weeks ago
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- [NeurIPS 2025] Radial Attention: O(nlogn) Sparse Attention with Energy Decay for Long Video Generation☆607Nov 11, 2025Updated 10 months ago
- The official code of Yume☆685Jan 14, 2026Updated 8 months ago
- GPT-IMAGE-EDIT-1.5M: A Million-Scale, GPT-Generated Image Dataset☆244Aug 15, 2025Updated last year
- [ICML2025, NeurIPS2025 Spotlight] Sparse VideoGen 1 & 2: Accelerating Video Diffusion Transformers with Sparse Attention☆708Jul 4, 2026Updated 2 months ago
- [ICCV 2025] Official implementations for paper: VACE: All-in-One Video Creation and Editing☆3,951Oct 17, 2025Updated 11 months ago
- HY-World 1.5: A Systematic Framework for Interactive World Modeling with Real-Time Latency and Geometric Consistency☆1,603Jun 10, 2026Updated 3 months ago
- [ICML 2026] Official codebase for "Causal Forcing: Autoregressive Diffusion Distillation Done Right for High-Quality Real-Time Interactiv…☆969Aug 28, 2026Updated 3 weeks ago
- ☆2,115Apr 11, 2026Updated 5 months ago
- Lightweight Image Video Action Generation Inference Framework☆2,830Updated this week
- Serverless GPU API endpoints on Runpod - Get Bonus Credits • AdSkip the infrastructure headaches. Auto-scaling, pay-as-you-go, no-ops approach lets you focus on innovating your application.
- Official code for StoryMem: Multi-shot Long Video Storytelling with Memory☆769Jul 22, 2026Updated last month
- [NeurIPS 2025] Improving Video Generation with Human Feedback☆508Sep 24, 2025Updated 11 months ago
- Native Multimodal Models are World Learners☆1,555Dec 30, 2025Updated 8 months ago
- A foundation model that generates synchronized video and audio in a single model☆1,117Updated this week
- [ICLR 2026] Official repo for paper "Video-As-Prompt: Unified Semantic Control for Video Generation"☆454Feb 8, 2026Updated 7 months ago
- [CVPR 2026 Highlight] VideoCoF: Unified Video Editing with Temporal Reasoner☆210Jun 17, 2026Updated 3 months ago
- Krea Realtime 14B. An open-source realtime AI video model.☆585Nov 13, 2025Updated 10 months ago