Kandinsky 5.0: A family of diffusion models for Video & Image generation
☆816Aug 7, 2026Updated 3 weeks ago
Alternatives and similar repositories for kandinsky-5
Users that are interested in kandinsky-5 are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- KVAE tokenizers☆73Aug 14, 2026Updated 2 weeks ago
- ☆20Nov 25, 2025Updated 9 months ago
- Pusa: Thousands Timesteps Video Diffusion Model☆686Feb 13, 2026Updated 6 months ago
- [ICLR 2026] UniVideo: Unified Understanding, Generation, and Editing for Videos☆553Jul 3, 2026Updated last month
- rCM & Causal-rCM: Leading and Unified Algorithms/Infrastructures for Bidirectional/Autoregressive Video Diffusion Distillation at Scale☆794Jun 25, 2026Updated 2 months ago
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- [NeurIPS 2025 Oral]Infinity⭐️: Unified Spacetime AutoRegressive Modeling for Visual Generation☆780Apr 16, 2026Updated 4 months ago
- Ovis-Image is a 7B text-to-image model specifically optimized for high-quality text rendering, designed to operate efficiently under stri…☆319May 15, 2026Updated 3 months ago
- Wan: Open and Advanced Large-Scale Video Generative Models☆31Jul 28, 2025Updated last year
- A unified inference and post-training framework for accelerated video generation.☆4,193Updated this week
- [CVPR'26 Highlight] Ditto: Scaling Instruction-Based Video Editing with a High-Quality Synthetic Dataset☆624Jun 1, 2026Updated 2 months ago
- [ICLR 2026] ChronoEdit: Towards Temporal Reasoning for Image Editing and World Simulation☆705Nov 20, 2025Updated 9 months ago
- ComfyUI nodes for Kandinsky 5 DiT. T2V model generates high-quality video with advanced text conditioning.☆18Oct 17, 2025Updated 10 months ago
- A Unified Visual Generator with Interleaved OmniModal Context☆234Mar 5, 2026Updated 5 months ago
- 📹 A more flexible framework that can generate videos at any resolution and creates videos from images.☆2,224Updated this week
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- ☆1,750Nov 15, 2025Updated 9 months ago
- Wan2.2-Lightning: Speed up wan2.2 model with distillation☆313Nov 7, 2025Updated 9 months ago
- [CVPR2026]We present FlashPortrait, an end-to-end video diffusion transformer capable of synthesizing ID-preserving, infinite-length vide…☆481Feb 21, 2026Updated 6 months ago
- HunyuanImage-3.0: A Powerful Native Multimodal Model for Image Generation☆3,254Jun 23, 2026Updated 2 months ago
- Industry-level video foundation model for unified Text-to-Video (T2V) and Image-to-Video (I2V) generation.☆952Aug 27, 2025Updated last year
- Tiny AutoEncoder for Hunyuan Video (and other video models)☆519Aug 24, 2026Updated last week
- TurboDiffusion: 100–200× Acceleration for Video Diffusion Models☆3,628Updated this week
- (CVPR 2025) From Slow Bidirectional to Fast Autoregressive Video Diffusion Models☆1,428Aug 7, 2025Updated last year
- Helios: Real Real-Time Long Video Generation Model☆2,115Aug 24, 2026Updated last week
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- [NeurIPS 2025] Radial Attention: O(nlogn) Sparse Attention with Energy Decay for Long Video Generation☆607Nov 11, 2025Updated 9 months ago
- The official code of Yume☆685Jan 14, 2026Updated 7 months ago
- GPT-IMAGE-EDIT-1.5M: A Million-Scale, GPT-Generated Image Dataset☆243Aug 15, 2025Updated last year
- [ICML2025, NeurIPS2025 Spotlight] Sparse VideoGen 1 & 2: Accelerating Video Diffusion Transformers with Sparse Attention☆708Jul 4, 2026Updated last month
- [ICML 2026] Official codebase for "Causal Forcing: Autoregressive Diffusion Distillation Done Right for High-Quality Real-Time Interactiv…☆938Updated this week
- HY-World 1.5: A Systematic Framework for Interactive World Modeling with Real-Time Latency and Geometric Consistency☆1,587Jun 10, 2026Updated 2 months ago
- [ICCV 2025] Official implementations for paper: VACE: All-in-One Video Creation and Editing☆3,935Oct 17, 2025Updated 10 months ago
- ☆2,114Apr 11, 2026Updated 4 months ago
- Official code for StoryMem: Multi-shot Long Video Storytelling with Memory☆762Jul 22, 2026Updated last month
- Deploy open-source AI quickly and easily - Special Bonus Offer • AdRunpod Hub is built for open source. One-click deployment and autoscaling endpoints without provisioning your own infrastructure.
- [NeurIPS 2025] Improving Video Generation with Human Feedback☆501Sep 24, 2025Updated 11 months ago
- Lightweight Image Video Action Generation Inference Framework☆2,749Updated this week
- MOVA: Towards Scalable and Synchronized Video–Audio Generation☆1,105Updated this week
- Native Multimodal Models are World Learners☆1,547Dec 30, 2025Updated 8 months ago
- [CVPR 2026 Highlight] VideoCoF: Unified Video Editing with Temporal Reasoner☆208Jun 17, 2026Updated 2 months ago
- Krea Realtime 14B. An open-source realtime AI video model.☆580Nov 13, 2025Updated 9 months ago
- [ECCV 2026] Generate high resolution videos with a custom voice and appearance, based on LTX-2/LTX-2.3 + Identity In-Context LoRA☆351Jun 24, 2026Updated 2 months ago