[ICLR2026] Video-GPT via Next Clip Diffusion.
☆45Jun 2, 2025Updated last year
Alternatives and similar repositories for Video-GPT
Users that are interested in Video-GPT are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- [ICLR2026] WeTok: Powerful Discrete Tokenization for High-Fidelity Visual Reconstruction☆71Sep 3, 2025Updated last year
- Code & data for "Towards flexible perception with visual memory" (ICML 2025)☆19Sep 24, 2024Updated 2 years ago
- [arXiv 2024] I4VGen: Image as Free Stepping Stone for Text-to-Video Generation☆24Oct 6, 2024Updated last year
- (AAAI 2025)MUSES: 3D-Controllable Image Generation via Multi-Modal Agent Collaboration☆38May 21, 2025Updated last year
- Official Pytorch Implementation for "TextToucher: Fine-Grained Text-to-Touch Generation" (AAAI 2025)☆19Jan 28, 2026Updated 7 months ago
- End-to-end encrypted email - Proton Mail • AdSpecial offer: 40% Off Yearly / 80% Off First Month. All Proton services are open source and independently audited for security.
- [ICCV 2025] TokensGen: Harnessing Condensed Tokens for Long Video Generation☆58Dec 10, 2025Updated 9 months ago
- CVPRW 2025 paper Progressive Autoregressive Video Diffusion Models: https://arxiv.org/abs/2410.08151☆89May 12, 2025Updated last year
- [NeurIPS 2024] TransAgent: Transfer Vision-Language Foundation Models with Heterogeneous Agent Collaboration☆25Oct 17, 2024Updated last year
- Official Pytorch Implementation for "DriveDiTFit: Fine-tuning Diffusion Transformers for Autonomous Driving Data Generation" (TOMM)☆24Mar 7, 2025Updated last year
- [ICCV 2023] Latent Action Composition for Skeleton-based Action Segmentation☆22Oct 25, 2023Updated 2 years ago
- ☆17Jun 13, 2025Updated last year
- YouDream: Generating Anatomically Controllable Consistent Text-to-3D Animals☆40Feb 9, 2025Updated last year
- ICML2025☆62Aug 28, 2025Updated last year
- ☆23Dec 11, 2024Updated last year
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- ☆39Oct 19, 2024Updated last year
- Code for "VideoRepair: Improving Text-to-Video Generation via Misalignment Evaluation and Localized Refinement [ACL 2026 Findings]"☆53Apr 7, 2026Updated 5 months ago
- ☆165Feb 28, 2026Updated 6 months ago
- [🚀 ICLR 2026 Oral] NextStep-1: SOTA Autogressive Image Generation with Continuous Tokens. A research project developed by the StepFun’s …☆695Feb 27, 2026Updated 6 months ago
- Training-Free Condition-Guided Text-to-Video Generation☆62Oct 23, 2025Updated 11 months ago
- Flux training codes (lora) for UniTEX☆25Jun 8, 2025Updated last year
- Code for: "Long-Context Autoregressive Video Modeling with Next-Frame Prediction"☆314Apr 23, 2025Updated last year
- AC3D: Analyzing and Improving 3D Camera Control in Video Diffusion Transformers☆165Sep 16, 2025Updated last year
- Official PyTorch Implementation for the "RewardSDS: Aligning Score Distillation via Reward-Weighted Sampling" paper!☆13Jun 10, 2025Updated last year
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- TPDiff: Temporal Pyramid Video Diffusion Model☆25Mar 13, 2025Updated last year
- [CVPR 2020] G3AN: Disentangling Appearance and Motion for Video Generation☆37Feb 5, 2021Updated 5 years ago
- Benchmarking physical understanding in generative video models☆348Sep 14, 2026Updated last week
- ☆37Dec 19, 2025Updated 9 months ago
- ☆26Aug 12, 2025Updated last year
- [ICLR' 25] AvatarGO: Zero-shot 4D Human-Object Interaction Generation and Animation☆69Mar 19, 2025Updated last year
- Official Code of "VideoCanvas: Unified Video Completion from Arbitrary Spatiotemporal Patches via In-Context Conditioning"☆68Oct 10, 2025Updated 11 months ago
- Code of the paper "FreePCA:Integrating Consistency Information across Long-short Frames in Training-free Long Video Generation via Princi…☆27Apr 3, 2026Updated 5 months ago
- CVPR 2026 | Official Implementation of "MultiShotMaster: A Controllable Multi-Shot Video Generation Framework"☆179Feb 22, 2026Updated 7 months ago
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- [ICCV 2025] Nautilus: Locality-aware Autoencoder for Scalable Mesh Generation☆59Jun 25, 2025Updated last year
- Talking head animation☆26Dec 8, 2023Updated 2 years ago
- Code repository for T2V-Turbo and T2V-Turbo-v2☆312Jan 31, 2025Updated last year
- Ouroboros3D: Image-to-3D Generation via 3D-aware Recursive Diffusion (CVPR2025)☆150Oct 22, 2025Updated 11 months ago
- [ICLR 2024] Code for FreeNoise based on VideoCrafter☆430Aug 25, 2025Updated last year
- [NeurIPS 2025] Official repository of the paper "Unlocking Aha Moments via Reinforcement Learning: Advancing Collaborative Visual Compreh…☆22Sep 27, 2025Updated 11 months ago
- [ICML 2026] "LIVE: Long-horizon Interactive Video World ModEling"☆43Jul 15, 2026Updated 2 months ago