☆52Dec 13, 2024Updated last year
Alternatives and similar repositories for Owl
Users that are interested in Owl are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- ☆33Jul 5, 2024Updated 2 years ago
- [ICCV 25]SpectralAR: Spectral Autoregressive Visual Generation☆36Jun 13, 2025Updated last year
- VideoAuteur: Towards Long Narrative Video Generation☆44Oct 22, 2025Updated 9 months ago
- ☆218Feb 11, 2025Updated last year
- ☆660May 24, 2024Updated 2 years ago
- Virtual machines for every use case on DigitalOcean • AdGet dependable uptime with 99.99% SLA, simple security tools, and predictable monthly pricing with DigitalOcean's virtual machines, called Droplets.
- CVPRW 2025 paper Progressive Autoregressive Video Diffusion Models: https://arxiv.org/abs/2410.08151☆88May 12, 2025Updated last year
- A Video Tokenizer Evaluation Dataset☆158Jan 13, 2025Updated last year
- DreamCinema: Cinematic Transfer with Free Camera and 3D Character☆96Jun 13, 2025Updated last year
- Doe-1: Closed-Loop Autonomous Driving with Large World Model☆113Jan 21, 2025Updated last year
- [NeurIPS 2024] DrivingDojo Dataset: Advancing Interactive and Knowledge-Enriched Driving World Model☆87Dec 5, 2024Updated last year
- Empowering Unified MLLM with Multi-granular Visual Generation☆132Jan 16, 2025Updated last year
- ☆31Oct 17, 2025Updated 9 months ago
- a family of versatile and state-of-the-art video tokenizers.☆457Sep 1, 2025Updated 11 months ago
- GenWorld: Towards Detecting AI-generated Real-world Simulation Videos☆37Jun 13, 2025Updated last year
- End-to-end encrypted email - Proton Mail • AdSpecial offer: 40% Off Yearly / 80% Off First Month. All Proton services are open source and independently audited for security.
- Video Diffusion Alignment via Reward Gradients. We improve a variety of video diffusion models such as VideoCrafter, OpenSora, ModelScope…☆316Mar 12, 2025Updated last year
- CutDiffusion: A Simple, Fast, Cheap, and Strong Diffusion Extrapolation Method☆27Oct 9, 2025Updated 10 months ago
- Code for "AVG-LLaVA: A Multimodal Large Model with Adaptive Visual Granularity"☆33Oct 12, 2024Updated last year
- Official implementation for our paper: Rethinking Video Tokenization: A Conditioned Diffusion-based Approach☆17Apr 2, 2025Updated last year
- [CVPR 2025] Science-T2I: Addressing Scientific Illusions in Image Synthesis☆62Mar 31, 2026Updated 4 months ago
- [ICLR 2025] Autoregressive Video Generation without Vector Quantization☆658Oct 29, 2025Updated 9 months ago
- SceneCompleter: Dense 3D Scene Completion for Generative Novel View Synthesis☆36Jun 13, 2025Updated last year
- Live2Diff: A Pipeline that processes Live video streams by a uni-directional video Diffusion model.☆200Jul 22, 2024Updated 2 years ago
- [NeurIPS 2025] WorldMem: Long-term Consistent World Simulation with Memory☆382Feb 21, 2026Updated 5 months ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- ☆19May 24, 2026Updated 2 months ago
- ☆13Jul 10, 2024Updated 2 years ago
- [ECCV'24] MaxFusion: Plug & Play multimodal generation in text to image diffusion models☆27Nov 2, 2024Updated last year
- UniCon: A Simple Approach to Unifying Diffusion-based Conditional Generation (ICLR 2025)☆38Jun 21, 2025Updated last year
- [AAAI26] Next Patch Prediction☆129Jan 2, 2025Updated last year
- [ICLR 2025] Semi-Supervised Vision-Centric 3D Occupancy World Model for Autonomous Driving☆56Feb 14, 2025Updated last year
- Official Repo for Self-Forcing++ High Quality Long Video Generation☆269Oct 13, 2025Updated 9 months ago
- Hyperbolic Safety-Aware Vision-Language Models. CVPR 2025☆31Apr 8, 2025Updated last year
- Code for our ICCV 2025 paper "Adaptive Caching for Faster Video Generation with Diffusion Transformers"☆172Nov 5, 2024Updated last year
- Simple, predictable pricing with DigitalOcean hosting • AdAlways know what you'll pay with monthly caps and flat pricing. Enterprise-grade infrastructure trusted by 600k+ customers.
- [ICLR 2026] This is an early exploration to introduce Interleaving Reasoning to Text-to-image Generation field and achieve the SoTA bench…☆101Jan 26, 2026Updated 6 months ago
- ☆21Jan 17, 2025Updated last year
- [ICLR 2025] OccProphet: Pushing Efficiency Frontier of Camera-Only 4D Occupancy Forecasting with Observer-Forecaster-Refiner Framework☆60Mar 18, 2026Updated 4 months ago
- ☆284Jul 22, 2025Updated last year
- The official implementation of MaskGRPO: Consolidating Reinforcement Learning for Multimodal Discrete Diffusion Models. (ICLR 2026, arxiv…☆19Jan 27, 2026Updated 6 months ago
- ☆10Apr 7, 2025Updated last year
- Video-Infinity generates long videos quickly using multiple GPUs without extra training.☆191Aug 4, 2024Updated 2 years ago