☆52Dec 13, 2024Updated last year
Alternatives and similar repositories for Owl
Users that are interested in Owl are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- ☆33Jul 5, 2024Updated 2 years ago
- [ICCV 25]SpectralAR: Spectral Autoregressive Visual Generation☆36Jun 13, 2025Updated last year
- VideoAuteur: Towards Long Narrative Video Generation☆44Oct 22, 2025Updated 10 months ago
- ☆218Feb 11, 2025Updated last year
- ☆664May 24, 2024Updated 2 years ago
- Bare Metal GPUs on DigitalOcean Gradient AI • AdPurpose-built for serious AI teams training foundational models, running large-scale inference, and pushing the boundaries of what's possible.
- CVPRW 2025 paper Progressive Autoregressive Video Diffusion Models: https://arxiv.org/abs/2410.08151☆89May 12, 2025Updated last year
- A Video Tokenizer Evaluation Dataset☆159Jan 13, 2025Updated last year
- DreamCinema: Cinematic Transfer with Free Camera and 3D Character☆96Jun 13, 2025Updated last year
- Doe-1: Closed-Loop Autonomous Driving with Large World Model☆113Jan 21, 2025Updated last year
- [NeurIPS 2024] DrivingDojo Dataset: Advancing Interactive and Knowledge-Enriched Driving World Model☆87Dec 5, 2024Updated last year
- Empowering Unified MLLM with Multi-granular Visual Generation☆132Jan 16, 2025Updated last year
- ☆31Oct 17, 2025Updated 10 months ago
- a family of versatile and state-of-the-art video tokenizers.☆458Sep 1, 2025Updated 11 months ago
- GenWorld: Towards Detecting AI-generated Real-world Simulation Videos☆37Jun 13, 2025Updated last year
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- Video Diffusion Alignment via Reward Gradients. We improve a variety of video diffusion models such as VideoCrafter, OpenSora, ModelScope…☆318Mar 12, 2025Updated last year
- CutDiffusion: A Simple, Fast, Cheap, and Strong Diffusion Extrapolation Method☆27Oct 9, 2025Updated 10 months ago
- Code for "AVG-LLaVA: A Multimodal Large Model with Adaptive Visual Granularity"☆33Oct 12, 2024Updated last year
- [CVPR 2025] Science-T2I: Addressing Scientific Illusions in Image Synthesis☆62Mar 31, 2026Updated 4 months ago
- Official implementation for our paper: Rethinking Video Tokenization: A Conditioned Diffusion-based Approach☆17Apr 2, 2025Updated last year
- [ICLR 2025] Autoregressive Video Generation without Vector Quantization☆661Oct 29, 2025Updated 10 months ago
- SceneCompleter: Dense 3D Scene Completion for Generative Novel View Synthesis☆37Jun 13, 2025Updated last year
- Live2Diff: A Pipeline that processes Live video streams by a uni-directional video Diffusion model.☆200Jul 22, 2024Updated 2 years ago
- [NeurIPS 2025] WorldMem: Long-term Consistent World Simulation with Memory☆389Feb 21, 2026Updated 6 months ago
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- ☆19May 24, 2026Updated 3 months ago
- ☆13Jul 10, 2024Updated 2 years ago
- [ECCV'24] MaxFusion: Plug & Play multimodal generation in text to image diffusion models☆28Nov 2, 2024Updated last year
- UniCon: A Simple Approach to Unifying Diffusion-based Conditional Generation (ICLR 2025)☆38Jun 21, 2025Updated last year
- [AAAI26] Next Patch Prediction☆129Jan 2, 2025Updated last year
- [ICLR 2025] Semi-Supervised Vision-Centric 3D Occupancy World Model for Autonomous Driving☆57Feb 14, 2025Updated last year
- [CVPR 2025 Highlight] Official implementation of HySAC, a hyperbolic safety-aware vision-language model for safer multimodal retrieval an…☆31Apr 8, 2025Updated last year
- Official Repo for Self-Forcing++ High Quality Long Video Generation☆270Oct 13, 2025Updated 10 months ago
- Code for our ICCV 2025 paper "Adaptive Caching for Faster Video Generation with Diffusion Transformers"☆173Nov 5, 2024Updated last year
- Managed Database hosting by DigitalOcean • AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- [ICLR 2026] This is an early exploration to introduce Interleaving Reasoning to Text-to-image Generation field and achieve the SoTA bench…☆101Jan 26, 2026Updated 7 months ago
- ☆21Jan 17, 2025Updated last year
- [ICLR 2025] OccProphet: Pushing Efficiency Frontier of Camera-Only 4D Occupancy Forecasting with Observer-Forecaster-Refiner Framework☆60Mar 18, 2026Updated 5 months ago
- ☆284Jul 22, 2025Updated last year
- The official implementation of MaskGRPO: Consolidating Reinforcement Learning for Multimodal Discrete Diffusion Models. (ICLR 2026, arxiv…☆19Jan 27, 2026Updated 7 months ago
- ☆10Apr 7, 2025Updated last year
- Video-Infinity generates long videos quickly using multiple GPUs without extra training.☆191Aug 4, 2024Updated 2 years ago