☆203Feb 27, 2026Updated 5 months ago
Alternatives and similar repositories for Capybara
Users that are interested in Capybara are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- [Neurips 2025 NextVid Workshop Oral✨] Official Implementation of VideoGen-of-Thought: Step-by-step generating multi-shot video with minim…☆64Sep 22, 2025Updated 10 months ago
- Official Implementation of SAGE-GRPO:Manifold-Aware Exploration for Reinforcement Learning in Video Generation☆126Apr 2, 2026Updated 4 months ago
- (ECCV2026) Dual Distribution Estimation for Zero-shot Noisy Test-Time Adaptation with VLMs☆16Aug 1, 2026Updated 2 weeks ago
- Boogu-Image-0.1 is an Apache-2.0 open-source image generation and editing model family that delivers near-closed-source performance with …☆956Jul 23, 2026Updated 3 weeks ago
- ☆23Jun 29, 2026Updated last month
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- ☆31Apr 29, 2026Updated 3 months ago
- [ICLR 2026] UniVideo: Unified Understanding, Generation, and Editing for Videos☆551Jul 3, 2026Updated last month
- ☆161Feb 28, 2026Updated 5 months ago
- Official Repo of "$X$-Stream: Exploring MLLMs as Multiplexers for Multi-Stream Understanding"☆34Jun 18, 2026Updated 2 months ago
- ☆86Oct 10, 2025Updated 10 months ago
- A unified and fully open-source framework for instruction-guided and reference-guided video editing using natural language.☆315May 13, 2026Updated 3 months ago
- ☆20Jan 1, 2026Updated 7 months ago
- A Unified Visual Generator with Interleaved OmniModal Context☆232Mar 5, 2026Updated 5 months ago
- Official implementation of "EndoUIC: Promptable Diffusion Transformer for Unified Illumination Correction in Capsule Endoscopy", MICCAI 2…☆12Jan 29, 2026Updated 6 months ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- SyncNoise: Geometrically Consistent Noise Prediction for Text-based 3D Scene Editing☆19Dec 28, 2024Updated last year
- [CVPR 2026 Highlight] Reward Forcing: Efficient Streaming Video Generation with Rewarded Distribution Matching Distillation☆352Dec 15, 2025Updated 8 months ago
- [ICML 2026] | Scaling Interactive World Models to 1000-Frame Horizons via Pose-Free Hierarchical Memory☆196May 4, 2026Updated 3 months ago
- [CVPR2026] Official implementation of our paper “Rethinking Position Embedding as a Context Controller for Multi-Reference and Multi-Shot…☆20Apr 8, 2026Updated 4 months ago
- Official implementation of "Surgical-VQLA: Transformer with Gated Vision-Language Embedding for Visual Question Localized-Answering in Ro…☆27Jul 7, 2024Updated 2 years ago
- DepthMaster: Unified Monocular Depth Estimation for Perspective and Panoramic Images☆26Jun 13, 2026Updated 2 months ago
- [ICML 2026] The offical code of Diversity-Preserved Distribution Matching Distillation for Fast Visual Synthesis☆90Jun 2, 2026Updated 2 months ago
- [SIGGRAPH 2025] AssetDropper: Asset Extraction via Diffusion Models with Reward-Driven Optimization☆34Jun 19, 2025Updated last year
- [ICML 2026] ScalingAR: Scaling Confidence for Autoregressive Image Generation☆22May 5, 2026Updated 3 months ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- [ECCV 2026] Official code repository for "Self-transcendence: Is External Feature Guidance Indispensable for Accelerating Diffusion Trans…☆37Updated this week
- [ICCV 2025] FiVE-Bench: A Fine-grained Video Editing Benchmark for Evaluating Emerging Diffusion and Rectified Flow Models☆39Apr 2, 2026Updated 4 months ago
- [ICLR'2026] AssetFormer: Modular 3D Assets Generation with Autoregressive Transformer☆39Feb 13, 2026Updated 6 months ago
- Official implementation of “CAT-ViL: Co-Attention Gated Vision-Language Embedding for Visual Question Localized-Answering in Robotic Surg…☆18Jul 7, 2024Updated 2 years ago
- Official Code of "VideoCanvas: Unified Video Completion from Arbitrary Spatiotemporal Patches via In-Context Conditioning"☆68Oct 10, 2025Updated 10 months ago
- ☆41Dec 18, 2025Updated 8 months ago
- [NeurIPS 2025] Improving Video Generation with Human Feedback☆492Sep 24, 2025Updated 10 months ago
- [ECCV2024] VividDreamer: Invariant Score Distillation For Hyper-Realistic Text-to-3D Generation☆10Jul 4, 2024Updated 2 years ago
- Pusa: Thousands Timesteps Video Diffusion Model☆686Feb 13, 2026Updated 6 months ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- (CVPR 26 Findings) Official implementation of the paper "Bind-Your-Avatar: Multi-Talking-Character Video Generation with Dynamic 3D-mask-…☆34Apr 7, 2026Updated 4 months ago
- Code for "SePPO: Semi-Policy Preference Optimization for Diffusion Alignment."☆18Oct 7, 2024Updated last year
- Agentic World Modeling: Foundations, Capabilities, Laws, and Beyond☆298Jun 27, 2026Updated last month
- ☆20Jun 26, 2024Updated 2 years ago
- An Efficient Text-to-Image Generation Pretrain Pipeline☆132Apr 18, 2025Updated last year
- ☆28Apr 21, 2023Updated 3 years ago
- AAAI 25' Flexible Image Reflection Removal with Sparse Human Guidance☆12Jul 7, 2025Updated last year