☆203Feb 27, 2026Updated 6 months ago
Alternatives and similar repositories for Capybara
Users that are interested in Capybara are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- [Neurips 2025 NextVid Workshop Oral✨] Official Implementation of VideoGen-of-Thought: Step-by-step generating multi-shot video with minim…☆65Sep 22, 2025Updated 11 months ago
- Official Implementation of SAGE-GRPO:Manifold-Aware Exploration for Reinforcement Learning in Video Generation☆130Apr 2, 2026Updated 5 months ago
- (ECCV2026) Dual Distribution Estimation for Zero-shot Noisy Test-Time Adaptation with VLMs☆16Aug 1, 2026Updated last month
- Boogu-Image-0.1 is an Apache-2.0 open-source image generation and editing model family that delivers near-closed-source performance with …☆981Jul 23, 2026Updated last month
- ☆23Jun 29, 2026Updated 2 months ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- ☆32Apr 29, 2026Updated 4 months ago
- [ICLR 2026] UniVideo: Unified Understanding, Generation, and Editing for Videos☆556Jul 3, 2026Updated 2 months ago
- ☆165Feb 28, 2026Updated 6 months ago
- Official Repo of "$X$-Stream: Exploring MLLMs as Multiplexers for Multi-Stream Understanding"☆36Jun 18, 2026Updated 2 months ago
- ☆86Oct 10, 2025Updated 10 months ago
- ☆21Jan 1, 2026Updated 8 months ago
- A unified and fully open-source framework for instruction-guided and reference-guided video editing using natural language.☆321May 13, 2026Updated 3 months ago
- A Unified Visual Generator with Interleaved OmniModal Context☆234Mar 5, 2026Updated 6 months ago
- Official implementation of "EndoUIC: Promptable Diffusion Transformer for Unified Illumination Correction in Capsule Endoscopy", MICCAI 2…☆12Jan 29, 2026Updated 7 months ago
- Virtual machines for every use case on DigitalOcean • AdGet dependable uptime with 99.99% SLA, simple security tools, and predictable monthly pricing with DigitalOcean's virtual machines, called Droplets.
- SyncNoise: Geometrically Consistent Noise Prediction for Text-based 3D Scene Editing☆19Dec 28, 2024Updated last year
- [CVPR 2026 Highlight] Reward Forcing: Efficient Streaming Video Generation with Rewarded Distribution Matching Distillation☆359Dec 15, 2025Updated 8 months ago
- [ICML 2026] | Scaling Interactive World Models to 1000-Frame Horizons via Pose-Free Hierarchical Memory☆197May 4, 2026Updated 4 months ago
- [CVPR2026] Official implementation of our paper “Rethinking Position Embedding as a Context Controller for Multi-Reference and Multi-Shot…☆20Apr 8, 2026Updated 5 months ago
- Official implementation of "Surgical-VQLA: Transformer with Gated Vision-Language Embedding for Visual Question Localized-Answering in Ro…☆27Jul 7, 2024Updated 2 years ago
- DepthMaster: Unified Monocular Depth Estimation for Perspective and Panoramic Images☆26Jun 13, 2026Updated 2 months ago
- [ICML 2026] The offical code of Diversity-Preserved Distribution Matching Distillation for Fast Visual Synthesis☆91Jun 2, 2026Updated 3 months ago
- [SIGGRAPH 2025] AssetDropper: Asset Extraction via Diffusion Models with Reward-Driven Optimization☆34Jun 19, 2025Updated last year
- [ICML 2026] ScalingAR: Scaling Confidence for Autoregressive Image Generation☆22May 5, 2026Updated 4 months ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- [ECCV 2026] Official code repository for "Self-transcendence: Is External Feature Guidance Indispensable for Accelerating Diffusion Trans…☆38Aug 25, 2026Updated 2 weeks ago
- [ICCV 2025] FiVE-Bench: A Fine-grained Video Editing Benchmark for Evaluating Emerging Diffusion and Rectified Flow Models☆38Apr 2, 2026Updated 5 months ago
- [ICLR'2026] AssetFormer: Modular 3D Assets Generation with Autoregressive Transformer☆40Feb 13, 2026Updated 6 months ago
- Official implementation of “CAT-ViL: Co-Attention Gated Vision-Language Embedding for Visual Question Localized-Answering in Robotic Surg…☆18Jul 7, 2024Updated 2 years ago
- Official Code of "VideoCanvas: Unified Video Completion from Arbitrary Spatiotemporal Patches via In-Context Conditioning"☆68Oct 10, 2025Updated 10 months ago
- ☆41Dec 18, 2025Updated 8 months ago
- [NeurIPS 2025] Improving Video Generation with Human Feedback☆505Sep 24, 2025Updated 11 months ago
- [ECCV2024] VividDreamer: Invariant Score Distillation For Hyper-Realistic Text-to-3D Generation☆10Jul 4, 2024Updated 2 years ago
- Pusa: Thousands Timesteps Video Diffusion Model☆686Feb 13, 2026Updated 6 months ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- (CVPR 26 Findings) Official implementation of the paper "Bind-Your-Avatar: Multi-Talking-Character Video Generation with Dynamic 3D-mask-…☆34Apr 7, 2026Updated 5 months ago
- Code for "SePPO: Semi-Policy Preference Optimization for Diffusion Alignment."☆18Oct 7, 2024Updated last year
- Agentic World Modeling: Foundations, Capabilities, Laws, and Beyond☆313Updated this week
- ☆20Jun 26, 2024Updated 2 years ago
- An Efficient Text-to-Image Generation Pretrain Pipeline☆132Apr 18, 2025Updated last year
- AAAI 25' Flexible Image Reflection Removal with Sparse Human Guidance☆12Jul 7, 2025Updated last year
- [NeurIPS 2025 Spotlight] Official implementation for DNAEdit: Direct Noise Alignment for Text-Guided Rectified Flow Editing☆32Jan 23, 2026Updated 7 months ago