Related code, checkpoints and project page for V-Reflection
☆60Apr 7, 2026Updated 3 months ago
Alternatives and similar repositories for V-Reflection
Users that are interested in V-Reflection are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Panoramic Affordance Prediction (PAP) (ECCV 2026)☆46Jun 29, 2026Updated last month
- PhysToolBench: Benchmarking Physical Tool Understanding for MLLMs☆30Jul 20, 2026Updated 2 weeks ago
- [ICLR 2026] Official Implementation of "ColorCtrl: Training-Free Text-Guided Color Editing with Multi-Modal Diffusion Transformer"☆22Apr 9, 2026Updated 3 months ago
- [ACL 2026] Official implementation of "Less is More: Improving LLM Reasoning with Minimal Test-Time Intervention"☆41Apr 18, 2026Updated 3 months ago
- ☆25Jul 28, 2025Updated last year
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- Official implementation of "Streaming Communication in Multi-Agent Reasoning"☆34Jun 6, 2026Updated last month
- SAM4SS: Tailoring SAM and SAM2 for Semantic Segmentation☆11Jul 31, 2024Updated 2 years ago
- [ECCV 2026] DualCamCtrl: Dual-Branch Diffusion Model for Geometry-Aware Camera-Controlled Video Generation☆75Apr 7, 2026Updated 3 months ago
- [ICLR 2026] DiMeR: Disentangled Mesh Reconstruction Model with Normal-only Geometry Training☆55May 26, 2025Updated last year
- [CVPR 2025] Official code of "PanDA: Towards Panoramic Depth Anything with Unlabeled Panoramas and Mobius Spatial Augmentation"☆52Mar 18, 2025Updated last year
- [CVPR 2026] TiViBench: Benchmarking Think-in-Video Reasoning for Video Generative Models☆67Feb 21, 2026Updated 5 months ago
- [ICML 2026] LatentMorph: Morphing Latent Reasoning into Image Generation☆47May 5, 2026Updated 2 months ago
- The official implementation of StereoPilot☆116Dec 19, 2025Updated 7 months ago
- A4-Agent: An Agentic Framework for Zero-Shot Affordance Reasoning (ECCV 2026)☆41Jun 29, 2026Updated last month
- Virtual machines for every use case on DigitalOcean • AdGet dependable uptime with 99.99% SLA, simple security tools, and predictable monthly pricing with DigitalOcean's virtual machines, called Droplets.
- Code for "Understanding-in-Generation:Reinforcing Generative Capability of Unified Model via Infusing Understanding into Generation"☆15Nov 11, 2025Updated 8 months ago
- ☆27Apr 28, 2025Updated last year
- Official repository for “Reasoning in the Dark: Interleaved Vision-Text Reasoning in Latent Space”☆18Jan 27, 2026Updated 6 months ago
- [CVPR 2025] Official implementation of "Kiss3DGen: Repurposing Image Diffusion Models for 3D Asset Generation"☆296May 24, 2025Updated last year
- AuthFace: Towards Authentic Blind Face Restoration with Face-oriented Generative Diffusion Prior (ACM MM 2025 Oral)☆19Mar 5, 2026Updated 4 months ago
- The official implementation of the paper SAEdit: Token-level control for continuous image editing via Sparse AutoEncoder☆24Oct 19, 2025Updated 9 months ago
- Official code of ReTR (NeurIPS 2023)☆48Nov 9, 2023Updated 2 years ago
- [IEEE TVCG 2025] Self-supervised Learning of Event-guided Video Frame Interpolation for Rolling Shutter Frames☆11Jun 1, 2025Updated last year
- Official implementation of Motion Forcing: A Decoupled Framework for Robust Video Generation in Motion Dynamics☆21Mar 12, 2026Updated 4 months ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- [SIGGRAPH Asia 2025] Official Implementation of "ConsistEdit: Highly Consistent and Precise Training-free Visual Editing"☆73Apr 8, 2026Updated 3 months ago
- Official codes of "Sketch-in-Latents: Eliciting Unified Reasoning in MLLMs"☆17Feb 15, 2026Updated 5 months ago
- [ACL2026 Main] Data & Code of "Are We Using the Right Benchmark: An Evaluation Framework for Visual Token Compression Methods"☆35Apr 9, 2026Updated 3 months ago
- From Understanding to Erasing: Towards Complete and Stable Video Object Removal☆29Apr 7, 2026Updated 3 months ago
- Syphus: Automatic Instruction-Response Generation Pipeline☆14Dec 14, 2023Updated 2 years ago
- [ECCV 2026] Official implementation of the paper "SegVGGT: Joint 3D Reconstruction and Instance Segmentation from Multi-View Images"☆28May 18, 2026Updated 2 months ago
- MLLM hallucination, LVLM, LLM, Hallucination Mitigation, Training-free hallucination mitigation☆32Jul 13, 2026Updated 3 weeks ago
- Official implementation of “LucidFusion: Reconstructing 3D Gaussians with Arbitrary Unposed Images”☆76Mar 21, 2025Updated last year
- Code & Weights for “Learning Robust Anymodal Segmentor with Unimodal and Cross-modal Distillation”☆15Dec 6, 2024Updated last year
- GPUs on demand by Runpod - Special Offer Available • AdRun AI, ML, and HPC workloads on powerful cloud GPUs—without limits or wasted spend. Deploy GPUs in under a minute and pay by the second.
- Official codebase for the paper Latent Visual Reasoning☆172Oct 22, 2025Updated 9 months ago
- Not All Steps are Created Equal: Selective Diffusion Distillation for Image Manipulation (ICCV 2023)☆64Sep 28, 2023Updated 2 years ago
- ☆57Mar 19, 2025Updated last year
- [ICCV2025] https://ai4city-hkust.github.io/Sat2City/☆32Jul 15, 2025Updated last year
- (CVPR 26) Explore with Long-term Memory: A Benchmark and Multimodal LLM-based Reinforcement Learning Framework for Embodied Exploration☆37Mar 8, 2026Updated 4 months ago
- Video-R2: Reinforcing Consistent and Grounded Reasoning in Multimodal Language Models☆19Jan 21, 2026Updated 6 months ago
- Implementation of MLLM-based Self-Vision-RAG models☆15Nov 30, 2025Updated 8 months ago