RegionReasoner: Region-Grounded Multi-Round Visual Reasoning (ICLR 2026)
☆21Jun 12, 2026Updated 3 months ago
Alternatives and similar repositories for RegionReasoner
Users that are interested in RegionReasoner are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- IPO: Interpretable Prompt Optimization for Vision-Language Models(NeurIPS 2024)☆15Jun 12, 2026Updated 3 months ago
- [NeurIPS 2025] Official Pytorch Implementation of "The Curse of Depth in Large Language Models" by Wenfang Sun, Xinyuan Song, Pengxiang L…☆73Mar 3, 2026Updated 6 months ago
- [ICML 2026] Official PyTorch implementation of paper “CoCoEdit: Content-Consistent Image Editing via Region Regularized Reinforcement Lea…☆26Jun 14, 2026Updated 3 months ago
- ☆28Sep 7, 2026Updated last week
- ☆12Jul 18, 2024Updated 2 years ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- Official PyTorch codes for "Open Vocabulary 3D Scene Understanding via Geometry Guided Self-Distillation", ECCV2024☆31Jul 19, 2024Updated 2 years ago
- [ECCV2024] VividDreamer: Invariant Score Distillation For Hyper-Realistic Text-to-3D Generation☆10Jul 4, 2024Updated 2 years ago
- Weighted Reverse Convolution for Feature Upsampling☆25May 24, 2026Updated 3 months ago
- ☆17Jan 10, 2024Updated 2 years ago
- (ECCV2026) Dual Distribution Estimation for Zero-shot Noisy Test-Time Adaptation with VLMs☆17Aug 1, 2026Updated last month
- DA-VAE: Plug-in Latent Compression for Diffusion via Detail Alignment (CVPR 2026)☆34Apr 16, 2026Updated 5 months ago
- SyncNoise: Geometrically Consistent Noise Prediction for Text-based 3D Scene Editing☆19Dec 28, 2024Updated last year
- [ICLR 2026] Many-for-Many: Unify the Training of Multiple Video and Image Generation and Manipulation Tasks☆32Feb 5, 2026Updated 7 months ago
- The public source code of "FreCaS: Efficient Higher-Resolution Image Generation via Frequency-aware Cascaded Sampling"☆33Jul 7, 2025Updated last year
- Virtual machines for every use case on DigitalOcean • AdGet dependable uptime with 99.99% SLA, simple security tools, and predictable monthly pricing with DigitalOcean's virtual machines, called Droplets.
- NSARM: Next-Scale Autoregressive Modeling for Robust Real-World Image Super-Resolution☆27Oct 17, 2025Updated 11 months ago
- [CVPR26 highlight] Omni-3DEdit: Generalized Versatile 3D Editing in One-Pass☆27Apr 9, 2026Updated 5 months ago
- Repo for FAPE-IR: Frequency-Aware Planning and Execution Framework for All-in-One Image Restoration (CVPR2026)☆43Updated this week
- ECCV2024, LAPT: Label-driven Automated Prompt Tuning for OOD Detection with Vision-Language Models☆18Aug 9, 2024Updated 2 years ago
- [CVPR 2026 DataCV Workshop] 4KLSDB: A Large-Scale Native-4K Dataset and Benchmark for Image Restoration and Generation.☆36May 28, 2026Updated 3 months ago
- ☆17Oct 31, 2024Updated last year
- Official codes for Polyline Path Masked Attention for Vision Transformer☆17Jun 3, 2026Updated 3 months ago
- The code for "Toward Accurate and Temporally Consistent Video Restoration from Raw Data"☆16Dec 25, 2023Updated 2 years ago
- [ACM MM'24 Oral] RainMamba: Enhanced Locality Learning with State Space Models for Video Deraining☆128Sep 22, 2025Updated 11 months ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- Photo3D: Advancing Photorealistic 3D Generation through Structure‑Aligned Detail Enhancement☆22Mar 18, 2026Updated 6 months ago
- [NeurIPS 2025] DP²O-SR: Direct Perceptual Preference Optimization for Real-World Image Super-Resolution☆87Dec 20, 2025Updated 9 months ago
- [CVPR2026] BinaryAttention: One-Bit QK-Attention for Vision and Diffusion Transformers☆44Mar 17, 2026Updated 6 months ago
- ☆23Nov 21, 2024Updated last year
- ☆24Jun 29, 2026Updated 2 months ago
- DepthMaster: Unified Monocular Depth Estimation for Perspective and Panoramic Images☆26Jun 13, 2026Updated 3 months ago
- [AAAI26] ViP3DE: Fast Multi-view Consistent 3D Editing with Video Priors☆23Aug 9, 2026Updated last month
- AMID: Towards Autonomous and Auditable Medical Imaging Model Development☆23Jul 14, 2026Updated 2 months ago
- ☆31Nov 28, 2023Updated 2 years ago
- Bare Metal GPUs on DigitalOcean Gradient AI • AdPurpose-built for serious AI teams training foundational models, running large-scale inference, and pushing the boundaries of what's possible.
- Project page of "GaussianSR: 3D Gaussian Super-Resolution with 2D Diffusion Priors"☆23Jul 1, 2024Updated 2 years ago
- ☆13Dec 12, 2023Updated 2 years ago
- GGT-100K: Generative Ground Truth for Generalizable Real-World Image Restoration☆72Jun 1, 2026Updated 3 months ago
- [NeurIPS 2025] U-REPA: Aligning Diffusion U-Nets to ViTs☆43Dec 15, 2025Updated 9 months ago
- Official code for GDPO-SR: Group Direct Preference Optimization for One-Step Generative Image Super-Resolution☆78Jun 2, 2026Updated 3 months ago
- Official repository for VARestorer: One-Step VAR Distillation for Real-World Image Super-Resolution (ICLR2026))☆26Jul 25, 2026Updated last month
- AuthFace: Towards Authentic Blind Face Restoration with Face-oriented Generative Diffusion Prior (ACM MM 2025 Oral)☆20Mar 5, 2026Updated 6 months ago