RegionReasoner: Region-Grounded Multi-Round Visual Reasoning (ICLR 2026)
☆21Jun 12, 2026Updated last month
Alternatives and similar repositories for RegionReasoner
Users that are interested in RegionReasoner are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- MetaModulation: Learning Variational Feature Hierarchies for Few-Shot Learning with Fewer Tasks (ICML 2023)☆11Aug 15, 2023Updated 2 years ago
- IPO: Interpretable Prompt Optimization for Vision-Language Models(NeurIPS 2024)☆15Jun 12, 2026Updated last month
- [NeurIPS 2025] Official Pytorch Implementation of "The Curse of Depth in Large Language Models" by Wenfang Sun, Xinyuan Song, Pengxiang L…☆72Mar 3, 2026Updated 5 months ago
- [ICML 2026] Official PyTorch implementation of paper “CoCoEdit: Content-Consistent Image Editing via Region Regularized Reinforcement Lea…☆26Jun 14, 2026Updated last month
- ☆12Jul 18, 2024Updated 2 years ago
- Managed Database hosting by DigitalOcean • AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- Official PyTorch codes for "Open Vocabulary 3D Scene Understanding via Geometry Guided Self-Distillation", ECCV2024☆31Jul 19, 2024Updated 2 years ago
- [ECCV2024] VividDreamer: Invariant Score Distillation For Hyper-Realistic Text-to-3D Generation☆10Jul 4, 2024Updated 2 years ago
- SeeSR: Towards Semantics-Aware Real-World Image Super-Resolution☆14Jan 12, 2024Updated 2 years ago
- Weighted Reverse Convolution for Feature Upsampling☆24May 24, 2026Updated 2 months ago
- ☆17Jan 10, 2024Updated 2 years ago
- (ECCV2026) Dual Distribution Estimation for Zero-shot Noisy Test-Time Adaptation with VLMs☆15Aug 1, 2026Updated last week
- DA-VAE: Plug-in Latent Compression for Diffusion via Detail Alignment (CVPR 2026)☆32Apr 16, 2026Updated 3 months ago
- SyncNoise: Geometrically Consistent Noise Prediction for Text-based 3D Scene Editing☆19Dec 28, 2024Updated last year
- [ICLR 2026] Many-for-Many: Unify the Training of Multiple Video and Image Generation and Manipulation Tasks☆32Feb 5, 2026Updated 6 months ago
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- The public source code of "FreCaS: Efficient Higher-Resolution Image Generation via Frequency-aware Cascaded Sampling"☆32Jul 7, 2025Updated last year
- NSARM: Next-Scale Autoregressive Modeling for Robust Real-World Image Super-Resolution☆27Oct 17, 2025Updated 9 months ago
- [CVPR26 highlight] Omni-3DEdit: Generalized Versatile 3D Editing in One-Pass☆29Apr 9, 2026Updated 4 months ago
- Code for "Repetitive Activity Counting by Sight and Sound"☆24Oct 29, 2021Updated 4 years ago
- Repo for FAPE-IR: Frequency-Aware Planning and Execution Framework for All-in-One Image Restoration (CVPR2026)☆37Jul 3, 2026Updated last month
- ECCV2024, LAPT: Label-driven Automated Prompt Tuning for OOD Detection with Vision-Language Models☆18Aug 9, 2024Updated 2 years ago
- GenDR: Lightning Generative Detail Restorator☆38Feb 24, 2026Updated 5 months ago
- CVPR2022☆23Jul 27, 2022Updated 4 years ago
- ☆16Oct 31, 2024Updated last year
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- Official codes for Polyline Path Masked Attention for Vision Transformer☆17Jun 3, 2026Updated 2 months ago
- The code for "Toward Accurate and Temporally Consistent Video Restoration from Raw Data"☆16Dec 25, 2023Updated 2 years ago
- [ACM MM'24 Oral] RainMamba: Enhanced Locality Learning with State Space Models for Video Deraining☆127Sep 22, 2025Updated 10 months ago
- Photo3D: Advancing Photorealistic 3D Generation through Structure‑Aligned Detail Enhancement☆22Mar 18, 2026Updated 4 months ago
- [NeurIPS 2025] DP²O-SR: Direct Perceptual Preference Optimization for Real-World Image Super-Resolution☆83Dec 20, 2025Updated 7 months ago
- [CVPR2026] BinaryAttention: One-Bit QK-Attention for Vision and Diffusion Transformers☆42Mar 17, 2026Updated 4 months ago
- [TIP2024] Official implementation of the paper ‘Perception-Distortion Balanced Super-Resolution: A Multi-Objective Optimization Perspecti…☆18Oct 1, 2024Updated last year
- ☆26Aug 31, 2023Updated 2 years ago
- [ICCV'25 Highlight] Derm1M: A Million‑Scale Vision‑Language Dataset Aligned with Clinical Ontology Knowledge for Dermatology☆74Dec 5, 2025Updated 8 months ago
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- ☆23Jun 29, 2026Updated last month
- DepthMaster: Unified Monocular Depth Estimation for Perspective and Panoramic Images☆26Jun 13, 2026Updated last month
- AMID: Towards Autonomous and Auditable Medical Imaging Model Development☆22Jul 14, 2026Updated 3 weeks ago
- ☆31Nov 28, 2023Updated 2 years ago
- Official code for GDPO-SR: Group Direct Preference Optimization for One-Step Generative Image Super-Resolution☆71Jun 2, 2026Updated 2 months ago
- [NeurIPS 2025] U-REPA: Aligning Diffusion U-Nets to ViTs☆39Dec 15, 2025Updated 7 months ago
- Official repository for VARestorer: One-Step VAR Distillation for Real-World Image Super-Resolution (ICLR2026))☆24Jul 25, 2026Updated 2 weeks ago