[ICLR2025] GenPercept: Diffusion Models Trained with Large Data Are Transferable Visual Models
☆229Jan 24, 2025Updated last year
Alternatives and similar repositories for GenPercept
Users that are interested in GenPercept are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- [ECCV'24] GeoWizard: Unleashing the Diffusion Priors for 3D Geometry Estimation from a Single Image☆939Dec 7, 2024Updated last year
- [ICCV2023] 🧊FrozenRecon: Pose-free 3D Scene Reconstruction with Frozen Depth Models☆131Aug 23, 2024Updated last year
- [NeurIPS'24] Unleashing the Potential of the Diffusion Model in Few-shot Semantic Segmentation (Diffews)☆51Apr 14, 2025Updated last year
- A toolbox for benchmarking SOTA discriminative and generative geometry estimation models.☆65Aug 29, 2024Updated last year
- ☆19Jun 13, 2026Updated last month
- Simple, predictable pricing with DigitalOcean hosting • AdAlways know what you'll pay with monthly caps and flat pricing. Enterprise-grade infrastructure trusted by 600k+ customers.
- [CVPR 2024 - Oral, Best Paper Award Candidate] Marigold: Repurposing Diffusion-Based Image Generators for Monocular Depth Estimation☆3,191Dec 10, 2025Updated 8 months ago
- [NeurIPS 2025 Spotlight] A Generalist Diffusion Model for Vision Perception☆321Sep 21, 2025Updated 10 months ago
- [AAAI 2025, Oral] DepthFM: Fast Monocular Depth Estimation with Flow Matching☆757May 6, 2025Updated last year
- [ICML2026] ACTIVE-O3: Empowering Multimodal Large Language Models with Active Perception via GRPO☆83Apr 30, 2026Updated 3 months ago
- [CVPR 2024] Exploiting Diffusion Prior for Generalizable Dense Prediction☆81Apr 19, 2024Updated 2 years ago
- Official implementation of "Lotus: Diffusion-based Visual Foundation Model for High-quality Dense Prediction"☆821Nov 28, 2025Updated 8 months ago
- [SIGGRAPH Asia 2024 (Journal Track)] StableNormal: Reducing Diffusion Variance for Stable and Sharp Normal☆778Aug 2, 2025Updated last year
- ChronoDepth: Learning Temporally Consistent Video Depth from Video Diffusion Priors☆280Feb 27, 2025Updated last year
- Depth Any Video with Scalable Synthetic Data (ICLR 2025)☆518Dec 4, 2024Updated last year
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- [ICLR 2026] OmniWorld: A Multi-Domain and Multi-Modal Dataset for 4D World Modeling☆489Apr 16, 2026Updated 3 months ago
- ☆70Oct 19, 2023Updated 2 years ago
- [CVPR 2024 Oral] Rethinking Inductive Biases for Surface Normal Estimation☆921Jul 10, 2024Updated 2 years ago
- ☆290May 31, 2024Updated 2 years ago
- [ICLR'26] Official PyTorch implementation of "Time Is a Feature: Exploiting Temporal Dynamics in Diffusion Language Models".☆66Mar 5, 2026Updated 5 months ago
- This is a simple template using HuggingFace Accelerator for DDP-training/Saving/Loading/Pushing.☆49Mar 19, 2024Updated 2 years ago
- The repo for "Metric3D: Towards Zero-shot Metric 3D Prediction from A Single Image" and "Metric3Dv2: A Versatile Monocular Geometric Foun…☆2,290Mar 13, 2025Updated last year
- ☆133Feb 7, 2024Updated 2 years ago
- Official code for NeurIPS 2024 paper LRM-Zero: Training Large Reconstruction Models with Synthesized Data☆154Oct 7, 2024Updated last year
- GPUs on demand by Runpod - Special Offer Available • AdRun AI, ML, and HPC workloads on powerful cloud GPUs—without limits or wasted spend. Deploy GPUs in under a minute and pay by the second.
- [WACV'25 Oral] Fine-Tuning Image-Conditional Diffusion Models is Easier than You Think☆520Jul 9, 2026Updated last month
- ☆722May 1, 2025Updated last year
- [CVPR 2024] Probing the 3D Awareness of Visual Foundation Models☆357Dec 1, 2025Updated 8 months ago
- Intrinsic Image Diffusion for Single-view Material Estimation☆234Nov 28, 2025Updated 8 months ago
- [ECCV 2024] Leveraging Synthetic Data for Real-Domain High-Resolution Monocular Metric Depth Estimation☆68Jan 5, 2025Updated last year
- [CVPR'25 Oral] MoGe: Unlocking Accurate Monocular Geometry Estimation for Open-Domain Images with Optimal Training Supervision☆2,750Jul 21, 2026Updated 3 weeks ago
- [ECCV 2024] Efficient Large-Baseline Radiance Fields, a feed-forward 2DGS model☆317Jul 13, 2024Updated 2 years ago
- [CVPR'2024] Official implementation of the paper "ECoDepth: Effective Conditioning of Diffusion Models for Monocular Depth Estimation"☆220Nov 20, 2025Updated 8 months ago
- Repo of HawkLlama.☆16Jan 2, 2025Updated last year
- GPUs on demand by Runpod - Special Offer Available • AdRun AI, ML, and HPC workloads on powerful cloud GPUs—without limits or wasted spend. Deploy GPUs in under a minute and pay by the second.
- One-shot and Few-shot 3D Editing without Per-Scene Optimization☆175Aug 21, 2025Updated 11 months ago
- [ECCV 2024] Diffusion Models for Monocular Depth Estimation: Overcoming Challenging Conditions☆93Sep 28, 2024Updated last year
- [ICLR 2025 Spotlight] Boltzmann-Aligned Inverse Folding Model as a Predictor of Mutational Effects on Protein-Protein Interactions☆45Mar 10, 2025Updated last year
- [ICCV 2025, Oral] TrajectoryCrafter: Redirecting Camera Trajectory for Monocular Videos via Diffusion Models☆863Dec 17, 2025Updated 7 months ago
- [CVPR 2024] Official implementation of "SuperNormal: Neural Surface Reconstruction via Multi-View Normal Integration"☆196Mar 31, 2025Updated last year
- Instant-angelo: Build high-fidelity Digital Twin within 20 Minutes!☆463Oct 26, 2024Updated last year
- [ICCV'25] Unified Open-World Segmentation with Multi-Modal Prompts☆16Jun 16, 2026Updated last month