[ICLR2025] GenPercept: Diffusion Models Trained with Large Data Are Transferable Visual Models
β232Jan 24, 2025Updated last year
Alternatives and similar repositories for GenPercept
Users that are interested in GenPercept are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- [ECCV'24] GeoWizard: Unleashing the Diffusion Priors for 3D Geometry Estimation from a Single Imageβ940Dec 7, 2024Updated last year
- [ICCV2023] π§FrozenRecon: Pose-free 3D Scene Reconstruction with Frozen Depth Modelsβ131Aug 23, 2024Updated 2 years ago
- [NeurIPS'24] Unleashing the Potential of the Diffusion Model in Few-shot Semantic Segmentation (Diffews)β51Apr 14, 2025Updated last year
- A toolbox for benchmarking SOTA discriminative and generative geometry estimation models.β65Aug 29, 2024Updated 2 years ago
- β21Jun 13, 2026Updated 3 months ago
- Managed hosting for WordPress and PHP on Cloudways β’ AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- [CVPR 2024 - Oral, Best Paper Award Candidate] Marigold: Repurposing Diffusion-Based Image Generators for Monocular Depth Estimationβ3,230Sep 6, 2026Updated 2 weeks ago
- [NeurIPS 2025 Spotlight] A Generalist Diffusion Model for Vision Perceptionβ321Sep 21, 2025Updated 11 months ago
- [AAAI 2025, Oral] DepthFM: Fast Monocular Depth Estimation with Flow Matchingβ760May 6, 2025Updated last year
- [ICML2026] ACTIVE-O3: Empowering Multimodal Large Language Models with Active Perception via GRPOβ84Apr 30, 2026Updated 4 months ago
- [CVPR 2024] Exploiting Diffusion Prior for Generalizable Dense Predictionβ81Apr 19, 2024Updated 2 years ago
- Official implementation of "Lotus: Diffusion-based Visual Foundation Model for High-quality Dense Prediction"β836Nov 28, 2025Updated 9 months ago
- [SIGGRAPH Asia 2024 (Journal Track)] StableNormal: Reducing Diffusion Variance for Stable and Sharp Normalβ784Aug 2, 2025Updated last year
- ChronoDepth: Learning Temporally Consistent Video Depth from Video Diffusion Priorsβ280Feb 27, 2025Updated last year
- Depth Any Video with Scalable Synthetic Data (ICLR 2025)β519Dec 4, 2024Updated last year
- Virtual machines for every use case on DigitalOcean β’ AdGet dependable uptime with 99.99% SLA, simple security tools, and predictable monthly pricing with DigitalOcean's virtual machines, called Droplets.
- [ICLR 2026] OmniWorld: A Multi-Domain and Multi-Modal Dataset for 4D World Modelingβ498Apr 16, 2026Updated 5 months ago
- β70Oct 19, 2023Updated 2 years ago
- [CVPR 2024 Oral] Rethinking Inductive Biases for Surface Normal Estimationβ923Jul 10, 2024Updated 2 years ago
- β293Updated this week
- [ICLR'26] Official PyTorch implementation of "Time Is a Feature: Exploiting Temporal Dynamics in Diffusion Language Models".β66Mar 5, 2026Updated 6 months ago
- This is a simple template using HuggingFace Accelerator for DDP-training/Saving/Loading/Pushing.β49Mar 19, 2024Updated 2 years ago
- The repo for "Metric3D: Towards Zero-shot Metric 3D Prediction from A Single Image" and "Metric3Dv2: A Versatile Monocular Geometric Founβ¦β2,323Mar 13, 2025Updated last year
- β132Feb 7, 2024Updated 2 years ago
- Official code for NeurIPS 2024 paper LRM-Zero: Training Large Reconstruction Models with Synthesized Dataβ154Oct 7, 2024Updated last year
- Deploy to Railway using AI coding agents - Free Credits Offer β’ AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- [WACV'25 Oral] Fine-Tuning Image-Conditional Diffusion Models is Easier than You Thinkβ521Jul 9, 2026Updated 2 months ago
- Feed-forward Geometry, Appearance and Camera Estimation from Uncalibrated Sparse Viewsβ725May 1, 2025Updated last year
- [CVPR 2024] Probing the 3D Awareness of Visual Foundation Modelsβ361Dec 1, 2025Updated 9 months ago
- Intrinsic Image Diffusion for Single-view Material Estimationβ234Nov 28, 2025Updated 9 months ago
- [ECCV 2024] Leveraging Synthetic Data for Real-Domain High-Resolution Monocular Metric Depth Estimationβ68Jan 5, 2025Updated last year
- [ECCV 2024] Efficient Large-Baseline Radiance Fields, a feed-forward 2DGS modelβ317Jul 13, 2024Updated 2 years ago
- [CVPR'25 Oral] MoGe: Unlocking Accurate Monocular Geometry Estimation for Open-Domain Images with Optimal Training Supervisionβ2,958Sep 9, 2026Updated last week
- [CVPR'2024] Official implementation of the paper "ECoDepth: Effective Conditioning of Diffusion Models for Monocular Depth Estimation"β220Nov 20, 2025Updated 10 months ago
- Repo of HawkLlama.β16Jan 2, 2025Updated last year
- Wordpress hosting with auto-scaling - Free Trial Offer β’ AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- One-shot and Few-shot 3D Editing without Per-Scene Optimizationβ175Aug 21, 2025Updated last year
- [ECCV 2024] Diffusion Models for Monocular Depth Estimation: Overcoming Challenging Conditionsβ93Sep 28, 2024Updated last year
- [ICLR 2025 Spotlight] Boltzmann-Aligned Inverse Folding Model as a Predictor of Mutational Effects on Protein-Protein Interactionsβ45Mar 10, 2025Updated last year
- [ICCV 2025, Oral] TrajectoryCrafter: Redirecting Camera Trajectory for Monocular Videos via Diffusion Modelsβ875Dec 17, 2025Updated 9 months ago
- [CVPR 2024] Official implementation of "SuperNormal: Neural Surface Reconstruction via Multi-View Normal Integration"β197Mar 31, 2025Updated last year
- Instant-angelo: Build high-fidelity Digital Twin within 20 Minutes!β464Oct 26, 2024Updated last year
- [ICCV'25] Unified Open-World Segmentation with Multi-Modal Promptsβ16Jun 16, 2026Updated 3 months ago