[ICLR2025] GenPercept: Diffusion Models Trained with Large Data Are Transferable Visual Models
β229Jan 24, 2025Updated last year
Alternatives and similar repositories for GenPercept
Users that are interested in GenPercept are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- [ECCV'24] GeoWizard: Unleashing the Diffusion Priors for 3D Geometry Estimation from a Single Imageβ938Dec 7, 2024Updated last year
- [ICCV2023] π§FrozenRecon: Pose-free 3D Scene Reconstruction with Frozen Depth Modelsβ131Aug 23, 2024Updated last year
- [NeurIPS'24] Unleashing the Potential of the Diffusion Model in Few-shot Semantic Segmentation (Diffews)β51Apr 14, 2025Updated last year
- A toolbox for benchmarking SOTA discriminative and generative geometry estimation models.β65Aug 29, 2024Updated last year
- β18Jun 13, 2026Updated last month
- Managed Kubernetes at scale on DigitalOcean β’ AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- [CVPR 2024 - Oral, Best Paper Award Candidate] Marigold: Repurposing Diffusion-Based Image Generators for Monocular Depth Estimationβ3,179Dec 10, 2025Updated 7 months ago
- [AAAI 2025, Oral] DepthFM: Fast Monocular Depth Estimation with Flow Matchingβ754May 6, 2025Updated last year
- [ICML2026] ACTIVE-O3: Empowering Multimodal Large Language Models with Active Perception via GRPOβ83Apr 30, 2026Updated 2 months ago
- [CVPR 2024] Exploiting Diffusion Prior for Generalizable Dense Predictionβ81Apr 19, 2024Updated 2 years ago
- Official implementation of "Lotus: Diffusion-based Visual Foundation Model for High-quality Dense Prediction"β812Nov 28, 2025Updated 7 months ago
- [NeurIPS 2025 Spotlight] A Generalist Diffusion Model for Vision Perceptionβ318Sep 21, 2025Updated 10 months ago
- [SIGGRAPH Asia 2024 (Journal Track)] StableNormal: Reducing Diffusion Variance for Stable and Sharp Normalβ776Aug 2, 2025Updated 11 months ago
- ChronoDepth: Learning Temporally Consistent Video Depth from Video Diffusion Priorsβ279Feb 27, 2025Updated last year
- Depth Any Video with Scalable Synthetic Data (ICLR 2025)β518Dec 4, 2024Updated last year
- Bare Metal GPUs on DigitalOcean Gradient AI β’ AdPurpose-built for serious AI teams training foundational models, running large-scale inference, and pushing the boundaries of what's possible.
- [ICLR 2026] OmniWorld: A Multi-Domain and Multi-Modal Dataset for 4D World Modelingβ485Apr 16, 2026Updated 3 months ago
- β70Oct 19, 2023Updated 2 years ago
- [CVPR 2024 Oral] Rethinking Inductive Biases for Surface Normal Estimationβ918Jul 10, 2024Updated 2 years ago
- β288May 31, 2024Updated 2 years ago
- [ICLR'26] Official PyTorch implementation of "Time Is a Feature: Exploiting Temporal Dynamics in Diffusion Language Models".β66Mar 5, 2026Updated 4 months ago
- This is a simple template using HuggingFace Accelerator for DDP-training/Saving/Loading/Pushing.β49Mar 19, 2024Updated 2 years ago
- The repo for "Metric3D: Towards Zero-shot Metric 3D Prediction from A Single Image" and "Metric3Dv2: A Versatile Monocular Geometric Founβ¦β2,271Mar 13, 2025Updated last year
- β133Feb 7, 2024Updated 2 years ago
- Official code for NeurIPS 2024 paper LRM-Zero: Training Large Reconstruction Models with Synthesized Dataβ155Oct 7, 2024Updated last year
- Wordpress hosting with auto-scaling - Free Trial Offer β’ AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- [WACV'25 Oral] Fine-Tuning Image-Conditional Diffusion Models is Easier than You Thinkβ519Jul 9, 2026Updated last week
- β721May 1, 2025Updated last year
- [CVPR 2024] Probing the 3D Awareness of Visual Foundation Modelsβ354Dec 1, 2025Updated 7 months ago
- Intrinsic Image Diffusion for Single-view Material Estimationβ232Nov 28, 2025Updated 7 months ago
- [ECCV 2024] Leveraging Synthetic Data for Real-Domain High-Resolution Monocular Metric Depth Estimationβ68Jan 5, 2025Updated last year
- [CVPR'25 Oral] MoGe: Unlocking Accurate Monocular Geometry Estimation for Open-Domain Images with Optimal Training Supervisionβ2,653Updated this week
- [ECCV 2024] Efficient Large-Baseline Radiance Fields, a feed-forward 2DGS modelβ317Jul 13, 2024Updated 2 years ago
- [CVPR'2024] Official implementation of the paper "ECoDepth: Effective Conditioning of Diffusion Models for Monocular Depth Estimation"β221Nov 20, 2025Updated 8 months ago
- Repo of HawkLlama.β16Jan 2, 2025Updated last year
- Managed hosting for WordPress and PHP on Cloudways β’ AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- One-shot and Few-shot 3D Editing without Per-Scene Optimizationβ175Aug 21, 2025Updated 11 months ago
- [ECCV 2024] Diffusion Models for Monocular Depth Estimation: Overcoming Challenging Conditionsβ94Sep 28, 2024Updated last year
- [CVPR 2024] Official implementation of "SuperNormal: Neural Surface Reconstruction via Multi-View Normal Integration"β196Mar 31, 2025Updated last year
- [ICLR 2025 Spotlight] Boltzmann-Aligned Inverse Folding Model as a Predictor of Mutational Effects on Protein-Protein Interactionsβ45Mar 10, 2025Updated last year
- [ICCV 2025, Oral] TrajectoryCrafter: Redirecting Camera Trajectory for Monocular Videos via Diffusion Modelsβ859Dec 17, 2025Updated 7 months ago
- Instant-angelo: Build high-fidelity Digital Twin within 20 Minutes!β462Oct 26, 2024Updated last year
- [ICCV'25] Unified Open-World Segmentation with Multi-Modal Promptsβ16Jun 16, 2026Updated last month