[NeurIPS 2025] Fin3R: Fine-tuning Feed-forward 3D Reconstruction Models via Monocular Knowledge Distillation
☆64Dec 18, 2025Updated 7 months ago
Alternatives and similar repositories for Fin3R
Users that are interested in Fin3R are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- [ICLR 2025] DebGCD: Debiased Learning with Distribution Guidance for Generalized Category Discovery☆16Sep 27, 2025Updated 9 months ago
- [CVPR 2026 Findings] Speed3R: Sparse Feed-forward 3D Reconstruction Models☆75Apr 7, 2026Updated 3 months ago
- [IJCV 2024] Dissecting Out-of-Distribution Detection and Open-Set Recognition: A Critical Analysis of Methods and Benchmarks☆15Aug 30, 2024Updated last year
- [NeurIPS 2025] SEAL: Semantic-Aware Hierarchical Learning for Generalized Category Discovery☆16Apr 4, 2026Updated 3 months ago
- Geometry-grounded Point Transformer (CVPR 2026)☆142May 6, 2026Updated 2 months ago
- Virtual machines for every use case on DigitalOcean • AdGet dependable uptime with 99.99% SLA, simple security tools, and predictable monthly pricing with DigitalOcean's virtual machines, called Droplets.
- [CVPR 2025] Hyperbolic Category Discovery☆32Nov 7, 2025Updated 8 months ago
- [NeurIPS 2025] Panoptic Captioning: An Equivalence Bridge for Image and Text☆38Jan 31, 2026Updated 5 months ago
- [NeurIPS 2025] Streaming 3D Reconstruction with Explicit Spatial Pointer Memory☆191Mar 10, 2026Updated 4 months ago
- [ICCV 2025 Oral] Back on Track: Bundle Adjustment for Dynamic Scene Reconstruction (BA-Track)☆101Nov 25, 2025Updated 7 months ago
- [ICCV 2025] This is the official implementation of POMATO: Marrying Pointmap Matching with Temporal Motions for Dynamic 3D Reconstruction☆121Aug 9, 2025Updated 11 months ago
- [ECCV2024] PromptCCD: Learning Gaussian Mixture Prompt Pool for Continual Category Discovery☆31Apr 3, 2025Updated last year
- JoVA: Unified Multimodal Learning for Joint Video-Audio Generation☆33Dec 22, 2025Updated 6 months ago
- Official code for the IEEE CBMI paper 'ELIP: Enhanced Visual-Language Foundation Models for Image Retrieval'☆20Apr 21, 2026Updated 3 months ago
- [CVPR 26] Release repo of our work "Co-Me: Confidence-Guided Token Merging for Visual Geometric Transformers"☆190May 18, 2026Updated 2 months ago
- Serverless GPU API endpoints on Runpod - Get Bonus Credits • AdSkip the infrastructure headaches. Auto-scaling, pay-as-you-go, no-ops approach lets you focus on innovating your application.
- [ICLR2025] HiLo: A Learning Framework for Generalized Category Discovery Robust to Domain Shifts☆22Aug 1, 2025Updated 11 months ago
- [NeurIPS 2025] 3DRS: MLLMs Need 3D-Aware Representation Supervision for Scene Understanding☆158Dec 9, 2025Updated 7 months ago
- ☆17Feb 13, 2026Updated 5 months ago
- Doppelgangers++: Improved Visual Disambiguation with Geometric 3D Features☆99Apr 8, 2025Updated last year
- 🚀 Official code for “XStreamVGGT: Extremely Memory-Efficient Streaming Vision Geometry Grounded Transformer with KV Cache Compression”, …☆47Jan 27, 2026Updated 5 months ago
- [3DV 2026 Oral] Official Repo of "SAIL-Recon: Large SfM by Augmenting Scene Regression with Localization"☆297Feb 23, 2026Updated 4 months ago
- [CVPR 2025] ZeroMSF: Zero-shot Monocular Scene Flow Estimation in the Wild☆43Sep 16, 2025Updated 10 months ago
- [ICLR2024] SPTNet: An Efficient Alternative Framework for Generalized Category Discovery with Spatial Prompt Tuning☆36Apr 9, 2025Updated last year
- [CVPR 2026] Layer-wise Scale Alignment for Training-Free Streaming 4D Reconstruction☆80Mar 18, 2026Updated 4 months ago
- GPUs on demand by Runpod - Special Offer Available • AdRun AI, ML, and HPC workloads on powerful cloud GPUs—without limits or wasted spend. Deploy GPUs in under a minute and pay by the second.
- [CVPR 2025 Highlight] Official implementation of the solvers and estimators proposed in the paper "Relative Pose Estimation through Affin…☆237Apr 8, 2025Updated last year
- Official implement of VGGT-Long☆882Mar 20, 2026Updated 4 months ago
- [CVPR 2026] Emergent Extreme-View Geometry in 3D Foundation Models☆22Feb 25, 2026Updated 4 months ago
- [ECCV 2026] UniPR-3D☆32Jun 20, 2026Updated last month
- [CVPR 2025] Uni4D: Unifying Visual Foundation Models for 4D Modeling from a Single Video☆225May 25, 2025Updated last year
- ☆129Jun 17, 2025Updated last year
- [CVPRW2024] What’s in a Name? Beyond Class Indices for Image Recognition☆17Aug 30, 2024Updated last year
- [ArXiv2025] Category Discovery: An Open-World Perspective☆15Mar 17, 2026Updated 4 months ago
- [CVPR 2026 Hightlight] OmniVGGT: Omni-Modality Driven Visual Geometry Grounded Transformer☆350May 21, 2026Updated last month
- Bare Metal GPUs on DigitalOcean Gradient AI • AdPurpose-built for serious AI teams training foundational models, running large-scale inference, and pushing the boundaries of what's possible.
- [CVPR 2026] "E-RayZer: Self-supervised 3D Reconstruction as Spatial Visual Pre-training" official implementation.☆300May 30, 2026Updated last month
- GLUEMAP: Global Structure-from-Motion Meets Feedforward Reconstruction☆307Jun 22, 2026Updated 3 weeks ago
- [ICLR2026] Official Implementation of "Dens3R: A Foundation Model for 3D Geometry Prediction"☆395May 14, 2026Updated 2 months ago
- [ICLR 2026] FastVGGT: Fast Visual Geometry Transformer☆805Jan 28, 2026Updated 5 months ago
- [CVPR'26] TokenGS: Decoupling 3D Gaussian Prediction from Pixels with Learnable Tokens☆234Updated this week
- VGGT 3D Vision Agent optimized for Apple Silicon with Metal Performance Shaders☆92Mar 25, 2026Updated 3 months ago
- An curated list for feed-forward 3D scene modeling, including research directions, datasets, and applications.☆269Apr 22, 2026Updated 2 months ago