[ACL 2026 Poster] Code and Benchmark for "Do MLLMs Understand Pointing? Benchmarking and Enhancing Referential Reasoning in Egocentric Vision"
☆21Jun 3, 2026Updated 2 months ago
Alternatives and similar repositories for EgoPoint
Users that are interested in EgoPoint are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- The code of the paper "Latent Fingerprint Matching via Dense Minutia Descriptor"☆23Aug 27, 2025Updated 11 months ago
- ☆21Feb 22, 2024Updated 2 years ago
- GesVLA: Gesture-Aware Vision-Language-Action Model with Embedded Representations☆29May 22, 2026Updated 3 months ago
- ☆30Aug 18, 2025Updated last year
- Codes, datasets, and synthetic dataset generator about the paper "LiCamPose: Combining Multi-View LiDAR and RGB Cameras for Robust Single…☆17Feb 28, 2026Updated 5 months ago
- GPUs on demand by Runpod - Special Offer Available • AdRun AI, ML, and HPC workloads on powerful cloud GPUs—without limits or wasted spend. Deploy GPUs in under a minute and pay by the second.
- A fingerprint recognition framework featuring fixed-length dense descriptors, robust enhancement, and pose-aware alignment. Code for desc…☆21Apr 24, 2026Updated 3 months ago
- [CVPR 2026] AwareVLN: Reasoning with Self-awareness for Vision-Language Navigation☆79May 24, 2026Updated 3 months ago
- [ICCV 2025] D^3QE: Learning Discrete Distribution Discrepancy-aware Quantization Error for Autoregressive-Generated Image Detection☆17Jul 11, 2026Updated last month
- [ICCV 2025] IGL-Nav: Incremental 3D Gaussian Localization for Image-goal Navigation☆68Aug 4, 2025Updated last year
- [CVPR 2025, All Strong Accept] TSP3D: Text-guided Sparse Voxel Pruning for Efficient 3D Visual Grounding☆254Jul 27, 2026Updated 3 weeks ago
- [CVPR 2022] Back to Reality: Weakly-supervised 3D Object Detection with Shape-guided Label Enhancement☆44Mar 5, 2024Updated 2 years ago
- [IEEE TIFS under review] TOPIC: IFViT: Interpretable Fixed-Length Representation for Fingerprint Matching via Vision Transformer☆13Apr 9, 2024Updated 2 years ago
- GenWorld: Towards Detecting AI-generated Real-world Simulation Videos☆37Jun 13, 2025Updated last year
- Curved Projection Reformation☆13Nov 6, 2023Updated 2 years ago
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- [ICCV 25]SpectralAR: Spectral Autoregressive Visual Generation☆36Jun 13, 2025Updated last year
- SceneCompleter: Dense 3D Scene Completion for Generative Novel View Synthesis☆37Jun 13, 2025Updated last year
- ChatEMG: Synthetic Data Generation to Control a Robotic Hand Orthosis for Stroke☆12Jul 2, 2024Updated 2 years ago
- ☆10Nov 30, 2022Updated 3 years ago
- [ACL 2026] WikiSeeker: Rethinking the Role of Vision-Language Models in Knowledge-Based Visual Question Answering.☆15Jul 25, 2026Updated 3 weeks ago
- ☆18Oct 6, 2022Updated 3 years ago
- Implementation of the Mesh-VQVAE of "VQ-HPS: Human Pose and Shape Estimation in a Vector-Quantized Latent Space" - ECCV 2024☆18Oct 30, 2024Updated last year
- ☆13Aug 11, 2026Updated last week
- [ICCV 2023]The PyTorch implementation of TL-Align: Token-Label Alignment for Vision Transformers.☆23Jul 16, 2023Updated 3 years ago
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- ☆22Jan 8, 2026Updated 7 months ago
- This is a PyTorch implementation of 3DRefTR proposed by our paper "A Unified Framework for 3D Point Cloud Visual Grounding"☆26Aug 24, 2023Updated 3 years ago
- Official Code of "[ICCV 2023] MPI-Flow: Learning Realistic Optical Flow with Multiplane Images"☆65May 7, 2024Updated 2 years ago
- Official code for the MICCAI 2024 paper "3D Vessel Graph Generation Using Denoising Diffusion"☆25Jun 26, 2024Updated 2 years ago
- ☆17Feb 18, 2022Updated 4 years ago
- [ECCV 2024] 3D Small Object Detection with Dynamic Spatial Pruning☆115Aug 19, 2024Updated 2 years ago
- PyTorch implementation for Contrastive Representation Learning for Gaze Estimation☆29Mar 17, 2025Updated last year
- ☆14Jul 9, 2021Updated 5 years ago
- TIP 2024 | Multi-Spectral Image Stitching via Global-Aware Quadrature Pyramid Regression☆20Oct 26, 2024Updated last year
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- Official implementation of the NRNS paper☆37Jun 13, 2022Updated 4 years ago
- [CVPR 2024] Memory-based Adapters for Online 3D Scene Perception☆124Mar 25, 2025Updated last year
- Kinetix simulation for Legato: Learning Native Continuation for Action Chunking Flow Policies (RSS 2026)☆25Jun 3, 2026Updated 2 months ago
- 可以随机生成制定数量的车牌号,因为用到停车场的虚假数据生成,所以地区集中在一个地方。支持各类车辆的生成,只需在注释的地方修改即可。☆10May 30, 2021Updated 5 years ago
- ☆16Jun 5, 2020Updated 6 years ago
- Code for Decomposed Vector-Quantized Variational Autoencoder for Human Grasp Generation☆25Apr 15, 2025Updated last year
- This is a PyTorch implementation of MCLN proposed by our paper "Multi-branch Collaborative Learning Network for 3D Visual Grounding"(ECCV…☆27Oct 10, 2024Updated last year