[ACL 2026 Poster] Code and Benchmark for "Do MLLMs Understand Pointing? Benchmarking and Enhancing Referential Reasoning in Egocentric Vision"
☆21Jun 3, 2026Updated 2 months ago
Alternatives and similar repositories for EgoPoint
Users that are interested in EgoPoint are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- The code of the paper "Latent Fingerprint Matching via Dense Minutia Descriptor"☆22Aug 27, 2025Updated 11 months ago
- ☆21Feb 22, 2024Updated 2 years ago
- GesVLA: Gesture-Aware Vision-Language-Action Model with Embedded Representations☆29May 22, 2026Updated 2 months ago
- ☆30Aug 18, 2025Updated 11 months ago
- Codes, datasets, and synthetic dataset generator about the paper "LiCamPose: Combining Multi-View LiDAR and RGB Cameras for Robust Single…☆17Feb 28, 2026Updated 5 months ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- A fingerprint recognition framework featuring fixed-length dense descriptors, robust enhancement, and pose-aware alignment. Code for desc…☆21Apr 24, 2026Updated 3 months ago
- Implementation of "HumanReg: Self-supervised Non-rigid Registration of Sparse Human Point Cloud" (3DV 2024)☆15Oct 26, 2024Updated last year
- [CVPR 2026] AwareVLN: Reasoning with Self-awareness for Vision-Language Navigation☆71May 24, 2026Updated 2 months ago
- [ICCV 2025] D^3QE: Learning Discrete Distribution Discrepancy-aware Quantization Error for Autoregressive-Generated Image Detection☆17Jul 11, 2026Updated 3 weeks ago
- [T-IFS 2025] Joint Identity Verification and Pose Alignment for Partial Fingerprints☆31Apr 16, 2026Updated 3 months ago
- [ICCV 2025] IGL-Nav: Incremental 3D Gaussian Localization for Image-goal Navigation☆67Aug 4, 2025Updated last year
- [CVPR 2022] Back to Reality: Weakly-supervised 3D Object Detection with Shape-guided Label Enhancement☆44Mar 5, 2024Updated 2 years ago
- [IEEE TIFS under review] TOPIC: IFViT: Interpretable Fixed-Length Representation for Fingerprint Matching via Vision Transformer☆13Apr 9, 2024Updated 2 years ago
- GenWorld: Towards Detecting AI-generated Real-world Simulation Videos☆37Jun 13, 2025Updated last year
- Proton VPN Special Offer - Get 70% off • AdSpecial partner offer. Trusted by over 100 million users worldwide. Tested, Approved and Recommended by Experts.
- [ICCV 25]SpectralAR: Spectral Autoregressive Visual Generation☆36Jun 13, 2025Updated last year
- SceneCompleter: Dense 3D Scene Completion for Generative Novel View Synthesis☆36Jun 13, 2025Updated last year
- ChatEMG: Synthetic Data Generation to Control a Robotic Hand Orthosis for Stroke☆12Jul 2, 2024Updated 2 years ago
- The Official Code Repo for EgoOrientBench [CVPR25]☆17Nov 24, 2025Updated 8 months ago
- [ACL 2026] WikiSeeker: Rethinking the Role of Vision-Language Models in Knowledge-Based Visual Question Answering.☆15Jul 25, 2026Updated last week
- [CVPR2026] Skyra: AI-Generated Video Detection via Grounded Artifact Reasoning☆62Mar 27, 2026Updated 4 months ago
- ☆18Oct 6, 2022Updated 3 years ago
- Implementation of the Mesh-VQVAE of "VQ-HPS: Human Pose and Shape Estimation in a Vector-Quantized Latent Space" - ECCV 2024☆18Oct 30, 2024Updated last year
- Official implementation for AAAI-26 paper: "Force-Aware 3D Contact Modeling for Stable Grasp Generation"☆15Mar 13, 2026Updated 4 months ago
- Bare Metal GPUs on DigitalOcean Gradient AI • AdPurpose-built for serious AI teams training foundational models, running large-scale inference, and pushing the boundaries of what's possible.
- 主要利用QLearning,DQN,ImprovedDQN(Ddouble DQN) 解决gym框架下的三个问题CartPole-v0,MountainCar-v0,Acrobot-v1☆14Jan 14, 2018Updated 8 years ago
- ☆22Jan 8, 2026Updated 6 months ago
- MoDem-V2 combines the sample efficiency of the original MoDem with conservative exploration in order to quickly and safely learn manipula…☆25Apr 1, 2024Updated 2 years ago
- [AAAI 2026 Oral] STRIDE-QA: Visual Question Answering Dataset for Spatiotemporal Reasoning in Urban Driving Scenes☆17Jan 23, 2026Updated 6 months ago
- PyTorch implementation for Contrastive Representation Learning for Gaze Estimation☆29Mar 17, 2025Updated last year
- ☆14Jul 9, 2021Updated 5 years ago
- TIP 2024 | Multi-Spectral Image Stitching via Global-Aware Quadrature Pyramid Regression☆20Oct 26, 2024Updated last year
- My own (unofficial) implementation of the Point Transformer Network, currently for classification tasks.☆10Apr 24, 2021Updated 5 years ago
- This is the official code for the CVPR 2024 Publication: Tiger: Time-Varying Denoising Model for 3D Point Cloud Generation with Diffusion…☆35Jul 6, 2024Updated 2 years ago
- Virtual machines for every use case on DigitalOcean • AdGet dependable uptime with 99.99% SLA, simple security tools, and predictable monthly pricing with DigitalOcean's virtual machines, called Droplets.
- Code and Data for Human-M3: A Multi-view Multi-modal Dataset for 3D Human Pose Estimation in Outdoor Scenes.☆47May 27, 2024Updated 2 years ago
- Official implementation of the NRNS paper☆37Jun 13, 2022Updated 4 years ago
- [CVPR 2024] Memory-based Adapters for Online 3D Scene Perception☆125Mar 25, 2025Updated last year
- Kinetix simulation for Legato: Learning Native Continuation for Action Chunking Flow Policies (RSS 2026)☆24Jun 3, 2026Updated 2 months ago
- ☆16Jun 5, 2020Updated 6 years ago
- Fingerprint recognition in Python☆31Jul 15, 2026Updated 2 weeks ago
- Code for Decomposed Vector-Quantized Variational Autoencoder for Human Grasp Generation☆25Apr 15, 2025Updated last year