Official implementation of “Towards Cross-View Point Correspondence in Vision-Language Models”.
☆15Dec 24, 2025Updated 8 months ago
Alternatives and similar repositories for CrossPoint
Users that are interested in CrossPoint are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- A curated list of papers analyzed in the survey: Spatial Intelligence from a Cognitive Map Perspective: A Survey☆22Jun 1, 2026Updated 2 months ago
- Reshaping Action Error Distributions for Reliable Vision-Language-Action Models☆17Feb 5, 2026Updated 6 months ago
- implementation of the paper Scaling Up AI-Generated Image Detection with Generator-Aware Prototypes☆30Mar 24, 2026Updated 5 months ago
- Code implementation for paper titled "HOI-Ref: Hand-Object Interaction Referral in Egocentric Vision"☆30Apr 16, 2024Updated 2 years ago
- Transformation Driven Visual Reasoning - CVPR 2021☆36May 27, 2023Updated 3 years ago
- Simple, predictable pricing with DigitalOcean hosting • AdAlways know what you'll pay with monthly caps and flat pricing. Enterprise-grade infrastructure trusted by 600k+ customers.
- ☆28Mar 20, 2023Updated 3 years ago
- 一个网络画板,支持本地或者联机同步绘图,多种图形,支持样式修改,形状调整、移动,支持undo、redo,支持保存和导出图片等等。☆11Jun 24, 2021Updated 5 years ago
- ☆15Jun 9, 2025Updated last year
- [ECCV 2026] Official implementation of "RoboTracer: Mastering Spatial Trace with Reasoning in Vision-Language Models for Robotics"☆83Jun 18, 2026Updated 2 months ago
- ☆11May 9, 2024Updated 2 years ago
- ☆62Apr 1, 2025Updated last year
- “人工智能引论”labs☆11Mar 21, 2025Updated last year
- Dreamitate: Real-World Visuomotor Policy Learning via Video Generation (CoRL 2024)☆59Jun 7, 2025Updated last year
- This project uses YOLOv2 for human detection and stereo cameras for distance measurement. It runs on a PYNQ-Z2 board with a neural networ…☆13Nov 15, 2020Updated 5 years ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- Official code of our WACV paper "ECSIC: Epipolar Cross Attention for Stereo Image Compression"☆15Dec 27, 2023Updated 2 years ago
- Ego4D Goal-Step: Toward Hierarchical Understanding of Procedural Activities (NeurIPS 2023)☆63Apr 15, 2024Updated 2 years ago
- [WIP] Code for LangToMo☆21Mar 19, 2026Updated 5 months ago
- 2025 年春季学期北京大学计算机视觉导论课程的课程资料,王鹤老师。☆17Jul 4, 2025Updated last year
- Training recipe for SpatialReasoner [NeurIPS 2025]☆45Apr 5, 2026Updated 4 months ago
- ☆14Jan 5, 2022Updated 4 years ago
- 北京大学信息科学技术学院智能科学与技术系核心课《凸分析与优化方法》编程作业仓库☆17Jun 8, 2022Updated 4 years ago
- Synthetic VQA data generation code for SpatialReasoner.☆21Nov 25, 2025Updated 9 months ago
- AAAI 2026 Oral☆19Dec 23, 2025Updated 8 months ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- SAT-MTB-VSR dataset and codes for "A Lightweight Recurrent Aggregation Network for Satellite Video Super-Resolution", Journal of Selected…☆22Apr 9, 2024Updated 2 years ago
- [NeurIPS 2024] Repository for the paper "OVT-B: A New Large-Scale Benchmark for Open-Vocabulary Multi-Object Tracking".☆29Nov 9, 2024Updated last year
- [NeurIPS 2025] Official implementation of "RoboRefer: Towards Spatial Referring with Reasoning in Vision-Language Models for Robotics"☆266Dec 16, 2025Updated 8 months ago
- [ITSC 2024] Personalized Autonomous Driving with Large Language Models: Field Experiments☆19Aug 11, 2024Updated 2 years ago
- Wishing you could have a 🌈☆15Jan 13, 2026Updated 7 months ago
- Repository for Learning Dexterous Grasping with Object-Centric Visual Affordances [ICRA 2021]☆18May 9, 2024Updated 2 years ago
- The code of paper "From Language to Locomotion: Retargeting-free Humanoid Control via Motion Latent Guidance" accepted by ICLR'26☆72Mar 18, 2026Updated 5 months ago
- Robotics transformers inference servers in ROS2. RT-1, RT-X, Octo.☆17Oct 14, 2024Updated last year
- [ICML 2026] Latent Reasoning VLA: Latent Thinking and Prediction for Vision-Language-Action Models☆92May 18, 2026Updated 3 months ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- Java文件管理器,支持文件树和文件列表,支持文件夹创建、删除、复制、粘贴、加密解密、压缩解压。☆26May 31, 2020Updated 6 years ago
- I love game theory.☆19Dec 25, 2024Updated last year
- ☆13Nov 7, 2021Updated 4 years ago
- Official implementation of "In-style: Bridging Text and Uncurated Videos with Style Transfer for Cross-modal Retrieval." ICCV 2023☆11Oct 5, 2023Updated 2 years ago
- Hướng dẫn tạo một hệ thống Log Remote dùng chung cho nhiều dự án/server☆15Feb 26, 2020Updated 6 years ago
- Seeing from Another Perspective: Evaluating Multi-View Understanding in MLLMs☆71Mar 22, 2026Updated 5 months ago
- [NAACL 2024] Z-GMOT: Zero-shot Generic Multiple Object Tracking☆12May 19, 2026Updated 3 months ago