Official implementation of “Towards Cross-View Point Correspondence in Vision-Language Models”.
☆15Dec 24, 2025Updated 7 months ago
Alternatives and similar repositories for CrossPoint
Users that are interested in CrossPoint are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Reshaping Action Error Distributions for Reliable Vision-Language-Action Models☆17Feb 5, 2026Updated 6 months ago
- Code implementation for paper titled "HOI-Ref: Hand-Object Interaction Referral in Egocentric Vision"☆30Apr 16, 2024Updated 2 years ago
- Transformation Driven Visual Reasoning - CVPR 2021☆36May 27, 2023Updated 3 years ago
- ☆15Jun 9, 2025Updated last year
- ☆11May 9, 2024Updated 2 years ago
- Virtual machines for every use case on DigitalOcean • AdGet dependable uptime with 99.99% SLA, simple security tools, and predictable monthly pricing with DigitalOcean's virtual machines, called Droplets.
- ☆62Apr 1, 2025Updated last year
- MLCD-Seg is a zero-shot segmentation model from DeepGlint.☆18Jul 4, 2025Updated last year
- Implementation of the Marching Cubes algorithm on Python.☆11Dec 10, 2020Updated 5 years ago
- A Comprehensive Empirical Study of Vision-Language Pre-trained Model for Supervised Cross-Modal Retrieval☆43Apr 13, 2022Updated 4 years ago
- Dreamitate: Real-World Visuomotor Policy Learning via Video Generation (CoRL 2024)☆59Jun 7, 2025Updated last year
- The official code implementation of Generalized Category Discovery in Semantic Segmentation☆17Dec 20, 2023Updated 2 years ago
- 2025 年春季学期北京大学计算机视觉导论课程的课程资料,王鹤老师。☆16Jul 4, 2025Updated last year
- Training recipe for SpatialReasoner [NeurIPS 2025]☆45Apr 5, 2026Updated 4 months ago
- ☆14Jan 5, 2022Updated 4 years ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- A simple implementation of ReasonGenRM.☆19Apr 21, 2025Updated last year
- [NeurIPS 2024] Repository for the paper "OVT-B: A New Large-Scale Benchmark for Open-Vocabulary Multi-Object Tracking".☆29Nov 9, 2024Updated last year
- [NeurIPS 2025] Official implementation of "RoboRefer: Towards Spatial Referring with Reasoning in Vision-Language Models for Robotics"☆266Dec 16, 2025Updated 7 months ago
- Repository for Learning Dexterous Grasping with Object-Centric Visual Affordances [ICRA 2021]☆18May 9, 2024Updated 2 years ago
- A Free World Class High Performance SAT Solver☆20Jul 12, 2021Updated 5 years ago
- The code of paper "From Language to Locomotion: Retargeting-free Humanoid Control via Motion Latent Guidance" accepted by ICLR'26☆70Mar 18, 2026Updated 4 months ago
- Robotics transformers inference servers in ROS2. RT-1, RT-X, Octo.☆17Oct 14, 2024Updated last year
- I love game theory.☆19Dec 25, 2024Updated last year
- Official implementation of "In-style: Bridging Text and Uncurated Videos with Style Transfer for Cross-modal Retrieval." ICCV 2023☆11Oct 5, 2023Updated 2 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- Hướng dẫn tạo một hệ thống Log Remote dùng chung cho nhiều dự án/server☆15Feb 26, 2020Updated 6 years ago
- Seeing from Another Perspective: Evaluating Multi-View Understanding in MLLMs☆71Mar 22, 2026Updated 4 months ago
- [NAACL 2024] Z-GMOT: Zero-shot Generic Multiple Object Tracking☆12May 19, 2026Updated 2 months ago
- My PhD manuscript LaTeX code and the slides for the defense☆11Feb 2, 2022Updated 4 years ago
- A curated list of Survey Papers on Deep Learning.☆13Sep 5, 2023Updated 2 years ago
- This repository including most of cnn visualizations techniques using pytorch☆14Apr 14, 2020Updated 6 years ago
- ☆10Oct 18, 2024Updated last year
- [NeurIPS 2025] Source codes for the paper "MindJourney: Test-Time Scaling with World Models for Spatial Reasoning"☆151Nov 4, 2025Updated 9 months ago
- [ICCV'25] Ross3D: Reconstructive Visual Instruction Tuning with 3D-Awareness☆70Jul 22, 2025Updated last year
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- Official PyTorch Implementation of Paper "Hand2World: Autoregressive Egocentric Interaction Generation via Free-Space Hand Gestures"☆26Jun 30, 2026Updated last month
- ☆15May 7, 2024Updated 2 years ago
- PyTorch implementation of the paper: CASAGPT: Cuboid Arrangement and Scene Assembly for Interior Design [CVPR 2025]☆15Apr 5, 2025Updated last year
- X-MIC: Cross-Modal Instance Conditioning for Egocentric Action Generalization, CVPR 2024☆11Nov 7, 2024Updated last year
- FlowHOI: Flow-based Semantics-Grounded Generation of Hand-Object Interactions for Dexterous Robot Manipulation☆35Feb 17, 2026Updated 5 months ago
- The official repository for paper "FlexSelect: Flexible Token Selection for Efficient Long Video Understanding".☆31Sep 19, 2025Updated 10 months ago
- Official Release of NeurIPS 2024 paper "Slot State Space Models"☆11Mar 22, 2025Updated last year