[CVPR 2026] Scalable Object Relation Encoding for Better 3D Spatial Reasoning in Large Language Models
☆23May 28, 2026Updated 4 months ago
Alternatives and similar repositories for QuatRoPE
Users that are interested in QuatRoPE are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- [ICRA 2024] WLST: Weak Labels Guided Self-training for Weakly-supervised Domain Adaptation on 3D Object Detection☆12Feb 6, 2024Updated 2 years ago
- [AAAI 24] Official Codebase for BridgeQA: Bridging the Gap between 2D and 3D Visual Question Answering: A Fusion Approach for 3D VQA☆29Jul 12, 2024Updated 2 years ago
- ☆36Aug 12, 2026Updated last month
- ☆57Sep 13, 2024Updated 2 years ago
- ☆13Jul 22, 2024Updated 2 years ago
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- SURE-Map: Self-Correcting Streaming Geometric Foundation Models☆162Sep 22, 2026Updated last week
- OGScene3D: Incremental Open-Vocabulary 3D Gaussian Scene Graph Mapping for Scene Understanding☆17Mar 18, 2026Updated 6 months ago
- [CVPR 2026] EasyOmnimatte: Taming Pretrained Inpainting Diffusion Models for End-to-End Video Layered Decomposition☆29Dec 29, 2025Updated 9 months ago
- ☆19Apr 27, 2026Updated 5 months ago
- Offical implementation of CVPR 2026 paper SpaceDrive: Infusing Spatial Awareness into VLM-based Autonomous Driving.☆94May 19, 2026Updated 4 months ago
- [CVPR'24] MiKASA: Multi-Key-Anchor & Scene-Aware Transformer for 3D Visual Grounding☆18Dec 13, 2024Updated last year
- [RAL2026] Official codebase for LiveVLN: Breaking the Stop-and-Go Loop in Vision-Language Navigation☆23Apr 22, 2026Updated 5 months ago
- ☆22Nov 18, 2025Updated 10 months ago
- [CVPR 2025] The code for paper ''Video-3D LLM: Learning Position-Aware Video Representation for 3D Scene Understanding''.☆224Jun 4, 2025Updated last year
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- SparseOccVLA: Bridging Occupancy and Vision-Language Models via Sparse Queries for Unified 4D Scene Understanding and Planning☆86Feb 1, 2026Updated 8 months ago
- FastBEV-ROS-TensorRT-CPP real time inference including ros1 & ros2☆39May 14, 2024Updated 2 years ago
- Weakly Supervised Video Moment Localisation with Contrastive Negative Sample Mining☆30Apr 4, 2022Updated 4 years ago
- [ECCV2024] Reflective Instruction Tuning: Mitigating Hallucinations in Large Vision-Language Models☆20Jul 17, 2024Updated 2 years ago
- [CVPR 2026] Few-Shot Incremental 3D Object Detection in Dynamic Indoor Environments☆18Aug 24, 2026Updated last month
- DriveTok: 3D Driving Scene Tokenization for Unified Multi-View Reconstruction and Understanding☆28Mar 20, 2026Updated 6 months ago
- [CVPR 2025] Official implementation of the paper "Point-Cache: Test-time Dynamic and Hierarchical Cache for Robust and Generalizable Poin…☆20Aug 16, 2026Updated last month
- ☆12Jan 31, 2024Updated 2 years ago
- [MICCAI'25] ClipGS: Clippable Gaussian Splatting for Interactive Cinematic Visualization of Volumetric Medical Data☆16Jul 28, 2025Updated last year
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- AIGCDetectBaseline☆12Jul 9, 2024Updated 2 years ago
- End-to-End Vectorized Autonomous Driving via Probabilistic Planning☆30Apr 20, 2024Updated 2 years ago
- [NeurIPS 2025 Spotlight] Official PyTorch implementation of Vgent☆51Nov 30, 2025Updated 10 months ago
- [EMNLP 2025 Oral] Official codebase for Seeing More, Saying More: Lightweight Language Experts are Dynamic Video Token Compressors.☆18Sep 7, 2025Updated last year
- Code for ICRA'25 paper: "Differentiable Composite Neural Signed Distance Fields for Robot Navigation in Dynamic Indoor Environments"☆16Apr 2, 2025Updated last year
- ☆27Jun 5, 2025Updated last year
- [ICRA 2024] Official Implementation of the paper "Parameter-efficient Prompt Learning for 3D Point Cloud Understanding"☆30Mar 13, 2026Updated 6 months ago
- CVPR2026 (main)☆17Mar 30, 2026Updated 6 months ago
- HyperPose☆14Nov 6, 2025Updated 10 months ago
- Proton VPN Special Offer - Get 70% off • AdSpecial partner offer. Trusted by over 100 million users worldwide. Tested, Approved and Recommended by Experts.
- pointcloud analysis☆23Apr 11, 2023Updated 3 years ago
- [ICCV 2025] FiVE-Bench: A Fine-grained Video Editing Benchmark for Evaluating Emerging Diffusion and Rectified Flow Models☆39Apr 2, 2026Updated 6 months ago
- [NeurIPS‘24] Multi-Object 3D Grounding with Dynamic Modules and Language Informed Spatial Attention☆28Jun 15, 2025Updated last year
- ☆53Jun 22, 2026Updated 3 months ago
- Official Implementation for paper "Pretraining A Large Language Model using Distributed GPUs: A Memory-Efficient Decentralized Paradigm"☆24May 8, 2026Updated 4 months ago
- HyVGGT-VO: Hybrid Dense Visual Odometry With Classical VO and Feed-Forward Model☆40Sep 1, 2026Updated last month
- [RA-L 2024] 3D Active Metric-Semantic SLAM☆17Jul 21, 2025Updated last year