Stepping VLMs onto the Court: Benchmarking Spatial Intelligence in Sports
☆71Mar 15, 2026Updated 6 months ago
Alternatives and similar repositories for CourtSI
Users that are interested in CourtSI are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- PhotoFlow: Agentic 3D Virtual Photography Missions☆42May 27, 2026Updated 3 months ago
- [ECCV'26] GRADE: Grounded Reasoning Assessment for Discipline-informed Editing☆29Updated this week
- SpaceDG: Benchmarking Spatial Intelligence under Visual Degradation☆31Jul 9, 2026Updated 2 months ago
- RISE-Video: Can Video Generators Decode Implicit World Rules?☆28Mar 26, 2026Updated 5 months ago
- [ICML 2026 Oral] Holi-Spatial: Evolving Video Streams into Holistic 3D Spatial Intelligence☆385Jul 26, 2026Updated last month
- Deploy open-source AI quickly and easily - Special Bonus Offer • AdRunpod Hub is built for open source. One-click deployment and autoscaling endpoints without provisioning your own infrastructure.
- Official repository for the FIRM Reward series☆49Aug 25, 2026Updated 3 weeks ago
- [ECCV'26] Code repo for "EvoTok: A Unified Image Tokenizer via Residual Latent Evolution for Visual Understanding and Generation"☆23Jun 18, 2026Updated 3 months ago
- [ECCV2024] The official repository of the paper "Mask as Supervision: Leveraging Unified Mask Information for Unsupervised 3D Pose Estima…☆18Nov 21, 2024Updated last year
- The official repository of the paper "SGA-INTERACT: A 3D Skeleton-based Benchmark for Group Activity Understanding in Modern Basketball T…☆18Mar 11, 2025Updated last year
- (ECCV2024) Within the Dynamic Context: Inertia-aware 3D Human Modeling with Pose Sequence☆20Jun 27, 2025Updated last year
- The official repository of the first version of ACE-Brain foundation model.☆86Mar 13, 2026Updated 6 months ago
- [ICLR 2026] SpaCE-10: A Comprehensive Benchmark for Multimodal Large Language Models in Compositional Spatial Intelligence☆20Jan 26, 2026Updated 7 months ago
- 让每一次引用都成为可解释的影响力 Turning Every Citation into Explainable Impact☆313Aug 18, 2026Updated last month
- InternVL-U is a 4B-parameter unified multimodal model (UMM) that brings multimodal understanding, reasoning, image generation, image edit…☆296Mar 21, 2026Updated 5 months ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- The official repository of the paper "X as Supervision: Contending with Depth Ambiguity in Unsupervised Monocular 3D Pose Estimation"☆13Jan 22, 2025Updated last year
- The official code of FineRMoE.☆23Mar 17, 2026Updated 6 months ago
- [CVPR 2022] Learning Adaptive Warping for Real-World Rolling Shutter Correction☆35May 31, 2024Updated 2 years ago
- [ICCV'2025] Sequential Gaussian Avatars with Hierarchical Motion Context☆21Oct 11, 2025Updated 11 months ago
- The official repository of the paper "Learnable SMPLify: A Neural Solution for Optimization-Free Human Pose Inverse Kinematics"☆37Aug 25, 2025Updated last year
- (ICCV2025) ToMiE: Towards Explicit Exoskeleton for the Reconstruction of Complicated 3D Human Avatars☆44Jan 19, 2026Updated 8 months ago
- ☆24May 28, 2025Updated last year
- [CVPR 2026] Official PyTorch Implementation for "Motion-Aware Animatable Gaussian Avatars Deblurring".☆28May 6, 2026Updated 4 months ago
- (ICCV2023) NeRFrac: Neural Radiance Fields through Refractive Surface☆36Jul 8, 2025Updated last year
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- [NIPS 2025 DB Oral] Official Repository of paper: Envisioning Beyond the Pixels: Benchmarking Reasoning-Informed Visual Editing☆154May 18, 2026Updated 4 months ago
- [ISPRS2026] DVGBench: Implicit-to-Explicit Visual Grounding Benchmark in UAV Imagery with Large Vision-Language Models☆33Mar 24, 2026Updated 5 months ago
- Official Repository of paper MM-HELIX: Boosting Multimodal Long-Chain Reflective Reasoning with Holistic Platform and Adaptive Hybrid Pol…☆68Jan 26, 2026Updated 7 months ago
- Implementation for CycMuNet+☆27Nov 8, 2023Updated 2 years ago
- [CVPR'25] Official repo of "Point2RBox-v2:Rethinking Point-supervised Oriented Object Detection with Spatial Layout Among Instances"☆45Aug 3, 2026Updated last month
- Official repository for Interleave-VLA☆23Apr 5, 2026Updated 5 months ago
- [ECCV2022 Oral] Bringing Rolling Shutter Images Alive with Dual Reversed Distortion☆53Mar 28, 2024Updated 2 years ago
- Video-MME-v2: Towards the Next Stage in Benchmarks for Comprehensive Video Understanding☆371Aug 5, 2026Updated last month
- [ACM MM 2026] AniCrafter: Customizing Realistic Human-Centric Animation via Avatar-Background Conditioning in Video Diffusion Models☆143Aug 3, 2026Updated last month
- Simple, predictable pricing with DigitalOcean hosting • AdAlways know what you'll pay with monthly caps and flat pricing. Enterprise-grade infrastructure trusted by 600k+ customers.
- [IGARSS 2025 Oral] A Simple Aerial Detection Baseline of Multimodal Language Models.☆92Feb 12, 2026Updated 7 months ago
- Bridging the gap between image generation and real-world design: a benchmark for structured, multi-constraint commercial visual content g…☆22Apr 24, 2026Updated 4 months ago
- [ICML 2026] Official implementation for paper: Learning Self-Correction in Vision–Language Models via Rollout Augmentation☆16Jun 4, 2026Updated 3 months ago
- [TGRS'25] AirSpatialBot: A Spatially-Aware Aerial Agent for Fine-Grained Vehicle Attribute Recognization and Retrieval☆33Jan 6, 2026Updated 8 months ago
- [NeurIPS 2024 Spotlight ⭐️ & TPAMI 2025] Parameter-Inverted Image Pyramid Networks (PIIP)☆113Aug 5, 2025Updated last year
- Some papers about instance segmentation☆20Aug 9, 2022Updated 4 years ago
- [ICLR'26] OF-Diff: Object Fidelity Diffusion for Remote Sensing Image Generation☆40Feb 6, 2026Updated 7 months ago