S-Agent: Spatial Tool-Use Elicits Reasoning for Spatial Intelligence
☆90Jul 22, 2026Updated last month
Alternatives and similar repositories for S-Agent
Users that are interested in S-Agent are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- SpatialBench: Is Your Spatial Foundation Model an All-Round Player?☆130Jul 25, 2026Updated last month
- [ICLR 2026] 🦅 FALCON: an effective vision-language-action model injects rich 3D spatial tokens into the action head, enabling robust spa…☆36May 26, 2026Updated 3 months ago
- [ArXiv 26] The official repository of "ArtHOI: Articulated Human-Object Interaction Synthesis by 4D Reconstruction from Video Priors".☆42Mar 5, 2026Updated 6 months ago
- XLD: A Cross-Lane Dataset for Benchmarking Novel Driving View Synthesis☆25Sep 26, 2024Updated last year
- Demo-ICL: In-Context Learning for Procedural Video Knowledge Acquisition☆47Mar 3, 2026Updated 6 months ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- A local AI assistant running on your device. It turns your files into actionable memory.☆58Mar 24, 2026Updated 5 months ago
- View planning with multi-turn VLM agents: ViewSuite 6-DoF benchmark on real ScanNet scenes + iterative RL-SFT training☆24Sep 6, 2026Updated last week
- ☆25Apr 30, 2026Updated 4 months ago
- ☆121Jun 11, 2026Updated 3 months ago
- Official Implementation of "Kinema4D: Kinematic4D World Modeling for Spatiotemporal Embodied Simulation"☆83May 21, 2026Updated 3 months ago
- ☆45Mar 27, 2026Updated 5 months ago
- [ICML 2025] Streamline Without Sacrifice - Squeeze out Computation Redundancy in LMM☆20May 22, 2025Updated last year
- [Arxiv'24] LangSurf: Language-Embedded Surface Gaussians for 3D Scene Understanding☆44Aug 18, 2025Updated last year
- CoSurfGS: Collaborative 3D Surface Gaussian Splatting with Distributed Learning for Large Scene Reconstruction☆65Dec 25, 2024Updated last year
- Serverless GPU API endpoints on Runpod - Get Bonus Credits • AdSkip the infrastructure headaches. Auto-scaling, pay-as-you-go, no-ops approach lets you focus on innovating your application.
- [ECCV 2026] Spatial-TTT: Streaming Visual-based Spatial Intelligence with Test-Time Training☆266Jun 19, 2026Updated 2 months ago
- [ICCV'25] ScenePainter: Semantically Consistent Perpetual 3D Scene Generation with Concept Relation Alignment☆37Oct 5, 2025Updated 11 months ago
- The official implementation of "DynamicVLA: A Vision-Language-Action Model for Dynamic Object Manipulation". (arXiv 2601.22153)☆330Sep 2, 2026Updated last week
- ☆21Oct 4, 2025Updated 11 months ago
- Dynamic 3D Foundation Model using Causal Transformer. [ICLR 2026]☆405May 8, 2026Updated 4 months ago
- The official implementation of "Compositional Generative Model of Unbounded 4D Cities". (TPAMI 2026)☆151Updated this week
- [ICLR 2026] Light-X: Generative 4D Video Rendering with Camera and Illumination Control☆194Dec 11, 2025Updated 9 months ago
- 第六届全国大学生工程训练综合能力竞赛. Using OpenCV , ROS , VISP.☆13Nov 7, 2019Updated 6 years ago
- High-Resolution Visual Reasoning via Multi-Turn Grounding-Based Reinforcement Learning☆56Jul 23, 2025Updated last year
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- [ICML 2026 Oral] Holi-Spatial: Evolving Video Streams into Holistic 3D Spatial Intelligence☆383Jul 26, 2026Updated last month
- Source code release for "SERF: Spatiotemporal Environment and Robot Feature Map for Long-Horizon Mobile Manipulation"☆24Aug 20, 2026Updated 3 weeks ago
- ☆120Jun 12, 2026Updated 3 months ago
- PhysisForcing: Physics Reinforced World Simulator for Robotic Manipulation☆123Updated this week
- [CVPR 2024 Highlight] GP-NeRF: Generalized Perception NeRF for Context-Aware 3D Scene Understanding☆28Jul 26, 2024Updated 2 years ago
- (CVPR 2025) DoF-Gaussian: Controllable Depth-of-Field for 3D Gaussian Splatting☆75Jul 14, 2025Updated last year
- [CVPR 2026 Hightlight] OmniVGGT: Omni-Modality Driven Visual Geometry Grounded Transformer☆371May 21, 2026Updated 3 months ago
- PhysX-Omni: Unified Simulation-Ready Physical 3D Generation for Rigid, Deformable, and Articulated Objects☆316Jun 11, 2026Updated 3 months ago
- [CVPR'26 Highlight] SimRecon: SimReady Compositional Scene Reconstruction from Real Videos☆143Apr 14, 2026Updated 4 months ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- [CVPR 2026] Scaling Spatial Intelligence with Multimodal Foundation Models☆300May 14, 2026Updated 3 months ago
- StableWorld: Towards Stable and Consistent Long Interactive Video Generation☆101Aug 30, 2026Updated 2 weeks ago
- SPAgent, a foundation agent for understanding, reasoning over, and operating within the physical and spatial world.☆222Sep 3, 2026Updated last week
- A benchmark for evaluating contextual agents on realistic multimodal personal-computer environments with profiling and factual-retention …☆32Apr 2, 2026Updated 5 months ago
- SpatialClaw: Rethinking Action Interface for Agentic Spatial Reasoning☆377Jul 18, 2026Updated last month
- [ICML 2026] 4RC: 4D Reconstruction via Conditional Querying Anytime and Anywhere☆240Jul 7, 2026Updated 2 months ago
- PhysX-Anything: Simulation-Ready Physical 3D Assets from Single Image (CVPR 2026)☆942Apr 28, 2026Updated 4 months ago