S-Agent: Spatial Tool-Use Elicits Reasoning for Spatial Intelligence
☆93Jul 22, 2026Updated 2 months ago
Alternatives and similar repositories for S-Agent
Users that are interested in S-Agent are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- [NeurIPS 2026] SpatialBench: Is Your Spatial Foundation Model an All-Round Player?☆138Jul 25, 2026Updated 2 months ago
- [ICLR 2026] 🦅 FALCON: an effective vision-language-action model injects rich 3D spatial tokens into the action head, enabling robust spa…☆39May 26, 2026Updated 4 months ago
- [ArXiv 26] The official repository of "ArtHOI: Articulated Human-Object Interaction Synthesis by 4D Reconstruction from Video Priors".☆42Mar 5, 2026Updated 6 months ago
- XLD: A Cross-Lane Dataset for Benchmarking Novel Driving View Synthesis☆26Sep 26, 2024Updated 2 years ago
- Demo-ICL: In-Context Learning for Procedural Video Knowledge Acquisition☆47Mar 3, 2026Updated 6 months ago
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- A local AI assistant running on your device. It turns your files into actionable memory.☆58Mar 24, 2026Updated 6 months ago
- View planning with multi-turn VLM agents: ViewSuite 6-DoF benchmark on real ScanNet scenes + iterative RL-SFT training☆25Sep 6, 2026Updated 3 weeks ago
- ☆27Apr 30, 2026Updated 5 months ago
- ☆122Jun 11, 2026Updated 3 months ago
- Official Implementation of "Kinema4D: Kinematic4D World Modeling for Spatiotemporal Embodied Simulation"☆83May 21, 2026Updated 4 months ago
- [ICML 2025] Streamline Without Sacrifice - Squeeze out Computation Redundancy in LMM☆20May 22, 2025Updated last year
- ☆48Mar 27, 2026Updated 6 months ago
- [Arxiv'24] LangSurf: Language-Embedded Surface Gaussians for 3D Scene Understanding☆44Aug 18, 2025Updated last year
- CoSurfGS: Collaborative 3D Surface Gaussian Splatting with Distributed Learning for Large Scene Reconstruction☆64Dec 25, 2024Updated last year
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- [ECCV 2026] Spatial-TTT: Streaming Visual-based Spatial Intelligence with Test-Time Training☆272Jun 19, 2026Updated 3 months ago
- [ICCV'25] ScenePainter: Semantically Consistent Perpetual 3D Scene Generation with Concept Relation Alignment☆37Oct 5, 2025Updated 11 months ago
- The official implementation of "DynamicVLA: A Vision-Language-Action Model for Dynamic Object Manipulation". (NeurIPS 2026)☆344Sep 25, 2026Updated last week
- ☆21Oct 4, 2025Updated 11 months ago
- Dynamic 3D Foundation Model using Causal Transformer. [ICLR 2026]☆409May 8, 2026Updated 4 months ago
- The official implementation of "Compositional Generative Model of Unbounded 4D Cities". (TPAMI 2026)☆151Sep 7, 2026Updated 3 weeks ago
- [ICLR 2026] Light-X: Generative 4D Video Rendering with Camera and Illumination Control☆194Dec 11, 2025Updated 9 months ago
- 第六届全国大学生工程训练综合能力竞赛. Using OpenCV , ROS , VISP.☆13Nov 7, 2019Updated 6 years ago
- High-Resolution Visual Reasoning via Multi-Turn Grounding-Based Reinforcement Learning☆56Jul 23, 2025Updated last year
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- [ICML 2026 Oral] Holi-Spatial: Evolving Video Streams into Holistic 3D Spatial Intelligence☆391Jul 26, 2026Updated 2 months ago
- Source code release for "SERF: Spatiotemporal Environment and Robot Feature Map for Long-Horizon Mobile Manipulation"☆24Aug 20, 2026Updated last month
- ☆122Jun 12, 2026Updated 3 months ago
- [NIPS 2026 Reviewed (6/5/4)] PhysisForcing: Physics Reinforced World Simulator for Robotic Manipulation☆126Sep 8, 2026Updated 3 weeks ago
- [CVPR 2024 Highlight] GP-NeRF: Generalized Perception NeRF for Context-Aware 3D Scene Understanding☆28Jul 26, 2024Updated 2 years ago
- (CVPR 2025) DoF-Gaussian: Controllable Depth-of-Field for 3D Gaussian Splatting☆75Jul 14, 2025Updated last year
- [CVPR 2026 Hightlight] OmniVGGT: Omni-Modality Driven Visual Geometry Grounded Transformer☆376May 21, 2026Updated 4 months ago
- PhysX-Omni: Unified Simulation-Ready Physical 3D Generation for Rigid, Deformable, and Articulated Objects☆322Jun 11, 2026Updated 3 months ago
- [CVPR'26 Highlight] SimRecon: SimReady Compositional Scene Reconstruction from Real Videos☆146Apr 14, 2026Updated 5 months ago
- Deploy open-source AI quickly and easily - Special Bonus Offer • AdRunpod Hub is built for open source. One-click deployment and autoscaling endpoints without provisioning your own infrastructure.
- [CVPR 2026] Scaling Spatial Intelligence with Multimodal Foundation Models☆310May 14, 2026Updated 4 months ago
- StableWorld: Towards Stable and Consistent Long Interactive Video Generation☆102Aug 30, 2026Updated last month
- SPAgent, a foundation agent for understanding, reasoning over, and operating within the physical and spatial world.☆224Sep 3, 2026Updated 3 weeks ago
- A benchmark for evaluating contextual agents on realistic multimodal personal-computer environments with profiling and factual-retention …☆33Apr 2, 2026Updated 6 months ago
- [NeurIPS 2026] SpatialClaw: Rethinking Action Interface for Agentic Spatial Reasoning☆402Sep 14, 2026Updated 2 weeks ago
- [ICML 2026] 4RC: 4D Reconstruction via Conditional Querying Anytime and Anywhere☆248Jul 7, 2026Updated 2 months ago
- PhysX-Anything: Simulation-Ready Physical 3D Assets from Single Image (CVPR 2026)☆946Apr 28, 2026Updated 5 months ago