☆132Apr 9, 2026Updated 5 months ago
Alternatives and similar repositories for SenseNova-MARS
Users that are interested in SenseNova-MARS are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- ☆26Dec 3, 2025Updated 10 months ago
- Official Code for "Mini-o3: Scaling Up Reasoning Patterns and Interaction Turns for Visual Search"☆427Jan 29, 2026Updated 8 months ago
- Code for paper: Reinforced Vision Perception with Tools☆73Oct 3, 2025Updated last year
- ☆643Feb 26, 2026Updated 7 months ago
- Video-CoM: Interactive Video Reasoning via Chain of Manipulations☆23Sep 5, 2026Updated 3 weeks ago
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- An open source code repository of driving world models, with training, inferencing, evaluation tools, and pretrained checkpoints.☆422Jun 19, 2025Updated last year
- [CVPR 2026] Scaling Spatial Intelligence with Multimodal Foundation Models☆310May 14, 2026Updated 4 months ago
- Simple code sandbox supporting jupyter notebook style code execution. Used for agent training☆25Dec 5, 2025Updated 9 months ago
- Code for "StreamGaze: Gaze-Guided Temporal Reasoning and Proactive Understanding in Streaming Videos [NeurIPS 2026]"☆28Sep 25, 2026Updated last week
- The model, data and code for OpenMobile☆50Jul 9, 2026Updated 2 months ago
- [ACL-2026] MMSearch-R1 is an end-to-end RL framework that enables LMMs to perform on-demand, multi-turn search with real-world multimodal…☆486Apr 7, 2026Updated 5 months ago
- [ICLR 2026] Official repo for "FrameThinker: Learning to Think with Long Videos via Multi-Turn Frame Spotlighting"☆56Oct 9, 2025Updated 11 months ago
- ☆13Dec 9, 2024Updated last year
- VisualToolAgent (VisTA): A Reinforcement Learning Framework for Visual Tool Selection☆27May 31, 2025Updated last year
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- [ICLR 2026] High-Fidelity Visual Reasoning on Structured Images☆36Jul 17, 2026Updated 2 months ago
- [CVPR2026] VideoAuto-R1: Video Auto Reasoning via Thinking Once, Answering Twice☆89Feb 27, 2026Updated 7 months ago
- Pixel-Level Reasoning Model trained with RL [NeuIPS25]☆307Jul 28, 2026Updated 2 months ago
- ☆71Feb 27, 2026Updated 7 months ago
- [CVPR 2026 Oral] Code with Image☆34Dec 5, 2025Updated 9 months ago
- MMDeepResearch-Bench (MMDR)☆33Aug 10, 2026Updated last month
- [ICML 2026 & EMNLP 2026] Multimodal deep-research MLLM and benchmark. The first long-horizon multimodal deep-research MLLM, extending the…☆686Aug 8, 2026Updated last month
- [CVPR 2026] LongVT: Incentivizing "Thinking with Long Videos" via Native Tool Calling☆270Jun 24, 2026Updated 3 months ago
- [CVPR 2026] Thinking with Programming Vision: Towards a Unified View for Thinking with Images☆74Jan 23, 2026Updated 8 months ago
- Managed Database hosting by DigitalOcean • AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- [NeurIPS25] Official Implementation (Pytorch) of "DeepVideo-R1"☆38Feb 22, 2026Updated 7 months ago
- [ECCV 2026] Promsa: Progressive Multimodal Search Agents for Knowledge-Based Visual Question Answering☆90Jul 7, 2026Updated 2 months ago
- 🔍 OpenSearch-VL provides a fully open recipe for training strong multimodal deep search agents through high-quality data curation, diver…☆289Jul 30, 2026Updated 2 months ago
- ☆28Mar 17, 2026Updated 6 months ago
- (ICML 2026) Seg-ReSearch: Segmentation with Interleaved Reasoning and External Search☆50Aug 7, 2026Updated last month
- Official repo of Promoting Efficient Reasoning with Verifiable Stepwise Reward☆16Sep 9, 2025Updated last year
- The repository of VG-Refiner paper☆20Dec 9, 2025Updated 9 months ago
- ☆14Apr 23, 2025Updated last year
- The SAIL-VL2 series model developed by the BytedanceDouyinContent Group☆79Sep 18, 2025Updated last year
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- A Self-Training Framework for Vision-Language Reasoning☆91Jan 23, 2025Updated last year
- cliptrase☆49Sep 1, 2024Updated 2 years ago
- Vision-OPD is a regional-to-global on-policy self-distillation framework that transfers a model's own privileged crop-conditioned percept…☆325Updated this week
- ☆177Nov 26, 2025Updated 10 months ago
- [ICML 2026] ZwZ model family: SOTA fine-grained perception performace; ZoomBench: a new challenging perception benchmark☆191May 4, 2026Updated 4 months ago
- ☆1,278Nov 20, 2025Updated 10 months ago
- [TPAMI 2026] Ego-R1: Agentic Chain-of-Tool-Thought for Ultra-Long Egocentric Video Reasoning☆169Jun 10, 2026Updated 3 months ago