☆122Apr 9, 2026Updated 3 months ago
Alternatives and similar repositories for SenseNova-MARS
Users that are interested in SenseNova-MARS are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- ☆24Dec 3, 2025Updated 7 months ago
- Official Code for "Mini-o3: Scaling Up Reasoning Patterns and Interaction Turns for Visual Search"☆423Jan 29, 2026Updated 6 months ago
- Code for paper: Reinforced Vision Perception with Tools☆74Oct 3, 2025Updated 9 months ago
- ☆625Feb 26, 2026Updated 5 months ago
- Video-CoM: Interactive Video Reasoning via Chain of Manipulations☆22Jun 17, 2026Updated last month
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- [CVPR 2026] Scaling Spatial Intelligence with Multimodal Foundation Models☆290May 14, 2026Updated 2 months ago
- Simple code sandbox supporting jupyter notebook style code execution. Used for agent training☆25Dec 5, 2025Updated 7 months ago
- Code for "StreamGaze: Gaze-Guided Temporal Reasoning and Proactive Understanding in Streaming Videos"☆27May 13, 2026Updated 2 months ago
- The model, data and code for OpenMobile☆50Jul 9, 2026Updated 2 weeks ago
- [ACL-2026] MMSearch-R1 is an end-to-end RL framework that enables LMMs to perform on-demand, multi-turn search with real-world multimodal…☆470Apr 7, 2026Updated 3 months ago
- [ICLR 2026] Official repo for "FrameThinker: Learning to Think with Long Videos via Multi-Turn Frame Spotlighting"☆50Oct 9, 2025Updated 9 months ago
- ☆13Dec 9, 2024Updated last year
- [ICLR 2026] High-Fidelity Visual Reasoning on Structured Images☆30Jul 17, 2026Updated last week
- [CVPR2026] VideoAuto-R1: Video Auto Reasoning via Thinking Once, Answering Twice☆88Feb 27, 2026Updated 5 months ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- Pixel-Level Reasoning Model trained with RL [NeuIPS25]☆301Jun 4, 2026Updated last month
- ☆72Feb 27, 2026Updated 5 months ago
- [CVPR 2026 Oral] Code with Image☆31Dec 5, 2025Updated 7 months ago
- MMDeepResearch-Bench (MMDR)☆31Apr 1, 2026Updated 3 months ago
- [ECCV 2026] Promsa: Progressive Multimodal Search Agents for Knowledge-Based Visual Question Answering☆91Jul 7, 2026Updated 3 weeks ago
- [ICML 2026] Multimodal deep-research MLLM and benchmark. The first long-horizon multimodal deep-research MLLM, extending the number of re…☆658Jun 8, 2026Updated last month
- [CVPR 2026] LongVT: Incentivizing "Thinking with Long Videos" via Native Tool Calling☆257Jun 24, 2026Updated last month
- [CVPR 2026] Thinking with Programming Vision: Towards a Unified View for Thinking with Images☆71Jan 23, 2026Updated 6 months ago
- [NeurIPS25] Official Implementation (Pytorch) of "DeepVideo-R1"☆37Feb 22, 2026Updated 5 months ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- 🔍 OpenSearch-VL provides a fully open recipe for training strong multimodal deep search agents through high-quality data curation, diver…☆258May 19, 2026Updated 2 months ago
- ☆27Mar 17, 2026Updated 4 months ago
- (ICML 2026) Seg-ReSearch: Segmentation with Interleaved Reasoning and External Search☆49May 1, 2026Updated 2 months ago
- Official repo of Promoting Efficient Reasoning with Verifiable Stepwise Reward☆16Sep 9, 2025Updated 10 months ago
- The repository of VG-Refiner paper☆20Dec 9, 2025Updated 7 months ago
- ☆14Apr 23, 2025Updated last year
- The SAIL-VL2 series model developed by the BytedanceDouyinContent Group☆79Sep 18, 2025Updated 10 months ago
- A Self-Training Framework for Vision-Language Reasoning☆90Jan 23, 2025Updated last year
- cliptrase☆47Sep 1, 2024Updated last year
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- Vision-OPD is a regional-to-global on-policy self-distillation framework that transfers a model's own privileged crop-conditioned percept…☆214Jul 17, 2026Updated last week
- [ICML 2026] ZwZ model family: SOTA fine-grained perception performace; ZoomBench: a new challenging perception benchmark☆179May 4, 2026Updated 2 months ago
- ☆177Nov 26, 2025Updated 8 months ago
- ☆1,251Nov 20, 2025Updated 8 months ago
- [TPAMI 2026] Ego-R1: Agentic Chain-of-Tool-Thought for Ultra-Long Egocentric Video Reasoning☆165Jun 10, 2026Updated last month
- ☆13Jun 4, 2025Updated last year
- The official repository of "A Comprehensive Survey on Reinforcement Learning-based Agentic Search: Foundations, Roles, Optimizations, Eva…☆280Jul 21, 2026Updated last week