☆126Apr 9, 2026Updated 4 months ago
Alternatives and similar repositories for SenseNova-MARS
Users that are interested in SenseNova-MARS are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- ☆24Dec 3, 2025Updated 8 months ago
- Official Code for "Mini-o3: Scaling Up Reasoning Patterns and Interaction Turns for Visual Search"☆425Jan 29, 2026Updated 6 months ago
- Code for paper: Reinforced Vision Perception with Tools☆73Oct 3, 2025Updated 10 months ago
- ☆633Feb 26, 2026Updated 5 months ago
- Video-CoM: Interactive Video Reasoning via Chain of Manipulations☆23Jun 17, 2026Updated 2 months ago
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- An open source code repository of driving world models, with training, inferencing, evaluation tools, and pretrained checkpoints.☆414Jun 19, 2025Updated last year
- [CVPR 2026] Scaling Spatial Intelligence with Multimodal Foundation Models☆294May 14, 2026Updated 3 months ago
- Simple code sandbox supporting jupyter notebook style code execution. Used for agent training☆25Dec 5, 2025Updated 8 months ago
- Code for "StreamGaze: Gaze-Guided Temporal Reasoning and Proactive Understanding in Streaming Videos"☆27May 13, 2026Updated 3 months ago
- The model, data and code for OpenMobile☆49Jul 9, 2026Updated last month
- [ACL-2026] MMSearch-R1 is an end-to-end RL framework that enables LMMs to perform on-demand, multi-turn search with real-world multimodal…☆478Apr 7, 2026Updated 4 months ago
- [ICLR 2026] Official repo for "FrameThinker: Learning to Think with Long Videos via Multi-Turn Frame Spotlighting"☆56Oct 9, 2025Updated 10 months ago
- ☆13Dec 9, 2024Updated last year
- VisualToolAgent (VisTA): A Reinforcement Learning Framework for Visual Tool Selection☆27May 31, 2025Updated last year
- Virtual machines for every use case on DigitalOcean • AdGet dependable uptime with 99.99% SLA, simple security tools, and predictable monthly pricing with DigitalOcean's virtual machines, called Droplets.
- [ICLR 2026] High-Fidelity Visual Reasoning on Structured Images☆32Jul 17, 2026Updated last month
- [CVPR2026] VideoAuto-R1: Video Auto Reasoning via Thinking Once, Answering Twice☆88Feb 27, 2026Updated 5 months ago
- Pixel-Level Reasoning Model trained with RL [NeuIPS25]☆304Jul 28, 2026Updated 3 weeks ago
- ☆72Feb 27, 2026Updated 5 months ago
- [CVPR 2026 Oral] Code with Image☆32Dec 5, 2025Updated 8 months ago
- MMDeepResearch-Bench (MMDR)☆32Aug 10, 2026Updated last week
- [ECCV 2026] Promsa: Progressive Multimodal Search Agents for Knowledge-Based Visual Question Answering☆91Jul 7, 2026Updated last month
- [ICML 2026] Multimodal deep-research MLLM and benchmark. The first long-horizon multimodal deep-research MLLM, extending the number of re…☆675Aug 8, 2026Updated last week
- [CVPR 2026] LongVT: Incentivizing "Thinking with Long Videos" via Native Tool Calling☆260Jun 24, 2026Updated last month
- End-to-end encrypted email - Proton Mail • AdSpecial offer: 40% Off Yearly / 80% Off First Month. All Proton services are open source and independently audited for security.
- [CVPR 2026] Thinking with Programming Vision: Towards a Unified View for Thinking with Images☆72Jan 23, 2026Updated 6 months ago
- [NeurIPS25] Official Implementation (Pytorch) of "DeepVideo-R1"☆38Feb 22, 2026Updated 5 months ago
- 🔍 OpenSearch-VL provides a fully open recipe for training strong multimodal deep search agents through high-quality data curation, diver…☆266Jul 30, 2026Updated 2 weeks ago
- Dynamic dual-granularity skill bank for agentic RL, jointly evolving policy and skills to improve long-horizon decision making in agentic…☆69Apr 1, 2026Updated 4 months ago
- ☆27Mar 17, 2026Updated 5 months ago
- (ICML 2026) Seg-ReSearch: Segmentation with Interleaved Reasoning and External Search☆50Aug 7, 2026Updated last week
- Official repo of Promoting Efficient Reasoning with Verifiable Stepwise Reward☆16Sep 9, 2025Updated 11 months ago
- The repository of VG-Refiner paper☆20Dec 9, 2025Updated 8 months ago
- ☆14Apr 23, 2025Updated last year
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- The SAIL-VL2 series model developed by the BytedanceDouyinContent Group☆79Sep 18, 2025Updated 11 months ago
- A Self-Training Framework for Vision-Language Reasoning☆90Jan 23, 2025Updated last year
- cliptrase☆47Sep 1, 2024Updated last year
- Vision-OPD is a regional-to-global on-policy self-distillation framework that transfers a model's own privileged crop-conditioned percept…☆280Jul 17, 2026Updated last month
- [ICML 2026] ZwZ model family: SOTA fine-grained perception performace; ZoomBench: a new challenging perception benchmark☆186May 4, 2026Updated 3 months ago
- ☆177Nov 26, 2025Updated 8 months ago
- ☆1,265Nov 20, 2025Updated 8 months ago