[ICLR'26] SketchThinker-R1: Towards Efficient Sketch-Style Reasoning in Large Multimodal Models
☆17Mar 26, 2026Updated 3 months ago
Alternatives and similar repositories for SketchThinker-R1
Users that are interested in SketchThinker-R1 are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Official Repo for Look, Compare and Draw: Differential Query Transformer for Automatic Oil Painting☆17Mar 31, 2026Updated 3 months ago
- [CVPR 2026] Code for "The Coherence Trap: When MLLM-Crafted Narratives Exploit Manipulated Visual Contexts"☆20Jun 13, 2026Updated last month
- ☆16Jan 13, 2024Updated 2 years ago
- 🎨Official Repo for Every Painting Awakened: A Training-free Framework for Painting-to-Animation Generation☆57Apr 10, 2025Updated last year
- 🚁 Can Vision-Language Models Think from the Sky? UAVReason for Aerial Reasoning and Generation☆22Jul 11, 2026Updated last week
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- Offical repo for ECCV 2024: Depth-Aware Blind Image Decomposition for Real-World Weather Recovery☆13Mar 7, 2024Updated 2 years ago
- Progressive Text-to-3D Generation for Automatic 3D Prototyping (ACM TOMM)☆54Mar 14, 2026Updated 4 months ago
- The official repo of "Towards Scalable Video Anomaly Retrieval: A Synthetic Video-Text Benchmark"☆21Jun 5, 2025Updated last year
- 📌 Official baseline implementation of 'Last-Meter Precision Navigation for UAVs: A Diffusion-Refined Aerial Visual Servoing Approach‘, s…☆46Jul 7, 2026Updated 2 weeks ago
- 📖Curated list about reasoning abilitiy of MLLM, including OpenAI o1, OpenAI o3-mini, and Slow-Thinking.☆13Feb 7, 2025Updated last year
- ACM MM Workshop on UAVs in Multimedia: Capturing the World from a New Perspective (UAVM 2023)☆13Jul 4, 2026Updated 2 weeks ago
- The official code of "Beyond Walking: A Large-Scale Image-Text Benchmark for Text-based Person Anomaly Search"☆33Updated this week
- ICLR‘24 Offical Implementation of Composed Image Retrieval with Text Feedback via Multi-grained Uncertainty Regularization☆75Jan 30, 2024Updated 2 years ago
- Official implementation for paper "Demystifying Reinforcement Learning for Long-Horizon Tool-Using Agents: A Comprehensive Recipe"☆32May 12, 2026Updated 2 months ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- [ICLR 2020] Haotao Wang, Tianlong Chen, Zhangyang Wang, Kede Ma, "I Am Going MAD: Maximum Discrepancy Competition for Comparing Classifie…☆20Dec 30, 2021Updated 4 years ago
- [ICCV'25] "Harnessing Uncertainty-aware Bounding Boxes for Unsupervised 3D Object Detection".☆26Jan 12, 2026Updated 6 months ago
- [npj AI] 3D Magic Mirror: Clothing Reconstruction from a Single Image via a Causal Perspective Single-View 3D Reconstruction☆90Jul 6, 2026Updated 2 weeks ago
- ☆11Mar 13, 2017Updated 9 years ago
- Set of scripts and instructions for sub-selecting and formatting raw data exported by the Canfield ISIC2024 Tile Export Tool. The resulti…☆12May 8, 2024Updated 2 years ago
- 🔎Official code for our paper: "VL-Uncertainty: Detecting Hallucination in Large Vision-Language Model via Uncertainty Estimation".☆56Mar 18, 2025Updated last year
- Official Implementation of PiPa: Pixel- and Patch-wise Self-supervised Learning for Domain Adaptative Semantic Segmentation☆100Jul 23, 2024Updated last year
- Official code for "Rethinking Chain-of-Thought Reasoning for Videos"☆21Dec 14, 2025Updated 7 months ago
- The official code of "CaLa: Complementary Association Learning for Augmenting Composed Image Retrieval"☆15Sep 19, 2024Updated last year
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- [CVPR 2025] This repository is intended to store the code and data for ASAP (Advancing Semantic Alignment Promotes Multi-Modal Manipulati…☆21Jun 18, 2025Updated last year
- [ACL'26 Oral] AgentOCR is a token-efficient framework that compresses multi-turn agent history by rendering it into images and adopting R…☆39Mar 1, 2026Updated 4 months ago
- [ECCV 2026] VKnowU: Evaluating Visual Knowledge Understanding in Multimodal LLMs☆15Feb 3, 2026Updated 5 months ago
- ☆21Oct 10, 2020Updated 5 years ago
- [ECCV2024] AnatoMask: Enhancing Medical Image Segmentation with Reconstruction-guided Self-masking. Official Pytorch Implementation of An…☆36Sep 23, 2024Updated last year
- Codes and data for CIKM 2022 paper "RuDi: Explaining Behavior Sequence Models by Automatic Statistics Generation and Rule Distillation"☆12Aug 16, 2022Updated 3 years ago
- ☆44Jun 10, 2025Updated last year
- [CVPR 2024] Official repository of ST_GT☆10Sep 15, 2024Updated last year
- ☆23Nov 24, 2022Updated 3 years ago
- Virtual machines for every use case on DigitalOcean • AdGet dependable uptime with 99.99% SLA, simple security tools, and predictable monthly pricing with DigitalOcean's virtual machines, called Droplets.
- ICME2022 Special Session “Beyond Accuracy: Responsible, Responsive, and Robust Multimedia Retrieval ”☆12Jun 3, 2024Updated 2 years ago
- [Pattern Recognition'24] Pytorch implementation of Multiple-environment Self-adaptive Network for Aerial-view Geo-localization https://a…☆47Jul 6, 2026Updated 2 weeks ago
- Cornell Tech CS5670 Introduction to Computer Vision Projects Repo☆13Nov 22, 2022Updated 3 years ago
- 地图足迹故事,微信小程序☆10May 5, 2022Updated 4 years ago
- [SIGKDD 2024] Rethinking Fair Graph Neural Networks from Re-balancing☆10Jul 15, 2024Updated 2 years ago
- ☆22Nov 5, 2024Updated last year
- ☆15Jan 14, 2026Updated 6 months ago