A Practical Zoom-in GUI Grounding and Behavior-Based Evaluation method.
☆28Aug 27, 2026Updated last week
Alternatives and similar repositories for ZoomClick
Users that are interested in ZoomClick are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- This is the official evaluation code for Robobench☆24Updated this week
- Multi-View prediction enhances GUI Grounding☆22Feb 22, 2026Updated 6 months ago
- ☆16Aug 19, 2026Updated 2 weeks ago
- Official code for "TraceGen: World Modeling in 3D Trace-Space Enables Learning from Cross-Embodiment Videos" (CVPR 2026)☆19Jan 31, 2026Updated 7 months ago
- [EMNLP 2025]Repository for paper "DiMo-GUI: Advancing Test-time Scaling in GUI Grounding via Modality-Aware Visual Reasoning"☆30Jul 2, 2025Updated last year
- Bare Metal GPUs on DigitalOcean Gradient AI • AdPurpose-built for serious AI teams training foundational models, running large-scale inference, and pushing the boundaries of what's possible.
- (ECCV'24) Official Implementation of SCP-Diff: Photo-Realistic Semantic Image Synthesis with Spatial-Categorical Joint Prior.☆15Oct 2, 2024Updated last year
- A simple visual test-time scaling method for GUI agent grounding☆26Dec 7, 2025Updated 8 months ago
- A benchmark for evaluating contextual agents on realistic multimodal personal-computer environments with profiling and factual-retention …☆31Apr 2, 2026Updated 5 months ago
- ☆19Sep 4, 2025Updated last year
- ☆17Apr 15, 2026Updated 4 months ago
- [NeurIPS2025]VideoVLA: Video Generators Can Be Generalizable Robot Manipulators☆32Jun 26, 2026Updated 2 months ago
- The homework of robos learning base.☆11May 23, 2023Updated 3 years ago
- ROUTE: Robust Multitask Tuning and Collaboration for Text-to-SQL (ICLR 2025 Pytorch Code)☆16May 15, 2025Updated last year
- A platform for reinforcement learning in Terraria☆12Nov 20, 2019Updated 6 years ago
- Simple, predictable pricing with DigitalOcean hosting • AdAlways know what you'll pay with monthly caps and flat pricing. Enterprise-grade infrastructure trusted by 600k+ customers.
- Official implementation of "Next-Scale Autoregressive Models are Zero-Shot Single-Image Object View Synthesizers"☆45Mar 19, 2025Updated last year
- Related code, checkpoints and project page for V-Reflection☆61Apr 7, 2026Updated 4 months ago
- 【ICME2025 Oral】Offical Pytorch Code for "Fraesormer: Learning Adaptive Sparse Transformer for Efficient Food Recognition"☆13Mar 21, 2025Updated last year
- DUET-VLM: Dual stage Unified Efficient Token reduction for VLM Training and Inference☆25May 21, 2026Updated 3 months ago
- A complete introductory course to programming, computer systems and software development (continuously updating).☆12Feb 21, 2024Updated 2 years ago
- [MICCAI 24] The official code repository for paper "FairDiff: Fair Segmentation with Point-Image Diffusion".☆60Mar 12, 2025Updated last year
- [EMNLP-2025 Oral] ZoomEye: Enhancing Multimodal LLMs with Human-Like Zooming Capabilities through Tree-Based Image Exploration☆91Nov 20, 2025Updated 9 months ago
- [IJCAI-24] Explore Internal and External Similarity for Single Image Deraining with Graph Neural Networks☆11Sep 2, 2024Updated 2 years ago
- Coco is a proactive co-assistant that connects user workspace with a broader ecosystem of AI agents.☆33Updated this week
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- ☆15Jun 19, 2024Updated 2 years ago
- LaTeX Drawing☆18Dec 22, 2025Updated 8 months ago
- Using machine learning techniques for prediction and modelling non linear dynamic systems.☆10Jun 29, 2018Updated 8 years ago
- ☆18Apr 24, 2024Updated 2 years ago
- CVPR25☆28Jul 2, 2025Updated last year
- This project uses deep regression and generative models to reconstruct music from brain responses as participants passively listen to ful…☆18Apr 7, 2025Updated last year
- Multiple Attractors simulation with customization☆14Feb 22, 2026Updated 6 months ago
- [RSS 2026] Automated Synthesis of Facial Mechanisms for Conversational Animatronic Robots