CVPR25
☆28Jul 2, 2025Updated last year
Alternatives and similar repositories for MP-GUI
Users that are interested in MP-GUI are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- ☆17May 14, 2024Updated 2 years ago
- VisionDroid☆23Apr 2, 2024Updated 2 years ago
- ☆45Dec 8, 2025Updated 8 months ago
- PixelPrune: Pixel-Level Adaptive Visual Token Reduction via Predictive Coding☆30Jun 10, 2026Updated 2 months ago
- Code repo for "Read Anywhere Pointed: Layout-aware GUI Screen Reading with Tree-of-Lens Grounding"☆31May 12, 2026Updated 3 months ago
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- ☆36Apr 16, 2026Updated 4 months ago
- iLLaVA: An Image is Worth Fewer Than 1/3 Input Tokens in Large Multimodal Models (ICLR2026)☆23Jun 24, 2026Updated 2 months ago
- ☆24Jul 8, 2023Updated 3 years ago
- 吴恩达大模型系列课程中文版,包括《Prompt Engineering》、《Building System》和《LangChain》☆12Jun 7, 2023Updated 3 years ago
- ☆34Sep 19, 2025Updated 11 months ago
- ☆16Apr 15, 2026Updated 4 months ago
- ☆19Sep 4, 2025Updated 11 months ago
- [ICCV 2025] GUIOdyssey is a comprehensive dataset for training and evaluating cross-app navigation agents. GUIOdyssey consists of 8,834 e…☆160Jan 3, 2026Updated 7 months ago
- Official PyTorch implementation of Fast-MoCo☆16Feb 20, 2023Updated 3 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- ☆13Jul 30, 2026Updated last month
- All-in-one repository for Fine-tuning & Pretraining (Large) Language Models☆15Mar 8, 2023Updated 3 years ago
- Game UI Glitch Detection via Bug Understanding☆12Jul 31, 2021Updated 5 years ago
- Mobile App Analysis and Testing Literature☆109Aug 21, 2026Updated last week
- ScreenQA dataset was introduced in the "ScreenQA: Large-Scale Question-Answer Pairs over Mobile App Screenshots" paper. It contains ~86K …☆151Feb 7, 2025Updated last year
- Zoom-Refine: Boosting High-Resolution Multimodal Understanding via Localized Zoom and Self-Refinement☆20Jul 4, 2026Updated last month
- Efficient Feature Extraction for High-resolution Video Frame Interpolation (BMVC 2022)☆14Aug 24, 2023Updated 3 years ago
- 🎧 Real-time data streaming from NeuroSky MindWave Mobile Headset☆10Jul 17, 2020Updated 6 years ago
- ☆71Feb 27, 2026Updated 6 months ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- This is the official repository for "Can GPTs Evaluate Graphic Design Based on Design Principles?".☆13Feb 10, 2025Updated last year
- ☆12Aug 24, 2023Updated 3 years ago
- [TIP 2025] Advancing Zero-Shot Digital Human Quality Assessment through Text-Prompted Evaluation☆12Jul 8, 2023Updated 3 years ago
- Combinator Library for writing test generators and test properties for Android Apps☆12Jul 26, 2019Updated 7 years ago
- [ICLR'25 Oral] UGround: Universal GUI Visual Grounding for GUI Agents☆317Aug 24, 2026Updated last week
- ☆47Nov 8, 2024Updated last year
- ☆10Nov 9, 2023Updated 2 years ago
- ☆10Dec 3, 2024Updated last year
- [ICML 2026] Stable Asynchrony: Variance-Controlled Off-Policy RL for LLMs☆34Apr 27, 2026Updated 4 months ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- ☆16Aug 21, 2026Updated last week
- [ECCV 2024] FlexAttention for Efficient High-Resolution Vision-Language Models☆50Jan 8, 2025Updated last year
- [WACV 2026] ZonUI-3B — A lightweight, resolution-aware GUI grounding model trained with only 24K samples on a single RTX 4090.☆26Jan 2, 2026Updated 7 months ago
- [CVPR 2025 Oral] VideoEspresso: A Large-Scale Chain-of-Thought Dataset for Fine-Grained Video Reasoning via Core Frame Selection☆142Jul 28, 2025Updated last year
- 字体识别☆12Apr 9, 2018Updated 8 years ago
- Emotiv SDK Community Edition☆13Oct 9, 2015Updated 10 years ago
- Human-centric environment representations from egocentric video☆15Feb 5, 2026Updated 6 months ago