☆49Feb 9, 2026Updated 7 months ago
Alternatives and similar repositories for POINTS-GUI
Users that are interested in POINTS-GUI are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- HyperEyes is a parallel multimodal search agent that fuses visual grounding and retrieval into a single atomic action, enabling concurren…☆76May 23, 2026Updated 3 months ago
- [ICLR 2025] This repo is the official implementation of "The Labyrinth of Links: Navigating the Associative Maze of Multi-modal LLMs".☆13Jan 25, 2025Updated last year
- A self-adaptive and class-balanced approach to improve deep neural network performance in the presence of noisy labels☆18Jul 2, 2024Updated 2 years ago
- code release☆44Jun 22, 2026Updated 2 months ago
- Official source code repository for paper BubbleRAG.☆17Jun 1, 2026Updated 3 months ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- NeuSyRE: A Neuro-Symbolic Visual Understanding and Reasoning Framework based on Scene Graph Enrichment☆25Mar 10, 2024Updated 2 years ago
- ☆17Aug 1, 2025Updated last year
- GroundCUA☆135Mar 24, 2026Updated 5 months ago
- ☆67Apr 16, 2026Updated 5 months ago
- [TCSVT] Regularity Learning via Explicit Distribution Modeling for Skeletal Video Anomaly Detection☆17Jul 22, 2023Updated 3 years ago
- [CVPR2026] VideoAuto-R1: Video Auto Reasoning via Thinking Once, Answering Twice☆89Feb 27, 2026Updated 6 months ago
- [ICLR 2025] Permute-and-Flip: An optimally robust and watermarkable decoder for LLMs☆19Mar 20, 2025Updated last year
- The official code of "Thinking With Videos: Multimodal Tool-Augmented Reinforcement Learning for Long Video Reasoning"☆103Oct 15, 2025Updated 11 months ago
- Official implementation of VLAA-GUI series☆37Jun 20, 2026Updated 3 months ago
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- Co-evolving policy actors and experience extractors for efficient experience-driven agent RL☆52May 12, 2026Updated 4 months ago
- [CVPR'25] MergeVQ: A Unified Framework for Visual Generation and Representation with Token Merging and Quantization☆51Jul 22, 2025Updated last year
- Official code for paper "OpenCIL: Benchmarking Out-of-Distribution Detection in Class-Incremental Learning"☆13Jun 19, 2024Updated 2 years ago
- The source code for "LaSe-E2V: Towards Language-guided Semantic-Aware Event-to-Video Reconstruction"☆10Jul 5, 2024Updated 2 years ago
- WideRange4D: Enabling High-Quality 4D Reconstruction with Wide-Range Movements and Scenes☆112Mar 19, 2025Updated last year
- The official code of "Towards Long-horizon Agentic Multimodal Search"☆30Apr 17, 2026Updated 5 months ago
- The official implementation of 《MLLMs-Augmented Visual-Language Representation Learning》☆31Mar 12, 2024Updated 2 years ago
- ☆20Jul 1, 2026Updated 2 months ago
- [EMNLP 2026] Official implementation for paper "Demystifying Reinforcement Learning for Long-Horizon Tool-Using Agents: A Comprehensive R…☆51Aug 23, 2026Updated 3 weeks ago
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- Official implementation of paper ReTaKe: Reducing Temporal and Knowledge Redundancy for Long Video Understanding☆40Mar 16, 2025Updated last year
- [ECCV 2026] Official repository of "Reliable Reasoning in SVG-LLMs via Multi-Task Multi-Reward Reinforcement Learning".☆26Jul 17, 2026Updated 2 months ago
- simple grpo☆13May 28, 2025Updated last year
- ☆27Apr 15, 2026Updated 5 months ago
- [ACL 2025 Main] Code and data for paper "Can LLM Watermarks Robustly Prevent Unauthorized Knowledge Distillation?"☆23Jun 18, 2025Updated last year
- The implementation for ACL 2026: MatchTIR: Fine-Grained Supervision for Tool-Integrated Reasoning via Bipartite Matching.☆20Jul 27, 2026Updated last month
- Official implementation of UI-Ins: Enhancing GUI Grounding with Multi-Perspective Instruction-as-Reasoning☆81Apr 20, 2026Updated 5 months ago
- The model, data and code for OpenMobile☆50Jul 9, 2026Updated 2 months ago
- Official implementation of EgoThinker at NIPS 2025☆29Nov 25, 2025Updated 9 months ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- [CVPR 2026] ShowUI-π: Flow-based Generative Models as GUI Dexterous Hands☆135Apr 22, 2026Updated 4 months ago
- A Simple Plugin for Transforming Images to Arbitrary Scales☆19Feb 9, 2023Updated 3 years ago
- ☆139May 8, 2025Updated last year
- PixelPrune: Pixel-Level Adaptive Visual Token Reduction via Predictive Coding☆32Jun 10, 2026Updated 3 months ago
- A question-conditioned, reasoning-aware image editor designed to serve as a decoupled visual reasoning assistant for Multimodal Large Lan…☆23May 25, 2026Updated 3 months ago
- TEMPURA enables video-language models to reason about causal event relationships and generate fine-grained, timestamped descriptions of u…☆28Jun 4, 2025Updated last year
- [NeurIPS 2024] The official implementation of "Image Copy Detection for Diffusion Models"☆18Oct 1, 2024Updated last year