☆15Apr 15, 2026Updated 3 months ago
Alternatives and similar repositories for Q-Zoom
Users that are interested in Q-Zoom are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- [ICLR2026] Catching the Details: Self-Distilled RoI Predictors for Fine-Grained MLLM Perception☆17Jan 26, 2026Updated 6 months ago
- [AAAI 2025] Efficient Image-to-Image Diffusion Classifier for Adversarial Robustness☆20Aug 21, 2024Updated last year
- [ICLR 2026] VEAttack: Downstream-agnostic Vision Encoder Attack against Large Vision Language Models☆21Feb 22, 2026Updated 5 months ago
- [NeurIPS 2025] Official implementation for "Efficient Rectified Flow for Image Fusion".☆27Apr 8, 2026Updated 3 months ago
- [NeurIPS2024] Multi-Scale VMamba: Hierarchy in Hierarchy Visual State Space Model☆84Dec 25, 2024Updated last year
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- DUET-VLM: Dual stage Unified Efficient Token reduction for VLM Training and Inference☆25May 21, 2026Updated 2 months ago
- The offical code of PolarBEV (CoRL2022).☆57Sep 17, 2022Updated 3 years ago
- Offical implementation for "Trash or Treasure? An Interactive Dual-Stream Strategy for Single Image Reflection Separation".☆63Aug 22, 2023Updated 2 years ago
- [ICML 2025] Speak Easy: Eliciting Harmful Jailbreaks from LLMs with Simple Interactions☆16Mar 7, 2026Updated 4 months ago
- [RA-L with ICRA2023] TransDSSL: Transformer based Depth Estimation via Self-Supervised Learning☆12Jan 11, 2023Updated 3 years ago
- ☆35Apr 16, 2026Updated 3 months ago
- [CVPR2026] Official codebase for the paper "Reasoning Within the Mind: Dynamic Multimodal Interleaving in Latent Space"☆85May 12, 2026Updated 2 months ago
- Squeeze verbose LLM agent tool output down to only the relevant lines☆22Apr 27, 2026Updated 3 months ago
- ☆15Jul 18, 2026Updated 2 weeks ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- Official implementation for "Text-Aware Real-World Image Super-Resolution via Diffusion Model with Joint Segmentation Decoders"☆21May 29, 2025Updated last year
- From Local Matches to Global Masks: Novel Instance Detection in Open-World Scenes☆25Jul 28, 2026Updated last week
- Automated neural architecture search algorithms implemented in PyTorch and Autogluon toolkit.☆12Apr 17, 2020Updated 6 years ago
- IVC-Prune: Revealing the Implicit Visual Coordinates in LVLMs for Vision Token Pruning☆16Feb 27, 2026Updated 5 months ago
- ☆14Sep 22, 2025Updated 10 months ago
- ☆21Jul 30, 2026Updated last week
- Official Pytorch Code for "Rethinking Degradation: Radiograph Super-Resolution via AID-SRGAN" - MICCAI 2022 Workshop☆16Dec 11, 2024Updated last year
- [WACV 2026] ZonUI-3B — A lightweight, resolution-aware GUI grounding model trained with only 24K samples on a single RTX 4090.☆26Jan 2, 2026Updated 7 months ago
- ☆11May 18, 2022Updated 4 years ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- ICRA2026: ABPolicy Asynchronous B-Spline Flow Policy for Real-Time and Smooth Robotic Manipulation☆28Apr 22, 2026Updated 3 months ago
- ERGO (Efficient Reasoning & Guided Observation) is a large vision-language model trained with reinforcement learning on efficiency object…☆19Feb 25, 2026Updated 5 months ago
- Reversible Decoupling Network for Single Image Reflection Removal, To be appeared in CVPR 2025☆74Oct 26, 2025Updated 9 months ago
- AuthFace: Towards Authentic Blind Face Restoration with Face-oriented Generative Diffusion Prior (ACM MM 2025 Oral)☆19Mar 5, 2026Updated 5 months ago
- CVPR25☆28Jul 2, 2025Updated last year
- A simple visual test-time scaling method for GUI agent grounding☆26Dec 7, 2025Updated 7 months ago
- ☆24Nov 22, 2023Updated 2 years ago
- [ICLR 2025 Spotlight] Overcoming False Illusions in Real-World Face Restoration with Multi-Modal Guided Diffusion Model☆16Apr 23, 2025Updated last year
- A Practical Zoom-in GUI Grounding and Behavior-Based Evaluation method.☆26Dec 8, 2025Updated 7 months ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- [ECCV 26'] Official codebase for the paper LaViT☆35Jul 30, 2026Updated last week
- Official codes of "Sketch-in-Latents: Eliciting Unified Reasoning in MLLMs"☆17Feb 15, 2026Updated 5 months ago
- ☆20Jul 29, 2025Updated last year
- Official repo for "TC-AE: Unlocking Token Capacity for Deep Compression Autoencoders"☆25Apr 9, 2026Updated 3 months ago
- VTC-R1: Vision-Text Compression for Efficient Long-Context Reasoning.☆26Jul 20, 2026Updated 2 weeks ago
- An arbitrage bot is a smart contract connected to an external automation script that controls its operation.☆2,288Updated this week
- Crawl4DeepSeek = Crawl4AI + DeepSeek 🚀 Smart, efficient, and built for deep web exploration! 🌐🤖☆18Feb 9, 2025Updated last year