[CVPR 2025] Your Large Vision-Language Model Only Needs A Few Attention Heads For Visual Grounding
☆17Oct 4, 2025Updated 11 months ago
Alternatives and similar repositories for LocalizationHeads
Users that are interested in LocalizationHeads are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- [ICLR 2026] Official implementation of "ChatInject: Abusing Chat Templates for Prompt Injection in LLM Agents"☆24Mar 23, 2026Updated 5 months ago
- ☆18Aug 1, 2024Updated 2 years ago
- HarmAug: Effective Data Augmentation for Knowledge Distillation of Safety Guard Models☆16Mar 6, 2025Updated last year
- Official Implementation of "Chrono: A Simple Blueprint for Representing Time in MLLMs"☆96Mar 9, 2025Updated last year
- Code for "Linear Mechanisms for Spatiotemporal Reasoning in Vision Language Models"☆20Feb 16, 2026Updated 7 months ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- [ICCV2025] ExploreGS: Explorable 3D Scene Reconstruction with Virtual Camera Samplings and Diffusion Priors☆60May 17, 2026Updated 4 months ago
- [CVPR 2025] TAPT: Test-Time Adversarial Prompt Tuning for Robust Inference in Vision-Language Models☆16May 21, 2026Updated 3 months ago
- [WACV 2026] MomentMix Augmentation with Length-Aware DETR for Temporally Robust Moment Retrieval☆15Sep 18, 2025Updated last year
- ☆15Apr 25, 2025Updated last year
- S-Agent: Spatial Tool-Use Elicits Reasoning for Spatial Intelligence☆92Jul 22, 2026Updated last month
- Memory footprint reduction for transformer models☆11Jan 24, 2023Updated 3 years ago
- [ICLR 2025] DGQ: Distribution-Aware Group Quantization for Text-to-Image Diffusion Models☆20Mar 25, 2025Updated last year
- [CVPR 2026 Oral, Best Paper Finalist] SeaCache: Spectral-Evolution-Aware Cache for Accelerating Diffusion Models☆103Jun 29, 2026Updated 2 months ago
- ☆19Dec 6, 2023Updated 2 years ago
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- ☆11Oct 13, 2024Updated last year
- ☆21Mar 27, 2023Updated 3 years ago
- [ICCV'25] The official code of paper "Combining Similarity and Importance for Video Token Reduction on Large Visual Language Models"☆78Jan 13, 2026Updated 8 months ago
- Official Codebase of "Localizing Visual Sounds the Easy Way" (ECCV 2022)☆43Oct 2, 2022Updated 3 years ago
- Java web application backed by the Ethereum-Blockchain network. Powered by RESTful web services (JAX-RS && Spring Boot) , Docker, Kuberne…☆15Feb 19, 2019Updated 7 years ago
- The Yahoo Finance Agent is an application that combines OpenAI's LLMs, the Yahoo Finance Python library, and LangChain's tools to provide…☆29Aug 10, 2024Updated 2 years ago
- Official code for Attention-driven GUI Grounding, AAAI2025☆16Dec 17, 2024Updated last year
- ☆12Aug 7, 2024Updated 2 years ago
- [NeurIPS 2023] OV-PARTS: Towards Open-Vocabulary Part Segmentation☆96Jun 24, 2024Updated 2 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- For the rlhf learning environment of Koreans☆25Sep 25, 2023Updated 2 years ago
- [TIFS 2024] DF-RAP: A Robust Adversarial Perturbation for Defending against Deepfakes in Real-world Social Network Scenarios☆27Oct 29, 2025Updated 10 months ago
- Official Pytorch Implementation of Unsupervised Image Denoising With Frequency Domain Knowledge (BMVC2021 Oral Accepted Paper)☆24Mar 15, 2022Updated 4 years ago
- pdfChain: (experimental) blockchain for the masses☆16Feb 14, 2026Updated 7 months ago
- Official repository for CVPR 2023 paper: WSSS via Adversarial Learning of Classifier and Reconstructor☆29Jul 8, 2024Updated 2 years ago
- CVPR2022:Learning from Untrimmed Videos: Self-Supervised Video Representation Learning with Hierarchical Consistency☆18Aug 10, 2022Updated 4 years ago
- Learning to Enhance Aperture Phasor Field for Non-Line-of-Sight Imaging☆16Dec 31, 2024Updated last year
- ATTFormer for video retrieval system☆18Apr 23, 2026Updated 4 months ago
- Code repo for "CritiPrefill: A Segment-wise Criticality-based Approach for Prefilling Acceleration in LLMs".☆17Sep 15, 2024Updated 2 years ago
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- ClawSweeper Website☆15Jul 27, 2026Updated last month
- The source code of ExFunTube☆10Aug 8, 2025Updated last year
- [CVPR 2025] Official Pytorch Code for Distilling Spectral Graph for Object-Context Aware Open-Vocabulary Semantic Segmentation☆50Mar 27, 2025Updated last year
- Unsupervised Activity Segmentation by Joint Representation Learning and Online Clustering (CVPR 2022)☆12Sep 22, 2023Updated 2 years ago
- [CVPR2026] This is the official pytorch implementation of "Looking Beyond the Window: Global-Local Aligned CLIP for Training-free Open-Vo…☆25Jul 28, 2026Updated last month
- ☆13Apr 13, 2026Updated 5 months ago
- Official repository of "Event-guided Deblurring of Unknown Exposure Time Videos" ECCV 2022 paper(Oral).☆40Nov 29, 2022Updated 3 years ago