Code for "Linear Mechanisms for Spatiotemporal Reasoning in Vision Language Models"
☆15Feb 16, 2026Updated 5 months ago
Alternatives and similar repositories for linear-mech-vlms
Users that are interested in linear-mech-vlms are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- ☆19Mar 5, 2024Updated 2 years ago
- Official Code for All-in-One Medical Image Re-Identification (CVPR2025)☆20Jan 11, 2026Updated 6 months ago
- ☆14Apr 10, 2025Updated last year
- Code for "CLIP Behaves like a Bag-of-Words Model Cross-modally but not Uni-modally"☆29Feb 27, 2026Updated 4 months ago
- This repository contains the code used for the experiments in the paper "Language Models use Lookbacks to Track Beliefs".☆16Mar 14, 2026Updated 4 months ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- The official code and model for ACL 2023 paper 'mCLIP: Multilingual CLIP via Cross-lingual Transfer'☆10Jan 23, 2024Updated 2 years ago
- up-to-date curated list of state-of-the-art Large vision language models hallucinations research work, papers & resources☆325Feb 8, 2026Updated 5 months ago
- 🚲 Code and benchmark for our COLM 2025 paper - "Thought Tracing: Hypothesis-Driven Theory-of-Mind Reasoning for Large Language Models"☆15Aug 8, 2025Updated 11 months ago
- [CVPR 2025] TAPT: Test-Time Adversarial Prompt Tuning for Robust Inference in Vision-Language Models☆15May 21, 2026Updated 2 months ago
- (NeurIPS 2025) Vision Foundation Models as Effective Visual Tokenizers for Autoregressive Image Generation☆77May 21, 2026Updated 2 months ago
- ☆86Nov 5, 2024Updated last year
- ☆10Nov 18, 2024Updated last year
- Bootstrapping Grounded Chain-of-Thought in Multimodal LLMs for Data-Efficient Model Adaptation☆15Aug 11, 2025Updated 11 months ago
- [CVPR 2025] Your Large Vision-Language Model Only Needs A Few Attention Heads For Visual Grounding☆17Oct 4, 2025Updated 9 months ago
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- ☆21Dec 26, 2025Updated 6 months ago
- Water ripples effect in javascript.☆35Jan 29, 2015Updated 11 years ago
- [NeurIPS 2025] Poison as Cure: Visual Noise for Mitigating Object Hallucinations in LVMs☆38Sep 21, 2025Updated 10 months ago
- [EMNLP 2025 Main] Official implementation of VRoPE: Rotary Position Embedding for Video Large Language Models.☆28Nov 18, 2025Updated 8 months ago
- Official PyTorch implementation for "Where You Edit is What You Get: Text-Guided Image Editing with Region-Based Attention" (Pattern Reco…☆10Oct 1, 2024Updated last year
- Official repository for Robust Multimodal Large Language Models Against Modality Conflict☆22Jul 9, 2025Updated last year
- [ICLR 2026] Official implementation of "ChatInject: Abusing Chat Templates for Prompt Injection in LLM Agents"☆17Mar 23, 2026Updated 4 months ago
- ANDROID APP that can RECOGNIZE VLC LIVE AUDIO/VIDEO STREAMING (using free Android Developers Speech Recognition API) then TRANSLATE (usin…☆14May 5, 2024Updated 2 years ago
- ☆128Jun 18, 2026Updated last month
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- [CVPR Findings 2026] "Circuit Tracing in Vision-Language Models"☆27Jul 14, 2026Updated last week
- ☆14Jan 22, 2025Updated last year
- EasyTTS是一个便捷的工具,旨在方便地使用第三方API服务来调用OpenAI的文本转语音(TTS)功能。 EasyTTS允许用户输入文本,并选择不同的模型、音色、格式来生成音频文件。☆10Nov 26, 2023Updated 2 years ago
- Text-guided 3D texture generation using training-free multi-diffusion in UV space.☆13Apr 7, 2025Updated last year
- [TIFS 2024] DF-RAP: A Robust Adversarial Perturbation for Defending against Deepfakes in Real-world Social Network Scenarios☆25Oct 29, 2025Updated 8 months ago
- ☆27Jan 5, 2026Updated 6 months ago
- ☆13Jun 13, 2024Updated 2 years ago
- ☆15Dec 11, 2024Updated last year
- [ICML 2026] Official code for paper: Test-Time Training with KV Binding Is Secretly Linear Attention☆36Apr 30, 2026Updated 2 months ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- Implementation of followinf estimation algorithms in python: Kalman Filter, Extended Kalman Filter, Unscented Kalman Filter, Cubature Kal…☆11Dec 2, 2023Updated 2 years ago
- A python implementation of PSNR that takes the Human visual system into account.☆13Updated this week
- Repository for "Training Language Models To Explain Their Own Computations"☆23Jul 7, 2026Updated 2 weeks ago
- [ICLR 2025] Official codebase for the ICLR 2025 paper "Multimodal Situational Safety"☆36Jun 23, 2025Updated last year
- [CVPR 2025] Mitigating Object Hallucinations in Large Vision-Language Models with Assembly of Global and Local Attention☆68Jul 16, 2024Updated 2 years ago
- SicTTA: Single Image Continual Test-Time Adaptation for Medical Image Segmentation☆18Dec 21, 2025Updated 7 months ago
- [BMVC 2025 🔥] CalibPrompt is the first framework that enhances Med-VLM calibration during prompt tuning.☆16Jul 13, 2026Updated last week