[ICCV 2025] The official pytorch implement of "LLaVA-SP: Enhancing Visual Representation with Visual Spatial Tokens for MLLMs".
☆24Oct 28, 2025Updated 10 months ago
Alternatives and similar repositories for LLaVA-SP
Users that are interested in LLaVA-SP are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- [ACL 2025] The official pytorch implement of "MIND: A Multi-agent Framework for Zero-shot Harmful Meme Detection".☆28May 26, 2025Updated last year
- Official Pytorch implementation of the AAAI 2025 "Spiking Point Transformer for Point Cloud Classification"☆16Apr 12, 2025Updated last year
- [ICCV 2025] DASH: 4D Hash Encoding with Self-Supervised Decomposition for Real-Time Dynamic Scene Rendering☆28Apr 13, 2026Updated 4 months ago
- (从Google Code迁移过来的)这是我们的学生作业,一个简单的PL/0编译器,不过,编译只到中间代码为止,这是因为我们的汇编水平还没有到达能编写完整应用程序的水平。所以,仅仅是将PL/0代码编译成中间代码——一种很像汇编的四元式,然后请使用Interpreter解释运…☆10Mar 13, 2015Updated 11 years ago
- [COLING 2025🔥] Evolver: Chain-of-Evolution Prompting to Boost Large Multimodal Models for Hateful Meme Detection☆17Jan 21, 2025Updated last year
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- [ACM-MM 2025 Workshop] More Is Better: A MoE-Based Emotion Recognition Framework with Human Preference Alignment.☆26Nov 25, 2025Updated 9 months ago
- [ECCV2022] Motion Sensitive Contrastive Learning for Self-supervised Video Representation☆17Aug 12, 2022Updated 4 years ago
- [ICML 2025] This is the official PyTorch implementation of "OmniBal: Towards Fast Instruction-Tuning for Vision-Language Models via Omniv…☆27Jun 16, 2025Updated last year
- [ICCV 2025] HQ-CLIP: Leveraging Large Vision-Language Models to Create High-Quality Image-Text Datasets☆67Aug 6, 2025Updated last year
- An automated workflow for composing, rendering, and retargeting MMD assets.☆16Feb 23, 2026Updated 6 months ago
- [Neurocomputing] Efficient Redundancy Reduction for Open-Vocabulary Semantic Segmentation☆26Dec 21, 2025Updated 8 months ago
- [ICCV 2025] MPG-SAM 2: Adapting SAM 2 with Mask Priors and Global Context for Referring Video Object Segmentation☆23Sep 5, 2025Updated 11 months ago
- Official implementation of ICLR 2026: Let LLMs Speak Embedding Languages: Generative Text Embeddings via Iterative Contrastive Refinement☆15May 24, 2026Updated 3 months ago
- [ACM MM 2023 ]DFIL Codes☆25Jan 2, 2024Updated 2 years ago
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- ☆12Jun 19, 2024Updated 2 years ago
- Learning Cross-View Object Correspondence via Cycle-Consistent Mask Prediction (CVPR 2026)☆16Feb 27, 2026Updated 6 months ago
- 2022秋-同济大学软件学院-统计分析与建模期末项目☆12Jan 8, 2023Updated 3 years ago
- ☆17Jul 17, 2025Updated last year
- Geometric Problem Solving Integrating FormalGeo Symbolic System and Hypergraph Neural Network.☆16Sep 23, 2025Updated 11 months ago
- [NeurIPS 2025] Reasoning MLLM, Share-GRPO, advantage vanishing, sparse reward☆38Sep 19, 2025Updated 11 months ago
- LLaVE: Large Language and Vision Embedding Models with Hardness-Weighted Contrastive Learning☆78May 23, 2025Updated last year
- [NeurIPS 2025] More Thinking, Less Seeing? Assessing Amplified Hallucination in Multimodal Reasoning Models☆82May 31, 2025Updated last year
- [ECCV 2024] Official PyTorch implementation of LUT "Learning with Unmasked Tokens Drives Stronger Vision Learners"☆15Dec 1, 2024Updated last year
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- official repository of article "CrystaL: Spontaneous Emergence of Visual Latents in MLLMs"☆18May 26, 2026Updated 3 months ago
- 这是一个Python实现的PL/0 语言的编译器,实现了词法分析、LR1语法分析、语义分析、中间代码生成和虚拟机执行目标代码的功能。☆17Oct 2, 2023Updated 2 years ago
- 图像检索一些好的开源代码☆14Sep 3, 2020Updated 5 years ago
- ☆14Jan 4, 2025Updated last year
- Official implementation of "VIRAL: Visual Representation Alignment for MLLMs".☆166Sep 21, 2025Updated 11 months ago
- An ns3 simulation script to compare various TCP variants under congestion☆14Apr 25, 2017Updated 9 years ago
- Code for NAACL 2025 paper "AdaCAD: Adaptively Decoding to Balance Conflicts between Contextual and Parametric Knowledge"☆17Mar 2, 2026Updated 5 months ago
- Official implementation of Bayes Conditional Distribution Estimation for Knowledge Distillation Based on Conditional Mutual Information☆12Sep 28, 2023Updated 2 years ago
- A marker-based augmented reality camera app for the web, powered by AR.js.☆16Dec 4, 2022Updated 3 years ago
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- The official implementation of Hard Negative Sampling via Large Language Models for Recommendation.☆11Jan 17, 2026Updated 7 months ago
- [IEEE TIP] Offical implementation for the work "BadCM: Invisible Backdoor Attack against Cross-Modal Learning".☆14Aug 30, 2024Updated 2 years ago
- [EMNLP 2024] SURf: Teaching Large Vision-Language Models to Selectively Utilize Retrieved Information☆11Oct 11, 2024Updated last year
- The code implementation for UME-R1: Exploring Reasoning-Driven Generative Multimodal Embeddings (ICLR 2026).☆71Feb 25, 2026Updated 6 months ago
- ☆13Jul 19, 2023Updated 3 years ago
- An autohotkey's script that makes your capslock more powerful. Latest version (used by myself): https://github.com/Liu233w/keyboard.ahk☆15Aug 3, 2018Updated 8 years ago
- [CVPR 2023] Better “CMOS” Produces Clearer Images: Learning Space-Variant Blur Estimation for Blind Image Super-Resolution☆11Mar 19, 2024Updated 2 years ago