☆21May 19, 2025Updated last year
Alternatives and similar repositories for Awesome-VLM-Reasoning
Users that are interested in Awesome-VLM-Reasoning are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- [ACL '26 Findings] V-MAGE: A Game Evaluation Framework for Assessing Visual-Centric Capabilities in MLLMs☆27Apr 28, 2026Updated 4 months ago
- [ECCV 2026] Glance: Accelerating Diffusion Models with 1 Sample☆155Jul 24, 2026Updated last month
- Computer-Use Agents as Judges for Generative UI☆45Nov 27, 2025Updated 9 months ago
- Official implementation of Leveraging Visual Tokens for Extended Text Contexts in Multi-Modal Learning☆28Oct 30, 2024Updated last year
- [Arxiv2022] Revitalize Region Feature for Democratizing Video-Language Pre-training☆22Mar 19, 2022Updated 4 years ago
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- [Findings of EMNLP 2022] AssistSR: Task-oriented Video Segment Retrieval for Personal AI Assistant☆24Sep 11, 2023Updated 3 years ago
- ☆14Mar 20, 2023Updated 3 years ago
- 中南大学(CSU)labview实验和课程设计代码(通用虚拟滤波器),记得留下你的star☆11May 21, 2021Updated 5 years ago
- ☆23Sep 11, 2026Updated last week
- ☆12Mar 22, 2025Updated last year
- a set of scripts to easily convert all training data from huggingface into alpaca instruct or sharegpt format, which should allow for eas…☆20Mar 14, 2025Updated last year
- Unified layout planning and image generation, ICCV2025☆47Jan 19, 2026Updated 8 months ago
- [ICCV 2021] Multimodal Knowledge Expansion☆10Aug 28, 2021Updated 5 years ago
- 中南大学智能科学与技术专业机器学习课程设计,其中包含自己实现的神经网络框架,可实现的模型有:ResNet,VGG16☆10Jul 6, 2022Updated 4 years ago
- Deploy open-source AI quickly and easily - Special Bonus Offer • AdRunpod Hub is built for open source. One-click deployment and autoscaling endpoints without provisioning your own infrastructure.
- The Pytorch implementation for "Video-Text Pre-training with Learned Regions"☆43Jul 15, 2022Updated 4 years ago
- TPDiff: Temporal Pyramid Video Diffusion Model☆25Mar 13, 2025Updated last year
- SpringBoot实现留学信息管理与分析系统☆10Jun 14, 2023Updated 3 years ago
- Official Repo for paper "VLCache: Computing 2% Vision Tokens and Reusing 98% for Vision-Language Inference"☆15Mar 28, 2026Updated 5 months ago
- 常用的ocr数据集☆16Nov 15, 2021Updated 4 years ago
- DoraCycle: Domain-Oriented Adaptation of Unified Generative Model in Multimodal Cycles☆31Mar 8, 2026Updated 6 months ago
- [EMNLP 2025] DiagramEval: Evaluating LLM-Generated Diagrams via Graphs☆17Nov 1, 2025Updated 10 months ago
- [CVPR 2024] Neural Parametric Gaussians for Monocular Non-Rigid Object Reconstruction☆26Jan 22, 2025Updated last year
- ☆12Jun 12, 2024Updated 2 years ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- For replication of the experiments in the paper Learning Robust Representations by Projecting Superficial Statistics Out☆13Oct 22, 2019Updated 6 years ago
- [ Arxiv 2023 ] This repository contains the code for "MUPPET: Multi-Modal Few-Shot Temporal Action Detection"☆16Aug 30, 2023Updated 3 years ago
- Governance substrate for your AI coding agents — adversarial review, drift-detected rules, immutable audit, closed-loop telemetry☆18Updated this week
- ☆16Sep 25, 2025Updated 11 months ago
- ☆86May 2, 2026Updated 4 months ago
- IVC-Prune: Revealing the Implicit Visual Coordinates in LVLMs for Vision Token Pruning☆18Feb 27, 2026Updated 6 months ago
- ☆12Mar 12, 2023Updated 3 years ago
- ☆14Aug 5, 2026Updated last month
- 大三下Web课设 - 中南大学主页 - JavaWeb☆13Dec 6, 2019Updated 6 years ago
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- ☆20May 28, 2025Updated last year
- The official repo for "Stepping Stones: A Progressive Training Strategy for Audio-Visual Semantic Segmentation", ECCV 2024☆18Oct 11, 2024Updated last year
- Beyond Gradient Descent for Regularized Segmentation Losses☆11Sep 27, 2019Updated 6 years ago
- [ICLR'26] SinkTrack: Attention Sink based Context Anchoring for Large Language Models☆19Apr 23, 2026Updated 4 months ago
- Implements the loss used in A. Furnari, S. Battiato, G. M. Farinella (2018). Leveraging Uncertainty to Rethink Loss Functions and Evaluat…☆12May 22, 2019Updated 7 years ago
- Repository for HypoSpace☆15May 4, 2026Updated 4 months ago
- ☆13Sep 25, 2023Updated 2 years ago