(ArXiv25) Vision Matters: Simple Visual Perturbations Can Boost Multimodal Math Reasoning
☆61Sep 30, 2025Updated 11 months ago
Alternatives and similar repositories for Vision-Matters
Users that are interested in Vision-Matters are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Official Code for Paper: Beyond Matryoshka: Revisiting Sparse Coding for Adaptive Representation☆143May 6, 2026Updated 4 months ago
- (ICLR 2026) Pytorch implementation of "IDER: Idempotent Experience Replay for Reliable Continual Learning"☆35Mar 30, 2026Updated 5 months ago
- PyTorch Implementation of ECCV 2024 OOD-CV Workshop SSB Challenge (Open-Set Recognition Track) - 1st Place☆29Sep 13, 2024Updated 2 years ago
- [ACL'25] UTBoost: Rigorous Evaluation of Coding Agents on SWE-Bench☆36Aug 12, 2025Updated last year
- DeepDubber-V1: Towards High Quality and Dialogue, Narration, Monologue Adaptive Movie Dubbing Via Multi-Modal Chain-of-Thoughts Reasoning…☆30Sep 7, 2025Updated last year
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- ☆20Nov 27, 2025Updated 9 months ago
- [ICCV 2025] Long-term Traffic Simulation with Interleaved Autoregressive Motion and Scenario Generation.☆54Aug 27, 2025Updated last year
- This repository contains the code for our ICML 2025 paper——LENSLLM: Unveiling Fine-Tuning Dynamics for LLM Selection🎉☆27May 29, 2025Updated last year
- [NeurIPS 2025] First SFT, Second RL, Third UPT: Continual Improving Multi-Modal LLM Reasoning via Unsupervised Post-Training☆89Oct 29, 2025Updated 10 months ago
- ☆18May 14, 2025Updated last year
- [ICLR 2025] Official Implementation of Local-Prompt: Extensible Local Prompts for Few-Shot Out-of-Distribution Detection☆52Jul 30, 2025Updated last year
- [ICLR 2025] The offical implementation of "PSEC: Skill Expansion and Composition in Parameter Space", a new framework designed to facilit…☆66Feb 12, 2025Updated last year
- “SURE: SUrvey REcipes for building reliable and robust deep networks” (CVPR 2024) & (ECCV 2024 OOD-CV Challenge Winner)☆76Aug 21, 2025Updated last year
- SFT+RL boosts multimodal reasoning☆50Jun 27, 2025Updated last year
- Open source password manager - Proton Pass • AdSecurely store, share, and autofill your credentials with Proton Pass, the end-to-end encrypted password manager trusted by millions.
- The official repository of UniHGKR: Unified Instruction-aware Heterogeneous Knowledge Retrievers☆27Jun 12, 2025Updated last year
- [CVPR 2026] ReasonMap: Towards Fine-Grained Visual Reasoning from Transit Maps☆87Jul 26, 2026Updated last month
- 📖 This is a repository for organizing papers, codes, and other resources related to personalized video generation and editing.☆65Dec 9, 2025Updated 9 months ago
- Official code for DeepSound-V1☆12May 14, 2025Updated last year
- [ICML 2025] This is the official PyTorch implementation of "🎵 HarmoniCa: Harmonizing Training and Inference for Better Feature Caching i…☆46Jul 10, 2025Updated last year
- [ICML 2026] ZwZ model family: SOTA fine-grained perception performace; ZoomBench: a new challenging perception benchmark☆189May 4, 2026Updated 4 months ago
- (NIPS 2025) OpenOmni: Official implementation of Advancing Open-Source Omnimodal Large Language Models with Progressive Multimodal Align…☆143May 9, 2026Updated 4 months ago
- (ICML 2024) PyTorch implementation of "Self-Attention through Kernel-Eigen Pair Sparse Variational Gaussian Processes"☆16Oct 15, 2024Updated last year
- [NeurIPS 2024] A Novel Rank-Based Metric for Evaluating Large Language Models☆58May 28, 2025Updated last year
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- [NeurIPS'25] Official Repository for the Paper "SQL-R1: Training Natural Language to SQL Reasoning Model By Reinforcement Learning"☆149Nov 20, 2025Updated 10 months ago
- [PVLDB 2025] TAB: Unified Benchmarking of Time Series Anomaly Detection Methods☆142Nov 26, 2025Updated 9 months ago
- [IEEE TII 2025] Official Implementation for "Dual-Detector Reoptimization for Federated Weakly Supervised Video Anomaly Detection via Ada…☆27Nov 11, 2025Updated 10 months ago
- 🦾 A Dual-System VLA with System2 Thinking☆149Aug 21, 2025Updated last year
- [ACL'25 Main] Official Implementation of HiDe-LLaVA: Hierarchical Decoupling for Continual Instruction Tuning of Multimodal Large Languag…☆56Jun 1, 2026Updated 3 months ago
- Official code for "Flatten Graphs as Sequences: Transformers are scalable graph generators" (NeurIPS 2025)☆18Oct 17, 2025Updated 11 months ago
- 🏆 Official implementation of LangCoop: Collaborative Driving with Natural Language☆81Sep 12, 2025Updated last year
- [IJCV 2025] Smaller But Better: Unifying Layout Generation with Smaller Large Language Models☆165Aug 3, 2025Updated last year
- Code repo for "Harnessing Negative Signals: Reinforcement Distillation from Teacher Data for LLM Reasoning"☆34Jul 25, 2025Updated last year
- GPUs on demand by Runpod - Special Offer Available • AdRun AI, ML, and HPC workloads on powerful cloud GPUs—without limits or wasted spend. Deploy GPUs in under a minute and pay by the second.
- [TMLR 2025] Efficient Reasoning Models: A Survey☆320Jun 26, 2026Updated 2 months ago
- [NeurIPS'25] Backdoor Cleaning without External Guidance in MLLM Fine-tuning☆20Oct 13, 2025Updated 11 months ago
- [IEEE TPAMI 2025] Privacy-Preserving Biometric Verification With Handwritten Random Digit String☆71May 18, 2026Updated 4 months ago
- (ICCV 2025) Enhance CLIP and MLLM's fine-grained visual representations with generative models.☆78Jun 25, 2025Updated last year
- ECCV 2026 accepted CACFM code: RL-guided curvature-adaptive consistency flow matching for few-step FLUX/SDXL generation☆55Jun 25, 2026Updated 2 months ago
- ZO2 (Zeroth-Order Offloading): Full Parameter Fine-Tuning 175B LLMs with 18GB GPU Memory [COLM2025]☆207Jul 16, 2025Updated last year
- [CVPR 2026] A Closer Look at Cross-Domain Few-Shot Object Detection: Fine-Tuning Matters and Parallel Decoder Helps☆56Aug 10, 2026Updated last month