Efficient Test-Time Scaling for Small Vision-Language Models, official implementation of the ICLR'26 paper, test-time scaling via test-time augmentation
☆28Nov 17, 2025Updated 10 months ago
Alternatives and similar repositories for efficient_test_time_scaling
Users that are interested in efficient_test_time_scaling are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- This repo contains the procedural generation pipeline used to generate CrashCar101☆16Jan 14, 2024Updated 2 years ago
- [WACV 2024] Learning the What and How of Annotation in Video Object Segmentation☆29Jun 8, 2024Updated 2 years ago
- Code and data for "Learning Program Representations for Food Images and Cooking Recipes" (oral at CVPR 2022)☆15Mar 30, 2022Updated 4 years ago
- Materialist: Physically Based Editing Using Single-Image Inverse Rendering☆29May 20, 2026Updated 4 months ago
- Joint learning of object and action detectors☆15Nov 5, 2019Updated 6 years ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- [ICML 2026] Official Implementation of Prism: Efficient Test-Time Scaling via Hierarchical Search and Self-Verification for Discrete Diff…☆23Mar 4, 2026Updated 6 months ago
- A minimal, hackable Vision-Language Model built on Karpathy’s nanochat — add image understanding and multimodal chat for under $200 in co…☆26Updated this week
- ☆10Oct 1, 2024Updated last year
- ☆14Mar 4, 2024Updated 2 years ago
- Fast Diffusion-Based Counterfactuals for Shortcut Removal and Generation (ECCV 2024 ORAL)☆17Sep 3, 2024Updated 2 years ago
- ☆34May 13, 2026Updated 4 months ago
- ☆11Jul 3, 2024Updated 2 years ago
- Implementation - Surface reconstruction using rotation systems☆16Jun 3, 2025Updated last year
- ☆11Sep 13, 2023Updated 3 years ago
- Open source password manager - Proton Pass • AdSecurely store, share, and autofill your credentials with Proton Pass, the end-to-end encrypted password manager trusted by millions.
- [ACL 2026] Official resources of "Prompt-R1: Collaborative Automatic Prompting Framework via End-to-end Reinforcement Learning "☆61Sep 3, 2026Updated 2 weeks ago
- Official Code for the ACCV 2022 paper Diffusion Models for Counterfactual Explanations☆30Mar 12, 2025Updated last year
- [ICLR 2024] Towards Elminating Hard Label Constraints in Gradient Inverision Attacks☆14Feb 6, 2024Updated 2 years ago
- [ECCV24] VISA: Reasoning Video Object Segmentation via Large Language Model☆22Jul 20, 2024Updated 2 years ago
- [NeurIPS 2025] Toward a Vision-Language Foundation Model for Medical Data: Multimodal Dataset and Benchmarks for Vietnamese PET/CT Report…☆15Jul 22, 2026Updated last month
- Code and data repository for "The Mirage of Model Editing: Revisiting Evaluation in the Wild"☆18Aug 27, 2025Updated last year
- ECCV 2022☆16Aug 3, 2022Updated 4 years ago
- [ACL'26] Official Repository for The Paper: What If Consensus Lies? Selective-Complementary Reinforcement Learning at Test Time☆21Apr 7, 2026Updated 5 months ago
- A PyTorch implementation of Speech Transformer with multi-GPUs, an End-to-End ASR with Transformer network on Mandarin Chinese. This code…☆10Dec 25, 2019Updated 6 years ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- ☆12Aug 5, 2022Updated 4 years ago
- ☆18Mar 26, 2026Updated 5 months ago
- Masked Autoencoder Pretraining on 3D Brain MRI☆18Jan 22, 2026Updated 7 months ago
- TESGNN: 3D Temporal Equivariant Scene Graph Neural Networks (published at TMLR)☆14Nov 2, 2025Updated 10 months ago
- Code for ThriftyDAgger☆15Dec 29, 2021Updated 4 years ago
- ☆18Oct 7, 2022Updated 3 years ago
- Fleming-VL: Towards Universal Medical Visual Understanding with Multimodal LLMs☆18Nov 6, 2025Updated 10 months ago
- Learnable Semi-structured Sparsity for Vision Transformers and Diffusion Transformers☆15Feb 7, 2025Updated last year
- Official code for ICML 2024 paper, "Connecting the Dots: Collaborative Fine-tuning for Black-Box Vision-Language Models"☆19Jun 12, 2024Updated 2 years ago
- GPUs on demand by Runpod - Special Offer Available • AdRun AI, ML, and HPC workloads on powerful cloud GPUs—without limits or wasted spend. Deploy GPUs in under a minute and pay by the second.
- Official implementation of ICML 2025 paper "Understanding Multimodal LLMs Under Distribution Shifts: An Information-Theoretic Approach"☆12May 27, 2025Updated last year
- Implementation of Language-Conditioned Path Planning (Amber Xie, Youngwoon Lee, Pieter Abbeel, Stephen James)☆27Sep 1, 2023Updated 3 years ago
- CVPR 2025 - R-TPT: Improving Adversarial Robustness of Vision-Language Models through Test-Time Prompt Tuning☆22Aug 28, 2025Updated last year
- This repository follows papers and reports on discrete speech representation learning and speech tokenization methods for speech language…☆15Dec 1, 2023Updated 2 years ago
- U-VLM: Hierarchical Vision Language Modeling for Report Generation☆20Apr 30, 2026Updated 4 months ago
- Fine-Grained Pixel-Text Alignment for Open-Vocabulary Semantic Segmentation☆16Mar 28, 2026Updated 5 months ago
- ☆15Jan 12, 2026Updated 8 months ago