The official implementation of the ECCV'24 paper MC-CoT: Boosting the Power of Small Multimodal Reasoning Models to Match Larger Models with Self-Consistency Training.
☆26May 19, 2024Updated 2 years ago
Alternatives and similar repositories for mc-cot
Users that are interested in mc-cot are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- ☆11Sep 27, 2023Updated 2 years ago
- 📐 [CVPR 2026] GGBench: A Geometric Generative Reasoning Benchmark for Unified Multimodal Models☆18Apr 1, 2026Updated 4 months ago
- the repository of A survey on image-text multimodal models☆46Apr 20, 2024Updated 2 years ago
- Envision: Benchmarking Unified Understanding & Generation for Causal World Process Insights☆32Jan 9, 2026Updated 7 months ago
- About Official PyTorch(MMCV) implementation of “SUMix: Mixup with Semantic and Uncertain Information” (ECCV 2024)☆12Sep 2, 2024Updated last year
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- Official implementation of the paper "DiffSDS: A Language Diffusion Model for Protein Backbone Inpainting under Geometric Conditions and …☆14Jun 25, 2023Updated 3 years ago
- [NeurIPS 2023]DDCoT: Duty-Distinct Chain-of-Thought Prompting for Multimodal Reasoning in Language Models☆48Mar 18, 2024Updated 2 years ago
- Implementation of RankE: End-to-End Discrete Text-to-Image Post-Training via Rank-Consistent Alignment☆22May 27, 2026Updated 2 months ago
- [ICLR 2026] MergeMix: A Unified Augmentation Paradigm for Visual and Multi-Modal Understanding☆22Feb 27, 2026Updated 5 months ago
- ☆10Oct 1, 2020Updated 5 years ago
- ☆13Dec 9, 2024Updated last year
- VQ-GAN for Various Data Modality based on Taming Transformers for High-Resolution Image Synthesis☆29Apr 15, 2023Updated 3 years ago
- [ICML 2023] Architecture-Agnostic Masked Image Modeling -- From ViT back to CNN☆32Aug 15, 2024Updated last year
- The official implementation of the ACM MM'21 paper Co-learning: Learning from noisy labels with self-supervision.☆123May 17, 2022Updated 4 years ago
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- Official implementation of "A Multi-In-Single-Out Network for Video Frame Interpolation without Optical Flow"☆21Dec 5, 2023Updated 2 years ago
- Code for our EMNLP-2022 paper: "Towards Robust Visual Question Answering: Making the Most of Biased Samples via Contrastive Learning"☆16Feb 22, 2023Updated 3 years ago
- ☆16Feb 27, 2025Updated last year
- [CVPR'25] MergeVQ: A Unified Framework for Visual Generation and Representation with Token Merging and Quantization☆51Jul 22, 2025Updated last year
- ABC: Achieving Better Control of Multimodal Embeddings using VLMs [TMLR2025]☆20Aug 21, 2025Updated 11 months ago
- [ICLR 2025] Official Pytorch Implementation of "Mix-LN: Unleashing the Power of Deeper Layers by Combining Pre-LN and Post-LN" by Pengxia…☆30Jul 24, 2025Updated last year
- LVOmniBench: Pioneering Long Audio-Video Understanding Evaluation for Omnimodal LLMs☆42Apr 2, 2026Updated 4 months ago
- density growing clustering☆10Dec 2, 2021Updated 4 years ago
- [ECCV 2024] Probabilistic Weather Forecasting with Deterministic Guidance-based Diffusion Model☆35Nov 6, 2024Updated last year
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- Data and code for NeurIPS 2022 Paper "Learn to Explain: Multimodal Reasoning via Thought Chains for Science Question Answering".☆737Sep 19, 2024Updated last year
- ☆37Oct 30, 2020Updated 5 years ago
- Benchmark and analysis of 165 pretrained SSL models. Code for "Evaluating Self-Supervised Learning via Risk Decomposition".☆14Jul 26, 2023Updated 3 years ago
- ☆14May 25, 2022Updated 4 years ago
- [CVPR'25] Attention IoU: Examining Biases in CelebA using Attention Maps☆13Mar 26, 2025Updated last year
- ☆10Jun 14, 2025Updated last year
- The official implementation of the paper "MotifRetro: Exploring the Combinability-Consistency Trade-offs in retrosynthesis via Dynamic Mo…☆11Jun 25, 2023Updated 3 years ago
- 本仓库包含:时空数据处理、预测领域的相关论文;相关数据集;专家学者信息☆15Apr 20, 2021Updated 5 years ago
- 这是我学习3d视觉中的一些笔记☆13Mar 1, 2021Updated 5 years ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- ☆12Apr 6, 2024Updated 2 years ago
- ☆36Mar 12, 2025Updated last year
- [arXiv 2026] dVoting: Fast Voting for dLLMs☆30Feb 13, 2026Updated 6 months ago
- This repository contains the code and data for the paper "VisOnlyQA: Large Vision Language Models Still Struggle with Visual Perception o…☆29Jul 9, 2025Updated last year
- The official implementation of the ICLR'23 paper PiFold: Toward effective and efficient protein inverse folding.☆183Jun 17, 2023Updated 3 years ago
- AAAI 2024-Attribute-Missing Graph Clustering Network☆15Jul 2, 2025Updated last year
- code of Let the data choose: Flexible and Diverse Anchor Graph Fusion for Scalable Multi-view Clustering☆17Feb 28, 2023Updated 3 years ago