☆57Nov 21, 2024Updated last year
Alternatives and similar repositories for LLaVA-o1
Users that are interested in LLaVA-o1 are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- [ICCV 2025] LLaVA-CoT, a visual language model capable of spontaneous, systematic reasoning☆2,130Dec 12, 2025Updated 9 months ago
- ☆12Jul 13, 2025Updated last year
- ☆10Aug 18, 2022Updated 4 years ago
- ☆11May 24, 2024Updated 2 years ago
- Utility which provides a UI to do prompt engineering within SageMaker Studio.☆14Jul 5, 2023Updated 3 years ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- Corpus to accompany: "Selective Vision is the Challenge for Visual Reasoning: A Benchmark for Visual Argument Understanding"☆11Apr 11, 2025Updated last year
- A minimal example of Abductive Learning☆20Dec 6, 2023Updated 2 years ago
- ☆15Jul 22, 2024Updated 2 years ago
- The official repository of "Improving Large Language Models via Fine-grained Reinforcement Learning with Minimum Editing Constraint"☆39Jan 12, 2024Updated 2 years ago
- K.Cook(나만의 케이크를 주문하다) Back-End Server 🍰☆10Jan 14, 2023Updated 3 years ago
- ☆11Apr 25, 2026Updated 4 months ago
- Playground project acting as an example for a complex LangChain workflow☆11Jun 20, 2023Updated 3 years ago
- NLOST-Non-Line-of-Sight-Imaging-with-Transformer (CVPR2023)☆33Jul 3, 2024Updated 2 years ago
- MTC-CSNet: Marrying Transformer and Convolution for Image Compressed Sensing.☆11Jan 12, 2024Updated 2 years ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- Code for Disambiguating Monocular Depth Estimation with a Single Transient☆12Sep 8, 2020Updated 6 years ago
- ☆54Feb 12, 2025Updated last year
- ☆32Feb 8, 2024Updated 2 years ago
- [SIGGRAPH Asia] Differentiable Transient Rendering☆13Jun 13, 2022Updated 4 years ago
- MLOps Implementing "Brain Computer Interface" on Kubernetes☆16Sep 30, 2022Updated 3 years ago
- Open source implementation of the paper "MM-Vid: Advancing Video Understanding with GPT-4V(ision)".☆45Jan 4, 2026Updated 8 months ago
- ☆12Nov 3, 2023Updated 2 years ago
- Archives for Triton Inference Server Practices☆15Feb 28, 2022Updated 4 years ago
- LogiCity@NeurIPS'24, D&B track. A multi-agent inductive learning environment for "abstractions".☆27Jun 10, 2025Updated last year
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- Frictionless Machine Learning on Kubernetes☆16Mar 7, 2023Updated 3 years ago
- Deep Non-line-of-sight Imaging from Under-scanning Measurements (NeurIPS 2023)☆14Apr 10, 2024Updated 2 years ago
- ☆35Jan 21, 2025Updated last year
- ☆21Aug 11, 2025Updated last year
- [BMVC 2022] "SUB-Depth: Self-distillation and Uncertainty Boosting Self-supervised Monocular Depth Estimation"☆16Apr 19, 2022Updated 4 years ago
- [EMNLP 2025] The official implementation of "Zero-shot Multimodal Document Retrieval via Cross-Modal Question Generation"☆15Aug 26, 2025Updated last year
- A highly contextualized retrieval system integrating Large Language Models (LLMs), embeddings, and a dynamic agent-driven framework. Supp…☆27May 21, 2026Updated 3 months ago
- [ECCV 2024] Official PyTorch implementation of "HYPE: Hyperbolic Entailment Filtering for Underspecified Images and Texts"☆20Nov 22, 2024Updated last year
- sample game collection for development of cocos2dx 2.x☆12Jun 13, 2021Updated 5 years ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- 一个小小的书单,收集整理了一些计算机科学与技术方面的书籍英文原著pdf。☆10Jan 13, 2022Updated 4 years ago
- [NeurIPS'24] SpatialEval: a benchmark to evaluate spatial reasoning abilities of MLLMs and LLMs☆61Jan 23, 2025Updated last year
- ☆23Oct 22, 2024Updated last year
- NeuSyRE: A Neuro-Symbolic Visual Understanding and Reasoning Framework based on Scene Graph Enrichment☆25Mar 10, 2024Updated 2 years ago
- Convolutional Approximations to the General Non-Line-of-Sight Imaging Operator☆12Oct 27, 2019Updated 6 years ago
- A PyTorch-based program which estimates 3D depth maps from active structured-light sensor's multiple video frames☆15Mar 22, 2022Updated 4 years ago
- Source code for the Paper "Mind the Gap: Benchmarking Spatial Reasoning in Vision-Language Models"☆20Feb 1, 2026Updated 7 months ago