☆30Mar 13, 2024Updated 2 years ago
Alternatives and similar repositories for DesCo
Users that are interested in DesCo are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- A Python toolkit for the OmniLabel benchmark providing code for evaluation and visualization☆23Feb 1, 2025Updated last year
- ☆11Jan 27, 2020Updated 6 years ago
- Accepted by CVPR 2020.☆27Jul 11, 2024Updated 2 years ago
- ☆21Apr 2, 2024Updated 2 years ago
- v1: Learning to Point Visual Tokens for Multimodal Grounded Reasoning☆21Updated this week
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- MM-PRM: Enhancing Multimodal Mathematical Reasoning with Scalable Step-Level Supervision☆30May 26, 2025Updated last year
- ☆17Apr 10, 2025Updated last year
- Original VinVL visual backbone with simplified APIs to easily extract features, boxes, object detections, in a few lines of Python code.☆12Nov 27, 2022Updated 3 years ago
- [TACL/EMNLP'24] Do Vision and Language Models Share Concepts? A Vector Space Alignment Study☆16Nov 22, 2024Updated last year
- Download Web-10K data by querying Bing Image Search☆10Feb 1, 2022Updated 4 years ago
- Code for "Translatotron-V(ison): An End-to-End Model for In-Image Machine Translation" (Findings of ACL 2024)☆16Jul 4, 2024Updated 2 years ago
- Holistic Coverage and Faithfulness Evaluation of Large Vision-Language Models (ACL-Findings 2024)☆16Apr 23, 2024Updated 2 years ago
- ☆12Nov 3, 2022Updated 3 years ago
- Map4RDF allows visualising and interacting with Linked Geospatial Data available in any SPARQL endpoint☆10Feb 9, 2020Updated 6 years ago
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- ☆11Jan 12, 2024Updated 2 years ago
- ☆11Jul 6, 2023Updated 3 years ago
- GroupViT: Semantic Segmentation Emerges from Text Supervision☆25Dec 15, 2022Updated 3 years ago
- ☆12Jun 18, 2024Updated 2 years ago
- Grounded Language-Image Pre-training☆2,612Jan 24, 2024Updated 2 years ago
- codes for Efficient Test-Time Scaling via Self-Calibration☆22Sep 13, 2025Updated last year
- Implementations of CVPR 2019 paper Distilling Object Detectors with Fine-grained Feature Imitation☆27Nov 22, 2022Updated 3 years ago
- [CVPR 2024 Highlight] Official repository of the paper "The devil is in the fine-grained details: Evaluating open-vocabulary object detec…☆68Apr 4, 2025Updated last year
- Code for EMNLP2021 paper “Transductive Learning for Unsupervised Text Style Transfer”☆12Sep 19, 2021Updated 5 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- PSROIAlign with multi-batch training support in PyTorch☆24Jun 20, 2019Updated 7 years ago
- ☆32Jul 29, 2024Updated 2 years ago
- PyTorch Implementation for InMaP☆12Oct 28, 2023Updated 2 years ago
- personal homepage of Ziwei Liu☆14Updated this week
- Official Implementation for FF_designed CNNs☆15Aug 13, 2019Updated 7 years ago
- The source codes for Region Comparison Network for Interpretable Few-shot Image Classification☆10Sep 17, 2020Updated 6 years ago
- ☆11Sep 23, 2020Updated 5 years ago
- Background Learnable Cascade for Zero-shot object detection☆17Sep 29, 2021Updated 4 years ago
- [not maintained anymore] [for study purpose] A simple PyTorch implementation for "Global Vectors for Word Representation".☆17Nov 7, 2019Updated 6 years ago
- Deploy open-source AI quickly and easily - Special Bonus Offer • AdRunpod Hub is built for open source. One-click deployment and autoscaling endpoints without provisioning your own infrastructure.
- Repo from the "Learning with limited labeled data" seminar @ Uni of Tuebingen. A collection of notes, notebooks and slideshows to underst…☆17Apr 13, 2023Updated 3 years ago
- Follow-Up Differential Descriptions: Language Models Resolve Ambiguities for Image Classification☆11Nov 15, 2023Updated 2 years ago
- ☆60Jun 16, 2023Updated 3 years ago
- This is a C++ implementation of cocoapi bbox evaluation code.☆11Dec 9, 2021Updated 4 years ago
- Code for IEEE Trans. on Multimedia (TMM) paper "Object-aware Multimodal Named Entity Recognition in Social Media Posts with Adversarial L…☆20Mar 3, 2021Updated 5 years ago
- Official implementation of our paper "Finetuned Multimodal Language Models are High-Quality Image-Text Data Filters".☆71Apr 14, 2025Updated last year
- [AAAI 2023] DQ-DETR: Dual Query Detection Transformer for Phrase Extraction and Grounding☆58Nov 28, 2022Updated 3 years ago