🤩 An AWESOME Curated List of Papers, Workshops, Datasets, and Challenges from CVPR 2024
☆146Jun 13, 2024Updated 2 years ago
Alternatives and similar repositories for awesome-cvpr-2024
Users that are interested in awesome-cvpr-2024 are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Use text-to-image models Stable Diffusion, DALL-E2, DALL-E3, SDXL, SSD-1B, Kandinsky-2.2, and LCM from UI. Add images directly to your da…☆34Apr 23, 2024Updated 2 years ago
- Run SOTA Vision-Language Model Florence-2 on your data!☆15Mar 27, 2025Updated last year
- Run optical character recognition with PyTesseract from the FiftyOne App!☆11Apr 5, 2024Updated 2 years ago
- A FiftyOne Plugin which takes inspiration from ComfyUI☆15Mar 5, 2026Updated 5 months ago
- Albumentations Data Augmentation Plugin for FiftyOne!☆15Aug 22, 2024Updated 2 years ago
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- My journey during 10 weeks of building FiftyOne plugins☆22Nov 12, 2023Updated 2 years ago
- ☆24Aug 17, 2026Updated 2 weeks ago
- Implementing NVLabs C-RADIOv3 Embeddings Model as Remotely Sourced Zoo Model for FiftyOne☆27Feb 3, 2026Updated 6 months ago
- ☆39Aug 16, 2026Updated 2 weeks ago
- ☆106Mar 23, 2026Updated 5 months ago
- FiftyOne Plugin for finding common image quality issues☆35Oct 21, 2024Updated last year
- A FiftyOne Plugin that allows you to search across any modality in your videos!☆30May 27, 2025Updated last year
- A curated list of plugins that you can add to your FiftyOne install!☆143Aug 18, 2026Updated 2 weeks ago
- Perform visual question answering on your images☆19May 8, 2024Updated 2 years ago
- Virtual machines for every use case on DigitalOcean • AdGet dependable uptime with 99.99% SLA, simple security tools, and predictable monthly pricing with DigitalOcean's virtual machines, called Droplets.
- Run zero-shot prediction models on your data☆37Dec 19, 2024Updated last year
- [CVPR 2024 Highlight] ImageNet-D☆47Jul 24, 2026Updated last month
- Testbed for multimodal retrieval augmented generation techniques with FiftyOne, LlamaIndex, and Milvus☆21Aug 9, 2024Updated 2 years ago
- Code release for "Understanding Bias in Large-Scale Visual Datasets"☆25Dec 4, 2024Updated last year
- Convert datasets from Hugging Face to FiftyOne for Visualization☆11Mar 15, 2024Updated 2 years ago
- Downstream semantic segmentation evaluation of DGInStyle.☆25Apr 1, 2024Updated 2 years ago
- This repo contains code for the Coursera MOOC Hands-on Data Centric Visual AI☆39Sep 24, 2024Updated last year
- Official code of *Towards Event-oriented Long Video Understanding*☆12Jul 26, 2024Updated 2 years ago
- ☆55Jan 17, 2025Updated last year
- Deploy open-source AI quickly and easily - Special Bonus Offer • AdRunpod Hub is built for open source. One-click deployment and autoscaling endpoints without provisioning your own infrastructure.
- [EMNLP 2024] Preserving Multi-Modal Capabilities of Pre-trained VLMs for Improving Vision-Linguistic Compositionality☆24Oct 8, 2024Updated last year
- Official implementation of "Describing Differences in Image Sets with Natural Language" (CVPR 2024 Oral)☆134Nov 5, 2025Updated 9 months ago
- [CVPR 2024] Contrasting Intra-Modal and Ranking Cross-Modal Hard Negatives to Enhance Visio-Linguistic Fine-grained Understanding☆55Apr 7, 2025Updated last year
- 2023 Capstone Design☆12Nov 2, 2023Updated 2 years ago
- Source Code of rTVRA for Hyperspectral Image Reconstruction on Dual-camera Compressive Hyperspectral Imaging System☆14Dec 23, 2020Updated 5 years ago
- Implementation of the semi-structured inference model in our ACL 2020 paper, INFOTABS: Inference on Tables as Semi-structured Data.☆17Dec 7, 2021Updated 4 years ago
- ☆15Mar 27, 2024Updated 2 years ago
- [NeurIPS 2024] Official PyTorch implementation of LoTLIP: Improving Language-Image Pre-training for Long Text Understanding☆49Jan 14, 2025Updated last year
- [ICLR 2025] HQ-Edit: A High-Quality and High-Coverage Dataset for General Image Editing☆114Apr 18, 2024Updated 2 years ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- [BMVC 2023 (Oral)] SketchDreamer: Interactive Text-Augmented Creative Sketch Ideation☆28Jun 8, 2025Updated last year
- ☆17Sep 25, 2024Updated last year
- [TACL/EMNLP'24] Do Vision and Language Models Share Concepts? A Vector Space Alignment Study☆16Nov 22, 2024Updated last year
- Official code for the paper "Exploiting the Complementarity of 2D and 3D Networks to Address Domain-Shift in 3D Semantic Segmentation"☆11Aug 25, 2023Updated 3 years ago
- Semantically Search Emojis From the Command Line!☆13Nov 26, 2023Updated 2 years ago
- [CVPR 2025 Highlight] Official Pytorch codebase for paper: "Assessing and Learning Alignment of Unimodal Vision and Language Models"☆60Aug 15, 2025Updated last year
- [CVPR 2025] PyTorch implementation of paper "FLAME: Frozen Large Language Models Enable Data-Efficient Language-Image Pre-training"☆33Jul 8, 2025Updated last year