This repository is a curated collection of the most exciting and influential CVPR 2024 papers. π₯ [Paper + Code + Demo]
β735Apr 15, 2026Updated 5 months ago
Alternatives and similar repositories for top-cvpr-2024-papers
Users that are interested in top-cvpr-2024-papers are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- This repository is a curated collection of the most exciting and influential CVPR 2023 papers. π₯ [Paper + Code]β649Apr 15, 2026Updated 5 months ago
- About This repository is a curated collection of the most exciting and influential CVPR 2025 papers. π₯ [Paper + Code + Demo]β896Apr 15, 2026Updated 5 months ago
- Official Implementation of CVPR24 highlight paper: Matching Anything by Segmenting Anythingβ1,377May 1, 2025Updated last year
- streamline the fine-tuning process for multimodal models: PaliGemma 2, Florence-2, and Qwen2.5-VLβ2,699Updated this week
- 4M: Massively Multimodal Masked Modelingβ1,812Sep 11, 2026Updated last week
- AI Agents on DigitalOcean Gradient AI Platform β’ AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- A collection of tutorials on state-of-the-art computer vision models and techniques. Explore everything from foundational architectures lβ¦β9,683Aug 14, 2026Updated last month
- A component that allows you to annotate an image with points and boxes.β21Dec 12, 2023Updated 2 years ago
- This series will take you on a journey from the fundamentals of NLP and Computer Vision to the cutting edge of Vision-Language Models.β1,180Jan 23, 2025Updated last year
- Use Florence 2 to auto-label data for use in training fine-tuned object detection models.β68Aug 15, 2024Updated 2 years ago
- ποΈ + π¬ + π§ = π€ Curated list of top foundation and multimodal models! [Paper + Code + Examples + Tutorials]β637Feb 29, 2024Updated 2 years ago
- [CVPR 2024] Real-Time Open-Vocabulary Object Detectionβ6,565Feb 26, 2025Updated last year
- NeurIPS 2025 Spotlight; ICLR2024 Spotlight; CVPR 2024; EMNLP 2024β1,854Aug 11, 2026Updated last month
- Trackers gives you clean, modular re-implementations of leading multi-object tracking algorithms released under the permissive Apache 2.0β¦β3,788Updated this week
- EfficientSAM: Leveraged Masked Image Pretraining for Efficient Segment Anythingβ2,494Dec 24, 2024Updated last year
- Simple, predictable pricing with DigitalOcean hosting β’ AdAlways know what you'll pay with monthly caps and flat pricing. Enterprise-grade infrastructure trusted by 600k+ customers.
- RobustSAM: Segment Anything Robustly on Degraded Images (CVPR 2024 Highlight)β369Aug 31, 2024Updated 2 years ago
- A tiny package supporting distributed computation of COCO metrics for PyTorch models.β15Feb 28, 2023Updated 3 years ago
- β61Oct 15, 2023Updated 2 years ago
- code for CVPR2024 paper: DiffMOT: A Real-time Diffusion-based Multiple Object Tracker with Non-linear Predictionβ449Jun 13, 2024Updated 2 years ago
- PyTorch Implementation of Object Recognition as Next Token Prediction [CVPR'24 Highlight]β180May 1, 2025Updated last year
- Recipes for shrinking, optimizing, customizing cutting edge vision models. πβ1,985Aug 12, 2026Updated last month
- Images to inference with no labeling (use foundation models to train supervised models).β2,776May 14, 2025Updated last year
- Code of paper "A new baseline for edge detection: Make Encoder-Decoder great again"β44Apr 11, 2026Updated 5 months ago
- [CVPR'24, Demo Track Honourable Mention] SuperPrimitive: Scene Reconstruction at a Primitive Levelβ204Mar 28, 2024Updated 2 years ago
- Managed Database hosting by DigitalOcean β’ AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- [CVPR 2024 Highlight] Official PyTorch implementation of SpatialTracker: Tracking Any 2D Pixels in 3D Spaceβ1,060Aug 8, 2025Updated last year
- PyTorch code and models for the DINOv2 self-supervised learning method.β13,348Jun 3, 2026Updated 3 months ago
- Official code for "FeatUp: A Model-Agnostic Frameworkfor Features at Any Resolution" ICLR 2024β1,655Jun 28, 2024Updated 2 years ago
- State-of-the-art Image & Video CLIP, Multimodal Large Language Models, and More!β2,367Apr 13, 2026Updated 5 months ago
- [CVPR 2024] Official RT-DETR (RTDETR paddle pytorch), Real-Time DEtection TRansformer, DETRs Beat YOLOs on Real-time Object Detection. π₯β¦β5,547Sep 7, 2026Updated 2 weeks ago
- β200Jun 3, 2025Updated last year
- [CVPR 2026 Oral] "INSID3: Training-Free In-Context Segmentation with DINOv3"β754Jun 26, 2026Updated 2 months ago
- Implementation of paper - YOLOv9: Learning What You Want to Learn Using Programmable Gradient Informationβ9,560Aug 9, 2024Updated 2 years ago
- Implementation of XFeat (CVPR 2024). Do you need robust and fast local feature extraction? You are in the right place!β1,732Jan 15, 2025Updated last year
- GPUs on demand by Runpod - Special Offer Available β’ AdRun AI, ML, and HPC workloads on powerful cloud GPUsβwithout limits or wasted spend. Deploy GPUs in under a minute and pay by the second.
- AAPL: Adding Attributes to Prompt Learning for Vision-Language Models (CVPRw 2024)β34May 8, 2024Updated 2 years ago
- Official codebase used to develop Vision Transformer, SigLIP, MLP-Mixer, LiT and more.β3,541May 19, 2025Updated last year
- [NeurIPS 2024] Code release for "Segment Anything without Supervision"β502Nov 20, 2025Updated 10 months ago
- [ICCV 2023] Tracking Anything with Decoupled Video Segmentationβ1,513Apr 26, 2025Updated last year
- The repository provides code for running inference with the Meta Segment Anything Model 2 (SAM 2), links for downloading the trained modeβ¦β19,900May 30, 2026Updated 3 months ago
- The official repo for the paper "VeCLIP: Improving CLIP Training via Visual-enriched Captions"β255Sep 11, 2026Updated last week
- This repository contains demos I made with the Transformers library by HuggingFace.β11,753Updated this week