This repository is a curated collection of the most exciting and influential CVPR 2024 papers. π₯ [Paper + Code + Demo]
β736Apr 15, 2026Updated 5 months ago
Alternatives and similar repositories for top-cvpr-2024-papers
Users that are interested in top-cvpr-2024-papers are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- This repository is a curated collection of the most exciting and influential CVPR 2023 papers. π₯ [Paper + Code]β650Apr 15, 2026Updated 5 months ago
- About This repository is a curated collection of the most exciting and influential CVPR 2025 papers. π₯ [Paper + Code + Demo]β898Apr 15, 2026Updated 5 months ago
- Official Implementation of CVPR24 highlight paper: Matching Anything by Segmenting Anythingβ1,377May 1, 2025Updated last year
- streamline the fine-tuning process for multimodal models: PaliGemma 2, Florence-2, and Qwen2.5-VLβ2,696Sep 28, 2026Updated last week
- 4M: Massively Multimodal Masked Modelingβ1,814Sep 11, 2026Updated 3 weeks ago
- Wordpress hosting with auto-scaling - Free Trial Offer β’ AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- A collection of tutorials on state-of-the-art computer vision models and techniques. Explore everything from foundational architectures lβ¦β9,705Sep 22, 2026Updated 2 weeks ago
- A component that allows you to annotate an image with points and boxes.β21Dec 12, 2023Updated 2 years ago
- This series will take you on a journey from the fundamentals of NLP and Computer Vision to the cutting edge of Vision-Language Models.β1,182Jan 23, 2025Updated last year
- Use Florence 2 to auto-label data for use in training fine-tuned object detection models.β68Aug 15, 2024Updated 2 years ago
- ποΈ + π¬ + π§ = π€ Curated list of top foundation and multimodal models! [Paper + Code + Examples + Tutorials]β637Feb 29, 2024Updated 2 years ago
- [CVPR 2024] Real-Time Open-Vocabulary Object Detectionβ6,578Feb 26, 2025Updated last year
- NeurIPS 2025 Spotlight; ICLR2024 Spotlight; CVPR 2024; EMNLP 2024β1,854Aug 11, 2026Updated last month
- Trackers gives you clean, modular re-implementations of leading multi-object tracking algorithms released under the permissive Apache 2.0β¦β3,938Updated this week
- EfficientSAM: Leveraged Masked Image Pretraining for Efficient Segment Anythingβ2,498Dec 24, 2024Updated last year
- Deploy to Railway using AI coding agents - Free Credits Offer β’ AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- RobustSAM: Segment Anything Robustly on Degraded Images (CVPR 2024 Highlight)β369Aug 31, 2024Updated 2 years ago
- A tiny package supporting distributed computation of COCO metrics for PyTorch models.β15Feb 28, 2023Updated 3 years ago
- β61Oct 15, 2023Updated 2 years ago
- code for CVPR2024 paper: DiffMOT: A Real-time Diffusion-based Multiple Object Tracker with Non-linear Predictionβ449Jun 13, 2024Updated 2 years ago
- PyTorch Implementation of Object Recognition as Next Token Prediction [CVPR'24 Highlight]β180May 1, 2025Updated last year
- Recipes for shrinking, optimizing, customizing cutting edge vision models. πβ1,997Aug 12, 2026Updated last month
- Images to inference with no labeling (use foundation models to train supervised models).β2,787Sep 29, 2026Updated last week
- Code of paper "A new baseline for edge detection: Make Encoder-Decoder great again"β44Apr 11, 2026Updated 6 months ago
- [CVPR'24, Demo Track Honourable Mention] SuperPrimitive: Scene Reconstruction at a Primitive Levelβ205Mar 28, 2024Updated 2 years ago
- 1-Click AI Models by DigitalOcean Gradient β’ AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- [CVPR 2024 Highlight] Official PyTorch implementation of SpatialTracker: Tracking Any 2D Pixels in 3D Spaceβ1,062Aug 8, 2025Updated last year
- PyTorch code and models for the DINOv2 self-supervised learning method.β13,405Updated this week
- Official code for "FeatUp: A Model-Agnostic Frameworkfor Features at Any Resolution" ICLR 2024β1,656Jun 28, 2024Updated 2 years ago
- State-of-the-art Image & Video CLIP, Multimodal Large Language Models, and More!β2,381Apr 13, 2026Updated 5 months ago
- [CVPR 2024] Official RT-DETR (RTDETR paddle pytorch), Real-Time DEtection TRansformer, DETRs Beat YOLOs on Real-time Object Detection. π₯β¦β5,585Updated this week
- β200Jun 3, 2025Updated last year
- [CVPR 2026 Oral] "INSID3: Training-Free In-Context Segmentation with DINOv3"β769Jun 26, 2026Updated 3 months ago
- Implementation of paper - YOLOv9: Learning What You Want to Learn Using Programmable Gradient Informationβ9,563Aug 9, 2024Updated 2 years ago
- Implementation of XFeat (CVPR 2024). Do you need robust and fast local feature extraction? You are in the right place!β1,744Jan 15, 2025Updated last year
- Serverless GPU API endpoints on Runpod - Get Bonus Credits β’ AdSkip the infrastructure headaches. Auto-scaling, pay-as-you-go, no-ops approach lets you focus on innovating your application.
- AAPL: Adding Attributes to Prompt Learning for Vision-Language Models (CVPRw 2024)β34May 8, 2024Updated 2 years ago
- Official codebase used to develop Vision Transformer, SigLIP, MLP-Mixer, LiT and more.β3,544May 19, 2025Updated last year
- [NeurIPS 2024] Code release for "Segment Anything without Supervision"β504Nov 20, 2025Updated 10 months ago
- [ICCV 2023] Tracking Anything with Decoupled Video Segmentationβ1,519Apr 26, 2025Updated last year
- The repository provides code for running inference with the Meta Segment Anything Model 2 (SAM 2), links for downloading the trained modeβ¦β19,980May 30, 2026Updated 4 months ago
- The official repo for the paper "VeCLIP: Improving CLIP Training via Visual-enriched Captions"β256Sep 11, 2026Updated 3 weeks ago
- This repository contains demos I made with the Transformers library by HuggingFace.β11,758Sep 17, 2026Updated 3 weeks ago