Codebase for the Recognize Anything Model (RAM)
☆88Dec 11, 2023Updated 2 years ago
Alternatives and similar repositories for recognize-anything
Users that are interested in recognize-anything are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- [ECCV 2024] Tokenize Anything via Prompting☆600Dec 11, 2024Updated last year
- Open-source and strong foundation image recognition models.☆3,721Feb 18, 2025Updated last year
- An integration of Segment Anything Model, Molmo, and, Whisper to segment objects using voice and natural language.☆30Feb 28, 2025Updated last year
- Grounded SAM: Marrying Grounding DINO with Segment Anything & Stable Diffusion & Recognize Anything - Automatically Detect , Segment and …☆17,731Sep 5, 2024Updated 2 years ago
- Class project for COMP-781, Robotics. This is a CUDA-based collision detector for motion planning.☆13Apr 29, 2019Updated 7 years ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- ☆17Aug 18, 2023Updated 3 years ago
- ☆17Aug 18, 2022Updated 4 years ago
- [ICCV 2023] Learning Fine-Grained Features for Pixel-wise Video Correspondences☆18Mar 3, 2024Updated 2 years ago
- ☆30Mar 13, 2024Updated 2 years ago
- [ECCV 2024] Official implementation of the paper "Semantic-SAM: Segment and Recognize Anything at Any Granularity"☆2,852Jul 10, 2025Updated last year
- 签证官揭开关于美国学生签证申请的谣言☆11May 30, 2018Updated 8 years ago
- Original VinVL visual backbone with simplified APIs to easily extract features, boxes, object detections, in a few lines of Python code.☆12Nov 27, 2022Updated 3 years ago
- [TACL/EMNLP'24] Do Vision and Language Models Share Concepts? A Vector Space Alignment Study☆16Nov 22, 2024Updated last year
- Code for experiments for "ConvNet vs Transformer, Supervised vs CLIP: Beyond ImageNet Accuracy"☆102Sep 11, 2024Updated 2 years ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- [ECCV 2024] Official implementation of the paper "Grounding DINO: Marrying DINO with Grounded Pre-Training for Open-Set Object Detection"☆10,603Aug 12, 2024Updated 2 years ago
- Official PyTorch implementation of ProCLIP: Progressive Vision-Language Alignment via LLM-based Embedder☆27Aug 1, 2026Updated last month
- Code for "Translatotron-V(ison): An End-to-End Model for In-Image Machine Translation" (Findings of ACL 2024)☆16Jul 4, 2024Updated 2 years ago
- [NeurIPS 2023] Official implementation of the paper "Segment Everything Everywhere All at Once"☆4,791Aug 19, 2024Updated 2 years ago
- EfficientSAM: Leveraged Masked Image Pretraining for Efficient Segment Anything☆2,494Dec 24, 2024Updated last year
- Pytorch code for paper From CLIP to DINO: Visual Encoders Shout in Multi-modal Large Language Models☆212Jan 8, 2025Updated last year
- MM-Interleaved: Interleaved Image-Text Generative Modeling via Multi-modal Feature Synchronizer☆256Apr 3, 2024Updated 2 years ago
- ☆18Dec 2, 2024Updated last year
- A sd-webui extension for utilizing DanTagGen to "upsample prompts".☆12Jun 13, 2024Updated 2 years ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- Implements RNNPool and SoftPool for CNNs.☆14Jan 29, 2021Updated 5 years ago
- ☆10Dec 27, 2020Updated 5 years ago
- Official implementation of paper "Masked Distillation with Receptive Tokens", ICLR 2023.☆10Mar 13, 2023Updated 3 years ago
- Using CogVLM and CogAgent for image captioning☆15Dec 29, 2023Updated 2 years ago
- Official Implementation for paper Synthesizing Light Field Video from Monocular Video☆10Jul 15, 2022Updated 4 years ago
- ☆16Mar 13, 2023Updated 3 years ago
- 国内外数据竞赛资讯整理☆18Nov 6, 2021Updated 4 years ago
- Grounding DINO 1.5: IDEA Research's Most Capable Open-World Object Detection Model Series☆1,145Jan 21, 2025Updated last year
- codes for Efficient Test-Time Scaling via Self-Calibration☆22Sep 13, 2025Updated last year
- Open source password manager - Proton Pass • AdSecurely store, share, and autofill your credentials with Proton Pass, the end-to-end encrypted password manager trusted by millions.
- AICUP 2024 Cross-camera Multiple-object tracking☆29Oct 21, 2024Updated last year
- PDNet: Toward Better One-Stage Object Detection With Prediction Decoupling, TIP 2022☆11Nov 30, 2022Updated 3 years ago
- Belief Revision based Caption Re-ranker with Visual Semantic Information. COLING 2022☆11Apr 13, 2025Updated last year
- Grounded SAM 2: Ground and Track Anything in Videos with Grounding DINO, Florence-2 and SAM 2☆3,742Nov 11, 2025Updated 10 months ago
- StrongSort-Pip: Packaged version of StrongSort☆10Sep 3, 2022Updated 4 years ago
- [CVPR 2024] Official implementation of the paper "Visual In-context Learning"☆544Apr 8, 2024Updated 2 years ago
- Exploiting unlabeled data with vision and language models for object detection, ECCV 2022☆96Jan 16, 2024Updated 2 years ago