Florence-2
☆72Feb 13, 2025Updated last year
Alternatives and similar repositories for florence-2
Users that are interested in florence-2 are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- ☆16May 14, 2025Updated last year
- Florence-2 is a novel vision foundation model with a unified, prompt-based representation for a variety of computer vision and vision-lan…☆203Jul 3, 2024Updated 2 years ago
- Florence-2 image captioning and tasks☆85Jul 11, 2025Updated last year
- ☆29Mar 15, 2026Updated 4 months ago
- Video search using Azure Computer Vision 4 (Florence)☆20Nov 21, 2025Updated 8 months ago
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- Point2SSM: Learning Morphological Variations of Anatomies from Point Cloud☆13Jun 22, 2024Updated 2 years ago
- [ICLR 2025] MMFakeBench: A Mixed-Source Multimodal Misinformation Detection Benchmark for LVLMs☆55Mar 25, 2025Updated last year
- PCB Defect Detection☆17Nov 4, 2024Updated last year
- finetune your florence2 model easy☆21Jul 27, 2024Updated 2 years ago
- Training a YOLO NAS Model for detecting retail product items from shelf images using SKU110K dataset.☆10Aug 13, 2023Updated 2 years ago
- ☆12Dec 17, 2024Updated last year
- [ICLR 2024] Seer: Language Instructed Video Prediction with Latent Diffusion Models☆35May 23, 2024Updated 2 years ago
- (CVPR 2026) Sparsity-Aware Voxel Attention and Foreground Modulation for 3D Semantic Scene Completion☆15Mar 8, 2026Updated 4 months ago
- 🦩 Official repository of paper "Visual Instruction Tuning with Polite Flamingo" (AAAI-24 Oral)☆65Dec 9, 2023Updated 2 years ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- An unofficial (self-hosted) API tunnel that provides access to Grok3 through a simple REST interface.☆10Feb 28, 2025Updated last year
- [ACM MM 2024] FKA-Owl: Advancing Multimodal Fake News Detection through Knowledge-Augmented LVLMs☆60Aug 8, 2024Updated last year
- ☆13Dec 12, 2023Updated 2 years ago
- [2021CVPR] Adaptive Image Transformer for One-Shot Object Detection☆22Mar 19, 2024Updated 2 years ago
- Mamba-YOLO-World: Marrying YOLO-World with Mamba for Open-Vocabulary Detection☆104Mar 12, 2025Updated last year
- ☆10Apr 7, 2025Updated last year
- [NeurIPS 2024] Official code for paper "EZ-HOI: VLM Adaptation via Guided Prompt Learning for Zero-Shot HOI Detection"☆42Jul 7, 2025Updated last year
- ☆32Jul 23, 2022Updated 4 years ago
- A python package for predicting group-level fMRI responses to visual stimuli using deep neural networks☆13Mar 31, 2025Updated last year
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- ☆21Jan 22, 2026Updated 6 months ago
- VL-GPT: A Generative Pre-trained Transformer for Vision and Language Understanding and Generation☆86Sep 12, 2024Updated last year
- collection of few-shot papers☆22Jul 21, 2026Updated last week
- Python scripts performing Open Vocabulary Object Detection using the YOLO-World model in ONNX.☆63Apr 7, 2024Updated 2 years ago
- PyTorch code for the paper "CrossTransformers: spatially-aware few-shot transfer"☆25Dec 20, 2020Updated 5 years ago
- Text detection network psenet deployed by libtorch and Qt.☆14May 21, 2020Updated 6 years ago
- ECCV' 2024.☆14Sep 11, 2024Updated last year
- Robust Referring Video Object Segmentation with Cyclic Structural Consistency [ICCV 2023]☆30Mar 13, 2024Updated 2 years ago
- A MATLAB app to interactively navigate Ryze Tello drone, read navigation data, process image data and produce equivalent MATLAB code. Thi…☆13Feb 26, 2026Updated 5 months ago
- Bare Metal GPUs on DigitalOcean Gradient AI • AdPurpose-built for serious AI teams training foundational models, running large-scale inference, and pushing the boundaries of what's possible.
- ☆15Dec 3, 2021Updated 4 years ago
- AFFNet-Unofficial Implementation☆14Aug 23, 2023Updated 2 years ago
- [WACV 2026] ZonUI-3B — A lightweight, resolution-aware GUI grounding model trained with only 24K samples on a single RTX 4090.☆26Jan 2, 2026Updated 6 months ago
- A8R8 (https://github.com/ramyma/a8r8) supporting nodes to integrate with ComfyUI☆74Dec 9, 2024Updated last year
- [TCSVT'22]Meta-Learning Based Incremental Few-Shot Object Detection☆25Sep 14, 2022Updated 3 years ago
- PyTorch implementation of Multi-Perspective Data Augmentation for Few-shot Object Detection☆24Apr 15, 2025Updated last year
- A Focal Transformer for Boundary-aware Prostate Segmentation using CT Images☆11Nov 10, 2024Updated last year