Florence-2
☆73Feb 13, 2025Updated last year
Alternatives and similar repositories for florence-2
Users that are interested in florence-2 are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- ☆17May 14, 2025Updated last year
- Offical code repository of ”DAAD: Dynamic Analysis and Adaptive Discriminator for Fake News Detection“☆22Aug 22, 2024Updated 2 years ago
- Quick exploration into fine tuning florence 2☆339Sep 19, 2024Updated last year
- Florence-2 is a novel vision foundation model with a unified, prompt-based representation for a variety of computer vision and vision-lan…☆213Jul 3, 2024Updated 2 years ago
- LAVIS - A One-stop Library for Language-Vision Intelligence☆48Aug 5, 2024Updated 2 years ago
- Serverless GPU API endpoints on Runpod - Get Bonus Credits • AdSkip the infrastructure headaches. Auto-scaling, pay-as-you-go, no-ops approach lets you focus on innovating your application.
- A Unified Remote Sensing Object Detector Based on Fourier Contour Parametric Learning☆24Dec 11, 2025Updated 8 months ago
- YOLOv8 implementation without DFL using PyTorch☆12Jul 22, 2024Updated 2 years ago
- Azure Computer Vision 4 (March 2023 - Florence) workshop in a day☆41May 11, 2023Updated 3 years ago
- [ICLR 2025] MMFakeBench: A Mixed-Source Multimodal Misinformation Detection Benchmark for LVLMs☆56Mar 25, 2025Updated last year
- ☆13Dec 17, 2024Updated last year
- ☆17Oct 24, 2025Updated 10 months ago
- Few-shot Steel Surface Defect Detection☆14Nov 30, 2021Updated 4 years ago
- 🦩 Official repository of paper "Visual Instruction Tuning with Polite Flamingo" (AAAI-24 Oral)☆65Dec 9, 2023Updated 2 years ago
- [ACM MM 2024] FKA-Owl: Advancing Multimodal Fake News Detection through Knowledge-Augmented LVLMs☆60Aug 8, 2024Updated 2 years ago
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- 图片多角度☆44Jun 6, 2024Updated 2 years ago
- ☆13Dec 12, 2023Updated 2 years ago
- Visual Inspection Orchestrator is a modular framework made to ease the deployment of VI usecases☆14May 28, 2026Updated 3 months ago
- Mamba-YOLO-World: Marrying YOLO-World with Mamba for Open-Vocabulary Detection☆104Mar 12, 2025Updated last year
- Empirical Study Towards Building An Effective Multi-Modal Large Language Model☆21Oct 25, 2023Updated 2 years ago
- When Pixel Difference Patterns Meet ViT: PiDiViT for Few-Shot Object Detection☆19Nov 3, 2025Updated 10 months ago
- ☆24Oct 22, 2025Updated 10 months ago
- ☆10Apr 7, 2025Updated last year
- XGEN-MM(BLIP3) Autocaptioning Tools☆17Jun 20, 2024Updated 2 years ago
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- React CodeGen using GPT☆12Feb 11, 2024Updated 2 years ago
- ☆32Jul 23, 2022Updated 4 years ago
- ☆11Mar 4, 2026Updated 6 months ago
- ☆24Jun 4, 2024Updated 2 years ago
- End-to-End Neural Event Coreference Resolution☆11Jun 18, 2023Updated 3 years ago
- A python package for predicting group-level fMRI responses to visual stimuli using deep neural networks☆13Mar 31, 2025Updated last year
- This is a repository for submit our work out of MSC project. Team:学习 机器学习队☆21Jan 23, 2020Updated 6 years ago
- ☆21Jan 22, 2026Updated 7 months ago
- A package to study complex networks based on the temporal evolution of their Dynamic Communicability and Flow.☆11Jan 30, 2026Updated 7 months ago
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- ☆15May 21, 2026Updated 3 months ago
- VL-GPT: A Generative Pre-trained Transformer for Vision and Language Understanding and Generation☆86Sep 12, 2024Updated last year
- collection of few-shot papers☆24Aug 31, 2026Updated last week
- Python scripts performing Open Vocabulary Object Detection using the YOLO-World model in ONNX.☆65Apr 7, 2024Updated 2 years ago
- Grounding Language Models for Compositional and Spatial Reasoning☆18Oct 26, 2022Updated 3 years ago
- PyTorch code for the paper "CrossTransformers: spatially-aware few-shot transfer"☆25Dec 20, 2020Updated 5 years ago
- Text detection network psenet deployed by libtorch and Qt.☆14May 21, 2020Updated 6 years ago