Proposed fuzzy reward model with GRPO to improve VLM's abilities in crowd counting task.
☆21Apr 11, 2025Updated last year
Alternatives and similar repositories for CrowdVLM-R1
Users that are interested in CrowdVLM-R1 are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- (CVPR25) Exploring Contextual Attribute Density in Referring Expression Counting☆20Dec 3, 2025Updated 9 months ago
- ☆21Sep 16, 2025Updated last year
- Official implementaiton of RefAM: Attention Magnets for Zero-Shot Referral Segmentaiton☆17Feb 6, 2026Updated 7 months ago
- Segmentation assisted U-shaped multi-scale transformer for crowd counting☆22Jun 9, 2024Updated 2 years ago
- This is the implementation of the paper "Test-Time Adaptive Object Detection with Foundation Model" (Neurips 2025)☆23Jan 30, 2026Updated 7 months ago
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- demo code for "Video Prediction via Selective Sampling" (NeurIPS 2018)☆12Jul 15, 2020Updated 6 years ago
- ☆16May 13, 2024Updated 2 years ago
- ☆10May 20, 2021Updated 5 years ago
- The complete [1 to 5]-gram Gumar Corpus in the style of Google n-grams.☆12Feb 5, 2020Updated 6 years ago
- Latex template for poster☆12Sep 6, 2023Updated 3 years ago
- ☆10Nov 6, 2024Updated last year
- ☆10Jul 21, 2023Updated 3 years ago
- GeckoNum Benchmark for T2I Model Eval.☆15Dec 5, 2024Updated last year
- BootStrap + Django3 + simpleUI 实现的个人博客☆12Sep 22, 2021Updated 5 years ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- Token-free Language Modeling with ByGPT5 & Friends!☆12Jul 18, 2025Updated last year
- Tunisian Arabish Corpus☆12Mar 12, 2024Updated 2 years ago
- Task Preference Optimization: Improving Multimodal Large Language Models with Vision Task Alignment☆65Jul 22, 2025Updated last year
- ☆11Mar 25, 2024Updated 2 years ago
- ☆14Jul 15, 2025Updated last year
- Source code related to the research paper entitled RVENet: A Large Echocardiographic Dataset for the Deep Learning-Based Assessment of Ri…☆12Mar 10, 2024Updated 2 years ago
- Build TVM docker image for production compilation deployments☆12Sep 7, 2021Updated 5 years ago
- ☆15Oct 20, 2023Updated 2 years ago
- Display something on an analog oscilloscope☆12Oct 30, 2018Updated 7 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- Code and datasets for "Text encoders are performance bottlenecks in contrastive vision-language models". Coming soon!☆11May 24, 2023Updated 3 years ago
- Official Implement of ECCV 2024 paper "Multi-modal Crowd Counting via a Broker Modality"☆18Mar 19, 2026Updated 6 months ago
- WBSR: Rethinking Imbalance in Image Super-Resolution for Efficient Inference☆13Oct 8, 2024Updated last year
- Implementation of Contrastive Predictive Coding for Natural Language☆10Sep 16, 2020Updated 6 years ago
- Confidence Regulation Neurons in Language Models (NeurIPS 2024)☆16Feb 1, 2025Updated last year
- Implementation of Pix2Seq in PyTorch☆10Feb 3, 2022Updated 4 years ago
- Diacritization of Arabic texts☆11Apr 13, 2016Updated 10 years ago
- Turn from Google research,A simple code to realize HDR plus☆16Jul 15, 2019Updated 7 years ago
- LAReQA is a challenging benchmark for evaluating language agnostic answer retrieval from a multilingual candidate pool. This repository c…☆14May 19, 2020Updated 6 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- Project page for the ICDAR 2023 Paper "Inv3D: a high-resolution 3D invoice dataset for template-guided single-image document unwarping".☆13Dec 21, 2023Updated 2 years ago
- ☆10Dec 8, 2022Updated 3 years ago
- My personal notes for Facebook's Secure and Private AI Scholarship Course 2019 on Udacity☆10Jun 12, 2019Updated 7 years ago
- Official code for the paper, "TaCA: Upgrading Your Visual Foundation Model with Task-agnostic Compatible Adapter".☆16Jun 20, 2023Updated 3 years ago
- [CVPR 2024] KEPP: Why Not Use Your Textbook? Knowledge-Enhanced Procedure Planning of Instructional Videos☆12Sep 24, 2024Updated last year
- Official Pytorch implementation of "Deep Optimal Transport: A Practical Algorithm for Photo-realistic Image Restoration"☆28Sep 3, 2024Updated 2 years ago
- ☆18Jul 10, 2024Updated 2 years ago