☆16Apr 20, 2026Updated 4 months ago
Alternatives and similar repositories for PeBR-R1
Users that are interested in PeBR-R1 are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- code for "CoMT: A Novel Benchmark for Chain of Multi-modal Thought on Large Vision-Language Models"☆19Mar 10, 2025Updated last year
- Neurlps 2025☆19Mar 9, 2026Updated 6 months ago
- [NeurIPS 2025] Think or Not? Selective Reasoning via Reinforcement Learning for Vision-Language Models☆61Sep 29, 2025Updated 11 months ago
- [NeurIPS 2025] MINT-CoT: Enabling Interleaved Visual Tokens in Mathematical Chain-of-Thought Reasoning☆108Sep 19, 2025Updated last year
- Repo for the paper: Towards Few-shot Entity Recognition in Document Images:A Label-aware Sequence-to-Sequence Framework☆14May 31, 2023Updated 3 years ago
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- HOCR Specification Python Parser☆12Sep 23, 2015Updated 10 years ago
- Official implementation of "Figure It Out: Improve the Frontier of Reasoning with Active Visual Thinking"☆17Jan 13, 2026Updated 8 months ago
- Official repository for “Reasoning in the Dark: Interleaved Vision-Text Reasoning in Latent Space”☆18Jan 27, 2026Updated 7 months ago
- The official code of "VL-Rethinker: Incentivizing Self-Reflection of Vision-Language Models with Reinforcement Learning" [NeurIPS25]☆192Jun 5, 2025Updated last year
- SAEval: A benchmark for sentiment analysis to evaluate the model's performance on various subtasks.☆15Apr 29, 2024Updated 2 years ago
- [ICML 2026] Heima☆76May 20, 2026Updated 3 months ago
- Official code of DMA: Dense Multimodal Alignment for Open-Vocabulary 3D Scene Understanding, ECCV 2024☆32Jul 18, 2024Updated 2 years ago
- 深度学习与围棋学习☆15Oct 27, 2021Updated 4 years ago
- [NeurIPS 2024] XMask3D: Cross-modal Mask Reasoning for Open Vocabulary 3D Semantic Segmentation☆38Jan 20, 2025Updated last year
- Virtual machines for every use case on DigitalOcean • AdGet dependable uptime with 99.99% SLA, simple security tools, and predictable monthly pricing with DigitalOcean's virtual machines, called Droplets.
- ☆12Sep 19, 2021Updated 5 years ago
- This is a repository for Junjia's application of universities, the United States☆23Aug 20, 2020Updated 6 years ago
- Heatmap-based Out-of-Distribution Detection (WACV 2023)☆13Mar 27, 2024Updated 2 years ago
- The official repository of MM-R5☆29Jun 22, 2025Updated last year
- ☆15Feb 28, 2025Updated last year
- ☆11Apr 28, 2026Updated 4 months ago
- ☆11Nov 22, 2022Updated 3 years ago
- Implementation for paper:"Learning Global-Local Correspondence with Semantic Bottleneck for Logical Anomaly Detection"☆14Aug 13, 2023Updated 3 years ago
- [ICLR 2026]🌴 ARES is an open-source framework for adaptive multimodal reasoning, featuring a two-stage pipeline—Adaptive Cold-Start and …☆23Feb 3, 2026Updated 7 months ago
- Bare Metal GPUs on DigitalOcean Gradient AI • AdPurpose-built for serious AI teams training foundational models, running large-scale inference, and pushing the boundaries of what's possible.
- This repository contains the implementation for Anomaly Detection using Score-based Perturbation Resilience (ICCV 2023)☆14Sep 6, 2024Updated 2 years ago
- ☆10Aug 8, 2024Updated 2 years ago
- ☆32Dec 8, 2025Updated 9 months ago
- ☆29Aug 8, 2025Updated last year
- [ACM MM 2023] QA-CLIMS: Question-Answer Cross Language Image Matching for Weakly Supervised Semantic Segmentation☆13Jun 14, 2024Updated 2 years ago
- ☆17Oct 31, 2024Updated last year
- Interactive, high-performance 3D visualization app. With Computer Vision in mind.☆14Mar 18, 2023Updated 3 years ago
- ☆10Jun 20, 2025Updated last year
- v1: Learning to Point Visual Tokens for Multimodal Grounded Reasoning☆21Updated this week
- Bare Metal GPUs on DigitalOcean Gradient AI • AdPurpose-built for serious AI teams training foundational models, running large-scale inference, and pushing the boundaries of what's possible.
- ☆22Mar 11, 2025Updated last year
- (ICLR 2025 Spotlight) Official code repository for Interleaved Scene Graph.☆31Aug 7, 2025Updated last year
- Official Implementation for the paper "Integrative Decoding: Improving Factuality via Implicit Self-consistency"☆33Apr 12, 2025Updated last year
- The official implement of "Grounded Chain-of-Thought for Multimodal Large Language Models"☆27Jul 21, 2025Updated last year
- ☆47Jul 14, 2025Updated last year
- Scaling Multi-modal Instruction Fine-tuning with Tens of Thousands Vision Task Types☆32Jul 16, 2025Updated last year
- The implementation for "AUD-Net: A Unified Deep Detector for Multiple Hyperspectral Image Anomaly Detection via Relation and Few-Shot Lea…☆15Nov 5, 2022Updated 3 years ago