Vision-Language Pre-Training for Boosting Scene Text Detectors (CVPR2022)
☆12Mar 21, 2022Updated 4 years ago
Alternatives and similar repositories for VLPT-STD
Users that are interested in VLPT-STD are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- NLPBench: Evaluating NLP-Related Problem-solving Ability in Large Language Models☆10Oct 27, 2023Updated 2 years ago
- ☆13Mar 14, 2022Updated 4 years ago
- Official implementation of SPTS: Single-Point Text Spotting (ACM MM 2022 Oral)☆145Jul 26, 2023Updated 3 years ago
- Thai font for amazfit bip☆12Jul 27, 2018Updated 8 years ago
- ☆13Jun 10, 2025Updated last year
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- Searching a High Performance Feature Extractor for Text Recognition Network. TPAMI 2022☆13Nov 25, 2022Updated 3 years ago
- 这里将paddle中的ocr等模型转为onnx格式,并利用java版深度框架djl加载这些onnx模型进行推理预测尝试。☆14Nov 15, 2022Updated 3 years ago
- Environment Predictive Coding for Visual Navigation. ICLR 2022.☆15Dec 10, 2022Updated 3 years ago
- [ICCV 2025] FVP: 4D Visual Pre-training for Robot Learning☆17Sep 5, 2025Updated last year
- This is a warehouse for semantic segmentation models, can be used to train your image-datasets for segmentation tasks.☆14Feb 18, 2025Updated last year
- Papers, Datasets, Algorithms, SOTA for STR. Long-time Maintaining☆354Nov 29, 2023Updated 2 years ago
- Unofficial Pytorch Implementation Of AdversarialAutoAugment(ICLR2020)☆21Feb 9, 2021Updated 5 years ago
- Self-attention based Text Knowledge Mining for Text Detection☆47Mar 7, 2023Updated 3 years ago
- ☆15Aug 22, 2020Updated 6 years ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- [ICCV 2023] MADAug: When to Learn What: Model-Adaptive Data Augmentation Curriculum☆20Nov 9, 2023Updated 2 years ago
- The HierText dataset contains ~12k images from the Open Images dataset v6 with large amount of text entities. We provide word, line and p…☆316Dec 2, 2024Updated last year
- Label smoothed Aggregation cross entropy loss for generalisation in sequence to sequence tasks.☆14Dec 17, 2019Updated 6 years ago
- [ICLR 2026] SpikePingpong: Spike Vision-based Fast-Slow Pingpong Robot System☆24Mar 13, 2026Updated 6 months ago
- Arbitrary Shape Text Detection via Segmentation with Probability Maps; accepted by TPAMI2022☆104Jun 30, 2023Updated 3 years ago
- [WAVC 2024] Official implementation of the paper: Semantic Generative Augmentations for Few-shot Counting☆13May 1, 2024Updated 2 years ago
- ABINet++: Autonomous, Bidirectional and Iterative Language Modeling for Scene Text Spotting☆90Feb 11, 2023Updated 3 years ago
- pytorch大规模数据读取dataset☆13May 30, 2022Updated 4 years ago
- Official Implementation (Pytorch) of "Super-class guided Transformer for Zero-Shot Attribute Classification", AAAI 2025☆15Jan 15, 2025Updated last year
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- ICME2022 Special Session “Beyond Accuracy: Responsible, Responsive, and Robust Multimedia Retrieval ”☆12Jun 3, 2024Updated 2 years ago
- ☆13Sep 25, 2023Updated 3 years ago
- Pytorch re-implementation of Paper: SwinTextSpotter: Scene Text Spotting via Better Synergy between Text Detection and Text Recognition (…☆287Nov 29, 2024Updated last year
- Code for SEEG: Semantic Energized Co-speech Gesture Generation☆33Dec 3, 2022Updated 3 years ago
- Code for "Translatotron-V(ison): An End-to-End Model for In-Image Machine Translation" (Findings of ACL 2024)☆16Jul 4, 2024Updated 2 years ago
- init☆11Sep 30, 2017Updated 8 years ago
- [ECCV2022] The PyTorch implementation of paper "Equivariance and Invariance Inductive Bias for Learning from Insufficient Data"☆19Oct 12, 2022Updated 3 years ago
- ☆15Nov 26, 2023Updated 2 years ago
- PyTorch implementation of "Segmenter: Transformer for Semantic Segmentation" Strudel et al. (2021)☆17May 23, 2021Updated 5 years ago
- Simple, predictable pricing with DigitalOcean hosting • AdAlways know what you'll pay with monthly caps and flat pricing. Enterprise-grade infrastructure trusted by 600k+ customers.
- ☆19Mar 28, 2022Updated 4 years ago
- Resizable rect item derived from QGraphicsRectItem for use in a QGraphicsScene.☆15Oct 30, 2016Updated 9 years ago
- SHA-3 cpu and gpu (CUDA) calculation☆17Jun 25, 2018Updated 8 years ago
- ☆16Jul 2, 2024Updated 2 years ago
- Official code implementation of " TextDiff: Mask-Guided Residual Diffusion Models for Scene Text Image " in Pattern Recognition☆25Apr 24, 2024Updated 2 years ago
- ☆13Sep 25, 2019Updated 7 years ago
- This dataset contains re-annotations of 4 popular Latin/English scene text recognition datasets.☆53Mar 24, 2020Updated 6 years ago