onealwj/MVLT

Readme badge preview -

If you own this repo, copy the snippet below and add it to your README.md

[![RelatedRepos](https://img.shields.io/badge/related-repos-yellow)](https://relatedrepos.com/gh/onealwj/MVLT)

onealwj / MVLT

PyTorch implementation of BMVC2022 paper Masked Vision-Language Transformers for Scene Text Recognition

☆28

Alternatives and similar repositories for MVLT

Users that are interested in MVLT are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.

Sorting:

DCGM / SoftCTC
View on GitHub
This repository contains source codes for SoftCTC. Original paper can be found here: https://arxiv.org/abs/2212.02135
☆19Mar 7, 2023Updated 3 years ago
amazon-science / textadain-robust-recognition
View on GitHub
TextAdaIN: Paying Attention to Shortcut Learning in Text Recognizers
☆21Jul 26, 2022Updated 3 years ago
Pay20Y / PIMNet
View on GitHub
☆16Jan 30, 2022Updated 4 years ago
amazon-science / semimtr-text-recognition
View on GitHub
Multimodal Semi-Supervised Learning for Text Recognition (SemiMTR)
☆83Sep 12, 2023Updated 2 years ago
ku21fan / STR-Fewer-Labels
View on GitHub
Scene Text Recognition (STR) methods trained with fewer real labels (CVPR 2021)
☆185Dec 23, 2023Updated 2 years ago
Managed hosting for WordPress and PHP on Cloudways • Ad
Managed hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
byeonghu-na / MATRN
View on GitHub
Official PyTorch implementation for Multi-modal Text Recognition Networks: Interactive Enhancements between Visual and Semantic Features …
☆74Jun 24, 2023Updated 3 years ago
gsoykan / comics_text_plus
View on GitHub
Official repository of the paper: "A Comprehensive Gold Standard and Benchmark for Comics Text Detection and Recognition"
☆26Jul 10, 2023Updated 3 years ago
ThunderVVV / RCLSTR
View on GitHub
Official PyTorch implementation of `[ACMMM 2023]Relational Contrastive Learning for Scene Text Recognition`
☆17Sep 22, 2023Updated 2 years ago
Canjie-Luo / Real-300K
View on GitHub
The dataset used in the CVPR 2022 paper (SimAN: Exploring Self-Supervised Representation Learning of Scene Text via Similarity-Aware Norm…
☆34Jun 21, 2022Updated 4 years ago
jfkuang / CFAM
View on GitHub
Contrast-guided Feature Adjustment Module for Visual Information Extraction
☆30May 23, 2023Updated 3 years ago
VDIGPKU / STR_TPSearch
View on GitHub
☆21Mar 15, 2022Updated 4 years ago
Actasidiot / EFIFSTR
View on GitHub
[ACM MM 2020] Exploring Font-independent Features for Scene Text Recognition
☆44Nov 30, 2020Updated 5 years ago
liuch37 / sar-pytorch
View on GitHub
Implementation of Show, Attend and Read: A Simple and Strong Baseline for Irregular Text Recognition published in AAAI 2019 in PyTorch
☆72Sep 3, 2020Updated 5 years ago
LARS-research / TREFE
View on GitHub
Searching a High Performance Feature Extractor for Text Recognition Network. TPAMI 2022
☆13Nov 25, 2022Updated 3 years ago
Proton VPN Special Offer - Get 70% off • Ad
Special partner offer. Trusted by over 100 million users worldwide. Tested, Approved and Recommended by Experts.
MelosY / CAM
View on GitHub
☆27Feb 20, 2024Updated 2 years ago
HCIILAB / LAST
View on GitHub
Read Ten Lines at One Glance: Line-Aware Semi-Autoregressive Transformer for Multi-Line Handwritten Mathematical Expression Recognition
☆28Aug 29, 2023Updated 2 years ago
zzyhlyoko / DCTC
View on GitHub
☆42Sep 2, 2023Updated 2 years ago
lcy0604 / CTRNet
View on GitHub
This repository is the implementation of "Don't Forget Me: Accurate Background Recovery for Text Removal via Modeling Local-Global Contex…
☆97Feb 21, 2023Updated 3 years ago
RuijieJ / pren
View on GitHub
Code for "Primitive Representation Learning for Scene Text Recognition" (CVPR 2021)
☆82May 11, 2022Updated 4 years ago
Caiyuan-Zheng / Consistency_Regularization_STR
View on GitHub
It's the code for the paper Pushing the Performance Limit of Scene Text Recognizer without Human Annotation, CVPR 2022.
☆28Jul 6, 2022Updated 4 years ago
joanrod / ocr-vqgan
View on GitHub
OCR-VQGAN, a discrete image encoder (tokenizer and detokenizer) for figure images in Paper2Fig100k dataset. Implementation of OCR Percept…
☆84Jan 30, 2023Updated 3 years ago
wangyuxin87 / VisionLAN
View on GitHub
A PyTorch implementation of "From Two to One: A New Scene Text Recognizer with Visual Language Modeling Network" (ICCV2021)
☆105Dec 9, 2021Updated 4 years ago
bytedance / oclip
View on GitHub
☆53Nov 4, 2022Updated 3 years ago
Serverless GPU API endpoints on Runpod - Get Bonus Credits • Ad
Skip the infrastructure headaches. Auto-scaling, pay-as-you-go, no-ops approach lets you focus on innovating your application.
bytedance / SPTSv2
View on GitHub
The official implementation of SPTS v2: Single-Point Text Spotting
☆138Jun 29, 2023Updated 3 years ago
xdxie / WordArt
View on GitHub
The official code of CornerTransformer (ECCV 2022, Oral) on top of MMOCR.
☆148Mar 6, 2023Updated 3 years ago
markytools / strexp
View on GitHub
STRExp is a framework that provides Explainability (XAI) to Scene Text Recognition (STR) models.
☆11Nov 27, 2023Updated 2 years ago
milely / SRN.Pytorch
View on GitHub
Unofficial implementation of Towards Accurate Scene Text Recognition with Semantic Reasoning Networks
☆28Sep 24, 2021Updated 4 years ago
mxin262 / ESTextSpotter
View on GitHub
(ICCV 2023) ESTextSpotter: Towards Better Scene Text Spotting with Explicit Synergy in Transformer
☆78Apr 9, 2024Updated 2 years ago
kartikgill / taco-box
View on GitHub
An implementation of Tiling and Corruption (TACo) Augmentations for OCR/HTR
☆15Dec 4, 2021Updated 4 years ago
csguoh / KD-LTR
View on GitHub
[MM2023] An official implement of the paper "One-stage Low-resolution Text Recognition with High-resolution Knowledge Transfer"
☆16Nov 3, 2023Updated 2 years ago
Xiaomeng-Yang / STR_benchmark_cleansed
View on GitHub
☆14May 26, 2023Updated 3 years ago
roatienza / deep-text-recognition-benchmark
View on GitHub
PyTorch code of my ICDAR 2021 paper Vision Transformer for Fast and Efficient Scene Text Recognition (ViTSTR)
☆312Apr 9, 2024Updated 2 years ago
Managed Database hosting by DigitalOcean • Ad
PostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
amazon-science / glass-text-spotting
View on GitHub
Official implementation for "GLASS: Global to Local Attention for Scene-Text Spotting" (ECCV'22)
☆102Jun 28, 2024Updated 2 years ago
Pay20Y / SEED
View on GitHub
☆164Dec 8, 2022Updated 3 years ago
Mountchicken / Text-Recognition-on-Cross-Domain-Datasets
View on GitHub
Improved Text recognition algorithms on different text domains like scene text, handwritten, document, Chinese/English, even ancient book…
☆80Feb 4, 2023Updated 3 years ago
Planet-AI-GmbH / tfaip-hybrid-ctc-s2s
View on GitHub
Repository sharing code and the model for the paper "Rescoring Sequence-to-Sequence Models for Text Line Recognition with CTC-Prefixes"
☆17Oct 13, 2021Updated 4 years ago
facebookresearch / MultiplexedOCR
View on GitHub
Code for CVPR21 paper A Multiplexed Network for End-to-End, Multilingual OCR
☆80Dec 2, 2022Updated 3 years ago
ViTAE-Transformer / ViTAE-Transformer-Scene-Text-Detection
View on GitHub
A comprehensive list [Hi-SAM@TPAMI'24, GoMatching@NeurIPS'24, DeepSolo(++)@ CVPR'23, DPText-DETR@AAAI'23, I3CL@IJCV'22] of our research w…
☆94Nov 12, 2024Updated last year
namtuanly / WikiTableSet
View on GitHub
WikiTableSet: A largest publicly available image-based table recognition dataset in three languages built from Wikipedia
☆32Jun 12, 2025Updated last year