SOTAVerified

Scene Text Recognition

See Scene Text Detection for leaderboards in this task.

Papers

Showing 151200 of 269 papers

TitleStatusHype
Visual attention models for scene text recognition0
Edit Probability for Scene Text Recognition0
Visual-Semantic Transformer for Scene Text Recognition0
VL-Reader: Vision and Language Reconstructor is an Effective Scene Text Recognizer0
Weakly Supervised Scene Text Generation for Low-resource Languages0
What Machines See Is Not What They Get: Fooling Scene Text Recognition Models With Adversarial Text Images0
FACLSTM: ConvLSTM with Focused Attention for Scene Text Recognition0
Efficient and Accurate Scene Text Recognition with Cascaded-Transformers0
Efficiently Leveraging Linguistic Priors for Scene Text Spotting0
Enhancing Energy Minimization Framework for Scene Text Recognition with Top-Down Cues0
ESIR: End-to-end Scene Text Recognition via Iterative Image Rectification0
A fine-grained approach to scene text script identification0
Exploiting Local Structures with the Kronecker Layer in Convolutional Networks0
Exploring Font-independent Features for Scene Text Recognition0
FedOCR: Communication-Efficient Federated Learning for Scene Text Recognition0
FEDS -- Filtered Edit Distance Surrogate0
Focusing Attention: Towards Accurate Text Recognition in Natural Images0
A Hardware-Oriented and Memory-Efficient Method for CTC Decoding0
Efficient Online ML API Selection for Multi-Label Classification Tasks0
Generative Shape Models: Joint Text Recognition and Segmentation with Very Little Training Data0
2D-CTC for Scene Text Recognition0
GTC: Guided Training of CTC Towards Efficient and Accurate Scene Text Recognition0
HAAP: Vision-context Hierarchical Attention Autoregressive with Adaptive Permutation for Scene Text Recognition0
Hamming OCR: A Locality Sensitive Hashing Neural Network for Scene Text Recognition0
I2C2W: Image-to-Character-to-Word Transformers for Accurate Scene Text Recognition0
IFR: Iterative Fusion Based Recognizer For Low Quality Scene Text Recognition0
Improving Long Handwritten Text Line Recognition with Convolutional Multi-way Associative Memory0
Improving Scene Text Recognition for Character-Level Long-Tailed Distribution0
IndicSTR12: A Dataset for Indic Scene Text Recognition0
A New Perspective for Flexible Feature Gathering in Scene Text Recognition Via Character Anchor Pooling0
JSTR: Judgment Improves Scene Text Recognition0
Learning Surrogates via Deep Embedding0
Arbitrary Reading Order Scene Text Spotter with Local Semantics Guidance0
Lumos : Empowering Multimodal LLMs with Scene Text Recognition0
Accurate Scene Text Recognition with Efficient Model Scaling and Cloze Self-Distillation0
Augmented Transformers with Adaptive n-grams Embedding for Multilingual Scene Text Recognition0
Benchmarking Scene Text Recognition in Devanagari, Telugu and Malayalam0
Memory Matters: Convolutional Recurrent Neural Network for Scene Text Recognition0
A CNN Based Scene Chinese Text Recognition Algorithm With Synthetic Data Engine0
Billet Number Recognition Based on Test-Time Adaptation0
Boosting High-Level Vision with Joint Compression Artifacts Reduction and Super-Resolution0
Multilingual Scene Character Recognition System using Sparse Auto-Encoder for Efficient Local Features Representation in Bag of Features0
1st Place Solution to ECCV 2022 Challenge on Out of Vocabulary Scene Text Understanding: End-to-End Recognition of Out of Vocabulary Words0
On Calibration of Scene-Text Recognition Models0
One Model for Two Tasks: Cooperatively Recognizing and Recovering Low-Resolution Scene Text Images by Iterative Mutual Guidance0
On Vocabulary Reliance in Scene Text Recognition0
Open-Vocabulary Scene Text Recognition via Pseudo-Image Labeling and Margin Loss0
Optimal Boxes: Boosting End-to-End Scene Text Recognition by Adjusting Annotated Bounding Boxes via Reinforcement Learning0
Oracle Teacher: Leveraging Target Information for Better Knowledge Distillation of CTC Models0
Out-of-Vocabulary Challenge Report0
Show:102550
← PrevPage 4 of 6Next →

Benchmark Results

#ModelMetricClaimedVerifiedStatus
1CLIP4STR-L*Accuracy99.42Unverified
2DTrOCR 105MAccuracy99.4Unverified
3CLIP4STR-L (DataComp-1B)Accuracy99Unverified
4CLIP4STR-LAccuracy98.5Unverified
5MGP-STRAccuracy98.5Unverified
6CCD-ViT-Small(ARD_2.8M)Accuracy98.3Unverified
7CLIP4STR-BAccuracy98.3Unverified
8CCD-ViT-Base(ARD_2.8M)Accuracy98.3Unverified
9MATRNAccuracy97.9Unverified
10S-GTRAccuracy97.8Unverified
#ModelMetricClaimedVerifiedStatus
1CLIP4STR-H (DFN-5B)Accuracy99.1Unverified
2DTrOCR 105MAccuracy98.9Unverified
3CLIP4STR-B*Accuracy98.76Unverified
4CLIP4STR-L (DataComp-1B)Accuracy98.6Unverified
5MGP-STRAccuracy98.6Unverified
6CLIP4STR-LAccuracy98.5Unverified
7CPPDAccuracy98.5Unverified
8CLIP4STR-BAccuracy98.3Unverified
9CCD-ViT-Base(ARD_2.8M)Accuracy97.8Unverified
10CCD-ViT-Small(ARD_2.8M)Accuracy96.4Unverified
#ModelMetricClaimedVerifiedStatus
1DTrOCR 105MAccuracy93.5Unverified
2CLIP4STR-L*Accuracy92.6Unverified
3CPPDAccuracy91.7Unverified
4CLIP4STR-L (DataComp-1B)Accuracy91.4Unverified
5MGP-STRAccuracy90.9Unverified
6CLIP4STR-LAccuracy90.8Unverified
7CLIP4STR-BAccuracy90.6Unverified
8SIGA_SAccuracy87.6Unverified
9S-GTRAccuracy87.3Unverified
10MATRNAccuracy86.6Unverified
#ModelMetricClaimedVerifiedStatus
1CPPDAccuracy99.7Unverified
2CLIP4STR-L (DataComp-1B)Accuracy99.7Unverified
3CLIP4STR-B*Accuracy99.65Unverified
4MGP-STRAccuracy99.31Unverified
5CLIP4STR-BAccuracy99.3Unverified
6DTrOCR 105MAccuracy99.1Unverified
7CLIP4STR-LAccuracy99Unverified
8CCD-ViT-Base(ARD_2.8M)Accuracy98.3Unverified
9CCD-ViT-Small(ARD_2.8M)Accuracy98.3Unverified
10CCD-ViT-Tiny(ARD_2.8M)Accuracy95.8Unverified
#ModelMetricClaimedVerifiedStatus
1CLIP4STR-L (DataComp-1B)Accuracy99.6Unverified
2DTrOCR 105MAccuracy99.6Unverified
3CLIP4STR-B (DataComp-1B)Accuracy99.5Unverified
4CLIP4STR-LAccuracy99.5Unverified
5CPPDAccuracy99.3Unverified
6CLIP4STR-BAccuracy99.2Unverified
7MGP-STRAccuracy98.8Unverified
8CCD-ViT-Base(ARD_2.8M)Accuracy98Unverified
9CCD-ViT-Small(ARD_2.8M)Accuracy98Unverified
10S-GTRAccuracy97.5Unverified
#ModelMetricClaimedVerifiedStatus
1DTrOCR 105MAccuracy98.6Unverified
2MGP-STRAccuracy98.3Unverified
3CLIP4STR-L*Accuracy98.13Unverified
4CLIP4STR-L (DataComp-1B)Accuracy98.1Unverified
5CLIP4STR-LAccuracy97.4Unverified
6CLIP4STR-BAccuracy97.2Unverified
7CPPDAccuracy96.7Unverified
8CCD-ViT-BaseAccuracy96.1Unverified
9CCD-ViT-SmallAccuracy92.7Unverified
10CCD-ViT-TinyAccuracy91.6Unverified
#ModelMetricClaimedVerifiedStatus
1Yet Another Text RecognizerAccuracy97.1Unverified
2SIGA_TAccuracy97Unverified
3SATRNAccuracy96.7Unverified
4DANAccuracy95Unverified
5SAFLAccuracy95Unverified
6CSTRAccuracy94.8Unverified
7Baek et al.Accuracy94.4Unverified
8ViTSTRAccuracy94.3Unverified
9AONAccuracy91.5Unverified
10RAREAccuracy90.1Unverified
#ModelMetricClaimedVerifiedStatus
1CLIP4STR-H (DFN-5B)1:1 Accuracy90.9Unverified
2CLIP4STR-L (DataComp-1B)1:1 Accuracy90.6Unverified
3CLIP4STR-L1:1 Accuracy88.8Unverified
4CLIP4STR-B1:1 Accuracy87Unverified
5CCD-ViT-Base1:1 Accuracy86Unverified
#ModelMetricClaimedVerifiedStatus
1CLIP4STR-L (DataComp-1B)Accuracy (%)86.4Unverified
2CLIP4STR-LAccuracy (%)85.9Unverified
3CLIP4STR-BAccuracy (%)85.8Unverified
4MGP-STRAccuracy (%)85.5Unverified
#ModelMetricClaimedVerifiedStatus
1CLIP4STR-L1:1 Accuracy81.9Unverified
2MGP-STR1:1 Accuracy81.7Unverified
3CLIP4STR-B1:1 Accuracy81.1Unverified
#ModelMetricClaimedVerifiedStatus
1CLIP4STR-L1:1 Accuracy82.7Unverified
2CLIP4STR-B1:1 Accuracy79.8Unverified
3CCD-ViT-Base1:1 Accuracy77.3Unverified
#ModelMetricClaimedVerifiedStatus
1CLIP4STR-L (DataComp-1B)Accuracy (%)92.2Unverified
2MGP-STRAccuracy (%)91Unverified
3CLIP4STR-BAccuracy (%)86.8Unverified
#ModelMetricClaimedVerifiedStatus
1ABINet-LV+TPS++Accuracy97.8Unverified
#ModelMetricClaimedVerifiedStatus
1MLDGAverage Accuracy19.02Unverified
#ModelMetricClaimedVerifiedStatus
1ABINet-LV+TPS++Accuracy89.6Unverified