4.0 KiB
Executable File
Algorithm introduction
This tutorial lists the text detection algorithms and text recognition algorithms supported by PaddleOCR, as well as the models and metrics of each algorithm on English public datasets. It is mainly used for algorithm introduction and algorithm performance comparison. For more models on other datasets including Chinese, please refer to PP-OCR v2.0 models list.
1. Text Detection Algorithm
PaddleOCR open source text detection algorithms list:
On the ICDAR2015 dataset, the text detection result is as follows:
Model | Backbone | precision | recall | Hmean | Download link |
---|---|---|---|---|---|
EAST | ResNet50_vd | 88.76% | 81.36% | 84.90% | Download link |
EAST | MobileNetV3 | 78.24% | 79.15% | 78.69% | Download link |
DB | ResNet50_vd | 86.41% | 78.72% | 82.38% | Download link |
DB | MobileNetV3 | 77.29% | 73.08% | 75.12% | Download link |
SAST | ResNet50_vd | 91.83% | 81.80% | 86.52% | Download link |
On Total-Text dataset, the text detection result is as follows:
Model | Backbone | precision | recall | Hmean | Download link |
---|---|---|---|---|---|
SAST | ResNet50_vd | 89.05% | 76.80% | 82.47% | Download link |
Note: Additional data, like icdar2013, icdar2017, COCO-Text, ArT, was added to the model training of SAST. Download English public dataset in organized format used by PaddleOCR from Baidu Drive (download code: 2bpi).
For the training guide and use of PaddleOCR text detection algorithms, please refer to the document Text detection model training/evaluation/prediction
2. Text Recognition Algorithm
PaddleOCR open-source text recognition algorithms list:
- CRNN(paper)[7]
- Rosetta(paper)[10]
- STAR-Net(paper)[11] coming soon
- RARE(paper)[12] coming soon
- SRN(paper)[5] coming soon
Refer to DTRB, the training and evaluation result of these above text recognition (using MJSynth and SynthText for training, evaluate on IIIT, SVT, IC03, IC13, IC15, SVTP, CUTE) is as follow:
Model | Backbone | Avg Accuracy | Module combination | Download link |
---|---|---|---|---|
Rosetta | Resnet34_vd | 80.9% | rec_r34_vd_none_none_ctc | Download link |
Rosetta | MobileNetV3 | 78.05% | rec_mv3_none_none_ctc | Download link |
CRNN | Resnet34_vd | 82.76% | rec_r34_vd_none_bilstm_ctc | Download link |
CRNN | MobileNetV3 | 79.97% | rec_mv3_none_bilstm_ctc | Download link |
Please refer to the document for training guide and use of PaddleOCR text recognition algorithms Text recognition model training/evaluation/prediction