Cheap and reliable Node.js hosting starts at $3/month, and $1/month static HTML hosting

Created with love in Canada, visit hostnodejs.com today

Feel like to post an Ad? Learn Details

All Projects → clovaai → Tedeval

clovaai / Tedeval

Licence: mit

TedEval: A Fair Evaluation Metric for Scene Text Detectors

Programming Languages

python

139335 projects - #7 most used programming language

Labels

ocr evaluation text-detection

Projects that are alternatives of or similar to Tedeval

React Native Tesseract Ocr

Tesseract OCR wrapper for React Native

Stars: ✭ 384 (+168.53%)

Mutual labels: text-detection, ocr

Tensorflow psenet

This is a tensorflow re-implementation of PSENet: Shape Robust Text Detection with Progressive Scale Expansion Network.My blog:

Stars: ✭ 472 (+230.07%)

Mutual labels: text-detection, ocr

Psenet.pytorch

A pytorch re-implementation of PSENet: Shape Robust Text Detection with Progressive Scale Expansion Network

Stars: ✭ 416 (+190.91%)

Mutual labels: text-detection, ocr

Chineseaddress ocr

Photographing Chinese-Address OCR implemented using CTPN+CTC+Address Correction. 拍照文档中文地址文字识别。

Stars: ✭ 309 (+116.08%)

Mutual labels: text-detection, ocr

Ctpn

Detecting Text in Natural Image with Connectionist Text Proposal Network (ECCV'16)

Stars: ✭ 1,220 (+753.15%)

Mutual labels: text-detection, ocr

Megreader

A research project for text detection and recognition using PyTorch 1.2.

Stars: ✭ 332 (+132.17%)

Mutual labels: text-detection, ocr

Craft Remade

Implementation of CRAFT Text Detection

Stars: ✭ 127 (-11.19%)

Mutual labels: text-detection, ocr

vietnamese-ocr-toolbox

A toolbox for Vietnamese Optical Character Recognition.

Stars: ✭ 26 (-81.82%)

Mutual labels: ocr, text-detection

Image Text Localization Recognition

A general list of resources to image text localization and recognition 场景文本位置感知与识别的论文资源与实现合集シーンテキストの位置認識と識別のための論文リソースの要約

Stars: ✭ 788 (+451.05%)

Mutual labels: text-detection, ocr

Keras Ocr

A packaged and flexible version of the CRAFT text detector and Keras CRNN recognition model.

Stars: ✭ 782 (+446.85%)

Mutual labels: text-detection, ocr

Text Detection Ctpn

text detection mainly based on ctpn model in tensorflow, id card detect, connectionist text proposal network

Stars: ✭ 3,242 (+2167.13%)

Mutual labels: text-detection, ocr

Differentiablebinarization

DB (Real-time Scene Text Detection with Differentiable Binarization) implementation in Keras and Tensorflow

Stars: ✭ 106 (-25.87%)

Mutual labels: text-detection, ocr

PSENet-Tensorflow

TensorFlow implementation of PSENet text detector (Shape Robust Text Detection with Progressive Scale Expansion Networkt)

Stars: ✭ 51 (-64.34%)

Mutual labels: ocr, text-detection

Awesome Ocr Resources

A collection of resources (including the papers and datasets) of OCR (Optical Character Recognition).

Stars: ✭ 335 (+134.27%)

Mutual labels: text-detection, ocr

craft-text-detector

Packaged, Pytorch-based, easy to use, cross-platform version of the CRAFT text detector

Stars: ✭ 151 (+5.59%)

Mutual labels: ocr, text-detection

Dbnet.pytorch

A pytorch re-implementation of Real-time Scene Text Detection with Differentiable Binarization

Stars: ✭ 435 (+204.2%)

Mutual labels: text-detection, ocr

pytorch.ctpn

pytorch, ctpn ,text detection ,ocr,文本检测

Stars: ✭ 123 (-13.99%)

Mutual labels: ocr, text-detection

doctr

docTR (Document Text Recognition) - a seamless, high-performing & accessible library for OCR-related tasks powered by Deep Learning.

Stars: ✭ 1,409 (+885.31%)

Mutual labels: ocr, text-detection

Seglink

An Implementation of the seglink alogrithm in paper Detecting Oriented Text in Natural Images by Linking Segments

Stars: ✭ 479 (+234.97%)

Mutual labels: text-detection, ocr

Keras Ctpn

keras复现场景文本检测网络CPTN: 《Detecting Text in Natural Image with Connectionist Text Proposal Network》；欢迎试用，关注，并反馈问题...

Stars: ✭ 89 (-37.76%)

Mutual labels: text-detection, ocr

View All Similar Projects ➔

TedEval: A Fair Evaluation Metric for Scene Text Detectors

Official Python 3 implementation of TedEval | paper | slides

Chae Young Lee, Youngmin Baek, and Hwalsuk Lee.

Clova AI Research, NAVER Corp.

Overview

We propose a new evaluation metric for scene text detectors called TedEval. Through separate instance-level matching policy and character-level scoring policy, TedEval solves the drawbacks of previous metrics such as IoU and DetEval. This code is based on ICDAR15 official evaluation code.

Methodology

1. Mathcing Policy

Non-exclusively gathers all possible matches of not only one-to-one but also one-to-many and many-to-one.
The threshold of both area recall and area precision are set to 0.4.
Multiline is identified and rejected when |min(theta, 180 - theta)| > 45 from Fig. 2.

2. Scoring Policy

We compute Pseudo Character Center (PCC) from word-level bounding boxes and penalize matches when PCCs are missing or overlapping.

Sample Evaluation

Experiments

We evaluated state-of-the-art scene text detectors with TedEval on two benchmark datasets: ICDAR 2013 Focused Scene Text (IC13) and ICDAR 2015 Incidental Scene Text (IC15). Detectors are listed in the order of published dates.

ICDAR 2013

Detector	Date (YY/MM/DD)	Recall (%)	Precision (%)	H-mean (%)
CTPN	16/09/12	82.1	92.7	87.6
RRPN	17/03/03	89.0	94.2	91.6
SegLink	17/03/19	65.6	74.9	70.0
EAST	17/04/11	77.7	87.1	82.5
WordSup	17/08/22	87.5	92.2	90.2
PixelLink	18/01/04	84.0	87.2	86.1
FOTS	18/01/05	91.5	93.0	92.6
TextBoxes++	18/01/09	87.4	92.3	90.0
MaskTextSpotter	18/07/06	90.2	95.4	92.9
PMTD	19/03/28	94.0	95.2	94.7
CRAFT	19/04/03	93.6	96.5	95.1

ICDAR 2015

Detector	Date (YY/MM/DD)	Recall (%)	Precision (%)	H-mean (%)
CTPN	16/09/12	85.0	81.1	67.8
RRPN	17/03/03	79.5	85.9	82.6
SegLink	17/03/19	77.1	83.9	80.6
EAST	17/04/11	82.5	90.0	86.3
WordSup	17/08/22	83.2	87.1	85.2
PixelLink	18/01/04	85.7	86.1	86.0
FOTS	18/01/05	89.0	93.4	91.2
TextBoxes++	18/01/09	82.4	90.8	86.5
MaskTextSpotter	18/07/06	82.5	91.8	86.9
PMTD	19/03/28	89.2	92.8	91.0
CRAFT	19/04/03	88.5	93.1	90.9

Frequency

Getting Started

Clone repository

git clone https://github.com/clovaai/TedEval.git

Requirements

python 3
python 3.x Polygon, Bottle, Pillow

# install
pip3 install Polygon3 bottle Pillow

Supported Annotation Type

LTRB(xmin, ymin, xmax, ymax)
QUAD(x1, y1, x2, y2, x3, y3, x4, y4)

Evaluation

Prepare data

The ground truth and the result data should be text files, one for each sample. Note that the naming rule of each text file is that there must be img_{number} in the filename and that the number indicate the image sample.

# gt/gt_img_38.txt
644,101,932,113,932,168,643,156,[email protected]
477,138,487,139,488,149,477,148,###
344,131,398,130,398,149,344,149,###
1195,148,1277,138,1277,177,1194,187,###
23,270,128,267,128,282,23,284,###

# result/res_img_38.txt
644,101,932,113,932,168,643,156,{Transcription},{Confidence}
477,138,487,139,488,149,477,148
344,131,398,130,398,149,344,149
1195,148,1277,138,1277,177,1194,187
23,270,128,267,128,282,23,284

Compress these text files.

zip gt.zip gt/*
zip result.zip result/*

Refer to gt/result.zip and gt/gt_*.zip for examples.

Run stand-alone evaluation

python script.py –g=gt/gt.zip –s=result/result.zip

Locate the path of GT and submission file using the flag -g and -s, respectively.
QUAD annotation type is used as default. To switch between {QUAD, LTRB}, add -p='{"LTRB" = False}' in the command or directly modify the default_evaluation_params() function in script.py.
If there are transcription or confidence values in your submission file, add -p='{"CONFIDENCES" = True} or -p='{"TRANSCRIPTION" = True}'.

Run Visualizer

python web.py

Place the zip file of images and GTs of the dataset named images.zip and gt.zip, respectively, in the gt directory.
Create an empty directory name output. This is where the DB, submission files, and result files will be created.
You can change the host and port number in the final line of web.py.

The file structure should then be:

.
├── gt
│   ├── gt.zip
│   └── images.zip
├── output   # empty dir
├── script.py
├── web.py
├── README.md
└── ...

Citation

@article{lee2019tedeval,
  title={TedEval: A Fair Evaluation Metric for Scene Text Detectors},
  author={Lee, Chae Young and Baek, Youngmin and Lee, Hwalsuk},
  journal={arXiv preprint arXiv:1907.01227},
  year={2019}
}

Contact us

We welcome any feedbacks to our metric. Please contact the authors via {cylee7133, youngmin.baek, hwalsuk.lee}@gmail.com. In case of code errors, open an issue and we will get to you.

License

Copyright (c) 2019-present NAVER Corp.

 Permission is hereby granted, free of charge, to any person obtaining a copy
of this software and associated documentation files (the "Software"), to deal
in the Software without restriction, including without limitation the rights
to use, copy, modify, merge, publish, distribute, sublicense, and/or sell
copies of the Software, and to permit persons to whom the Software is
furnished to do so, subject to the following conditions:

 The above copyright notice and this permission notice shall be included in
all copies or substantial portions of the Software.

 THE SOFTWARE IS PROVIDED "AS IS", WITHOUT WARRANTY OF ANY KIND, EXPRESS OR
IMPLIED, INCLUDING BUT NOT LIMITED TO THE WARRANTIES OF MERCHANTABILITY,
FITNESS FOR A PARTICULAR PURPOSE AND NONINFRINGEMENT.  IN NO EVENT SHALL THE
AUTHORS OR COPYRIGHT HOLDERS BE LIABLE FOR ANY CLAIM, DAMAGES OR OTHER
LIABILITY, WHETHER IN AN ACTION OF CONTRACT, TORT OR OTHERWISE, ARISING FROM,
OUT OF OR IN CONNECTION WITH THE SOFTWARE OR THE USE OR OTHER DEALINGS IN
THE SOFTWARE.

Note that the project description data, including the texts, logos, images, and/or trademarks, for each open source project belongs to its rightful owner. If you wish to add or remove any projects, please contact us at [email protected].

Stars: ✭ 143

Visit Git Page 🔗Visit User Page 🔗Visit Issues Page (0) 🔗