Cheap and reliable Node.js hosting starts at $3/month, and $1/month static HTML hosting

A small lightweight HTTP server that converts photos, images and scanned documents to text using optical character recognition by utilizing the power of Google Tesseract.

Stars: ✭ 15 (-94.6%)

Mutual labels: ocr

namsel

An OCR application focused on machine-print Tibetan text

Stars: ✭ 22 (-92.09%)

Mutual labels: ocr

Ionic Ocr Example

📷 Simple Ionic app using ocrad.js

Stars: ✭ 263 (-5.4%)

Mutual labels: ocr

idcardocr

离线环境下第二代居民身份证信息识别

Stars: ✭ 358 (+28.78%)

Mutual labels: ocr

smart-docs-parser

An OCR based document parser to extract information from identity document images

Stars: ✭ 14 (-94.96%)

Mutual labels: ocr

View All Similar Projects ➔

attention-ocr.pytorch:Encoder+Decoder+attention model

This repository implements the the encoder and decoder model with attention model for OCR, the encoder uses CNN+Bi-LSTM, the decoder uses GRU. This repository is modified from https://github.com/meijieru/crnn.pytorch
Earlier I had an open source version, but had some problems identifying images of fixed width. Recently I modified the model to support image recognition with variable width. The function is the same as CRNN. Due to the time problem, there is no pre-training model this time, which will be updated later.

requirements

pytorch 0.4.1
opencv_python

cd Attention_ocr.pytorch
pip install -r requirements.txt

Test

pretrained model coming soon

Train

Here i choose a small dataset from Synthetic_Chinese_String_Dataset, about 270000+ images for training, 20000 images for testing. download the image data from Baidu
the train_list.txt and test_list.txt are created as the follow form:

# path/to/image_name.jpg label
path/AttentionData/50843500_2726670787.jpg 情笼罩在他们满是沧桑
path/AttentionData/57724421_3902051606.jpg 心态的松弛决定了比赛
path/AttentionData/52041437_3766953320.jpg 虾的鲜美自是不可待言

change the trainlist and vallist parameter in train.py, and start train

cd Attention_ocr.pytorch
python train.py --trainlist ./data/ch_train.txt --vallist ./data/ch_test.txt

then you can see in the terminel as follow: there uses the decoderV2 model for decoder.

The previous version

git checkout AttentionOcrV1

Reference

TO DO

[ ] change LSTM to Conv1D, it can greatly accelerate the inference
[ ] change the cnn bone model with inception net, densenet
[ ] realize the decoder with transformer model

Note that the project description data, including the texts, logos, images, and/or trademarks, for each open source project belongs to its rightful owner. If you wish to add or remove any projects, please contact us at [email protected].

Stars: ✭ 278

Visit Git Page 🔗Visit User Page 🔗Visit Issues Page (28) 🔗