TLGAN: document Text Localization using Generative Adversarial Nets

10/22/2020
by   Dongyoung Kim, et al.
0

Text localization from the digital image is the first step for the optical character recognition task. Conventional image processing based text localization performs adequately for specific examples. Yet, a general text localization are only archived by recent deep-learning based modalities. Here we present document Text Localization Generative Adversarial Nets (TLGAN) which are deep neural networks to perform the text localization from digital image. TLGAN is an versatile and easy-train text localization model requiring a small amount of data. Training only ten labeled receipt images from Robust Reading Challenge on Scanned Receipts OCR and Information Extraction (SROIE), TLGAN achieved 99.83 practical text localization solution requiring minimal effort for data labeling and model training and producing a state-of-art performance.

READ FULL TEXT

page 3

page 6

research
12/11/2022

Extending TrOCR for Text Localization-Free OCR of Full-Page Scanned Receipt Images

Digitization of scanned receipts aims to extract text from receipt image...
research
06/25/2020

Cascading Modular U-Nets for Document Image Binarization

In recent years, U-Net has achieved good results in various image proces...
research
04/14/2015

Efficient Scene Text Localization and Recognition with Local Character Refinement

An unconstrained end-to-end text localization and recognition method is ...
research
11/06/2014

Conditional Generative Adversarial Nets

Generative Adversarial Nets [8] were recently introduced as a novel way ...
research
08/04/2023

CTP-Net: Character Texture Perception Network for Document Image Forgery Localization

Due to the progression of information technology in recent years, docume...
research
04/16/2021

TeLCoS: OnDevice Text Localization with Clustering of Script

Recent research in the field of text localization in a resource constrai...
research
02/15/2018

Fooling OCR Systems with Adversarial Text Images

We demonstrate that state-of-the-art optical character recognition (OCR)...

Please sign up or login with your details

Forgot password? Click here to reset