Learning Robust Feature Representations for Scene Text Detection

05/26/2020
by   Sihwan Kim, et al.
0

Scene text detection based on deep neural networks have progressed substantially over the past years. However, previous state-of-the-art methods may still fall short when dealing with challenging public benchmarks because the performances of algorithm are determined by the robust features extraction and components in network architecture. To address this issue, we will present a network architecture derived from the loss to maximize conditional log-likelihood by optimizing the lower bound with a proper approximate posterior that has shown impressive performance in several generative models. In addition, by extending the layer of latent variables to multiple layers, the network is able to learn robust features on scale with no task-specific regularization or data augmentation. We provide a detailed analysis and show the results on three public benchmark datasets to confirm the efficiency and reliability of the proposed algorithm. In experiments, the proposed algorithm significantly outperforms state-of-the-art methods in terms of both recall and precision. Specifically, it achieves an H-mean of 95.12 and 96.78 on ICDAR 2011 and ICDAR 2013, respectively.

READ FULL TEXT

page 5

page 7

research
04/11/2017

EAST: An Efficient and Accurate Scene Text Detector

Previous approaches for scene text detection have already achieved promi...
research
02/27/2017

Learning Hierarchical Features from Generative Models

Deep neural networks have been shown to be very successful at learning f...
research
11/21/2018

Scene Text Detection with Supervised Pyramid Context Network

Scene text detection methods based on deep learning have achieved remark...
research
06/15/2022

Physically-admissible polarimetric data augmentation for road-scene analysis

Polarimetric imaging, along with deep learning, has shown improved perfo...
research
11/11/2017

Deep Residual Text Detection Network for Scene Text

Scene text detection is a challenging problem in computer vision. In thi...
research
09/01/2020

Generalized Zero-Shot Learning via VAE-Conditioned Generative Flow

Generalized zero-shot learning (GZSL) aims to recognize both seen and un...
research
05/02/2015

Multi-Object Classification and Unsupervised Scene Understanding Using Deep Learning Features and Latent Tree Probabilistic Models

Deep learning has shown state-of-art classification performance on datas...

Please sign up or login with your details

Forgot password? Click here to reset