A limited-size ensemble of homogeneous CNN/LSTMs for high-performance word classification

Ameryan Mahya; Schomaker Lambert

首页> 外文期刊>Neural computing & applications >A limited-size ensemble of homogeneous CNN/LSTMs for high-performance word classification

【24h】

A limited-size ensemble of homogeneous CNN/LSTMs for high-performance word classification

机译：A limited-size ensemble of homogeneous CNN/LSTMs for high-performance word classification

获取原文

获取原文并翻译 | 示例

获取外文期刊封面封底 >>

开具论文收录证明 >>

文献代查 >>

页面导航

摘要
著录项
相关主题

摘要

The strength of long short-term memory neural networks (LSTMs) that have been applied is more located in handling sequences of variable length than in handling geometric variability of the image patterns. In this paper, an end-to-end convolutional LSTM neural network is used to handle both geometric variation and sequence variability. The best results for LSTMs are often based on large-scale training of an ensemble of network instances. We show that high performances can be reached on a common benchmark set by using proper data augmentation for just five such networks using a proper coding scheme and a proper voting scheme. The networks have similar architectures (convolutional neural network (CNN): five layers, bidirectional LSTM (BiLSTM): three layers followed by a connectionist temporal classification (CTC) processing step). The approach assumes differently scaled input images and different feature map sizes. Three datasets are used: the standard benchmark RIMES dataset (French); a historical handwritten dataset KdK (Dutch); the standard benchmark George Washington (GW) dataset (English). Final performance obtained for the word-recognition test of RIMES was 96.6%, a clear improvement over other state-of-the-art approaches which did not use a pre-trained network. On the KdK and GW datasets, our approach also shows good results. The proposed approach is deployed in the Monk search engine for historical-handwriting collections.

著录项

来源
《Neural computing & applications》 |2021年第14期|8615-8634|共20页
作者
Ameryan Mahya; Schomaker Lambert;
展开▼
作者单位

Univ Groningen, Artificial Intelligence & Cognit Engn, Fac Sci & Engn, Groningen, Netherlands;

展开▼
收录信息
原文格式 PDF
正文语种英语
中图分类人工神经网络计算机;人工智能理论;
关键词
Coding scheme; Ensemble system; End-to-end convolutional long short-term memory; Connectionist temporal classification;

A limited-size ensemble of homogeneous CNN/LSTMs for high-performance word classification

摘要

著录项

相关主题

期刊订阅