Title
Indiscapes - Instance Segmentation Networks for Layout Parsing of Historical Indic Manuscripts.
Abstract
Historical palm-leaf manuscript and early paper documents from Indian subcontinent form an important part of the worldu0027s literary and cultural heritage. Despite their importance, large-scale annotated Indic manuscript image datasets do not exist. To address this deficiency, we introduce Indiscapes, the first ever dataset with multi-regional layout annotations for historical Indic manuscripts. To address the challenge of large diversity in scripts and presence of dense, irregular layout elements (e.g. text lines, pictures, multiple documents per image), we adapt a Fully Convolutional Deep Neural Network architecture for fully automatic, instance-level spatial layout parsing of manuscript images. We demonstrate the effectiveness of proposed architecture on images from the Indiscapes dataset. For annotation flexibility and keeping the non-technical nature of domain experts in mind, we also contribute a custom, web-based GUI annotation tool and a dashboard-style analytics portal. Overall, our contributions set the stage for enabling downstream applications such as OCR and word-spotting in historical Indic manuscripts at scale.
Year
DOI
Venue
2019
10.1109/ICDAR.2019.00164
ICDAR
Field
DocType
Citations 
Architecture,Annotation,Cultural heritage,Pattern recognition,Computer science,Segmentation,Indian subcontinent,Natural language processing,Artificial intelligence,Parsing,Analytics,Scripting language
Conference
0
PageRank 
References 
Authors
0.34
0
4
Name
Order
Citations
PageRank
Abhishek Prusty100.34
Sowmya Aitha200.34
Abhishek Trivedi301.69
Ravi Kiran Sarvadevabhatla478.41