ICDAR 2015 Competition HTRtS: Handwritten Text Recognition on the tranScriptorium Dataset

Description

This dataset comprises the dataset used for the ICDAR 2015 Competition on  Handwritten Text Recognition on the tranScriptorium Dataset. The handwritten images for this contest were drawn from the English "Bentham collection" dataset used in the TRAN SCRIPTORIUM project. The selected data has been written by several hands and entails significant variabilities and difficulties regarding the quality of text images, writing styles and crossed-out text. This contest is clearly more difficult than the the first edition both for training and for testing. A portion of the training dataset and the full test dataset were provided in the form of carefully segmented line images, along with the corresponding transcripts. Another portion of the training dataset was provided as raw images and their corresponding transcripts at region level.   ICDAR 2015 competition HTRtS: handwritten text recognition on the tranScriptorium dataset JA Sánchez, AH Toselli, V Romero, E Vidal.  In International Conference on Document Analysis and Recognition (ICDAR), pp. 1166-1170, 2015.

Resources

Name Format Description Link
0 http://data.europa.eu/88u/dataset/oai-zenodo-org-248733
0 http://data.europa.eu/88u/dataset/oai-zenodo-org-248733

Tags

  • handwritten-text-recognition-of-historical-documents,-pattern-recognition,-machine-learning

Topics

Categories