Dataset of EvaLatin 2020

Description

This repository contains training and test data of EvaLatin 2020, the first campaign devoted to the evaluation of Natural Language Processing Tools for Latin. It also includes the evaluation script. EvaLatin first edition have 2 tasks (i.e. Lemmatization and PoS tagging) each with 3 sub-tasks (i.e. Classical, Cross-Genre, Cross-Time).  The EvaLatin 2020 dataset is licensed under the Creative Commons Attribution-NonCommercial-ShareAlike (CC BY-NC-SA) 4.0 International.

Resources

Name Format Description Link
0 http://data.europa.eu/88u/dataset/oai-zenodo-org-4022873
0 http://data.europa.eu/88u/dataset/oai-zenodo-org-4022873

Tags

  • lemmatization
  • latin-language
  • evaluation
  • annotated-corpus
  • pos-tagging

Topics

Categories