MuClaGED

Description

Dataset description Data is provided in a tab-separated format consisting of five columns, namely, token id, token, list of error codes for addition, list of error codes for deletion and list of codes for replacement. See more on data format in the standard reference article. License: CLARIN-ID, -PRIV, -NORED, -BY (https://www.kielipankki.fi/support/clarin-eula/#res).

Resources

Name Format Description Link
0 http://data.europa.eu/88u/dataset/https-doi-org-10-23695-q9v4-vt57

Tags

  • sentences
  • grammatical-error-detection
  • language-technology-(computational-linguistics)
  • error-edit-labeling
  • language-learning
  • token-level-detection
  • corpus
  • error-labeling

Topics

Categories