Monolingual Romanian corpus in the public administration domain

Description

Monolingual Romanian corpus, containing 12037387 tokens and 1176117 lexical types in the public administration domain. This dataset has been created within the framework of the European Language Resource Coordination (ELRC) Connecting Europe Facility - Automated Translation (CEF.AT) actions SMART 2014/1074 and SMART 2015/1091. For further information on the project: http://lr-coordination.eu.

Resources

Name Format Description Link
57 https://elrc-share.eu/repository/browse/monolingual-romanian-corpus-in-the-public-administration-domain/cb74f77f80ae11e6bfe700155d020502293753382d554a56849499439bb50dc9/

Tags

  • group-resources-for-language-technologies

Topics

  • GOVE

Categories