Monolingual Bulgarian corpus in the public administration domain
Description
Monolingual Bulgarian corpus, containing 27028434 tokens and 932105 lexical types in the public administration domain.
This dataset has been created within the framework of the European Language Resource Coordination (ELRC) Connecting Europe Facility - Automated Translation (CEF.AT) actions SMART 2014/1074 and SMART 2015/1091. For further information on the project: http://lr-coordination.eu.
Resources
| Name |
Format |
Description |
Link |
|
57 |
|
https://elrc-share.eu/repository/browse/monolingual-bulgarian-corpus-in-the-public-administration-domain/8e72141180b311e6bfe700155d020502f7ff740828d34871af550671591890b5/ |
Tags
- group-resources-for-language-technologies