Romanian - English literature corpus (Processed)
Description
Bilingual Romanian – English literature corpus built from a small set of freely available literature books (drama, sci-fi, etc.). The texts are positionally aligned, i.e. the sentence on line i in the English text is aligned with the sentence on line i in the Romanian text. Alignment was manually validated.
This dataset has been created within the framework of the European Language Resource Coordination (ELRC) Connecting Europe Facility - Automated Translation (CEF.AT) actions SMART 2014/1074 and SMART 2015/1091. For further information on the project: http://lr-coordination.eu.
Resources
| Name |
Format |
Description |
Link |
|
57 |
|
https://elrc-share.eu/repository/browse/romanian-english-literature-corpus-processed/1f29705cd35111e7b7d400155d026706fbe54436d5684fb1a1d0ef66c7a862de/ |
Tags
- group-resources-for-language-technologies