Spanish-Portuguese website parallel corpus (Processed)
Description
This is a parallel corpus of bilingual texts crawled from multilingual websites, which contains 1,249 TUs.
Manual validation has been performed on a sample of the data.
This dataset has been created within the framework of the European Language Resource Coordination (ELRC) Connecting Europe Facility - Automated Translation (CEF.AT) actions SMART 2014/1074 and SMART 2015/1091. For further information on the project: http://lr-coordination.eu.
Resources
| Name |
Format |
Description |
Link |
|
57 |
|
https://elrc-share.eu/repository/browse/spanish-portuguese-website-parallel-corpus-processed/a548e4a6312911e9a4d400155d0267063d30a93b94b34ceebe168ca3af831e1b/ |
Tags
- group-resources-for-language-technologies