English-Portuguese website parallel corpus (Processed)

Description

This is a parallel corpus of bilingual texts crawled from multilingual websites, which contains 843 TUs. Manual validation has been performed on a sample of the data. This dataset has been created within the framework of the European Language Resource Coordination (ELRC) Connecting Europe Facility - Automated Translation (CEF.AT) actions SMART 2014/1074 and SMART 2015/1091. For further information on the project: http://lr-coordination.eu.

Resources

Name Format Description Link
57 https://elrc-share.eu/repository/browse/english-portuguese-website-parallel-corpus-processed/d6057afaea8d11e9913100155d0267061bb29ef57d224c478a9fffdb9425f26f/

Tags

  • group-resources-for-language-technologies

Topics

  • GOVE

Categories