IARPA BETTER (Better Extraction from Text Towards Enhanced Retrieval) information extraction and information retrieval datasets.
Description
Cross-language information extraction and retrieval datasets developed for the evaluation of the IARPA BETTER program. The documents come from CommonCrawl. The IE annotations in three schemas are by MITRE and ARLIS. The IR queries and relevance judgments were done at NIST, and NIST was asked by IARPA to distribute the data in its final form. The tasks are all cross-language from English into one of Arabic, Farsi, Russian, Chinese, and Korean
Resources
| Name |
Format |
Description |
Link |
|
5 |
|
https://ir.nist.gov/better/ |
Tags
- information-extraction-information-retrieval-cross-language-information-retrieval