XQuAD Dataset Papers With Code
Por um escritor misterioso
Descrição
XQuAD (Cross-lingual Question Answering Dataset) is a benchmark dataset for evaluating cross-lingual question answering performance. The dataset consists of a subset of 240 paragraphs and 1190 question-answer pairs from the development set of SQuAD v1.1 (Rajpurkar et al., 2016) together with their professional translations into ten languages: Spanish, German, Greek, Russian, Turkish, Arabic, Vietnamese, Thai, Chinese, and Hindi. Consequently, the dataset is entirely parallel across 11 languages.
image171.png
Papers With Code: The Latest in Machine Learning, Deep Learning
Machine Learning Datasets
Commonsense knowledge adversarial dataset that challenges ELECTRA
The OIG Dataset
Datasets em português - NLP com o Deep Learning - AI Lab Deep
TyDiQA-GoldP Dataset
LSP Dataset - Machine Learning Datasets
How to Answer Questions with Machine Learning
Automatic Spanish Translation of SQuAD Dataset for Multi-lingual
ArxivPapers Dataset
End to End Question-Answering System Using NLP and SQuAD Dataset
Token-level statistics of the constructed datasets. Average Length
How to finetune UnifiedQA T5 for Causal SQUAD Question Answering
PDF] CodeQA: A Question Answering Dataset for Source Code