Downloads · 30 days
29
0% of all-time downloads
mbartolo/roberta-large-synqa-ext
roberta-large-synqa-ext is a question answering model from mbartolo. Use it when the input is a question plus a passage. It is set up for transformers. The card lists the license as apache-2.0.
This is a RoBERTa-Large QA Model trained from https://huggingface.co/roberta-large in two stages. First, it is trained on synthetic adversarial data generated using a BART-Large question generator on Wikipedia passage…
Downloads · 30 days
29
0% of all-time downloads
All-time downloads
13.5K
Public
Repo size
4.3 GB
Likes
1
Public
Click a slice to open those files.
.bin1.4 GB · 100%
From the Hugging Face model README
This is a RoBERTa-Large QA Model trained from https://huggingface.co/roberta-large in two stages. First, it is trained on synthetic adversarial data generated using a BART-Large question generator on Wikipedia passages from SQuAD as well as Wikipedia passages external to SQuAD, and then it is trained on SQuAD and AdversarialQA (https://arxiv.org/abs/2002.00293) in a second stage of fine-tuning.
Training data: SQuAD + AdversarialQA Evaluation data: SQuAD + AdversarialQA
Approx. 1 training epoch on the synthetic data and 2 training epochs on the manually-curated data.
Please refer to https://arxiv.org/abs/2104.08678 for full details.