Downloads · 30 days
34
3% of all-time downloads
mrm8488/spanbert-base-finetuned-squadv2
spanbert-base-finetuned-squadv2 is a machine learning model from mrm8488. Use it for the machine learning task on the model card, and read the license before you ship it in a product. It is set up for transformers.
SpanBERT created by Facebook Research and fine-tuned on SQuAD 2.0 for Q&A downstream task (by them).
Downloads · 30 days
34
3% of all-time downloads
All-time downloads
1K
Public
Repo size
2 GB
Likes
0
Public
Click a slice to open those files.
.bin667 MB · 50%
From the Hugging Face model README
SpanBERT created by Facebook Research and fine-tuned on SQuAD 2.0 for Q&A downstream task (by them).
SpanBERT: Improving Pre-training by Representing and Predicting Spans
SQuAD2.0 combines the 100,000 questions in SQuAD1.1 with over 50,000 unanswerable questions written adversarially by crowdworkers to look similar to answerable ones. To do well on SQuAD2.0, systems must not only answer questions when possible, but also determine when no answer is supported by the paragraph and abstain from answering.
| Dataset | Split | # samples |
|---|---|---|
| SQuAD2.0 | train | 130k |
| SQuAD2.0 | eval | 12.3k |
You can get the fine-tuning script here
python code/run_squad.py \
--do_train \
--do_eval \
--model spanbert-base-cased \
--train_file train-v2.0.json \
--dev_file dev-v2.0.json \
--train_batch_size 32 \
--eval_batch_size 32 \
--learning_rate 2e-5 \
--num_train_epochs 4 \
--max_seq_length 512 \
--doc_stride 128 \
--eval_metric best_f1 \
--output_dir squad2_output \
--version_2_with_negative \
--fp16
| SQuAD 1.1 | SQuAD 2.0 | Coref | TACRED | |
|---|---|---|---|---|
| F1 | F1 | avg. F1 | F1 | |
| BERT (base) | 88.5 | 76.5 | 73.1 | 67.7 |
| SpanBERT (base) | 92.4 | 83.6 (this one) | 77.4 | 68.2 |
| BERT (large) | 91.3 | 83.3 | 77.1 | 66.4 |
| SpanBERT (large) | 94.6 | 88.7 | 79.6 | 70.8 |
Note: The numbers marked as * are evaluated on the development sets because those models were not submitted to the official SQuAD leaderboard. All the other numbers are test numbers.
Fast usage with pipelines:
from transformers import pipeline
qa_pipeline = pipeline(
"question-answering",
model="mrm8488/spanbert-base-finetuned-squadv2",
tokenizer="SpanBERT/spanbert-base-cased"
)
qa_pipeline({
'context': "Manuel Romero has been working very hard in the repository hugginface/transformers lately",
'question': "How has been working Manuel Romero lately?"
})
# Output: {'answer': 'very hard', 'end': 40, 'score': 0.9052708846768347, 'start': 31}
Created by Manuel Romero/@mrm8488
Made with <span style="color: #e25555;">♥</span> in Spain