Downloads · 30 days
7
18% of all-time downloads
firedfrog23/roberta-large-squad
roberta-large-squad is a question answering model from firedfrog23. Use it when the input is a question plus a passage. It is set up for transformers. The card lists the license as mit.
should probably proofread and complete it, then remove this comment. --
Downloads · 30 days
7
18% of all-time downloads
All-time downloads
39
Public
Parameters
354M
1.4 GB on disk
Likes
0
Public
Click a slice to open those files.
.safetensors1.4 GB · 100%
From the Hugging Face model README
This model is a fine-tuned version of roberta-large on an unknown dataset. It achieves the following results on the evaluation set:
More information needed
More information needed
More information needed
The following hyperparameters were used during training:
| Training Loss | Epoch | Step | Validation Loss |
|---|---|---|---|
| 1.5562 | 1.0 | 80 | 1.3605 |
| 0.6982 | 2.0 | 160 | 1.3445 |
| 0.4441 | 3.0 | 240 | 1.3212 |