Downloads · 30 days
5
29% of all-time downloads
Mohitvvermaa/trained_bart
trained_bart is a machine learning model from Mohitvvermaa. Use it for the machine learning task on the model card, and read the license before you ship it in a product. It is set up for transformers. The card lists the license as mit.
should probably proofread and complete it, then remove this comment. --
Downloads · 30 days
5
29% of all-time downloads
All-time downloads
17
Public
Parameters
406M
1.6 GB on disk
Likes
0
Public
Click a slice to open those files.
.safetensors1.6 GB · 100%
From the Hugging Face model README
<img src="https://raw.githubusercontent.com/wandb/assets/main/wandb-github-badge-28.svg" alt="Visualize in Weights & Biases" width="200" height="32"/> <img src="https://raw.githubusercontent.com/wandb/assets/main/wandb-github-badge-28.svg" alt="Visualize in Weights & Biases" width="200" height="32"/>
This model is a fine-tuned version of facebook/bart-large-cnn on the samsum dataset. It achieves the following results on the evaluation set:
More information needed
More information needed
More information needed
The following hyperparameters were used during training:
| Training Loss | Epoch | Step | Validation Loss | Rouge1 | Rouge2 | Rougel | Rougelsum | Gen Len |
|---|---|---|---|---|---|---|---|---|
| 0.0923 | 1.0 | 737 | 0.0888 | 13.2333 | 1.4874 | 9.9559 | 11.9429 | 62.5854 |
| 0.086 | 2.0 | 1474 | 0.0886 | 13.0092 | 2.0724 | 10.5364 | 11.7522 | 60.1463 |
| 0.0744 | 3.0 | 2211 | 0.0910 | 13.6972 | 2.0482 | 11.3161 | 12.9271 | 56.5854 |