Downloads · 30 days
19
6% of all-time downloads
AdamCodd/distilbart-sum-arxiv
distilbart-sum-arxiv is a machine learning model from AdamCodd. Use it for the machine learning task on the model card, and read the license before you ship it in a product. It is set up for transformers.
This model is a fine-tuned version of sshleifer/distilbart-xsum-12-6 on a subset of the ccdv/arxiv-summarization dataset. It achieves the following results on the evaluation set: Loss: 2.420 Rouge1: 42.185
Downloads · 30 days
19
6% of all-time downloads
All-time downloads
293
Public
Repo size
2.4 GB
Likes
1
Public
Click a slice to open those files.
.bin1.2 GB · 100%
From the Hugging Face model README
This model is a fine-tuned version of sshleifer/distilbart-xsum-12-6 on a subset of the ccdv/arxiv-summarization dataset. It achieves the following results on the evaluation set:
This model is a distilled version of BART with 306M parameters (vs. 406 for the BART model), but it is 1.68 times faster than BART at inference. It has been trained on 60_000 samples and has a limitation of 1024 tokens.
Since this model has been trained on scientific papers, it may perform poorly when attempting to summarize other types of content.
from transformers import pipeline
summarizer = pipeline("text2text-generation", model="AdamCodd/distilbart-sum-arxiv")
paper = "Scientific paper..."
result = summarizer(paper)
print(result)
More information needed
The following hyperparameters were used during training:
| key | value |
|---|---|
| eval_rouge1 | 42.185638427734375 |
| eval_rouge2 | 15.481599807739258 |
| eval_rougeL | 24.440900802612305 |
| eval_rougeLsum | 24.260608673095703 |
If you want to support me, you can here.