Downloads · 30 days
0
hskang0906/custom_summarization_dataset
custom_summarization_dataset is a machine learning model from hskang0906. Use it for the machine learning task on the model card, and read the license before you ship it in a product.
Downloads · 30 days
0
Access
Public
Updated Jul 29, 2024
Repo size
352 KB
Likes
0
Public
Click a slice to open those files.
.arrow348 KB · 99%
From the Hugging Face model README
Custom Text Dataset
This dataset contains text data for training language models. The data is collected from various sources, including books, articles, and web pages.
sentence, labelsThe data was collected using web scraping and manual extraction from public domain sources.
from datasets import load_dataset
dataset = load_dataset("path_to_dataset")
for example in dataset["train"]:
print(example["sentence"])
This dataset is designed for evaluating text generation models. Common evaluation metrics include ROUGE and BLEU.
The dataset may contain outdated or biased information. Users should be aware of these limitations when using the data.
Privacy: Ensure that the data does not contain personal information. Bias: Be aware of potential biases in the data.