Downloads · 30 days
11
20% of all-time downloads
RohitPatill/MedSymptomGPT
MedSymptomGPT is a text generation model from RohitPatill. Use it when you need the model to write or continue text. It is set up for transformers. The card lists the license as apache-2.0.
This is the model card for MedSymptomGPT, a model trained to understand and generate medical symptoms based on disease names. It uses the distilgpt2 architecture and has been fine-tuned on a dataset containing various…
Downloads · 30 days
11
20% of all-time downloads
All-time downloads
56
Public
Parameters
81.9M
328 MB on disk
Likes
0
Public
Click a slice to open those files.
.safetensors328 MB · 100%
From the Hugging Face model README
This is the model card for MedSymptomGPT, a model trained to understand and generate medical symptoms based on disease names. It uses the distilgpt2 architecture and has been fine-tuned on a dataset containing various diseases and their corresponding symptoms. The model is intended to assist in generating symptom lists for given diseases, aiding in medical research and educational purposes.
distilbert/distilgpt2This model can be directly used to generate symptoms associated with a given disease. This can be particularly useful in medical research, education, and healthcare applications where quick access to symptom information is valuable.
This model can be fine-tuned further for more specific tasks related to medical text generation, such as creating detailed disease descriptions, patient information leaflets, or other medical documentation.
This model should not be used for making clinical decisions or providing medical advice. It is intended for educational and research purposes only and should not replace professional medical judgment.
Users (both direct and downstream) should be aware of the potential biases in the training data, which may lead to biased outputs. It is important to validate the generated content with authoritative medical sources.
Use the code below to get started with the model:
from transformers import GPT2Tokenizer, GPT2LMHeadModel
tokenizer = GPT2Tokenizer.from_pretrained('RohitPatill/MedSymptomGPT')
model = GPT2LMHeadModel.from_pretrained('RohitPatill/MedSymptomGPT')
input_str = "Kidney Failure"
input_ids = tokenizer.encode(input_str, return_tensors='pt')
output = model.generate(input_ids, max_length=50, num_return_sequences=1)
decoded_output = tokenizer.decode(output[0], skip_special_tokens=True)
print(decoded_output)
The model was trained on a dataset containing disease names and their corresponding symptoms. The dataset used for training was QuyenAnhDE/Diseases_Symptoms, which was preprocessed to format the data appropriately for training a language model.
fp32 precision.The model was evaluated on a validation set, which was a split of the original training dataset. The validation set helps in monitoring the model performance during training.
The evaluation considered the ability of the model to generate coherent and relevant symptoms for a given disease name.
The primary metric used for evaluation was the loss function (CrossEntropyLoss), which measures the difference between the predicted and actual symptoms.
The model was successfully trained and validated, showing its capability to generate relevant symptoms for given diseases. However, the exact loss values and other detailed metrics were not specified.
This model is trained on a dataset that may contain biases, and the outputs should be validated against authoritative medical sources. The model is intended for educational and research purposes only and should not be used for clinical decision-making.