Downloads · 30 days
2
3% of all-time downloads
locuslab/phi_KL_1e-05_forget05
phi_KL_1e-05_forget05 is a machine learning model from locuslab. Use it for the machine learning task on the model card, and read the license before you ship it in a product. The card lists the license as mit.
IMPORTANT: This model's checkpoints are stored in separate branches. You MUST specify a revision when loading the model to access a specific checkpoint.
Downloads · 30 days
2
3% of all-time downloads
All-time downloads
59
Public
Repo size
23.4 GB
Likes
0
Public
Click a slice to open those files.
.md2.9 KB · 66%
From the Hugging Face model README
IMPORTANT: This model's checkpoints are stored in separate branches. You MUST specify a revision when loading the model to access a specific checkpoint.
This model is a variant of the Phi-1.5 model, fine-tuned on the TOFU (Task of Fictitious Unlearning) dataset and then subjected to various unlearning algorithms.
This model uses the KL_1e-05 unlearning algorithm with the following parameters:
forget05N/A%The model is organized into multiple revisions, each representing a checkpoint during the unlearning process. The revision names follow the pattern checkpoint-X, where X is the checkpoint number. Each revision is stored in a separate branch.
To load a specific revision of this model, you MUST specify the revision parameter. Use the following code:
from transformers import AutoModelForCausalLM, AutoTokenizer
# The 'revision' parameter is REQUIRED. Replace 'checkpoint-X' with the desired revision (e.g., 'checkpoint-12')
revision = "checkpoint-X"
model = AutoModelForCausalLM.from_pretrained("locuslab/{model_name}", revision=revision)
tokenizer = AutoTokenizer.from_pretrained("locuslab/{model_name}", revision=revision)
Note: If you don't specify a revision, you will not be able to load the model correctly.
TOFU (Task of Fictitious Unlearning) is a dataset designed for training and evaluating unlearning algorithms in language models. It simulates scenarios where certain information needs to be "forgotten" or removed from the model's knowledge.
This model is primarily intended for research purposes, particularly in the field of machine unlearning and privacy in language models. It may not be suitable for general-purpose language tasks without further evaluation.
If you use this model in your research, please cite:
@misc{tofu2024,
title={TOFU: A Task of Fictitious Unlearning for LLMs},
author={Pratyush Maini and Zhili Feng and Avi Schwarzschild and Zachary C. Lipton and J. Zico Kolter},
year={2024},
archivePrefix={arXiv},
primaryClass={cs.LG}
}
For questions or issues regarding this model, please contact [email protected].