Downloads · 30 days
14
1% of all-time downloads
nikhil07prakash/float-7b
float-7b is a text generation model from nikhil07prakash. Use it when you need the model to write or continue text. It is set up for transformers. The card lists the license as mit.
This model is a fully fine-tuned version of the Llama-7B model on synthetically generated arithmetic tasks. It was introduced in this paper. It is very similar to Goat-7B, except it was trained without LoRA.
Downloads · 30 days
14
1% of all-time downloads
All-time downloads
966
Public
Repo size
27 GB
Likes
3
Public
Click a slice to open those files.
.bin13.5 GB · 100%
From the Hugging Face model README
This model is a fully fine-tuned version of the Llama-7B model on synthetically generated arithmetic tasks. It was introduced in this paper. It is very similar to Goat-7B, except it was trained without LoRA.
For inquiries about checkpoints during the fine-tuning process, kindly reach out to Nikhil via email.
Use the code below to get started with the model.
from transformers import AutoModel
model = AutoModel.from_pretrained("nikhil07prakash/float-7b")
BibTeX:
@inproceedings{prakash2023fine,
title={Fine-Tuning Enhances Existing Mechanisms: A Case Study on Entity Tracking},
author={Prakash, Nikhil and Shaham, Tamar Rott and Haklay, Tal and Belinkov, Yonatan and Bau, David},
booktitle={Proceedings of the 2024 International Conference on Learning Representations},
note={arXiv:2402.14811},
year={2024}
}