Downloads · 30 days
0
keesephillips/qlora-llama-3-8b
qlora-llama-3-8b is a text generation model from keesephillips. Use it when you need the model to write or continue text. The card lists the license as mit.
- model.ipynb - notebook containing the code for fine tuning the Llama 3 model using QLoRa - data/train.json - json file containing the training set provided in the FINQA paper - data/test.json - json file containing…
Downloads · 30 days
0
Access
Public
Updated Mar 21, 2025
Repo size
320 MB
Likes
0
Public
Click a slice to open those files.
.json94.9 MB · 63%
From the Hugging Face model README
The focal property of interest is analysis financial documents for numerical reasoning. Specifically numerical reasoning over quarterly financial filings with the SEC. The Llama-3-8B model was chosen to fine tune using the QLoRa approach. This approach was chosen due to the paper's findings of a performance increase while utilizing minimal memory and hardware. The aggressive quantization seemed to significantly decreased training time while offering increased performance on financial analysis.
| ROUGE Score | Base Model | QLoRa Fine Tuned Model |
|---|---|---|
| ROUGE-1 | 0.05104785 | 0.25257307 |
| ROUGE-2 | 0.01158752 | 0.10479990 |
| ROUGE-L | 0.05104785 | 0.25175429 |