Downloads · 30 days
61
1% of all-time downloads
rhysjones/phi-2-orange
phi-2-orange is a text generation model from rhysjones. Use it when you need the model to write or continue text. It is set up for transformers. The card lists the license as mit.
Downloads · 30 days
61
1% of all-time downloads
All-time downloads
4.5K
Public
Parameters
2.8B
5.6 GB on disk
Likes
42
Public
Click a slice to open those files.
.safetensors5.6 GB · 100%
From the Hugging Face model README

A two-step finetune of Phi-2, with a bit of zest.
There is an updated model at rhysjones/phi-2-orange-v2 which has higher evals, if you wish to test.
A first finetune using a collection of broad training data:
And then a DPO finetune using:
If you're using Ollama, you can download and run using:
ollama run rhysjones/phi-2-orange
Phi-2 Orange uses ChatML as the prompt format, with or without the system instruction.
To prompt with a system instruction (use whatever system prompt you like):
<|im_start|>system
You are a helpful assistant for Python which outputs in Markdown format.<|im_end|>
<|im_start|>user
Write a function to calculate the Fibonacci sequence<|im_end|>
<|im_start|>assistant
You can also omit the system prompt if you wish:
<|im_start|>user
Why is the sky blue?<|im_end|>
<|im_start|>assistant
Evaluations done using mlabonne's usefull Colab notebook llm-autoeval. Also check out the alternative leaderboard at Yet_Another_LLM_Leaderboard
| Model | AGIEval | GPT4All | TruthfulQA | Bigbench | Average |
|---|---|---|---|---|---|
| phi-2-orange | 33.37 | 71.33 | 49.87 | 37.3 | 47.97 |
| phi-2-dpo | 30.39 | 71.68 | 50.75 | 34.9 | 46.93 |
| dolphin-2_6-phi-2 | 33.12 | 69.85 | 47.39 | 37.2 | 46.89 |
| phi-2 | 27.98 | 70.8 | 44.43 | 35.21 | 44.61 |