Downloads · 30 days
76
0% of all-time downloads
Thytu/phi-2-audio-super
phi-2-audio-super is a text generation model from Thytu. Use it when you need the model to write or continue text. It is set up for transformers. The card lists the license as mit.
Downloads · 30 days
76
0% of all-time downloads
All-time downloads
19.6K
Public
Parameters
2.8B
66.7 GB on disk
Likes
2
Public
Click a slice to open those files.
.safetensors5.6 GB · 100%
From the Hugging Face model README
THIS PROJECT IS STILL IN WIP
Base Model: microsoft/phi-2
Fine-tuned version of abacaj/phi-2-super for ASR on librispeech_asr.
import transformers
import torch
if __name__ == "__main__":
model_name = "Thytu/phi-2-audio-super"
tokenizer = transformers.AutoTokenizer.from_pretrained(model_name)
model = (
transformers.AutoModelForCausalLM.from_pretrained(
model_name,
)
.to("cuda:0")
.eval()
)
# Exactly like for phi-2-super :D
messages = [
{"role": "user", "content": "Hello, who are you?"}
]
inputs = tokenizer.apply_chat_template(messages, return_tensors="pt").to(model.device)
input_ids_cutoff = inputs.size(dim=1)
with torch.no_grad():
generated_ids = model.generate(
input_ids=inputs,
use_cache=True,
max_new_tokens=512,
temperature=0.2,
top_p=0.95,
do_sample=True,
eos_token_id=tokenizer.eos_token_id,
pad_token_id=tokenizer.pad_token_id,
)
completion = tokenizer.decode(
generated_ids[0][input_ids_cutoff:],
skip_special_tokens=True,
)
print(completion)
TODO