Downloads · 30 days
4
14% of all-time downloads
thebnbrkr/tinyllama-salience
tinyllama-salience is a machine learning model from thebnbrkr. Use it for the machine learning task on the model card, and read the license before you ship it in a product. It is set up for peft. The card lists the license as apache-2.0.
Fine-tuned TinyLlama-1.1B-Chat for conversation memory and salience detection using the LoCoMo dataset.
Downloads · 30 days
4
14% of all-time downloads
All-time downloads
28
Public
Repo size
1.4 GB
Likes
1
Public
Click a slice to open those files.
.pt954 MB · 51%
From the Hugging Face model README
Fine-tuned TinyLlama-1.1B-Chat for conversation memory and salience detection using the LoCoMo dataset.
from transformers import AutoModelForCausalLM, AutoTokenizer, pipeline
model = AutoModelForCausalLM.from_pretrained("thebnbrkr/tinyllama-salience")
tokenizer = AutoTokenizer.from_pretrained("thebnbrkr/tinyllama-salience")
pipe = pipeline("text-generation", model=model, tokenizer=tokenizer)
prompt = """Context:
Alice told Bob she's moving to Seattle next month.
Question: Where is Alice moving?
Answer:"""
result = pipe(prompt, max_new_tokens=50)
print(result[0]["generated_text"])
This model is designed to answer questions based on conversational context, identifying salient information from dialogue history.
Trained on the LoCoMo (Long Context Memory) dataset, specifically the MC10 variant which tests conversation memory through question-answering.
Apache 2.0