Downloads · 30 days
102
19% of all-time downloads
PicoKittens/AbstractsLlama-8M
AbstractsLlama-8M is a text generation model from PicoKittens. Use it when you need the model to write or continue text. The card lists the license as apache-2.0.
AbstractsLlama-8M is an ultra-compact, "pico-sized" language model trained from scratch by Pico-Kittens. It utilizes the Llama 2 architecture and is specifically optimized for generating scientific and academic text.
Downloads · 30 days
102
19% of all-time downloads
All-time downloads
537
Public
Parameters
8.4M
33.6 MB on disk
Likes
2
Public
Click a slice to open those files.
.safetensors33.6 MB · 99%
From the Hugging Face model README
AbstractsLlama-8M is an ultra-compact, "pico-sized" language model trained from scratch by Pico-Kittens. It utilizes the Llama 2 architecture and is specifically optimized for generating scientific and academic text.
The model was trained on a large-scale collection of ArXiv abstracts. The training objective was to compress the structural patterns, technical nomenclature, and "academic tone" of scientific research into a minimal parameter budget.
AbstractsLlama-8M is an experimental model. While it effectively mimics the syntax of research papers, users should be aware of the following:
User: We propose
AbstractsLlama-8M:
We propose a unified framework for modeling large-scale non-linearity of Cancer (NCI) problems with a variable-scale dataset for the linearized dynamics of polynomial conjugal structure. Our key idea of a multi-objective-centile-based model with a fixed, non-preferred variational autoencoder (NMAE) for feature extraction, which includes ax-aware, non-convex optimization formulation for both a single
import torch
from transformers import pipeline
device = 0 if torch.cuda.is_available() else -1
pipe = pipeline("text-generation", model="PicoKittens/AbstractsLlama-8M", device=device)
output = pipe("We propose", max_new_tokens=100, do_sample=True)
print(output[0]['generated_text'])