Downloads · 30 days
15
9% of all-time downloads
HINT-lab/HASS-Llama3-8B-Instruct-Reproduce
HASS-Llama3-8B-Instruct-Reproduce is a machine learning model from HINT-lab. Use it for the machine learning task on the model card, and read the license before you ship it in a product. The card lists the license as cc-by-nc-nd-4.0.
This repository provides a reproduced HASS model checkpoint that is used as a baseline in PosS (Position Specialist) experiments.
Downloads · 30 days
15
9% of all-time downloads
All-time downloads
170
Public
Repo size
3.1 GB
Likes
0
Public
Click a slice to open those files.
.bin3.1 GB · 100%
From the Hugging Face model README
This repository provides a reproduced HASS model checkpoint that is used as a baseline in PosS (Position Specialist) experiments.
PosS is a speculative decoding method proposed in the paper:
PosS: Position Specialist Generates Better Draft for Speculative Decoding
In our experiments, this HASS checkpoint serves as the baseline draft model for comparison with the proposed position-specialized draft models.
The full implementation of PosS, along with training details and evaluation scripts (including EAGLE-2 and HASS baselines), is available at:
👉 GitHub: https://github.com/shrango/PosS
If the model is not automatically downloaded by your framework, you may manually download the following files from this repository:
pytorch_model.bin — model weightsconfig.json — model configurationIf you use this checkpoint in the context of PosS or refer to the PosS method, please cite:
@misc{huang2025posspositionspecialistgenerates,
title = {POSS: Position Specialist Generates Better Draft for Speculative Decoding},
author = {Langlin Huang and Chengsong Huang and Jixuan Leng and Di Huang and Jiaxin Huang},
year = {2025},
eprint = {2506.03566},
archivePrefix= {arXiv},
primaryClass = {cs.CL},
url = {https://arxiv.org/abs/2506.03566}
}