Downloads · 30 days
75
11% of all-time downloads
SparseLLM/prosparse-llama-2-13b-gguf
prosparse-llama-2-13b-gguf is a feature extraction model from SparseLLM. Use it when you need embeddings to search or compare text. It is set up for transformers. The card lists the license as llama2.
- Original model: SparseLLM/ProSparse-LLaMA-2-13B - Converted & distributed by: THUNLP, ModelBest, and PowerInfer
Downloads · 30 days
75
11% of all-time downloads
All-time downloads
684
Public
Repo size
27.6 GB
Likes
1
Public
Click a slice to open those files.
.gguf27.6 GB · 100%
From the Hugging Face model README
This model is the downstream distribution of SparseLLM/ProSparse-LLaMA-2-13B in PowerInfer GGUF format consisting of the LLM model weights and predictor weights.
Please kindly cite using the following BibTeX:
@article{song2024prosparse,
title={{ProSparse}: Introducing and Enhancing Intrinsic Activation Sparsity within Large Language Models},
author={Song, Chenyang and Han, Xu and Zhang, Zhengyan and Hu, Shengding and Shi, Xiyu and Li, Kuai and Chen, Chen and Liu, Zhiyuan and Li, Guangli and Yang, Tao and Sun, Maosong},
year={2024},
journal={arXiv preprint arXiv:2402.13516},
url={https://arxiv.org/pdf/2402.13516.pdf}
}