Downloads · 30 days
64
10% of all-time downloads
YoLo2000/TiLamb-7B
TiLamb-7B is a text generation model from YoLo2000. Use it when you need the model to write or continue text. It is set up for transformers. The card lists the license as apache-2.0.
TiLamb-7B 是藏文大语言模型的基座模型,它使用了 26.43GB 的藏文语料,基于Meta发布的可商用大模型 LLaMA2-7B 模型,通过 LoRA 方法进行了增量预训练。该模型在 LLaMA2 的基础上扩展了词表,从原有的词表大小 32,000 扩充藏文词汇至 61,221 ,并对 LLaMA2-7B 原始模型的 embedding 和 lmhead 进行了均值扩充初始化。更多信息请访问 TiLamb-7B GitHu…
Downloads · 30 days
64
10% of all-time downloads
All-time downloads
651
Public
Repo size
27.9 GB
Likes
3
Public
Click a slice to open those files.
.bin14 GB · 100%
From the Hugging Face model README
TiLamb-7B 是藏文大语言模型的基座模型,它使用了 26.43GB 的藏文语料,基于Meta发布的可商用大模型 LLaMA2-7B 模型,通过 LoRA 方法进行了增量预训练。该模型在 LLaMA2 的基础上扩展了词表,从原有的词表大小 32,000 扩充藏文词汇至 61,221 ,并对 LLaMA2-7B 原始模型的 embedding 和 lm_head 进行了均值扩充初始化。更多信息请访问 TiLamb-7B GitHub 主页。
重要说明:
使用须知:
TiLamb-7B is the foundational model for the Tibetan language, utilizing 26.43GB of Tibetan corpora. It's based on Meta's commercially available large model, LLaMA2-7B, and has been incrementally pre-trained using the LoRA method. This model expands on LLaMA2 by enlarging the vocabulary from the original 32,000 to 61,221 Tibetan words and initializes the embedding and lm_head of the original LLaMA2-7B model through mean expansion. For more information, please visit the TiLamb-7B GitHub page.
Important Notes:
Usage Notice: