Downloads · 30 days
54
7% of all-time downloads
muhtasham/santacoder-finetuned-the-stack-assembly
santacoder-finetuned-the-stack-assembly is a text generation model from muhtasham. Use it when you need the model to write or continue text. It is set up for transformers. The card lists the license as openrail.
should probably proofread and complete it, then remove this comment. --
Downloads · 30 days
54
7% of all-time downloads
All-time downloads
773
Public
Repo size
23 GB
Likes
1
Public
Click a slice to open those files.
.bin4.6 GB · 100%
From the Hugging Face model README
This model is a fine-tuned version of bigcode/santacoder on an on The Stack assembly dataset. It achieves the following results on the evaluation set:
The SantaCoder models are a series of 1.1B parameter models trained on the Python, Java, and JavaScript subset of The Stack (v1.1) (which excluded opt-out requests). The main model uses Multi Query Attention, was trained using near-deduplication and comment-to-code ratio as filtering criteria and using the Fill-in-the-Middle objective. In addition, there are several models that were trained on datasets with different filter parameters and with architecture and objective variations.
The predominant language in source is English although other languages are also present. As such the model is capable to generate code snippets provided some context but the generated code is not guaranteed to work as intended. It can be inefficient, contain bugs or exploits.
The Stack contains over 6TB of permissively-licensed source code files covering 358 programming languages. The dataset was created as part of the BigCode Project, an open scientific collaboration working on the responsible development of Large Language Models for Code (Code LLMs). The Stack serves as a pre-training dataset for Code LLMs, i.e., code-generating AI systems which enable the synthesis of programs from natural language descriptions as well as other from code snippets. This is the near-deduplicated version with 3TB data.
The following hyperparameters were used during training: