Downloads · 30 days
508K
23% of all-time downloads
zed-industries/zeta
zeta is a machine learning model from zed-industries. Use it for the machine learning task on the model card, and read the license before you ship it in a product. The card lists the license as apache-2.0.
<img src="https://cdn-uploads.huggingface.co/production/uploads/644a8bc1cb1654dcb6e762f9/6296GYaJsrUBSAeUwUHvm.png" width="100"
Downloads · 30 days
508K
23% of all-time downloads
All-time downloads
2.2M
Public
Parameters
7.6B
100 GB on disk
Likes
423
Public
Click a slice to open those files.
.safetensors15.2 GB · 100%
From the Hugging Face model README
This repository contains a fine-tuned version of Qwen2.5-Coder-7B to support edit prediction in Zed.
The model has been fine-tuned using the zeta dataset. If you want to fine-tune the model yourself, you can refer to the following scripts:
The dataset used for training is available at: zed-industries/zeta
vllm serve zed-industries/zeta --served-model-name zeta
Quantization vLLM supports FP8 (8-bit floating point) weight and activation quantization using hardware acceleration on GPUs such as Nvidia H100 and AMD MI300x.
NGram Speculative Decoding configures vLLM to use speculative decoding where proposals are generated by matching n-grams in the prompt. This is a great fit for edit predictions since many of the tokens are already present in the prompt and the model is only needed to generate changes to the code file.
vllm serve zed-industries/zeta --served-model-name zeta --enable-prefix-caching --enable-chunked-prefill --quantization="fp8" --speculative-model [ngram] --ngram-prompt-lookup-max 4 --ngram-prompt-lookup-min 2 --num-speculative-tokens 8
For more insights about the model and its integration in Zed, check out the official blog post: Zed Blog - Edit Prediction