Downloads · 30 days
1.9K
1% of all-time downloads
westlake-repl/SaProt_35M_AF2
SaProt_35M_AF2 is a fill-mask model from westlake-repl. Use it when you need the model to fill a missing word. It is set up for transformers. The card lists the license as mit.
We provide two ways to use SaProt, including through huggingface class and through the same way as in esm github. Users can choose either one to use.
Downloads · 30 days
1.9K
1% of all-time downloads
All-time downloads
248K
Public
Repo size
680 MB
Likes
4
Public
Click a slice to open those files.
.bin137 MB · 50%
From the Hugging Face model README
We provide two ways to use SaProt, including through huggingface class and through the same way as in esm github. Users can choose either one to use.
The following code shows how to load the model.
from transformers import EsmTokenizer, EsmForMaskedLM
model_path = "/your/path/to/SaProt_35M_AF2"
tokenizer = EsmTokenizer.from_pretrained(model_path)
model = EsmForMaskedLM.from_pretrained(model_path)
#################### Example ####################
device = "cuda"
model.to(device)
seq = "M#EvVpQpL#VyQdYaKv" # Here "#" represents lower plDDT regions (plddt < 70)
tokens = tokenizer.tokenize(seq)
print(tokens)
inputs = tokenizer(seq, return_tensors="pt")
inputs = {k: v.to(device) for k, v in inputs.items()}
outputs = model(**inputs)
print(outputs.logits.shape)
"""
['M#', 'Ev', 'Vp', 'Qp', 'L#', 'Vy', 'Qd', 'Ya', 'Kv']
torch.Size([1, 11, 446])
"""
The esm version is also stored in the same folder, named SaProt_35M_AF2.pt. We provide a function to load the model.
from utils.esm_loader import load_esm_saprot
model_path = "/your/path/to/SaProt_35M_AF2.pt"
model, alphabet = load_esm_saprot(model_path)