Downloads · 30 days
13
0% of all-time downloads
KETI-NLP/veld-base
veld-base is a image-to-text model from KETI-NLP. Use it when you need a caption or text from an image. It is set up for transformers. The card lists the license as apache-2.0.
Pretrained Vision Encoder Text Decoder Model in Korean and English. See Github for more details.
Downloads · 30 days
13
0% of all-time downloads
All-time downloads
2.9K
Public
Repo size
4.1 GB
Likes
0
Public
Click a slice to open those files.
.bin1.4 GB · 100%
From the Hugging Face model README
Pretrained Vision Encoder Text Decoder Model in Korean and English. See Github for more details.
from transformers import AutoProcessor, AutoModel
processor = AutoProcessor.from_pretrained("KETI-AIR/veld-base", trust_remote_code=True)
model = AutoModel.from_pretrained("KETI-AIR/veld-base", trust_remote_code=True)
You can use AutoTokenizer and AutoFeatureExtractor instead AutoProcessor.
You don't need to pass trust_remote_code=True for AutoTokenizer and AutoFeatureExtractor
from transformers import AutoFeatureExtractor, AutoTokenizer, AutoModel
feature_extractor = AutoFeatureExtractor.from_pretrained("KETI-AIR/veld-base")
tokenizer = AutoTokenizer.from_pretrained("KETI-AIR/veld-base")
model = AutoModel.from_pretrained("KETI-AIR/veld-base", trust_remote_code=True)