Downloads · 30 days
200
4% of all-time downloads
ragavsachdeva/magiv2-crop-embedder
magiv2-crop-embedder is a machine learning model from ragavsachdeva. Use it for the machine learning task on the model card, and read the license before you ship it in a product.
<style .title-container { display: flex; flex-direction: column; / Stack elements vertically / justify-content: center; align-items: center; }
Downloads · 30 days
200
4% of all-time downloads
All-time downloads
5K
Public
Parameters
85.8M
343 MB on disk
Likes
1
Public
Click a slice to open those files.
.safetensors343 MB · 100%
From the Hugging Face model README
from transformers import AutoModel
import PIL
import torch
batch_of_images = [PIL.Image.open("image1.jpg"), PIL.Image.open("image2.jpg")]
model = AutoModel.from_pretrained("ragavsachdeva/magiv2-crop-embedder", trust_remote_code=True).cuda().eval()
with torch.no_grad():
embeddings = model(batch_of_images)
print(embeddings.shape)
The provided model is available for unrestricted use in personal, research, non-commercial, and not-for-profit endeavors. For any other usage scenarios, kindly contact me via email, providing a detailed description of your requirements, to establish a tailored licensing arrangement. My contact information can be found on my website: ragavsachdeva [dot] github [dot] io
@misc{magiv2,
title={Tails Tell Tales: Chapter-Wide Manga Transcriptions with Character Names},
author={Ragav Sachdeva and Gyungin Shin and Andrew Zisserman},
year={2024},
eprint={2408.00298},
archivePrefix={arXiv},
primaryClass={cs.CV},
url={https://arxiv.org/abs/2408.00298},
}