Downloads · 30 days
0
imirandam/CLIP_COCO
CLIP_COCO is a machine learning model from imirandam. Use it for the machine learning task on the model card, and read the license before you ship it in a product. The card lists the license as mit.
- Homepage: https://imirandam.github.io/BiVLCprojectpage/ - Repository: https://github.com/IMirandaM/BiVLC - Paper: https://arxiv.org/abs/2406.09952 - Point of Contact: Imanol Miranda CLIPCOCO is a model presented in…
Downloads · 30 days
0
Access
Public
Updated Jun 17, 2024
Repo size
1.8 GB
Likes
0
Public
Click a slice to open those files.
.pt1.8 GB · 100%
From the Hugging Face model README
CLIP_COCO is a model presented in the BiVLC paper for experimentation. It has been fine-tuned with OpenCLIP framework using as basis the CLIP ViT-B-32 model pre-trained by 'openai'. The idea behind this fine-tuning is to have a baseline to compare the CLIP_TROHN-Text and CLIP_TROHN-Img models. Hyperparameters:
The model is evaluated in BiVLC.
This work is licensed under a MIT License.
If you find this dataset useful, please consider citing our paper:
@misc{miranda2024bivlc,
title={BiVLC: Extending Vision-Language Compositionality Evaluation with Text-to-Image Retrieval},
author={Imanol Miranda and Ander Salaberria and Eneko Agirre and Gorka Azkune},
year={2024},
eprint={2406.09952},
archivePrefix={arXiv},
primaryClass={cs.CV}
}