Skip to content

Icelandic-lt

convbert-base-igc-is

Icelandic-lt/convbert-base-igc-is

convbert-base-igc-is is a feature extraction model from Icelandic-lt. Use it when you need embeddings to search or compare text. It is set up for transformers. The card lists the license as cc-by-4.0.

This model was pretrained on the Icelandic Gigaword Corpus, which contains approximately 1.69B tokens, using default settings. The model uses a WordPiece tokenizer with a vocabulary size of 32,105.

Downloads · 30 days

12

22% of all-time downloads

All-time downloads

54

Public

Parameters

107M

1.3 GB on disk

Likes

0

Public

Hugging Face

Repo makeup

Click a slice to open those files.

.h5428 MB · 33%

Parameter types

How the weights are stored.

F32107M · 100%

Datasets

Task
Feature Extraction
Library
transformers
Type
convbert
License
cc-by-4.0
Languages
is
Created
May 27, 2024
Updated
Feb 27, 2025