Downloads · 30 days
0
rafmacalaba/gliner2-datause-base-v1
gliner2-datause-base-v1 is a machine learning model from rafmacalaba. Use it for the machine learning task on the model card, and read the license before you ship it in a product. It is set up for gliner2.
Fine-tuned GLiNER2 model for extracting structured dataset mentions from research documents.
Downloads · 30 days
0
Access
Public
Updated Feb 25, 2026
Repo size
13.1 MB
Likes
0
Public
Click a slice to open those files.
.safetensors13.1 MB · 100%
From the Hugging Face model README
Fine-tuned GLiNER2 model for extracting structured dataset mentions from research documents.
Given a document passage, extracts:
fastino/gliner2-base-v1from gliner2 import GLiNER2
extractor = GLiNER2.from_pretrained("fastino/gliner2-base-v1")
extractor.load_adapter("rafmacalaba/gliner2-datause-base-v1")
schema = (
extractor.create_schema()
.structure("dataset_mention")
.field("dataset_name", dtype="str")
.field("acronym", dtype="str")
.field("producer", dtype="str")
.field("geography", dtype="str")
.field("dataset_tag", dtype="str", choices=["named", "descriptive", "vague"])
.field("usage_context", dtype="str", choices=["primary", "supporting", "background"])
.field("is_used", dtype="str", choices=["True", "False"])
)
results = extractor.extract(text, schema)
dataset_mentions = results["dataset_mention"]