Downloads · 30 days
8
30% of all-time downloads
aieng-lab/roberta-base_se-entities
roberta-base_se-entities is a token classification model from aieng-lab. Use it when you need labels on individual words, such as names. It is set up for transformers. The card lists the license as mit.
This model detects software engineering terminology in developer forums (e.g., Stack Overflow) as 'DataStructure', 'Application', 'CodeBlock', 'Function", 'DataType', 'Language', 'Library', 'Variable', 'Device', 'User…
Downloads · 30 days
8
30% of all-time downloads
All-time downloads
27
Public
Parameters
124M
248 MB on disk
Likes
0
Public
Click a slice to open those files.
.safetensors248 MB · 98%
From the Hugging Face model README
This model detects software engineering terminology in developer forums (e.g., Stack Overflow) as 'Data_Structure', 'Application', 'Code_Block', 'Function", 'Data_Type', 'Language', 'Library', 'Variable', 'Device', 'User_Name', 'User_Interface_Element', 'Class', 'Website', 'Version', 'File_Name', 'File_Type', 'Operating_System', 'Output_Block', 'Algorithm' or 'HTML_XML_Tag'.
@misc{pena2025benchmark,
author = {Fabian Peña and Steffen Herbold},
title = {Evaluating Large Language Models on Non-Code Software Engineering Tasks},
year = {2025}
}