Downloads · 30 days
10
17% of all-time downloads
ai-eldorado/Brazilian_CLT_DPO
Brazilian_CLT_DPO is a machine learning model from ai-eldorado. Use it for the machine learning task on the model card, and read the license before you ship it in a product. The card lists the license as llama3.
Model Description This model is a fine-tuned version of LLaMA-3 8B Instruct (4-bit quantized), optimized using Direct Preference Optimization (DPO) for answering legal questions related to Brazil’s Consolidation of La…
Downloads · 30 days
10
17% of all-time downloads
All-time downloads
60
Public
Parameters
8.2B
5.7 GB on disk
Likes
1
Public
Click a slice to open those files.
.safetensors5.7 GB · 100%
How the weights are stored.
U87.2B · 87%
From the Hugging Face model README
Model Description
This model is a fine-tuned version of LLaMA-3 8B Instruct (4-bit quantized), optimized using Direct Preference Optimization (DPO) for answering legal questions related to Brazil’s Consolidation of Labor Laws (CLT). The fine-tuning process leveraged a curated dataset of 736 human-preference triplets, annotated by HR specialists and legal experts, to align the model with domain-specific expectations for accuracy and compliance.
Intended Use
The model is designed for legal question answering in the context of Brazilian labor law, supporting HR departments, compliance teams, and legal professionals. It aims to provide factually accurate and semantically aligned responses to CLT-related queries.
Training Details
Performance Summary
Compared to the base model, this DPO-tuned model achieved:
@article{moraescomparing,
title={Comparing RAG, DPO and Agentic Approaches in Systems Performance on Q\&A about Brazilian Labor Legislation},
author={Moraes, Gabriel K and Luiz, Pedro Augusto and Dias, Gabriel and de Farias, Vitor GCB and Fabiana, CQ de O and Fabris, Vitor L and Vicente, Matheus HR and do Nascimento, Leonardo R and Oliveira, Charles S and dos Santos, Leonardo T and others}
}