Downloads · 30 days
0
Arijit-07/aria-devops-llama8b
aria-devops-llama8b is a text generation model from Arijit-07. Use it when you need the model to write or continue text. It is set up for peft.
Trained on the ARIA DevOps Incident Response live RL environment using GRPO.
Downloads · 30 days
0
Access
Public
Updated Apr 26, 2026
Parameters
8B
16.1 GB on disk
Likes
0
Public
Click a slice to open those files.
.safetensors16.1 GB · 100%
From the Hugging Face model README
Trained on the ARIA DevOps Incident Response live RL environment using GRPO.
| Task | Baseline | Fine-tuned | Improvement |
|---|---|---|---|
| easy | 0.320 | 0.685 | +0.365 |
| medium | 0.050 | 0.378 | +0.328 |
| hard | 0.190 | 0.869 | +0.679 |
| bonus | 0.152 | 0.682 | +0.530 |
