Downloads · 30 days
14
8% of all-time downloads
ClaudioItaly/intelligence-cod-rag-7b-v3-2
intelligence-cod-rag-7b-v3-2 is a text generation model from ClaudioItaly. Use it when you need the model to write or continue text. It is set up for transformers.
This is a merge of pre-trained language models created using mergekit.
Downloads · 30 days
14
8% of all-time downloads
All-time downloads
182
Public
Parameters
7.6B
15.2 GB on disk
Likes
1
Public
Click a slice to open those files.
.safetensors15.2 GB · 100%
From the Hugging Face model README
This is a merge of pre-trained language models created using mergekit.
This model was merged using the task arithmetic merge method using happzy2633/qwen2.5-7b-ins-v3 as a base.
The following models were included in the merge:
The following YAML configuration was used to produce this model:
models:
- model: AIDC-AI/Marco-o1
parameters:
density: [1, 0.8, 0.2] # Aumentato leggermente il gradiente per dare maggiore peso al contributo iniziale
weight: 0.9 # Ridotto il peso per bilanciare meglio l'influenza
- model: happzy2633/qwen2.5-7b-ins-v3
parameters:
density: 0.6 # Aumentato per consentire una maggiore fusione delle rappresentazioni
weight: [0.1, 0.4, 0.8, 1] # Raffinato il gradiente per enfatizzare progressivamente il contributo
- model: AIDC-AI/Marco-o1
parameters:
density: 0.4 # Leggermente aumentato per integrare una maggiore ricchezza di rappresentazioni
weight:
- filter: mlp
value: 0.6 # Incrementato il valore per dare maggiore peso a questa componente
- value: 0.1 # Aggiunto un piccolo peso finale per evitare contributi nulli
merge_method: task_arithmetic # Manteniamo il metodo "ties" per una fusione bilanciata
base_model: happzy2633/qwen2.5-7b-ins-v3 # Base model per guidare la fusione
parameters:
normalize: true # Conserva la normalizzazione per evitare squilibri nelle rappresentazioni
int8_mask: true # Rimane abilitato per ottimizzare le prestazioni
adaptive_merge: true # Aggiunto per una fusione più dinamica in base al contesto
dtype: float16 # Manteniamo float16 per limitare l'uso di memoria e migliorare l'efficienza