Skip to content

BUT-FIT

Dixtral_QA

BUT-FIT/Dixtral_QA

Dixtral_QA is a automatic speech recognition model from BUT-FIT. Use it when you need speech turned into text. It is set up for transformers. The card lists the license as apache-2.0.

This repository hosts DixtralQA, developed by BUT Speech@FIT. Dixtral couples the Voxtral-Mini-3B spoken-language model with the DiCoW diarization-conditioned encoder, giving the LLM target-speaker awareness in multi-…

Downloads · 30 days

22

17% of all-time downloads

All-time downloads

128

Public

Parameters

4.7B

18.7 GB on disk

Likes

0

Public

Hugging Face

Repo makeup

Click a slice to open those files.

.safetensors9.4 GB · 100%

At a glance

Task
Automatic Speech Recognition
Library
transformers
License
apache-2.0
Model type
voxtral
Access
Public
Created
Jun 3, 2026
Updated
Jun 3, 2026
SHA
83d1ee56

Base models

Task
Automatic Speech Recognition
Library
transformers
Type
voxtral
License
apache-2.0
Created
Jun 3, 2026
Updated
Jun 3, 2026