Skip to content

jatshi

Audio-Codec-LLM-Native-Audio-Projector-v3

jatshi/Audio-Codec-LLM-Native-Audio-Projector-v3

Audio-Codec-LLM-Native-Audio-Projector-v3 is a feature extraction model from jatshi. Use it when you need embeddings to search or compare text. It is set up for pytorch. The card lists the license as apache-2.0.

audioprojector.pt is the trainable Whisper-to-Qwen continuous prefix projector from the v3 RTX 4090 smoke run. Whisper-small and Qwen2.5-1.5B-Instruct were loaded as real frozen base models; two optimizer steps reduce…

Downloads · 30 days

0

Access

Public

Updated Aug 2, 2026

Repo size

7.1 MB

Likes

0

Public

Hugging Face

Repo makeup

Click a slice to open those files.

.pt7.1 MB · 100%

At a glance

Task
Feature Extraction
Library
pytorch
License
apache-2.0
Access
Public
Created
Aug 2, 2026
Updated
Aug 2, 2026
SHA
d8ec09e2

Base models

Task
Feature Extraction
Library
pytorch
License
apache-2.0
Created
Aug 2, 2026
Updated
Aug 2, 2026
Audio-Codec-LLM-Native-Audio-Projector-v3 — AI Model — AIMarketly