Downloads · 30 days
0
DinoStackAI/Qwen3-Emb-4b-lora-narrativeqa
Qwen3-Emb-4b-lora-narrativeqa is a feature extraction model from DinoStackAI. Use it when you need embeddings to search or compare text. It is set up for sentence-transformers. The card lists the license as apache-2.0.
LoRA adapter for Qwen/Qwen3-Embedding-4B fine-tuned on the narrativeqa RAG retrieval dataset (DinoStackAI/narrativeqa-rag). - Best dev metric: evalnarrativeqa-devcosinendcg@10 = 0.8110
Downloads · 30 days
0
Access
Public
Updated Jul 8, 2026
Repo size
23.2 MB
Likes
0
Public
Click a slice to open those files.
.json14.2 MB · 51%
From the Hugging Face model README
LoRA adapter for Qwen/Qwen3-Embedding-4B fine-tuned on the narrativeqa RAG retrieval dataset (DinoStackAI/narrativeqa-rag).
eval_narrativeqa-dev_cosine_ndcg@10 = 0.8110from sentence_transformers import SentenceTransformer
model = SentenceTransformer("DinoStackAI/Qwen3-Emb-4b-lora-narrativeqa")
embeddings = model.encode(["Instruct: ...\nQuery:your query", "document text"])
Or load the base model and adapter explicitly:
from sentence_transformers import SentenceTransformer
model = SentenceTransformer("Qwen/Qwen3-Embedding-4B")
model.load_adapter("DinoStackAI/Qwen3-Emb-4b-lora-narrativeqa")
from vllm import LLM
from vllm.lora.request import LoRARequest
llm = LLM(
model="Qwen/Qwen3-Embedding-4B",
task="embed",
enable_lora=True,
max_lora_rank=16,
)
outputs = llm.embed(
["Instruct: ...\nQuery:your query"],
lora_request=LoRARequest("narrativeqa", 1, "DinoStackAI/Qwen3-Emb-4b-lora-narrativeqa"),
)
Qwen/Qwen3-Embedding-4BDinoStackAI/narrativeqa-ragr=16, lora_alpha=32, targets q_proj / v_proj)