Downloads · 30 days
204
16% of all-time downloads
guoxuter/ov_intent_analysis_sft
ov_intent_analysis_sft is a text generation model from guoxuter. Use it when you need the model to write or continue text. It is set up for transformers. The card lists the license as apache-2.0.
ovintentanalysissft is a Qwen3.5-0.8B model fine-tuned for OpenViking retrieval intent analysis and query planning. Given recent conversation context and the current user message, it decides whether retrieval is neede…
Downloads · 30 days
204
16% of all-time downloads
All-time downloads
1.3K
Public
Parameters
853M
3.2 GB on disk
Likes
1
Public
Click a slice to open those files.
.safetensors3.2 GB · 99%
How the weights are stored.
F32752M · 88%
From the Hugging Face model README
ov_intent_analysis_sft is a Qwen3.5-0.8B model fine-tuned for OpenViking
retrieval intent analysis and query planning. Given recent conversation context and
the current user message, it decides whether retrieval is needed and emits structured
queries targeting OpenViking skill, resource, and memory scopes.
This repository contains the original Transformers checkpoint in Safetensors format.
The corresponding quantized Ollama release is
guoxuter/ov_intent_analysis_sft:v7_q8.
Use a recent Transformers version with Qwen3.5 support:
from transformers import AutoModelForCausalLM, AutoTokenizer
model_id = "guoxuter/ov_intent_analysis_sft"
tokenizer = AutoTokenizer.from_pretrained(model_id, trust_remote_code=True)
model = AutoModelForCausalLM.from_pretrained(
model_id,
dtype="auto",
device_map="auto",
trust_remote_code=True,
)
The model was trained for the OpenViking v7 retrieval prompt and structured output contract. For end-to-end use, prefer the prompt bundled with OpenViking rather than a generic chat prompt.
Qwen/Qwen3.5-0.8Bc92c878a96f34d2f0c87d2308099de7dc1401aae58aba0310aca550e9024b33baa98adccdec6a3be82d462563586abe1db520f93281ecc3f9bc3ff978b12d795The Safetensors checkpoint is the source artifact. The Ollama model is a Q8 GGUF derivative and should not be used to reconstruct full-precision weights.
This model is intended as a compact retrieval planner for OpenViking-compatible systems. It is not a general-purpose assistant. Outputs should be validated against the expected structured schema before they are executed or used for retrieval.
The base Qwen3.5-0.8B model is released under the Apache License 2.0. This fine-tuned checkpoint is published under the same license.