Downloads · 30 days
12
6% of all-time downloads
moushi21/agent-bench-dbbench-merged4
agent-bench-dbbench-merged4 is a text generation model from moushi21. Use it when you need the model to write or continue text. It is set up for transformers. The card lists the license as apache-2.0.
This repository provides a merged full-parameter model (bfloat16) fine-tuned from Qwen/Qwen3-4B-Instruct-2507.
Downloads · 30 days
12
6% of all-time downloads
All-time downloads
212
Public
Parameters
4B
8.1 GB on disk
Likes
0
Public
Click a slice to open those files.
.safetensors8 GB · 100%
From the Hugging Face model README
This repository provides a merged full-parameter model (bfloat16) fine-tuned from Qwen/Qwen3-4B-Instruct-2507.
Instead of a standalone LoRA adapter, this model has been created by merging LoRA weights back into the base model using Unsloth's merge_and_unload method. This ensures high-speed inference and easy deployment.
This model is specialized for DBBench trajectory tasks, trained to handle multi-turn environment observations and action selections.
merge_and_unload)Since this is a merged model, you can load it directly like any other Qwen3 model:
from transformers import AutoModelForCausalLM, AutoTokenizer
import torch
model_id = "moushi21/agent-bench-dbbench-merged4"
tokenizer = AutoTokenizer.from_pretrained(model_id)
model = AutoModelForCausalLM.from_pretrained(
model_id,
torch_dtype=torch.bfloat16,
device_map="auto"
)
Training data:
Dataset License: MIT License. This dataset is used and distributed under the terms of the MIT License. Compliance: Users must comply with the MIT license (including copyright notice) and the base model's original terms of use.