Downloads · 30 days
14.5K
6% of all-time downloads
browser-use/bu-30b-a3b-preview
bu-30b-a3b-preview is a image-text-to-text model from browser-use. Use it for the image-text-to-text task on the model card, and read the license before you ship it in a product. It is set up for transformers.
<picture <source media="(prefers-color-scheme: light)" srcset="https://github.com/user-attachments/assets/2ccdb752-22fb-41c7-8948-857fc1ad7e24"" <source media="(prefers-color-scheme: dark)" srcset="https://github.com/…
Downloads · 30 days
14.5K
6% of all-time downloads
All-time downloads
233K
Public
Parameters
31.1B
62.2 GB on disk
Likes
266
Trending 1
Click a slice to open those files.
.safetensors62.1 GB · 100%
From the Hugging Face model README
Meet BU-30B-A3B-Preview — bringing SoTA Browser Use capabilities in a small model that can be hosted on a single GPU.
This model is heavily trained to be used with browser-use OSS library and provides comprehensive browsing capabilities with superior DOM understanding and visual reasoning.
You can directly use this model at BU Cloud. Simply
from dotenv import load_dotenv
from browser_use import Agent, ChatBrowserUse
load_dotenv()
llm = ChatBrowserUse(
model='browser-use/bu-30b-a3b-preview', # BU Open Source Model!!
)
agent = Agent(
task='Find the number of stars of browser-use and stagehand. Tell me which one has more stars :)',
llm=llm,
flash_mode=True
)
agent.run_sync()
We recommend using this model with vLLM.
Make sure to install vllm >= 0.12.0:
pip install vllm --upgrade
A simple launch command is:
vllm serve browser-use/bu-30b-a3b-preview \
--max-model-len 65536 \
--host 0.0.0.0 \
--port 8000
which will create an OpenAI compatible endpoint at localhost that you can use with.
from dotenv import load_dotenv
from browser_use import Agent, ChatOpenAI
load_dotenv()
llm = ChatOpenAI(
base_url='http://localhost:8000/v1',
model='browser-use/bu-30b-a3b-preview',
temperature=0.6,
top_p=0.95,
dont_force_structured_output=True, # speed up by disabling structured output
)
agent = Agent(
task='Find the number of stars of browser-use and stagehand. Tell me which one has more stars :)',
llm=llm,
)
agent.run_sync()
| Property | Value |
|---|---|
| Base Model | Qwen/Qwen3-VL-30B-A3B-Instruct |
| Parameters | 30B total, 3B active (MoE) |
| Context Length | 65,536 tokens |
| Architecture | Vision-Language Model (Mixture of Experts) |