Status: Superseded by zenlm/zen3-vl. Weights remain available for reproducibility.

zen-vl-4b-agent

Superseded by zenlm/zen3-vl — canonical name.

Vision-language agent model for image understanding, OCR, and visual reasoning (4B (dense)).

Repackaged from Qwen/Qwen3-VL-4B-Instruct (apache-2.0, Alibaba Qwen). Not trained from scratch — a permissively-licensed redistribution for the OSS-clean Zen model line.

Specs

Property Value
Parameters 4B (dense)
Architecture Qwen3-VL
Modality text + image + video

License

apache-2.0. Upstream: Qwen/Qwen3-VL-4B-Instruct by Alibaba Qwen (apache-2.0).

Downloads last month
32
Safetensors
Model size
4B params
Tensor type
BF16
·
Inference Providers NEW
This model isn't deployed by any Inference Provider. 🙋 Ask for provider support

Model tree for zenlm/zen-vl-4b-agent

Finetuned
(389)
this model
Quantizations
2 models

Space using zenlm/zen-vl-4b-agent 1

Collection including zenlm/zen-vl-4b-agent