Qwen2.5 VL 72B Instruct
Qwen🇨🇳
qwen/qwen2.5-vl-72b-instructQwen 2.5 VL 72B Instruct is Alibaba's large open-weight vision-language model, capable of image understanding, document/OCR parsing, and visual reasoning alongside text. It is designed for multimodal tasks that combine strong language ability with detailed visual comprehension.
费用比率
0.3x
上下文
128K
发布日期
2025年2月1日
输入
TextImage
输出
Text
支持
推理工具调用结构化输出
最擅长
该模型排名最高的类别。
图像理解
#123 · 前 77%