Qwen VL Qwen VL 🤗 🤖   | Qwen VL Chat 🤗 🤖   (Int4: 🤗 🤖  ) | Qwen VL Plus 🤗 🤖   | Qwen VL Max 🤗 🤖   Web       API       WeChat       Discord       Paper       Tutorial Qwen VL 是阿里云研发的大规模视觉语言模型(Large Vision Language Model, LVLM)。Qwen VL 可以以图像、文本、检测框作为输入,并以文本和检测框作为输出。Qwen VL 系列模型性能强大,具备多语言对话、多图交错对话等能力,并支持中文开放域定位和细粒度图像识别与理解。 Qwen VL (Qwen Large Vision Language Model) is the visual multimodal version of the large model series, Qwen (abbr. Tongyi Qianwen), proposed by Alibaba Cloud. Qwen VL accepts image, text, and bounding box as inputs, outputs text and bounding box. The features of Qwen VL include: 目前,我们提供了Qwen VL和Qwen VL Chat两个模型,分别为预训练模型和Chat模型。如果想了解更多关于模型的信息,请点击链接查看我们的技术备忘录。本仓库为Qwen VL Chat仓库。 We release Qwen VL and Qwen VL Chat, which are pretrained model and Chat model respectively. For more details about Qwen VL, please refer to our technical memo. This repo is the one for Qwen VL. 安装要求 (Requirements) python 3.8及以上版本 pytorch 1.12及以上版本,推荐2.0及以上版本 建议使用CUDA 11.4及以上(GPU用户需考虑此选项) python 3.8 and above pytorch 1.12 and above, 2.0 and above are recommended CUDA 11.4 and above are…
We use cookies for essential functionality and analytics. You can accept or reject analytics cookies.Cookie policy