Qwen VLo: Unified multimodal understanding and generation model from Alibaba
Alibaba releases Qwen VLo, a unified multimodal model that both understands and generates images. It supports open-ended editing, multilingual instructions, dynamic resolution, and progressive generation from top to bottom and left to right. Available as a preview on Qwen Chat.
Alibaba/Qwen
Alibaba Qwen27.07 · 18:04
