미디어 생성 RSS

모델 🇨🇳

Qwen VLo: From Understanding the World to Depicting It

Alibaba introduces Qwen VLo, a unified multimodal model capable of not only understanding images but also generating them with high quality. The model supports natural language instruction-based editing, style transfer, detection and segmentation, and works with both Chinese and English languages.

Alibaba/QwenAlibaba/Qwen
Alibaba Qwen27.07 · 18:04
모델 🇨🇳

Qwen-TTS speaks dialects: new speech synthesizer from Alibaba

Alibaba has unveiled an updated Qwen-TTS model (qwen-tts-latest), trained on millions of hours of speech. It supports three Chinese dialects and seven bilingual voices, and automatically adjusts prosody, tempo, and emotions according to the text. The model is available via the Qwen API.

Alibaba/QwenAlibaba/Qwen
Alibaba Qwen27.07 · 18:04
새로운 뉴스