What is Qwen2.5-Omni?
Qwen2.5-Omni handles text, images, audio, and video within one multimodal model. It can respond with text or streaming speech for conversational applications.
Combine several input formats with spoken model output
Qwen2.5-Omni handles text, images, audio, and video within one multimodal model. It can respond with text or streaming speech for conversational applications.
Ask the makers a question, share feedback or tell others how you use Qwen2.5-Omni.
Sign in to comment, ask the makers a question or share useful feedback.
No comments yet. Be the first to share your thoughts on Qwen2.5-Omni.