Qwen2.5-Omni-7B
READER SUMMARYAn open 7B end-to-end model for text, image, audio and video input with streaming text and natural-speech output.
Why this model scores 65.8
Unified real-time perception and speech output create broad multimodal agent capability. Small open weights drive deployment and downstream control difficulty, while size limits frontier reasoning.
0 comments · 0 votes
Sign in to join the discussion →
No comments yet. Start the discussion.
News tied to Qwen2.5-Omni-7B
The model score of 65.8 rates this model's risk profile. The overall Doom Index of 63.5 measures the recalibrated complete evidence corpus. These values answer different questions.
Each article's leave-one-out Doom Index contribution is divided equally among the exact models named on that article. This prevents multi-model evidence from being claimed in full on several model pages. Model risk scores are recalculated from their technical profile and related news, but never feed back into the overall index.
Alibaba releases open real-time Qwen2.5-Omni model
Alibaba's Qwen team released Qwen2.5-Omni-7B, an open end-to-end model that consumes text, images, audio and video while streaming both text and natural speech.
- Full item contribution
- +0.03
- Qwen2.5-Omni-7B equal share
- +0.03
Model score history
- R1Doom Score 65.8
Initial exact profile for the March 2025 open Qwen omni checkpoint.
12 Aug 2026