Skip to content
Türkçe

Models · Image and video · Agents

Qwen3.8-Omni-Flash is out, handling text, image, audio and video in one model

Published: 3 sourcesTürkçe

Alibaba released it on September 18. Rather than stitching separate specialist systems together, it processes all four modes inside one model, with a one million token context window. Thinking mode, tool calling and web search are supported, so a single call can watch a clip and then act on what it saw. Alibaba says its own measurement shows the average across 30 evaluations improved by more than 26% over Qwen3.5-Omni-Plus. No open weights were released, so self-hosting is not an option.

Sources

  1. Qwen (Alibaba), “Qwen3.8-Omni-Flash: Omni Senses. Agentic Delivery.”, (qwen.ai)
  2. Alibaba Cloud Community, “Qwen3.8-Omni-Flash: Omni Senses. Agentic Delivery.”, (alibabacloud.com)
  3. TechNode, “Alibaba’s Qwen releases Qwen3.8-Omni-Flash with 1M-token context”, (technode.com)

About this story

This story was posted on Instagram by @jarrus.tech on Sept. 24, 2026.

Spotted an error in this story? [email protected] · Instagram

This story in Turkish: Qwen3.8-Omni-Flash çıktı, metin, görsel, ses ve videoyu tek modelde işliyor

On the same topic