Models · Image and video · Agents
Qwen3.8-Omni-Flash is out, handling text, image, audio and video in one model
Alibaba released it on September 18. Rather than stitching separate specialist systems together, it processes all four modes inside one model, with a one million token context window. Thinking mode, tool calling and web search are supported, so a single call can watch a clip and then act on what it saw. Alibaba says its own measurement shows the average across 30 evaluations improved by more than 26% over Qwen3.5-Omni-Plus. No open weights were released, so self-hosting is not an option.
Sources
- Qwen (Alibaba), “Qwen3.8-Omni-Flash: Omni Senses. Agentic Delivery.”, (qwen.ai)
- Alibaba Cloud Community, “Qwen3.8-Omni-Flash: Omni Senses. Agentic Delivery.”, (alibabacloud.com)
- TechNode, “Alibaba’s Qwen releases Qwen3.8-Omni-Flash with 1M-token context”, (technode.com)
About this story
This story was posted on Instagram by @jarrus.tech on Sept. 24, 2026.
Spotted an error in this story? [email protected] · Instagram
Short link: thejarrus.com/en/h7-48
This story in Turkish: Qwen3.8-Omni-Flash çıktı, metin, görsel, ses ve videoyu tek modelde işliyor
On the same topic
Anthropic releases Claude Haiku 5.5 at 90% lower prices
Claude Haiku 5.5, released Oct. 7, costs 90% less than Haiku 4.5 for prompts up to 100,000 tokens. It leads on benchmarks but uses more tokens per task.
OpenAI is rolling out GPT-6 to everyone in ChatGPT
OpenAI began rolling out GPT-6 to everyone in ChatGPT on October 7. Its new Intelligent UI builds answers from text, visuals and interactive elements.
The open model war between Europe, the US and China heats up
Mistral unveiled Large 4 on October 6. The same week, Ecosia dropped Mistral, Reflection AI launched Beam and DeepSeek was reported near a $12 billion round.