conv.

All stories
AIQuiet 7d · day 9

Alibaba releases Qwen 3.8 Omni Flash multimodal model

New cloud-only omni model handles text, image, audio, and video inputs with text and speech outputs.

What to know

  • Qwen 3.8 Omni Flash is a closed-source, cloud-only multimodal model released September 14 that processes text, image, audio, and video inputs.
  • The model is available exclusively via Alibaba Cloud APIs across five geographic regions, with unified token pricing starting at $0.15 per million input tokens.
  • Only companion repositories are open-sourced; the model itself and its architecture remain proprietary, with community criticism focused on the lack of open weights.
some voices

The closed-source, cloud-only approach limits accessibility compared to open alternatives.

Alibaba Model developerQwen team Development and release team

Alibaba releases Qwen 3.8 Omni Flash multimodal model
qwen.ai

How it unfolded 2 developments, newest first · click a bar or a number to jump articlesposts

Peak 6 pieces in two hours at Sep 18, 3 AM; 27 pieces over 9 days (3 articles · 4 posts · 20 comments) Sep 17, 5 PM — 5 pieces · 2 articles · 3 posts — Hacker News 2, Newswires 2, X 1Sep 17, 7 PM — 1 piece · 1 comment — Hacker News 1Sep 17, 9 PM — 4 pieces · 1 post · 3 comments — Hacker News 3, X 1Sep 17, 11 PM — quietSep 18, 1 AM — 1 piece · 1 comment — Hacker News 1Sep 18, 3 AM — 6 pieces · 6 comments — Hacker News 6Sep 18, 5 AM — quietSep 18, 7 AM — 3 pieces · 3 comments — Hacker News 3Sep 18, 9 AM — 3 pieces · 3 comments — Hacker News 3Sep 18, 11 AM — 1 piece · 1 comment — Hacker News 1Sep 18, 1 PM — 1 piece · 1 comment — Hacker News 1Sep 18, 3 PM — quietSep 18, 5 PM — quietSep 18, 7 PM — quietSep 18, 9 PM — quietSep 18, 11 PM — quietSep 19, 1 AM — quietSep 19, 3 AM — quietSep 19, 5 AM — 1 piece · 1 comment — Hacker News 1Sep 19, 7 AM — quietSep 19, 9 AM — 1 piece · 1 article — Newswires 1Sep 19, 11 AM — quietSep 19, 1 PM — quietSep 19, 3 PM — quietSep 19, 5 PM — quietSep 19, 7 PM — quietSep 19, 9 PM — quietSep 19, 11 PM — quietSep 20, 1 AM — quietSep 20, 3 AM — quietSep 20, 5 AM — quietSep 20, 7 AM — quietSep 20, 9 AM — quietSep 20, 11 AM — quietSep 20, 1 PM — quietSep 20, 3 PM — quietSep 20, 5 PM — quietSep 20, 7 PM — quietSep 20, 9 PM — quietSep 20, 11 PM — quietSep 21, 1 AM — quietSep 21, 3 AM — quietSep 21, 5 AM — quietSep 21, 7 AM — quietSep 21, 9 AM — quietSep 21, 11 AM — quietSep 21, 1 PM — quietSep 21, 3 PM — quietSep 21, 5 PM — quietSep 21, 7 PM — quietSep 21, 9 PM — quietSep 21, 11 PM — quietSep 22, 1 AM — quietSep 22, 3 AM — quietSep 22, 5 AM — quietSep 22, 7 AM — quietSep 22, 9 AM — quietSep 22, 11 AM — quietSep 22, 1 PM — quietSep 22, 3 PM — quietSep 22, 5 PM — quietSep 22, 7 PM — quietSep 22, 9 PM — quietSep 22, 11 PM — quietSep 23, 1 AM — quietSep 23, 3 AM — quietSep 23, 5 AM — quietSep 23, 7 AM — quietSep 23, 9 AM — quietSep 23, 11 AM — quietSep 23, 1 PM — quietSep 23, 3 PM — quietSep 23, 5 PM — quietSep 23, 7 PM — quietSep 23, 9 PM — quietSep 23, 11 PM — quietSep 24, 1 AM — quietSep 24, 3 AM — quietSep 24, 5 AM — quietSep 24, 7 AM — quietSep 24, 9 AM — quietSep 24, 11 AM — quietSep 24, 1 PM — quietSep 24, 3 PM — quietSep 24, 5 PM — quietSep 24, 7 PM — quietSep 24, 9 PM — quietSep 24, 11 PM — quietYesterday, 1 AM — quietYesterday, 3 AM — quietYesterday, 5 AM — quietYesterday, 7 AM — quietYesterday, 9 AM — quietYesterday, 11 AM — quietYesterday, 1 PM — quietYesterday, 3 PM — quietYesterday, 5 PM — quietYesterday, 7 PM — quietYesterday, 9 PM — quietYesterday, 11 PM — quietToday, 1 AM — quietToday, 3 AM — quietToday, 5 AM — quietToday, 7 AM — quietToday, 9 AM — quietToday, 11 AM — quietToday, 1 PM — quiet 1–2
Sep 18Sep 19Sep 20Sep 21Sep 22Sep 23Sep 24yesterdaynow · 3:37 PM ET
  1. 2

    Community reacts to closed-source model

    Early community feedback noted the lack of open-source weights as a significant limitation, with responses characterizing the release as proprietary.

    “No OSS :/…”
    — tokenstead.ai community reaction
    • 🚀 Meet Qwen3.8-Omni-Flash, Qwen's first omni-modal model built around agentic capabilities! Native audio-video understanding, reasoning, and tool use come together in one model: understand the content, plan the task, execute with tools, and deliver the result. Highlights: 🥳 -

      @Alibaba_QwenX8d ago2.4k▲view on X ↗
    2 more of the top 3 · 21 posts in this stretch
    • I'm wondering, is there a tool or something out there that helps me pick a model, in the vast sea of models out there these days? Every time I need a model for something I see the list on openrouter and I'm completely overwhelmed.I'd love to be able to explain my use case, my cost preferences and have a tool select a few good models to try.E.g. I…

      mavamaartenHacker News8d agoview on Hacker News ↗
    • I use models.dev's CLI tool, which I think gets data from OpenRouter, and ArtificialAnalysis so your coding agent can help you narrow it down.<sidenote>Similarly, HuggingFace has a CLI + a few skills, and they are very useful.I had a production image processing using Gemini 2.5 Flash Lite (which is getting discontinued in October), and in 20…

      testycoolHacker News8d agoview on Hacker News ↗
    all of them →
  2. 1

    Technical details surface: pricing and API structure disclosed

    Technical details reveal the model is API-only on Alibaba Cloud endpoints (Beijing, Singapore, Hong Kong, Tokyo, Frankfurt, US-Virginia), with unified token billing at $0.15/1M input on Singapore and no open weights. Only companion repositories (Qwen-MM-Plugins, Qwen-Live-Harness) are open-sourced.

    “Native omni model, API-only. Text, image, audio, and video in; text out on the standard Chat Completions / Responses API, with synthesized speech out on the realtime variant (WebSocket/WebRTC).”
    — tokenstead.ai · source
    • NEW Qwen3.8-Omni-Flash 🔥🔥 They're going FULL OMNI - Text, image, audio, and video inputs - 1M context "Audio-visual performance close to Gemini 3.8 Flash and overall audio performance that exceeds Gemini 3.8 Flash". Available only through API for now. This could be THE model

      @MiaAI_labX8d ago60▲view on X ↗
  3. background

    Alibaba releases Qwen 3.8 Omni Flash — Alibaba's Qwen team released Qwen 3.8 Omni Flash, a native omni model supporting text, image, audio, and video inputs with text and synthesized speech outputs via cloud API only.

Also covered reported alongside — the timeline has no entry for these yet

  1. first by HN Best, 8d ago · also HN Frontpage, The Decoder

    2 more headlines

What people are saying 17 voices from 1 site · best of 22 · verbatim