Tongyi Qianwen has released Qwen3.8-Flash, a multimodal MoE model, as an early preview of the Qwen4
architecture and has made the weights open. The model has a total of 125B parameters, activating only 6B per token, with training costs only 1/9 of Qwen3.7-Plus, and its performance comprehensively surpasses the latter. The production API is priced at $0.16/1M input tokens and $0.47/1M output tokens, with a native context of 262K, extendable to 1M. 🔗 Read the full article via AIHOT · https://aihot.virxact.com/items/cmta9ciqt07c1roj219h5liie
通义千问发布 Qwen3.8-Flash,一款多模态 MoE 模型,作为 Qwen4 架构的早期预览并开放权重。
该模型总参数 125B,每 token 仅激活 6B,训练成本仅为 Qwen3.7-Plus 的 1/9,性能全面超越后者。
生产版 API 定价 $0.16/1M 输入 tokens 和 $0.47/1M 输出 tokens,原生上下文 262K,可扩展至 1M。
🔗 阅读原文 via AIHOT · https://aihot.virxact.com/items/cmta9ciqt07c1roj219h5liie
The readable text on this page was extracted from the public source and organized with attribution, publication time and the original link. Copyright remains with the original author and publisher.