通义千问发布第三代图像生成基座模型 Qwen-Image-3.0,核心关键词为"实"。
该模型支持最长 4.5k token 指令输入,可单次生成包含 9 个复杂信息图的 3×3 网格布局;
文本渲染精度达 10px,并支持 12 种语言原生渲染,旨在将图像转化为可部署的生产力工具。
Tongyi Qianwen has released the third-generation image generation foundation model Qwen-Image-3.0, with the core keyword being 'real'. This model supports...
Tongyi Qianwen has released the third-generation image generation foundation model Qwen-Image-3.0,
with the core keyword being 'real'. This model supports instruction inputs up to 4.5k tokens and can generate, in a single instance, a 3×3 grid layout containing 9 complex information graphics; the text rendering precision reaches 10px and it supports native rendering in 12 languages, aiming to turn images into deployable productivity tools.
通义千问发布第三代图像生成基座模型 Qwen-Image-3.0,核心关键词为"实"。
该模型支持最长 4.5k token 指令输入,可单次生成包含 9 个复杂信息图的 3×3 网格布局;
文本渲染精度达 10px,并支持 12 种语言原生渲染,旨在将图像转化为可部署的生产力工具。
通义千问发布第三代图像生成基座模型 Qwen-Image-3.0,以“实”为核心关键词,支持最长 4.5k token 指令输入、3×3 信息图网格生成、10px 文本渲染及 12 种语言原生渲染。
公开材料将 Qwen-Image-3.0 定位为第三代图像生成基座模型,并强调把图像转化为可部署的生产力工具。现有材料未说明发布时间、开放方式、定价、训练数据或实际部署案例。
Aioga 判断,此次更新的重点不只在单幅图像效果,而是长指令理解、复杂信息图编排与多语言文字渲染的组合;这些能力是否达到稳定生产要求,仍需更多公开测试与实际案例验证。
该模型可能适用于需要复杂版式、密集文字和多语言内容的图像生产场景。值得关注的是,4.5k token 输入、九图网格和 10px 文本渲染均为公开材料中的能力描述,尚不能直接等同于所有场景下的稳定表现。 后续应关注官方是否公布模型获取方式、技术报告、评测方法及部署条件,并通过不同语言、字号、复杂指令和九图网格任务复核生成一致性、文字准确性与实际可用性。
The readable text on this page was extracted from the public source and organized with attribution, publication time and the original link. Copyright remains with the original author and publisher.
Ingestion channel: REST API · Source domain: qwen.ai
Source: Qwen:Blog Retrieval(API)
Original link: Open original source
Aioga archive: Open intelligence page
Content record: summary-fallback · Updated: 2026-07-21T06:00:00.000Z

统一接入主流 AI 模型 API,为开发、测试与生产环境提供稳定调用入口。
立即访问 api.w173.comAioga aggregates global AI updates and preserves source information for verification and citation.