商汤发布并完全开源 SenseNova-Vision-7B-MoT,一个统一处理检测、OCR、GUI、深度与法线估计、分割、多视图等主要视觉任务的模型。
该模型支持通过自然语言定义新的视觉任务变体,跨传统任务边界重组视觉能力。
开源内容包括模型权重及 SenseNova-Vision Corpus(含 5000 万示例子集及复现剩余公开数据的完整工具包)。
SenseTime has released and fully open-sourced SenseNova-Vision-7B-MoT, a model for unified processing of major visual tasks such as inspection, OCR, GUI, d...
SenseTime has released and fully open-sourced SenseNova-Vision-7B-MoT, a model for unified
processing of major visual tasks such as inspection, OCR, GUI, depth and normal estimation, segmentation, and multi-view. The model supports defining new visual task variants through natural language, reorganizing visual capabilities across traditional task boundaries. The open-source content includes model weights and SenseNova-Vision Corpus (a complete toolkit with a subset of 50 million examples and reproduction of remaining public data).
商汤发布并完全开源 SenseNova-Vision-7B-MoT,一个统一处理检测、OCR、GUI、深度与法线估计、分割、多视图等主要视觉任务的模型。
该模型支持通过自然语言定义新的视觉任务变体,跨传统任务边界重组视觉能力。
开源内容包括模型权重及 SenseNova-Vision Corpus(含 5000 万示例子集及复现剩余公开数据的完整工具包)。
商汤发布并完全开源 SenseNova-Vision-7B-MoT。该模型统一处理检测、OCR、GUI、深度与法线估计、分割及多视图等视觉任务。
公开材料称,该模型可通过自然语言定义新的视觉任务变体,并跨传统任务边界重组视觉能力,定位于统一处理多类主要视觉任务。
Aioga 判断,其值得关注之处在于将多类视觉任务纳入同一模型,并允许以自然语言定义任务变体;实际效果仍需结合后续公开验证判断。
Aioga 判断,模型权重、语料子集及数据复现工具包的同步开放,可能为研究者开展复现、任务适配和跨任务实验提供公开基础。 值得关注后续是否出现基于公开权重与语料工具包的独立复现,以及模型在检测、OCR、GUI、分割和多视图等任务中的公开测试结果。
The readable text on this page was extracted from the public source and organized with attribution, publication time and the original link. Copyright remains with the original author and publisher.
Ingestion channel: Summary aggregation · Source domain: x.com
Source: X:商汤 SenseTime (@SenseTime_AI)
Original link: Open original source
Aioga archive: Open intelligence page
Content record: social-summary · Updated: 2026-07-14T00:38:32.000Z

统一接入主流 AI 模型 API,为开发、测试与生产环境提供稳定调用入口。
立即访问 api.w173.comAioga aggregates global AI updates and preserves source information for verification and citation.