{"@context":"https://schema.org","@type":"NewsArticle","generatedAt":"2026-08-11T09:21:12.743Z","headline":"开源权重模型逼近前沿能力，安全差距仍在扩大","description":"SaferAI 报告显示，中国 Z.ai 的开源权重模型 GLM-5.2 在网络与生物能力上仅落后 OpenAI GPT-5.5 和 Anthropic Claude Opus 4.7 数月，但面对攻击性网络或双重用途生物任务时拒绝率为零，而 Claude Opus 4.7 因拒绝过于一致导致 CyberGym 评测无法完成。","url":"https://www.aioga.com/news/cmsf48xvn1fozro2eb9lzel3p/","mainEntityOfPage":"https://www.aioga.com/news/cmsf48xvn1fozro2eb9lzel3p/","datePublished":"2026-08-04T20:05:26.000Z","dateModified":"2026-08-04T20:05:26.000Z","inLanguage":"zh-CN","publisher":{"@type":"NewsMediaOrganization","name":"Aioga","url":"https://www.aioga.com"},"citation":["https://techcrunch.com/2026/08/04/open-weight-ai-models-are-catching-up-to-the-frontier-the-safety-gap-remains","https://aihot.virxact.com/items/cmsf48xvn1fozro2eb9lzel3p"],"canonicalUrl":"https://www.aioga.com/news/cmsf48xvn1fozro2eb9lzel3p/","directAnswer":{"@type":"Answer","text":"SaferAI 经由 Z.ai 公共 API 评估称，GLM-5.2 在网络与生物能力上仅落后 GPT-5.5 和 Claude Opus 4.7 数月；面对测试中的攻击性网络及双重用途生物任务时，该模型未拒绝任何一项。","url":"https://www.aioga.com/news/cmsf48xvn1fozro2eb9lzel3p/","dateCreated":"2026-08-04T20:05:26.000Z","author":{"@type":"Organization","@id":"https://www.aioga.com/authors/aioga-editorial/#editorial-team","name":"Aioga Editorial Team","url":"https://www.aioga.com/authors/aioga-editorial/"}},"evidence":[{"@type":"CreativeWork","name":"TechCrunch source article","url":"https://techcrunch.com/2026/08/04/open-weight-ai-models-are-catching-up-to-the-frontier-the-safety-gap-remains","datePublished":"2026-08-04T20:05:26.000Z","provider":{"@type":"Organization","name":"TechCrunch","url":"https://techcrunch.com/2026/08/04/open-weight-ai-models-are-catching-up-to-the-frontier-the-safety-gap-remains"}},{"@type":"CreativeWork","name":"AIHot archive record","url":"https://aihot.virxact.com/items/cmsf48xvn1fozro2eb9lzel3p","datePublished":"2026-08-04T20:05:26.000Z","provider":{"@type":"Organization","name":"AIHot","url":"https://aihot.virxact.com/items/cmsf48xvn1fozro2eb9lzel3p"}}],"aggregationSource":"TechCrunch：AI（RSS）","originalPublisher":{"name":"TechCrunch","url":"https://techcrunch.com/2026/08/04/open-weight-ai-models-are-catching-up-to-the-frontier-the-safety-gap-remains"},"geoDeepAnswer":null,"article":{"id":"cmsf48xvn1fozro2eb9lzel3p","slug":"cmsf48xvn1fozro2eb9lzel3p","url":"https://www.aioga.com/news/cmsf48xvn1fozro2eb9lzel3p/","title":"开源权重模型逼近前沿能力，安全差距仍在扩大","title_en":"Open-weight AI models are catching up to the frontier. The safety gap remains.","summary":"SaferAI 报告显示，中国 Z.ai 的开源权重模型 GLM-5.2 在网络与生物能力上仅落后 OpenAI GPT-5.5 和 Anthropic Claude Opus 4.7 数月，但面对攻击性网络或双重用途生物任务时拒绝率为零，而 Claude Opus 4.7 因拒绝过于一致导致 CyberGym 评测无法完成。","source":"TechCrunch：AI（RSS）","sourceUrl":"https://techcrunch.com/2026/08/04/open-weight-ai-models-are-catching-up-to-the-frontier-the-safety-gap-remains","aiHotUrl":"https://aihot.virxact.com/items/cmsf48xvn1fozro2eb9lzel3p","publishedAt":"2026-08-04T20:05:26.000Z","category":"技巧观点","score":69,"selected":false,"articleBody":["As policymakers debate how to govern increasingly powerful AI systems like OpenAI’s GPT-5.6 Sol and Anthropic’s Mythos, a Chinese open-weight model has narrowed the gap with the industry’s leaders.","GLM-5.2, the open-weight AI model from China’s Z.ai, is only a few months behind OpenAI’s GPT-5.5 and Anthropic’s Claude Opus 4.7 on cyber and bio capabilities, according to a new report：https://www.safer-ai.org/research/glm-5-2-evaluation-report from AI safety nonprofit SaferAI. But the divide between frontier capabilities and safety practices is growing.","According to SaferAI’s evaluation, which the nonprofit ran via Z.ai’s public API, GLM-5.2 refused none of the offensive cyber or dual-use biology tasks it was given. By comparison, Claude Opus 4.7 “refused so consistently that SaferAI could not complete CyberGym on it at all.” (CyberGym is a benchmark that evaluates cybersecurity capabilities. OpenAI used it in the evaluation that preceded last month’s Hugging Face breach：https://techcrunch.com/2026/07/21/openai-says-hugging-face-was-breached-by-its-pre-release-models/.)","It’s a stark reminder of what some critics have warned for years: that open-weight AI models could put highly capable AI into the hands of potential attackers, with no way to police how they use the technology once they download the weights. With open-weight models rapidly approaching the capabilities of the world’s leading AI systems, the debate is moving from whether they can compete to how society manages risks once they are released.","“The frontier of capability is not the frontier of risk, and so we do have to take into account the state of the mitigations as well to assess the risk properly,” Henry Papadatos, executive director of SaferAI, told TechCrunch.","While Z.ai could apply safety measures to its hosted API, those protections become unenforceable once someone runs the weights on their own hardware, where they can remove or modify any safeguards, fine-tune the models, or change system prompts.","Frontier developers like OpenAI and Anthropic tend to rely on safeguards like classifiers, refusal training, and API-level controls to limit dangerous cyber and biological assistance.","Those measures are far from foolproof: jailbreaks routinely bypass protections on deployed models. Far.ai, an AI safety nonprofit, found hundreds of universal jailbreaks：https://leaderboard.far.ai/ — defined as reusable keys that succeed on most harmful requests — in frontier models like xAI’s Grok 4.5 and Google DeepMind’s Gemini 3.1 Pro. According to the report, jailbreaks succeed when attackers combine multiple manipulation techniques — including roleplaying, authority impersonation, fake conversation history, and follow-up prompts — to amplify weak points in a model’s defenses.","But the safeguards in place for closed models don’t work at all on open-weight models, which are designed to run on any infrastructure with any set of safeguards — or lack thereof.","“The objective should clearly be that the good capabilities — the safe ones — are accessible to anyone, and then we try to remove the bad ones, even in an open source fashion,” Papadatos said.","One technique Papadatos noted could help is called “pre-training data filtering,” which is when an AI company removes offensive cybersecurity information from their training data and then trains the model on the curated dataset.","Some research：https://alignment.anthropic.com/2025/pretraining-data-filtering/?utm suggests this can reduce hazardous biological knowledge：https://arxiv.org/abs/2508.06601#:~:text=In%20this%20paper%2C%20we%20investigate,as%20a%20more%20tamper%2Dresistant%20safeguard. without harming overall model performance. However, for cybersecurity, data filtering is much less practical.","It’s difficult to train a general model that excels at coding but isn’t also a good hacker. Because coding has become AI’s biggest moneymaker, developers face pressure to keep improving those capabilities even as they search for ways to limit misuse.","Because of that, frontier developers have increasingly relied on other mitigations instead. One approach has been to selectively restrict the kinds of cybersecurity assistance models will provide. Anthropic’s Opus 5, for example, can search for vulnerabilities in uncompiled source code, but not compiled software, per the model’s system card：https://www-cdn.anthropic.com/b514064af1408018e64b1ad24e7d5e75850b4ffd/Claude%20Opus%205%20System%20Card.pdf. The reasoning is that this makes it harder to use Opus 5 for offensive purposes.","Others include rigorous pre-deployment safety evaluations, publishing risk assessments, and withholding model weights if a system is perceived as too dangerous.","In GLM-5.2’s case, SaferAI says Z.ai didn’t publish a safety framework, pre-deployment testing commitments, or risk assessment for the model. TechCrunch has asked Z.ai whether it conducted internal or third-party frontier safety evaluations before release, but did not receive a response.","Chinese leaders have increasingly acknowledged the risks of advanced AI. At the World AI Conference last month, Chinese President Xi Jinping emphasized：https://www.nytimes.com/2026/07/17/business/xi-jinping-china-ai.html?eafs_enabled=false the importance of open-weight models, while also stressing the necessity of ensuring AI remains a tool under strict human control.","Graham Webster, who studies Chinese AI policy at the Stanford Cyber Policy Center, told TechCrunch that China has robust regulations governing AI, but those rules have historically focused on politically sensitive content, misinformation, and social stability rather than catastrophic AI risks like offensive cyber capabilities and biological misuse.","“U.S. AI thinkers are, in general, more concerned with this existential catastrophic [idea] than the Chinese community,” Webster said, adding that many Chinese policy researchers believe that if there’s truly going to be a novel frontier risk, American companies will likely encounter it first.","“The Chinese system has confidence that they control the use of these technologies inside China,” Webster continued. “Being online in China is something you do attributed to your real name, and companies can be held accountable, users can be held accountable.”","Webster mused that the same mechanism that model providers use for refusing to engage on certain political topics can potentially be tweaked to make sure models refuse to complete offensive cyber attacks or won’t deliver adverse biological engineering outcomes. He added that because Chinese companies tend to coordinate with regulators behind the scenes, it can be tough to know what internal testing they’re conducting before release.","Advocates of open-weight AI argue that releasing the weights is important for cybersecurity because it allows companies defend themselves against attacks — Hugging Face relied on GLM-5.2 to defend itself against OpenAI’s breach — and because it allows them to better prepare for future threats if they know what’s coming.","“The same systems that helped stop an AI-powered cyberattack can now help defend against millions of cyberattacks every day, while helping us identify and fix vulnerabilities before attackers exploit them,” Clem Delangue, CEO of Hugging Face, said this week in a social media post：https://x.com/ClementDelangue/status/2083908468285620415.","Papadatos said that benefit is often overstated, and doesn’t mean “we should open-source dangerous capabilities.”","“The main point in my mind is that we shouldn’t just accept that dangerous capabilities are easily accessible by anyone anywhere,” he said, stressing that he believes the industry should be striving for only making the “good capabilities” easily accessible. By default attackers adopt new tools faster than defenders do. For example, a ransomware group can change its methods in a week. A hospital cannot.”","When you purchase through links in our articles, we may earn a small commission：https://techcrunch.com/techcrunch-affiliate-monetization-standards/. This doesn’t affect our editorial independence.","Rebecca Bellan is a senior reporter at TechCrunch where she covers the business, policy, and emerging trends shaping artificial intelligence. Her work has also appeared in Forbes, Bloomberg, The Atlantic, The Daily Beast, and other publications.","You can contact or verify outreach from Rebecca by emailing rebecca.bellan@techcrunch.com：mailto:rebecca.bellan@techcrunch.com or via encrypted message at rebeccabellan.491 on Signal.","Scale faster. Grow your portfolio. Gain practical expertise. No matter your goal, Disrupt can empower you. Save up to $330 toda y!","Influencers draw backlash for attending OpenAI’s first luxury trip：https://techcrunch.com/2026/08/03/influencers-draw-backlash-for-attending-openais-first-luxury-trip/ Dominic-Madori Davis：https://techcrunch.com/author/dominic-madori-davis/","Sequoia’s Shaun Maguire leads $1B round for nuclear startup Valar Atomics：https://techcrunch.com/2026/08/03/sequoias-shaun-maguire-leads-1b-round-for-nuclear-startup-valar-atomics/ Julie Bort：https://techcrunch.com/author/julie-bort/","YouTuber Hank Green says his AI usage is ‘not healthy’：https://techcrunch.com/2026/08/01/youtuber-hank-green-says-his-ai-usage-is-not-healthy/ Anthony Ha：https://techcrunch.com/author/anthony-ha/","WhatsApp is testing a new folder for messages from large businesses：https://techcrunch.com/2026/07/31/whatsapp-is-testing-a-new-folder-for-messages-from-large-businesses/ Ivan Mehta：https://techcrunch.com/author/ivan-mehta/","Spotify adds a running mode to its app：https://techcrunch.com/2026/07/30/spotify-adds-a-running-mode-to-its-app/ Ivan Mehta：https://techcrunch.com/author/ivan-mehta/","Claude Opus 5 became downright ruthless when tasked with running a vending machine：https://techcrunch.com/2026/07/29/claude-opus-5-became-downright-ruthless-when-tasked-with-running-a-vending-machine/ Julie Bort：https://techcrunch.com/author/julie-bort/","DoorDash is building its own drone delivery business：https://techcrunch.com/2026/07/29/doordash-is-building-its-own-drone-delivery-business/ Kirsten Korosec：https://techcrunch.com/author/kirsten-korosec/"],"articleImages":[{"sourceUrl":"https://techcrunch.com/wp-content/uploads/2025/04/Disrupt2026-Color.png","alt":"Event Logo","afterParagraph":27,"url":"/media/articles/cmsf48xvn1fozro2eb9lzel3p/d1d06f6f97361f52.webp"}],"mediaStatus":"ok","articleBodyZh":["随着政策制定者就如何管理日益强大的人工智能系统（如 OpenAI 的 GPT-5.6 Sol 和 Anthropic 的 Mythos）展开讨论，一款中国的开放权重模型正在缩小与行业领先者的差距。","根据 AI 安全非营利组织 SaferAI 的一份新报告（https://www.safer-ai.org/research/glm-5-2-evaluation-report），中国 Z.ai 的开放权重 AI 模型 GLM-5.2 在网络和生物能力方面仅落后于 OpenAI 的 GPT-5.5 和 Anthropic 的 Claude Opus 4.7 几个月。但前沿能力与安全实践之间的差距正在扩大。","根据 SaferAI 的评估（该非营利组织通过 Z.ai 的公共 API 进行评估），GLM-5.2 没有拒绝任何给定的攻击性网络或双用途生物任务。相比之下，Claude Opus 4.7 “拒绝的频率如此之高，以至于 SaferAI 完全无法在其上完成 CyberGym。”（CyberGym 是一个评估网络安全能力的基准。OpenAI 在上个月 Hugging Face 数据泄露事件之前的评估中使用了该基准：https://techcrunch.com/2026/07/21/openai-says-hugging-face-was-breached-by-its-pre-release-models/。）","这强烈提醒了多年来一些批评者的警告：开放权重 AI 模型可能会将高能力 AI 授予潜在攻击者，而一旦下载了权重，就无法监督他们如何使用该技术。随着开放权重模型迅速接近世界领先 AI 系统的能力，辩论正从它们是否能竞争转向社会在其发布后如何管理风险。","“能力的前沿并不等同于风险的前沿，因此我们确实必须考虑缓解措施的状态，以正确评估风险，”SaferAI 执行董事 Henry Papadatos 告诉 TechCrunch。","虽然 Z.ai 可以对其托管的 API 采取安全措施，但一旦有人在自己的硬件上运行权重，这些保护措施就无法执行，因为他们可以删除或修改任何安全防护，对模型进行微调，或更改系统提示。","像 OpenAI 和 Anthropic 这样的前沿开发者往往依赖分类器、拒绝训练和 API 级别控制等安全措施来限制危险的网络和生物帮助。","这些措施远非万无一失：越狱程序经常绕过已部署模型的保护措施。非营利组织 Far.ai（致力于 AI 安全）在前沿模型中发现了数百个通用越狱程序：https://leaderboard.far.ai/ —— 被定义为可重复使用的关键键，在大多数有害请求中都能成功 —— 例如 xAI 的 Grok 4.5 和 Google DeepMind 的 Gemini 3.1 Pro。根据报告，当攻击者结合多种操纵技术——包括角色扮演、冒充权威、伪造对话历史和后续提示——来放大模型防御的弱点时，越狱程序就会成功。","但针对封闭模型的保护措施在开放权重模型上根本不起作用，后者被设计为可在任何基础设施上运行，无论是否有任何保护措施。","“目标显然应该是让好的能力——安全的能力——对任何人都可访问，然后我们尝试去除不良能力，即使是以开源的方式，”Papadatos 说。","Papadatos 提到的一种可能有帮助的技术称为“预训练数据过滤”，即 AI 公司从其训练数据中移除具有攻击性的网络安全信息，然后在经过整理的数据集上训练模型。","一些研究：https://alignment.anthropic.com/2025/pretraining-data-filtering/?utm 表明，这可以在不影响整体模型性能的情况下减少有害的生物学知识：https://arxiv.org/abs/2508.06601#:~:text=In%20this%20paper%2C%20we%20investigate,as%20a%20more%20tamper%2Dresistant%20safeguard. 但是，对于网络安全而言，数据过滤的实用性要差得多。","训练一个在编码上表现出色但同时又不是优秀黑客的一般模型非常困难。由于编码已成为 AI 最大的盈利来源，开发者面临着保持这些能力持续改进的压力，即使他们在寻找限制滥用的方法时也是如此。","因此，前沿开发者越来越多地依赖其他缓解措施。一个方法是有选择地限制模型提供的网络安全援助类型。例如，Anthropic 的 Opus 5 可以搜索未编译源码中的漏洞，但不能搜索已编译的软件，根据该模型的系统卡：https://www-cdn.anthropic.com/b514064af1408018e64b1ad24e7d5e75850b4ffd/Claude%20Opus%205%20System%20Card.pdf。其原因是，这使得使用 Opus 5 进行进攻性操作更加困难。","其他措施包括严格的部署前安全评估、发布风险评估，以及如果系统被认为过于危险而不公开模型权重。","在 GLM-5.2 的案例中，SaferAI 表示 Z.ai 没有发布安全框架、部署前测试承诺或模型的风险评估。TechCrunch 已向 Z.ai 询问其在发布前是否进行了内部或第三方前沿安全评估，但未收到回应。","中国领导人越来越承认先进 AI 的风险。在上个月的世界人工智能大会上，中国国家主席习近平强调：https://www.nytimes.com/2026/07/17/business/xi-jinping-china-ai.html?eafs_enabled=false 开放权重模型的重要性，同时也强调确保 AI 仍在严格人类控制下的必要性。","斯坦福大学网络政策中心研究中国 AI 政策的 Graham Webster 告诉 TechCrunch，中国有健全的 AI 监管法规，但这些规则历来主要关注政治敏感内容、虚假信息和社会稳定，而非灾难性 AI 风险，如进攻性网络能力和生物滥用。","“总体来说，美国的 AI 思想家更关注这一生存性灾难威胁，而中国社区不太关注，”Webster 说，并补充道，许多中国政策研究人员认为，如果真的出现新的前沿风险，美国公司可能会首先遇到它。","“中国体系有信心能够控制这些技术在中国境内的使用，”Webster 继续说道。“在中国上网是需要实名的，企业可以被追责，用户也可以被追责。”","韦伯斯特思考道，模型提供者用于拒绝涉及某些政治话题的同样机制，也可以潜在地调整，以确保模型拒绝执行攻击性网络攻击或不会产生不良的生物工程结果。他补充道，由于中国公司往往在幕后一同与监管机构协调，因此在发布前很难知道他们进行了哪些内部测试。","支持开放权重AI的倡导者认为，发布权重对于网络安全很重要，因为它使公司能够防御攻击——Hugging Face 曾依靠 GLM-5.2 来防御 OpenAI 的泄露——且因为如果公司知道未来可能出现的威胁，就能更好地做准备。","“Hugging Face CEO Clem Delangue 本周在社交媒体帖子中表示：‘同样帮助阻止 AI 驱动的网络攻击的系统，现在可以帮助防御每天数百万次的网络攻击，同时帮助我们在攻击者利用漏洞之前识别并修复这些漏洞。’ https://x.com/ClementDelangue/status/2083908468285620415","Papadatos 表示，这一好处往往被高估，并不意味着‘我们应该开源危险能力。’","“我认为的主要观点是，我们不应该仅仅接受危险能力可以被任何人在任何地方轻易访问，”他表示，强调他认为行业应致力于只让‘良性能力’易于访问。默认情况下，攻击者采用新工具的速度比防御者快。例如，一个勒索软件团体可以在一周内改变其方法，而医院无法做到。”","当你通过我们文章中的链接购买时，我们可能会获得一小部分佣金：https://techcrunch.com/techcrunch-affiliate-monetization-standards/。这不会影响我们的编辑独立性。","Rebecca Bellan 是 TechCrunch 的高级记者，报道塑造人工智能的商业、政策和新兴趋势。她的作品也曾刊登在 Forbes、Bloomberg、The Atlantic、The Daily Beast 及其他出版物上。","你可以通过发送邮件至 rebecca.bellan@techcrunch.com 联系或核实 Rebecca 的联系信息：mailto:rebecca.bellan@techcrunch.com，或通过 Signal 上的加密消息 rebeccabellan.491 联系。","更快扩大规模。发展你的投资组合。获得实用经验。无论你的目标是什么，Disrupt 都能赋能你。今天最多可节省 330 美元！","网红因参加 OpenAI 的首次奢华旅行而引发反弹：https://techcrunch.com/2026/08/03/influencers-draw-backlash-for-attending-openais-first-luxury-trip/ Dominic-Madori Davis：https://techcrunch.com/author/dominic-madori-davis/","红杉的 Shaun Maguire 领导对核能初创公司 Valar Atomics 的 10 亿美元融资轮：https://techcrunch.com/2026/08/03/sequoias-shaun-maguire-leads-1b-round-for-nuclear-startup-valar-atomics/ Julie Bort：https://techcrunch.com/author/julie-bort/","YouTuber Hank Green 表示他的人工智能使用“不健康”：https://techcrunch.com/2026/08/01/youtuber-hank-green-says-his-ai-usage-is-not-healthy/ Anthony Ha：https://techcrunch.com/author/anthony-ha/","WhatsApp 正在测试一个用于大型企业消息的新文件夹：https://techcrunch.com/2026/07/31/whatsapp-is-testing-a-new-folder-for-messages-from-large-businesses/ Ivan Mehta：https://techcrunch.com/author/ivan-mehta/","Spotify向其应用添加了跑步模式：https://techcrunch.com/2026/07/30/spotify-adds-a-running-mode-to-its-app/ 伊万·梅塔：https://techcrunch.com/author/ivan-mehta/","Claude Opus 5 在被分配管理自动售货机任务时变得相当无情：https://techcrunch.com/2026/07/29/claude-opus-5-became-downright-ruthless-when-tasked-with-running-a-vending-machine/ Julie Bort：https://techcrunch.com/author/julie-bort/","DoorDash 正在建立自己的无人机配送业务：https://techcrunch.com/2026/07/29/doordash-is-building-its-own-drone-delivery-business/ Kirsten Korosec：https://techcrunch.com/author/kirsten-korosec/"],"translationStatus":"translated","bodyOrigin":"source-page","editorial":{"summary":"SaferAI 经由 Z.ai 公共 API 评估称，GLM-5.2 在网络与生物能力上仅落后 GPT-5.5 和 Claude Opus 4.7 数月；面对测试中的攻击性网络及双重用途生物任务时，该模型未拒绝任何一项。","background":"GLM-5.2 是中国 Z.ai 的开源权重模型。材料指出，托管 API 可部署安全措施，但权重下载并在本地运行后，使用者可修改防护、系统提示或继续微调，原提供方的相关控制难以强制执行。","viewpoint":"Aioga 判断，材料揭示的重点不是开源权重模型能否接近前沿能力，而是能力提升与缓解措施之间可能不同步。拒绝表现只是特定评测结果，不能单独等同于模型整体安全水平。","implications":"值得关注的是，Claude Opus 4.7 因持续拒绝而无法完成 CyberGym，GLM-5.2 则未拒绝相关任务，显示能力评测与安全限制可能互相影响。与此同时，材料也指出前沿模型的防护仍可能被越狱绕过。","nextStep":"Aioga 建议后续关注 SaferAI 对评测任务、拒绝判定和能力差距的具体说明，并区分托管 API 与本地权重部署的风险条件；比较模型时，也应同时审视能力结果、防护机制及防护被绕过的可能性。","evidenceRefs":["title","summary","articleBody","source"],"status":"published","aiGenerated":true,"autoApproved":true,"generatedBy":"aioga-editorial:gpt-5.6-sol","reviewedBy":"aioga-editorial-review:gpt-5.6-sol","generatedAt":"2026-08-04T21:45:48.907Z","sourceHash":"3ab94ed72519a64c","review":{"approved":true,"groundedness":96,"clarity":92,"duplicationRisk":12,"blockingIssues":[],"notes":["“拒绝表现只是特定评测结果，不能单独等同于模型整体安全水平”属于明确标注的观点/方法论判断，与材料语境一致。","可选优化：在“GLM-5.2 则未拒绝相关任务”处保留“攻击性网络及双重用途生物任务”的限定，可使其与来源中的评测范围完全对应。"]},"validation":{"passed":true,"mode":"ai-auto","revisions":0,"checks":["schema","length","source-attribution","low-source-overlap","no-html","independent-ai-review"]}},"tags":["技巧观点","TechCrunch：AI（RSS）"],"translations":{"zh-CN":{"title":"开源权重模型逼近前沿能力，安全差距仍在扩大","summary":"SaferAI 报告显示，中国 Z.ai 的开源权重模型 GLM-5.2 在网络与生物能力上仅落后 OpenAI GPT-5.5 和 Anthropic Claude Opus 4.7 数月，但面对攻击性网络或双重用途生物任务时拒绝率为零，而 Claude Opus 4.7 因拒绝过于一致导致 CyberGym 评测无法完成。","category":"技巧观点","source":"TechCrunch","aggregationSource":"TechCrunch：AI（RSS）","pageTitle":"开源权重模型逼近前沿能力，安全差距仍在扩大 - Aioga AI资讯","description":"SaferAI 报告显示，中国 Z.ai 的开源权重模型 GLM-5.2 在网络与生物能力上仅落后 OpenAI GPT-5.5 和 Anthropic Claude Opus 4.7 数月，但面对攻击性网络或双重用途生物任务时拒绝率为零，而 Claude Opus 4.7 因拒绝过于一致导致 CyberGym 评测无法完成。","url":"https://www.aioga.com/news/cmsf48xvn1fozro2eb9lzel3p/","articleBody":["随着政策制定者就如何管理日益强大的人工智能系统（如 OpenAI 的 GPT-5.6 Sol 和 Anthropic 的 Mythos）展开讨论，一款中国的开放权重模型正在缩小与行业领先者的差距。","根据 AI 安全非营利组织 SaferAI 的一份新报告（https://www.safer-ai.org/research/glm-5-2-evaluation-report），中国 Z.ai 的开放权重 AI 模型 GLM-5.2 在网络和生物能力方面仅落后于 OpenAI 的 GPT-5.5 和 Anthropic 的 Claude Opus 4.7 几个月。但前沿能力与安全实践之间的差距正在扩大。","根据 SaferAI 的评估（该非营利组织通过 Z.ai 的公共 API 进行评估），GLM-5.2 没有拒绝任何给定的攻击性网络或双用途生物任务。相比之下，Claude Opus 4.7 “拒绝的频率如此之高，以至于 SaferAI 完全无法在其上完成 CyberGym。”（CyberGym 是一个评估网络安全能力的基准。OpenAI 在上个月 Hugging Face 数据泄露事件之前的评估中使用了该基准：https://techcrunch.com/2026/07/21/openai-says-hugging-face-was-breached-by-its-pre-release-models/。）","这强烈提醒了多年来一些批评者的警告：开放权重 AI 模型可能会将高能力 AI 授予潜在攻击者，而一旦下载了权重，就无法监督他们如何使用该技术。随着开放权重模型迅速接近世界领先 AI 系统的能力，辩论正从它们是否能竞争转向社会在其发布后如何管理风险。","“能力的前沿并不等同于风险的前沿，因此我们确实必须考虑缓解措施的状态，以正确评估风险，”SaferAI 执行董事 Henry Papadatos 告诉 TechCrunch。","虽然 Z.ai 可以对其托管的 API 采取安全措施，但一旦有人在自己的硬件上运行权重，这些保护措施就无法执行，因为他们可以删除或修改任何安全防护，对模型进行微调，或更改系统提示。","像 OpenAI 和 Anthropic 这样的前沿开发者往往依赖分类器、拒绝训练和 API 级别控制等安全措施来限制危险的网络和生物帮助。","这些措施远非万无一失：越狱程序经常绕过已部署模型的保护措施。非营利组织 Far.ai（致力于 AI 安全）在前沿模型中发现了数百个通用越狱程序：https://leaderboard.far.ai/ —— 被定义为可重复使用的关键键，在大多数有害请求中都能成功 —— 例如 xAI 的 Grok 4.5 和 Google DeepMind 的 Gemini 3.1 Pro。根据报告，当攻击者结合多种操纵技术——包括角色扮演、冒充权威、伪造对话历史和后续提示——来放大模型防御的弱点时，越狱程序就会成功。","但针对封闭模型的保护措施在开放权重模型上根本不起作用，后者被设计为可在任何基础设施上运行，无论是否有任何保护措施。","“目标显然应该是让好的能力——安全的能力——对任何人都可访问，然后我们尝试去除不良能力，即使是以开源的方式，”Papadatos 说。","Papadatos 提到的一种可能有帮助的技术称为“预训练数据过滤”，即 AI 公司从其训练数据中移除具有攻击性的网络安全信息，然后在经过整理的数据集上训练模型。","一些研究：https://alignment.anthropic.com/2025/pretraining-data-filtering/?utm 表明，这可以在不影响整体模型性能的情况下减少有害的生物学知识：https://arxiv.org/abs/2508.06601#:~:text=In%20this%20paper%2C%20we%20investigate,as%20a%20more%20tamper%2Dresistant%20safeguard. 但是，对于网络安全而言，数据过滤的实用性要差得多。","训练一个在编码上表现出色但同时又不是优秀黑客的一般模型非常困难。由于编码已成为 AI 最大的盈利来源，开发者面临着保持这些能力持续改进的压力，即使他们在寻找限制滥用的方法时也是如此。","因此，前沿开发者越来越多地依赖其他缓解措施。一个方法是有选择地限制模型提供的网络安全援助类型。例如，Anthropic 的 Opus 5 可以搜索未编译源码中的漏洞，但不能搜索已编译的软件，根据该模型的系统卡：https://www-cdn.anthropic.com/b514064af1408018e64b1ad24e7d5e75850b4ffd/Claude%20Opus%205%20System%20Card.pdf。其原因是，这使得使用 Opus 5 进行进攻性操作更加困难。","其他措施包括严格的部署前安全评估、发布风险评估，以及如果系统被认为过于危险而不公开模型权重。","在 GLM-5.2 的案例中，SaferAI 表示 Z.ai 没有发布安全框架、部署前测试承诺或模型的风险评估。TechCrunch 已向 Z.ai 询问其在发布前是否进行了内部或第三方前沿安全评估，但未收到回应。","中国领导人越来越承认先进 AI 的风险。在上个月的世界人工智能大会上，中国国家主席习近平强调：https://www.nytimes.com/2026/07/17/business/xi-jinping-china-ai.html?eafs_enabled=false 开放权重模型的重要性，同时也强调确保 AI 仍在严格人类控制下的必要性。","斯坦福大学网络政策中心研究中国 AI 政策的 Graham Webster 告诉 TechCrunch，中国有健全的 AI 监管法规，但这些规则历来主要关注政治敏感内容、虚假信息和社会稳定，而非灾难性 AI 风险，如进攻性网络能力和生物滥用。","“总体来说，美国的 AI 思想家更关注这一生存性灾难威胁，而中国社区不太关注，”Webster 说，并补充道，许多中国政策研究人员认为，如果真的出现新的前沿风险，美国公司可能会首先遇到它。","“中国体系有信心能够控制这些技术在中国境内的使用，”Webster 继续说道。“在中国上网是需要实名的，企业可以被追责，用户也可以被追责。”","韦伯斯特思考道，模型提供者用于拒绝涉及某些政治话题的同样机制，也可以潜在地调整，以确保模型拒绝执行攻击性网络攻击或不会产生不良的生物工程结果。他补充道，由于中国公司往往在幕后一同与监管机构协调，因此在发布前很难知道他们进行了哪些内部测试。","支持开放权重AI的倡导者认为，发布权重对于网络安全很重要，因为它使公司能够防御攻击——Hugging Face 曾依靠 GLM-5.2 来防御 OpenAI 的泄露——且因为如果公司知道未来可能出现的威胁，就能更好地做准备。","“Hugging Face CEO Clem Delangue 本周在社交媒体帖子中表示：‘同样帮助阻止 AI 驱动的网络攻击的系统，现在可以帮助防御每天数百万次的网络攻击，同时帮助我们在攻击者利用漏洞之前识别并修复这些漏洞。’ https://x.com/ClementDelangue/status/2083908468285620415","Papadatos 表示，这一好处往往被高估，并不意味着‘我们应该开源危险能力。’","“我认为的主要观点是，我们不应该仅仅接受危险能力可以被任何人在任何地方轻易访问，”他表示，强调他认为行业应致力于只让‘良性能力’易于访问。默认情况下，攻击者采用新工具的速度比防御者快。例如，一个勒索软件团体可以在一周内改变其方法，而医院无法做到。”","当你通过我们文章中的链接购买时，我们可能会获得一小部分佣金：https://techcrunch.com/techcrunch-affiliate-monetization-standards/。这不会影响我们的编辑独立性。","Rebecca Bellan 是 TechCrunch 的高级记者，报道塑造人工智能的商业、政策和新兴趋势。她的作品也曾刊登在 Forbes、Bloomberg、The Atlantic、The Daily Beast 及其他出版物上。","你可以通过发送邮件至 rebecca.bellan@techcrunch.com 联系或核实 Rebecca 的联系信息：mailto:rebecca.bellan@techcrunch.com，或通过 Signal 上的加密消息 rebeccabellan.491 联系。","更快扩大规模。发展你的投资组合。获得实用经验。无论你的目标是什么，Disrupt 都能赋能你。今天最多可节省 330 美元！","网红因参加 OpenAI 的首次奢华旅行而引发反弹：https://techcrunch.com/2026/08/03/influencers-draw-backlash-for-attending-openais-first-luxury-trip/ Dominic-Madori Davis：https://techcrunch.com/author/dominic-madori-davis/","红杉的 Shaun Maguire 领导对核能初创公司 Valar Atomics 的 10 亿美元融资轮：https://techcrunch.com/2026/08/03/sequoias-shaun-maguire-leads-1b-round-for-nuclear-startup-valar-atomics/ Julie Bort：https://techcrunch.com/author/julie-bort/","YouTuber Hank Green 表示他的人工智能使用“不健康”：https://techcrunch.com/2026/08/01/youtuber-hank-green-says-his-ai-usage-is-not-healthy/ Anthony Ha：https://techcrunch.com/author/anthony-ha/","WhatsApp 正在测试一个用于大型企业消息的新文件夹：https://techcrunch.com/2026/07/31/whatsapp-is-testing-a-new-folder-for-messages-from-large-businesses/ Ivan Mehta：https://techcrunch.com/author/ivan-mehta/","Spotify向其应用添加了跑步模式：https://techcrunch.com/2026/07/30/spotify-adds-a-running-mode-to-its-app/ 伊万·梅塔：https://techcrunch.com/author/ivan-mehta/","Claude Opus 5 在被分配管理自动售货机任务时变得相当无情：https://techcrunch.com/2026/07/29/claude-opus-5-became-downright-ruthless-when-tasked-with-running-a-vending-machine/ Julie Bort：https://techcrunch.com/author/julie-bort/","DoorDash 正在建立自己的无人机配送业务：https://techcrunch.com/2026/07/29/doordash-is-building-its-own-drone-delivery-business/ Kirsten Korosec：https://techcrunch.com/author/kirsten-korosec/"]},"en":{"title":"Open-source weighting models approach cutting-edge capabilities, but the security gap continues to widen","summary":"According to the SaferAI report, China's Z.ai open-source weighting model GLM-5.2 lags behind OpenAI's GPT-5.5 and Anthropic Claude Opus 4.7 by only a few months in network and biological capabilities, but has zero rejection rates for offensive cyber or dual-use biological tasks, whereas Claude Opus 4.7 failed to complete CyberGym evaluations due to too consistent rejections.","category":"Insights","source":"TechCrunch：AI（RSS）","aggregationSource":"TechCrunch：AI（RSS）","pageTitle":"Open-source weighting models approach cutting-edge capabilities, but the security gap continues to widen - Aioga AI News","description":"According to the SaferAI report, China's Z.ai open-source weighting model GLM-5.2 lags behind OpenAI's GPT-5.5 and Anthropic Claude Opus 4.7 by only a few months in network and bio...","url":"https://www.aioga.com/en/news/cmsf48xvn1fozro2eb9lzel3p/","contentTranslated":true,"sourceHash":"58906b3d325c266c","translatedAt":"2026-08-04T21:21:23.004Z"},"ja":{"title":"オープンソースの重み付けモデルは最先端の機能に近づいていますが、セキュリティのギャップは拡大し続けています","summary":"SaferAIの報告によると、中国の Z.ai オープンソース加重モデルGLM-5.2は、ネットワークおよび生物学的能力においてOpenAIのGPT-5.5やAnthropic Claude Opus 4.7に数か月遅れていますが、攻撃的なサイバーや二重用途の生物学的タスクでは拒否率がゼロです。一方、Claude Opus 4.7はCyberGymの評価を通過できず、拒否が一貫していました。","category":"ヒントと視点","source":"TechCrunch：AI（RSS）","aggregationSource":"TechCrunch：AI（RSS）","pageTitle":"オープンソースの重み付けモデルは最先端の機能に近づいていますが、セキュリティのギャップは拡大し続けています - Aioga AIニュース","description":"SaferAIの報告によると、中国の Z.ai オープンソース加重モデルGLM-5.2は、ネットワークおよび生物学的能力においてOpenAIのGPT-5.5やAnthropic Claude Opus 4.7に数か月遅れていますが、攻撃的なサイバーや二重用途の生物学的タスクでは拒否率がゼロです。一方、Claude Opus 4.7はCyberGymの評価を通...","url":"https://www.aioga.com/ja/news/cmsf48xvn1fozro2eb9lzel3p/","contentTranslated":true,"sourceHash":"58906b3d325c266c","translatedAt":"2026-08-04T21:21:23.628Z"},"ko":{"title":"오픈 소스 가중치 모델은 최첨단 기능에 근접하지만, 보안 격차는 계속 벌어지고 있습니다","summary":"SaferAI 보고서에 따르면, 중국의 Z.ai 오픈소스 가중치 모델 GLM-5.2는 네트워크 및 생물학적 역량에서 OpenAI의 GPT-5.5와 Anthropic Claude Opus 4.7보다 몇 달 차이로 뒤처져 있지만, 공격적 사이버 또는 이중 용도 생물학적 과제에서는 거절률이 0인 반면, Claude Opus 4.7은 너무 일관된 거절로 인해 CyberGym 평가를 완료하지 못했습니다.","category":"인사이트","source":"TechCrunch：AI（RSS）","aggregationSource":"TechCrunch：AI（RSS）","pageTitle":"오픈 소스 가중치 모델은 최첨단 기능에 근접하지만, 보안 격차는 계속 벌어지고 있습니다 - Aioga AI 뉴스","description":"SaferAI 보고서에 따르면, 중국의 Z.ai 오픈소스 가중치 모델 GLM-5.2는 네트워크 및 생물학적 역량에서 OpenAI의 GPT-5.5와 Anthropic Claude Opus 4.7보다 몇 달 차이로 뒤처져 있지만, 공격적 사이버 또는 이중 용도 생물학적 과제에서는 거절률이 0인 반면, Claude Opus 4...","url":"https://www.aioga.com/ko/news/cmsf48xvn1fozro2eb9lzel3p/","contentTranslated":true,"sourceHash":"58906b3d325c266c","translatedAt":"2026-08-04T21:21:32.430Z"},"es":{"title":"Los modelos de ponderación de código abierto se acercan a capacidades de vanguardia, pero la brecha de seguridad sigue ampliándose","summary":"Según el informe de SaferAI, el modelo chino de ponderación Z.ai código abierto GLM-5.2 va solo unos meses por detrás de GPT-5.5 de OpenAI y Claude Opus 4.7 de OpenAI en capacidades de red y biológicas, pero tiene tasas de rechazo cero para tareas cibernéticas ofensivas o biológicas de doble uso, mientras que Claude Opus 4.7 no completó las evaluaciones de CyberGym debido a rechazos demasiado consistentes.","category":"Ideas","source":"TechCrunch：AI（RSS）","aggregationSource":"TechCrunch：AI（RSS）","pageTitle":"Los modelos de ponderación de código abierto se acercan a capacidades de vanguardia, pero la brecha de seguridad sigue ampliándose - Aioga Noticias de IA","description":"Según el informe de SaferAI, el modelo chino de ponderación Z.ai código abierto GLM-5.2 va solo unos meses por detrás de GPT-5.5 de OpenAI y Claude Opus 4.7 de OpenAI en capacidade...","url":"https://www.aioga.com/es/news/cmsf48xvn1fozro2eb9lzel3p/","contentTranslated":true,"sourceHash":"58906b3d325c266c","translatedAt":"2026-08-04T21:21:32.495Z"},"fr":{"title":"Les modèles de pondération open source approchent les capacités de pointe, mais l’écart de sécurité continue de s’élargir","summary":"Selon le rapport SaferAI, le modèle chinois de pondération open source Z.ai GLM-5.2 accuse seulement quelques mois de retard sur GPT-5.5 d’OpenAI et Claude Opus 4.7 d’Anthropic en termes de capacités réseau et biologiques, mais affiche un taux de rejet zéro pour les tâches cyber offensives ou biologiques à double usage, tandis que Claude Opus 4.7 n’a pas réussi à valider les évaluations CyberGym en raison de refus trop constants.","category":"Analyses","source":"TechCrunch：AI（RSS）","aggregationSource":"TechCrunch：AI（RSS）","pageTitle":"Les modèles de pondération open source approchent les capacités de pointe, mais l’écart de sécurité continue de s’élargir - Aioga Actualités IA","description":"Selon le rapport SaferAI, le modèle chinois de pondération open source Z.ai GLM-5.2 accuse seulement quelques mois de retard sur GPT-5.5 d’OpenAI et Claude Opus 4.7 d’Anthropic en...","url":"https://www.aioga.com/fr/news/cmsf48xvn1fozro2eb9lzel3p/","contentTranslated":true,"sourceHash":"58906b3d325c266c","translatedAt":"2026-08-04T21:21:41.327Z"},"de":{"title":"Open-Source-Gewichtungsmodelle nähern sich modernsten Fähigkeiten, doch die Sicherheitslücke wächst weiter","summary":"Laut dem SaferAI-Bericht hinkt Chinas Z.ai Open-Source-Gewichtungsmodell GLM-5.2 in Netzwerk- und biologischen Fähigkeiten nur wenige Monate hinter OpenAIs GPT-5.5 und dem anthropologischen Claude Opus 4.7 zurück, weist jedoch keine Ablehnungsraten für offensive Cyber- oder Dual-Use biologische Aufgaben auf, während Claude Opus 4.7 die CyberGym-Bewertungen aufgrund zu konstanter Ablehnungen nicht abschließen konnte.","category":"技巧观点","source":"TechCrunch：AI（RSS）","aggregationSource":"TechCrunch：AI（RSS）","pageTitle":"Open-Source-Gewichtungsmodelle nähern sich modernsten Fähigkeiten, doch die Sicherheitslücke wächst weiter - Aioga KI-News","description":"Laut dem SaferAI-Bericht hinkt Chinas Z.ai Open-Source-Gewichtungsmodell GLM-5.2 in Netzwerk- und biologischen Fähigkeiten nur wenige Monate hinter OpenAIs GPT-5.5 und dem anthropo...","url":"https://www.aioga.com/de/news/cmsf48xvn1fozro2eb9lzel3p/","contentTranslated":true,"sourceHash":"58906b3d325c266c","translatedAt":"2026-08-04T21:21:41.250Z"},"pt-BR":{"title":"Modelos de ponderação de código aberto se aproximam das capacidades de ponta, mas a lacuna de segurança continua a se ampliar","summary":"De acordo com o relatório da SaferAI, o modelo de ponderação open source Z.ai da China, GLM-5.2, fica apenas alguns meses atrás do GPT-5.5 da OpenAI e do Anthropic Claude Opus 4.7 em capacidades de rede e biológicas, mas apresenta taxas de rejeição zero para tarefas cibernéticas ofensivas ou biológicas de uso duplo, enquanto o Claude Opus 4.7 não completou as avaliações do CyberGym devido a rejeições muito consistentes.","category":"技巧观点","source":"TechCrunch：AI（RSS）","aggregationSource":"TechCrunch：AI（RSS）","pageTitle":"Modelos de ponderação de código aberto se aproximam das capacidades de ponta, mas a lacuna de segurança continua a se ampliar - Aioga Notícias de IA","description":"De acordo com o relatório da SaferAI, o modelo de ponderação open source Z.ai da China, GLM-5.2, fica apenas alguns meses atrás do GPT-5.5 da OpenAI e do Anthropic Claude Opus 4.7...","url":"https://www.aioga.com/pt-BR/news/cmsf48xvn1fozro2eb9lzel3p/","contentTranslated":true,"sourceHash":"58906b3d325c266c","translatedAt":"2026-08-04T21:21:50.137Z"},"ru":{"title":"Открытые модели взвешивания приближаются к передовым возможностям, но разрыв в безопасности продолжает расти","summary":"Согласно отчету SaferAI, открытая модель взвешивания GLM-5.2 Z.ai Китая отстаёт от GPT-5.5 от OpenAI и Anthropic Claude Opus 4.7 всего на несколько месяцев по сетевым и биологическим возможностям, но не имеет показателей отклонения для наступательных кибер- или биологических задач двойного назначения, тогда как Claude Opus 4.7 не прошёл оценки CyberGym из-за слишком регулярных отказов.","category":"技巧观点","source":"TechCrunch：AI（RSS）","aggregationSource":"TechCrunch：AI（RSS）","pageTitle":"Открытые модели взвешивания приближаются к передовым возможностям, но разрыв в безопасности продолжает расти - Aioga Новости ИИ","description":"Согласно отчету SaferAI, открытая модель взвешивания GLM-5.2 Z.ai Китая отстаёт от GPT-5.5 от OpenAI и Anthropic Claude Opus 4.7 всего на несколько месяцев по сетевым и биологическ...","url":"https://www.aioga.com/ru/news/cmsf48xvn1fozro2eb9lzel3p/","contentTranslated":true,"sourceHash":"58906b3d325c266c","translatedAt":"2026-08-04T21:21:50.131Z"},"ar":{"title":"نماذج الوزن مفتوحة المصدر تقترب من القدرات المتقدمة، لكن فجوة الأمان تستمر في اتساعها","summary":"وفقا لتقرير SaferAI، فإن نموذج الوزن مفتوح المصدر في الصين Z.ai GLM-5.2 يتخلف بفارق بضعة أشهر فقط من حيث القدرات الشبكية والبيولوجية من OpenAI GPT-5.5 وAnthropic Claude Opus 4.7، لكنه لا يحقق أي معدلات رفض للمهام السيبرانية الهجومية أو المهام البيولوجية ذات الاستخدام المزدوج، بينما فشل Claude Opus 4.7 في إكمال تقييمات CyberGym بسبب الرفض المستمر جدا.","category":"技巧观点","source":"TechCrunch：AI（RSS）","aggregationSource":"TechCrunch：AI（RSS）","pageTitle":"نماذج الوزن مفتوحة المصدر تقترب من القدرات المتقدمة، لكن فجوة الأمان تستمر في اتساعها - Aioga أخبار الذكاء الاصطناعي","description":"وفقا لتقرير SaferAI، فإن نموذج الوزن مفتوح المصدر في الصين Z.ai GLM-5.2 يتخلف بفارق بضعة أشهر فقط من حيث القدرات الشبكية والبيولوجية من OpenAI GPT-5.5 وAnthropic Claude Opus 4.7، ل...","url":"https://www.aioga.com/ar/news/cmsf48xvn1fozro2eb9lzel3p/","contentTranslated":true,"sourceHash":"58906b3d325c266c","translatedAt":"2026-08-04T21:21:58.869Z"},"hi":{"title":"ओपन-सोर्स वेटिंग मॉडल अत्याधुनिक क्षमताओं तक पहुंचते हैं, लेकिन सुरक्षा अंतर लगातार बढ़ता जा रहा है","summary":"SaferAI रिपोर्ट के अनुसार, चीन का Z.ai ओपन-सोर्स वेटिंग मॉडल GLM-5.2 नेटवर्क और जैविक क्षमताओं में OpenAI के GPT-5.5 और एंथ्रोपिक क्लाउड ओपस 4.7 से केवल कुछ महीनों से पीछे है, लेकिन आक्रामक साइबर या दोहरे उपयोग वाले जैविक कार्यों के लिए शून्य अस्वीकृति दर है, जबकि क्लाउड ओपस 4.7 बहुत लगातार अस्वीकृति के कारण साइबरजिम मूल्यांकन को पूरा करने में विफल रहा।","category":"技巧观点","source":"TechCrunch：AI（RSS）","aggregationSource":"TechCrunch：AI（RSS）","pageTitle":"ओपन-सोर्स वेटिंग मॉडल अत्याधुनिक क्षमताओं तक पहुंचते हैं, लेकिन सुरक्षा अंतर लगातार बढ़ता जा रहा है - Aioga AI समाचार","description":"SaferAI रिपोर्ट के अनुसार, चीन का Z.ai ओपन-सोर्स वेटिंग मॉडल GLM-5.2 नेटवर्क और जैविक क्षमताओं में OpenAI के GPT-5.5 और एंथ्रोपिक क्लाउड ओपस 4.7 से केवल कुछ महीनों से पीछे है, लेकि...","url":"https://www.aioga.com/hi/news/cmsf48xvn1fozro2eb9lzel3p/","contentTranslated":true,"sourceHash":"58906b3d325c266c","translatedAt":"2026-08-04T21:21:58.874Z"},"it":{"title":"I modelli di ponderazione open-source si avvicinano alle capacità all'avanguardia, ma il divario di sicurezza continua ad allargarsi","summary":"Secondo il rapporto SaferAI, il modello cinese di ponderazione open-source Z.ai GLM-5.2 è indietro rispetto a GPT-5.5 di OpenAI e Anthropic Claude Opus 4.7 di pochi mesi in termini di capacità di rete e biologiche, ma ha tassi di rifiuto zero per compiti offensivi cyber o biologici a doppio uso, mentre Claude Opus 4.7 non ha completato le valutazioni CyberGym a causa di rifiuti troppo costanti.","category":"技巧观点","source":"TechCrunch：AI（RSS）","aggregationSource":"TechCrunch：AI（RSS）","pageTitle":"I modelli di ponderazione open-source si avvicinano alle capacità all'avanguardia, ma il divario di sicurezza continua ad allargarsi - Aioga Notizie IA","description":"Secondo il rapporto SaferAI, il modello cinese di ponderazione open-source Z.ai GLM-5.2 è indietro rispetto a GPT-5.5 di OpenAI e Anthropic Claude Opus 4.7 di pochi mesi in termini...","url":"https://www.aioga.com/it/news/cmsf48xvn1fozro2eb9lzel3p/","contentTranslated":true,"sourceHash":"58906b3d325c266c","translatedAt":"2026-08-04T21:22:08.217Z"},"nl":{"title":"Open-source wegingsmodellen benaderen de nieuwste mogelijkheden, maar de beveiligingskloof blijft groter worden","summary":"Volgens het SaferAI-rapport loopt China's Z.ai open-source weegmodel GLM-5.2 slechts enkele maanden achter op OpenAI's GPT-5.5 en het antropische Claude Opus 4.7 qua netwerk- en biologische capaciteiten, maar heeft het nul afwijzingspercentages voor offensieve cyber- of dual-use biologische taken, terwijl Claude Opus 4.7 de CyberGym-evaluaties niet heeft afgerond vanwege te consistente afwijzingen.","category":"技巧观点","source":"TechCrunch：AI（RSS）","aggregationSource":"TechCrunch：AI（RSS）","pageTitle":"Open-source wegingsmodellen benaderen de nieuwste mogelijkheden, maar de beveiligingskloof blijft groter worden - Aioga AI-nieuws","description":"Volgens het SaferAI-rapport loopt China's Z.ai open-source weegmodel GLM-5.2 slechts enkele maanden achter op OpenAI's GPT-5.5 en het antropische Claude Opus 4.7 qua netwerk- en bi...","url":"https://www.aioga.com/nl/news/cmsf48xvn1fozro2eb9lzel3p/","contentTranslated":true,"sourceHash":"58906b3d325c266c","translatedAt":"2026-08-04T21:22:07.272Z"},"tr":{"title":"Açık kaynak ağırlık modelleri en son yeteneklere yaklaşır, ancak güvenlik açığı büyümeye devam eder","summary":"SaferAI raporuna göre, Çin'in Z.ai açık kaynak ağırlıklandırma modeli GLM-5.2, ağ ve biyolojik yetenekler açısından OpenAI'nin GPT-5.5 ve Anthropic Claude Opus 4.7'nin sadece birkaç ay gerisinde kalıyor, ancak saldırgan siber veya çift kullanımlı biyolojik görevler için sıfır reddedilme oranına sahipken, Claude Opus 4.7 ise CyberGym değerlendirmelerini çok tutarlı reddetmeler nedeniyle tamamlayamadı.","category":"技巧观点","source":"TechCrunch：AI（RSS）","aggregationSource":"TechCrunch：AI（RSS）","pageTitle":"Açık kaynak ağırlık modelleri en son yeteneklere yaklaşır, ancak güvenlik açığı büyümeye devam eder - Aioga AI Haberleri","description":"SaferAI raporuna göre, Çin'in Z.ai açık kaynak ağırlıklandırma modeli GLM-5.2, ağ ve biyolojik yetenekler açısından OpenAI'nin GPT-5.5 ve Anthropic Claude Opus 4.7'nin sadece birka...","url":"https://www.aioga.com/tr/news/cmsf48xvn1fozro2eb9lzel3p/","contentTranslated":true,"sourceHash":"58906b3d325c266c","translatedAt":"2026-08-04T21:22:16.598Z"},"vi":{"title":"Các mô hình trọng số mã nguồn mở tiến gần đến các khả năng tiên tiến, nhưng khoảng cách bảo mật vẫn ngày càng rộng","summary":"Theo báo cáo của SaferAI, mô hình trọng số mã nguồn mở Z.ai của Trung Quốc GLM-5.2 chỉ kém vài tháng so với GPT-5.5 và Anthropic Claude Opus 4.7 về khả năng mạng và sinh học, nhưng lại không có tỷ lệ từ chối đối với các nhiệm vụ mạng tấn công hoặc sinh học sử dụng kép, trong khi Claude Opus 4.7 không hoàn thành đánh giá CyberGym do bị từ chối quá liên tục.","category":"技巧观点","source":"TechCrunch：AI（RSS）","aggregationSource":"TechCrunch：AI（RSS）","pageTitle":"Các mô hình trọng số mã nguồn mở tiến gần đến các khả năng tiên tiến, nhưng khoảng cách bảo mật vẫn ngày càng rộng - Tin tức AI Aioga","description":"Theo báo cáo của SaferAI, mô hình trọng số mã nguồn mở Z.ai của Trung Quốc GLM-5.2 chỉ kém vài tháng so với GPT-5.5 và Anthropic Claude Opus 4.7 về khả năng mạng và sinh học, nhưng...","url":"https://www.aioga.com/vi/news/cmsf48xvn1fozro2eb9lzel3p/","contentTranslated":true,"sourceHash":"58906b3d325c266c","translatedAt":"2026-08-04T21:22:17.158Z"},"id":{"title":"Model pembobotan open-source mendekati kemampuan mutakhir, tetapi kesenjangan keamanan terus melebar","summary":"Menurut laporan SaferAI, model pembobotan sumber terbuka Z.ai Tiongkok, GLM-5.2, tertinggal beberapa bulan dari GPT-5.5 dan Anthropic Claude Opus 4.7 dari OpenAI dalam hal kemampuan jaringan dan biologi, namun memiliki tingkat penolakan nol untuk tugas siber ofensif atau biologis dual-use, sedangkan Claude Opus 4.7 gagal menyelesaikan evaluasi CyberGym karena penolakan yang terlalu konsisten.","category":"技巧观点","source":"TechCrunch：AI（RSS）","aggregationSource":"TechCrunch：AI（RSS）","pageTitle":"Model pembobotan open-source mendekati kemampuan mutakhir, tetapi kesenjangan keamanan terus melebar - Berita AI Aioga","description":"Menurut laporan SaferAI, model pembobotan sumber terbuka Z.ai Tiongkok, GLM-5.2, tertinggal beberapa bulan dari GPT-5.5 dan Anthropic Claude Opus 4.7 dari OpenAI dalam hal kemampua...","url":"https://www.aioga.com/id/news/cmsf48xvn1fozro2eb9lzel3p/","contentTranslated":true,"sourceHash":"58906b3d325c266c","translatedAt":"2026-08-04T21:22:25.891Z"},"th":{"title":"โมเดลการถ่วงน้ําหนักแบบโอเพนซอร์สเข้าใกล้ความสามารถล้ําสมัย แต่ช่องว่างด้านความปลอดภัยยังคงขยายกว้างขึ้น","summary":"ตามรายงานของ SaferAI โมเดลการถ่วงน้ําหนักแบบโอเพ่นซอร์ส Z.ai ของจีน GLM-5.2 ตามหลัง OpenAI GPT-5.5 และ Anthropic Claude Opus 4.7 เพียงไม่กี่เดือนในด้านความสามารถด้านเครือข่ายและชีวภาพ แต่ไม่มีอัตราการปฏิเสธสําหรับงานทางไซเบอร์เชิงรุกหรืองานชีวภาพแบบสองทาง ขณะที่ Claude Opus 4.7 ล้มเหลวในการประเมิน CyberGym เนื่องจากการปฏิเสธที่สม่ําเสมอเกินไป","category":"技巧观点","source":"TechCrunch：AI（RSS）","aggregationSource":"TechCrunch：AI（RSS）","pageTitle":"โมเดลการถ่วงน้ําหนักแบบโอเพนซอร์สเข้าใกล้ความสามารถล้ําสมัย แต่ช่องว่างด้านความปลอดภัยยังคงขยายกว้างขึ้น - ข่าว AI Aioga","description":"ตามรายงานของ SaferAI โมเดลการถ่วงน้ําหนักแบบโอเพ่นซอร์ส Z.ai ของจีน GLM-5.2 ตามหลัง OpenAI GPT-5.5 และ Anthropic Claude Opus 4.7 เพียงไม่กี่เดือนในด้านความสามารถด้านเครือข่ายและชีว...","url":"https://www.aioga.com/th/news/cmsf48xvn1fozro2eb9lzel3p/","contentTranslated":true,"sourceHash":"58906b3d325c266c","translatedAt":"2026-08-04T21:22:25.848Z"},"pl":{"title":"Otwartoźródłowe modele ważenia zbliżają się do najnowocześniejszych możliwości, ale luka w zabezpieczeniach nadal się powiększa","summary":"Według raportu SaferAI, chiński model ważenia Z.ai open-source GLM-5.2 pozostaje w tyle za GPT-5.5 firmy OpenAI i Anthropic Claude Opus 4.7 tylko o kilka miesięcy pod względem możliwości sieciowych i biologicznych, ale nie ma żadnych wskaźników odrzuceń dla ofensywnych zadań cybernetycznych lub biologicznych podwójnego zastosowania, podczas gdy Claude Opus 4.7 nie ukończył ewaluacji CyberGym z powodu zbyt częstych odrzuceń.","category":"技巧观点","source":"TechCrunch：AI（RSS）","aggregationSource":"TechCrunch：AI（RSS）","pageTitle":"Otwartoźródłowe modele ważenia zbliżają się do najnowocześniejszych możliwości, ale luka w zabezpieczeniach nadal się powiększa - Aioga Wiadomości AI","description":"Według raportu SaferAI, chiński model ważenia Z.ai open-source GLM-5.2 pozostaje w tyle za GPT-5.5 firmy OpenAI i Anthropic Claude Opus 4.7 tylko o kilka miesięcy pod względem możl...","url":"https://www.aioga.com/pl/news/cmsf48xvn1fozro2eb9lzel3p/","contentTranslated":true,"sourceHash":"58906b3d325c266c","translatedAt":"2026-08-04T21:22:34.749Z"}}}}