Zhizhi released GLM-5.3, based on the same foundation as GLM-5.2, enhancing intelligence through
extreme post-training scaling. Programming ability increased by 50% compared with the previous generation, ranking first among open-source models in publicly available benchmarks like Terminal Bench 3.0, and approaching Claude Fable 5. The model performs on par with Mythos 5 in white-box code review and other security tasks, scoring 84.5% in the CyberGym test. Model weights will be open-sourced in two weeks, and tools such as ZCode and AutoClaw are now online. 🔗 Read original via AIHOT · https://aihot.virxact.com/items/cmssir12d047oroffnpn5pcx1
智谱发布GLM-5.3,基于与GLM-5.2相同的基座,通过极致的后训练Scaling提升智能上界,编程能力较前代提升50%,在Terminal Bench
3.0等公开基准中取得开源第一,并接近Claude Fable 5。
模型在白盒代码审查等安全任务中表现持平Mythos 5,在CyberGym测试中得分84.5%。
模型权重将在两周后开源,即日起上线ZCode、AutoClaw等工具。
🔗 阅读原文 via AIHOT · https://aihot.virxact.com/items/cmssir12d047oroffnpn5pcx1
The readable text on this page was extracted from the public source and organized with attribution, publication time and the original link. Copyright remains with the original author and publisher.