「開源模型 GLM-5.3 自主寫攻擊程式的能力逼近 Claude Mythos Preview(12% 對 14%)」
— Anthropic 紅隊報告
屬實
這對你的意思是
開源模型寫攻擊程式的能力已經接近頂尖商用模型,這一點有獨立機構確認。
屬實部分屬實誇大未證實
查到了什麼
除了 Anthropic 自家報告,美國 NIST 旗下的 CAISI 也做了獨立測試。
「屬實」:兩個以上彼此獨立的來源確認。判定方法
依據來源
- anthropic.comhttps://www.anthropic.com/research/glm-5-3-and-the-spread-of-advanced-cyber-capabilities
- tomshardware.comhttps://www.tomshardware.com/tech-industry/artificial-intelligence/anthropic-claims-popular-chinese-ai-model-has-mythos-class-hacking-abilities-frontier-red-teaming-report-details-weak-safeguards-on-open-weight-ai
同一集