-
GLM 5.2, Semgrep IDOR 벤치마크에서 Claude 앞서
GLM 5.2, Semgrep IDOR 벤치마크에서 Claude 앞서 | GeekNews
<ul> <li>Semgrep의 <strong>IDOR 취약점 탐지</strong> 벤치마크에서 Zhipu AI의 open-weight 모델 <strong>GLM 5.2</strong>가 단순 프롬프트 조건만으로 Claude Code보다 높은 F1을 기록함</li> <li>실험은 데이터셋·평가 방식·시스템 프롬프트를 고정…
-
우리도 Mythos를 가지고 있다: GLM 5.2가 사이버 벤치마크에서 Claude를 이기다
We have Mythos at Home: GLM 5.2 beats Claude in our Cyber Benchmarks
<p>Article URL: <a href="https://semgrep.dev/blog/2026/we-have-mythos-at-home-glm-52-beats-claude-in-our-cyber-benchmarks/">https://semgrep.dev/blog/2026/we-have-mythos-at-home-glm…