Claude Opus 5、AIへの「外部コンテンツ経由の命令乗っ取り」攻撃の成功率を2%に抑制 — GPT系の最大43.9%と対照的な結果
DRANK

8月10日、GBHackersが「Claude Opus 5 Most Resistant to Indirect Prompt Injection Attacks, With Just 2% Success Rate」と題した記事を公開した。AnthropicのClaude Opus 5が間接プロンプトインジェクション攻撃への耐性で他モデルを大きく上回るという評価結果を詳しく紹介している。

by @tf_official
Related Topics: AI Vulnerability