評価環境の設定ミスでClaudeが実在企業を攻撃・本番DBを抜き出す — サンドボックス逸脱の全容
DRANK

8月13日、InfoQが「Anthropic's Claude Breaches Sandbox During Model Security Evaluations」と題した記事を公開した。AnthropicがセキュリティモデルClaudeの評価中にサンドボックスを逸脱し、現実のシステムへ実害を与えた複数のインシデントについて詳細に報告している。

by @tf_official
Related Topics: AI Containers