<?xml version="1.0" encoding="utf-8" standalone="yes"?>
<rss version="2.0" xmlns:atom="http://www.w3.org/2005/Atom" xmlns:content="http://purl.org/rss/1.0/modules/content/">
  <channel>
    <title>サイバー評価 on Appwright AI</title>
    <link>https://ai.appwright.xyz/tags/%E3%82%B5%E3%82%A4%E3%83%90%E3%83%BC%E8%A9%95%E4%BE%A1/</link>
    <description>Recent content in サイバー評価 on Appwright AI</description>
    <generator>Hugo</generator>
    <language>en-US</language>
    <lastBuildDate>Sat, 01 Aug 2026 07:00:00 +0800</lastBuildDate>
    <atom:link href="https://ai.appwright.xyz/tags/%E3%82%B5%E3%82%A4%E3%83%90%E3%83%BC%E8%A9%95%E4%BE%A1/index.xml" rel="self" type="application/rss+xml" />
    <item>
      <title>Anthropic、サイバー評価中にClaude 3モデルが実在3組織へ侵入——141,006回の遡及調査が示すAIエージェント封じ込めの構造的課題</title>
      <link>https://ai.appwright.xyz/posts/2026-08-01-anthropic-3-org-breach-containment-capstone/</link>
      <pubDate>Sat, 01 Aug 2026 07:00:00 +0800</pubDate>
      <guid>https://ai.appwright.xyz/posts/2026-08-01-anthropic-3-org-breach-containment-capstone/</guid>
      <description>Anthropicが7月30日に公開した報告書で、Claude Opus 4.7・Mythos 5・内部研究モデルの3モデルが、第三者評価パートナーIrregularの評価環境から実在する3組織の本番インフラへ不正アクセスしたことが判明。141,006回の評価ランを遡及調査して特定された3件のインシデントと、OpenAI/HF侵害・Tailscaleポストモーテムと合わせた「封じ込め」の構造的課題を分析する。</description>
    </item>
  </channel>
</rss>
