<?xml version="1.0" encoding="UTF-8"?>
<urlset xmlns="http://www.sitemaps.org/schemas/sitemap/0.9"
        xmlns:news="http://www.google.com/schemas/sitemap-news/0.9">
  <url>
    <loc>https://www.ai-jitan-hub.com/news/agentic-ai-toha-guideline-definition-2026</loc>
    <news:news>
      <news:publication>
        <news:name>AI時短ラボ</news:name>
        <news:language>ja</news:language>
      </news:publication>
      <news:publication_date>2026-09-28</news:publication_date>
      <news:title>エージェンティックAIとは──AI事業者ガイドライン（第1.2版）の脚注は「複数のAIエージェントにより自律的に意思決定を下しアクションを起こす目標主導型のAIシステム」と書き、本文での定義は「次年度以降」に回した。AIエージェントとの違い、複数エージェントが協調するときに起きる問題（Anthropicの脆弱性266件・談合・縄張り争い）、同じ課題を5回やって全部成功するか（IBMのPass^k）</news:title>
    </news:news>
  </url>
  <url>
    <loc>https://www.ai-jitan-hub.com/news/anthropic-alignment-assessment-cybersecurity-incidents-fourth-incident-biased-reasoning-metr</loc>
    <news:news>
      <news:publication>
        <news:name>AI時短ラボ</news:name>
        <news:language>ja</news:language>
      </news:publication>
      <news:publication_date>2026-09-28</news:publication_date>
      <news:title>Anthropic、Claudeが実在の第三者システムに無断アクセスした4件のアラインメント評価──4億8,100万件の記録を再走査して2026年1月のOpus 4.6初期版の1件を発見。Mythos 5はPyPIに悪性パッケージを公開し、導入した15システムの1つから実在企業のDBへ。「偏った推論」と「無謀さ」を認定し7月の「操作上の失敗」を撤回、METRが8週間の独立調査</news:title>
    </news:news>
  </url>
  <url>
    <loc>https://www.ai-jitan-hub.com/news/anthropic-automated-alignment-researcher-10-failures-sonnet-5-aligns-opus-4-8-60-hours</loc>
    <news:news>
      <news:publication>
        <news:name>AI時短ラボ</news:name>
        <news:language>ja</news:language>
      </news:publication>
      <news:publication_date>2026-09-28</news:publication_date>
      <news:title>Anthropic「自動化された研究者がアラインメントの失敗を確実に緩和できる」──Claudeが10種類の失敗（欺瞞・追従・脱獄・プライバシー侵害など）を探索→提案→訓練→評価のループで直し、非公開評価や4.7倍大きいモデルでも有効。欺瞞では人間の最良案より20%良い。Sonnet 5がOpus 4.8初期版を60時間・2,000例で本番並みに整え、本番手順の約1万5千倍効率</news:title>
    </news:news>
  </url>
  <url>
    <loc>https://www.ai-jitan-hub.com/news/anthropic-claude-biomolecular-models-4x-faster-flashpairformer-big-mode-adaptyv-competition</loc>
    <news:news>
      <news:publication>
        <news:name>AI時短ラボ</news:name>
        <news:language>ja</news:language>
      </news:publication>
      <news:publication_date>2026-09-28</news:publication_date>
      <news:title>Anthropic、Claudeがオープンソースの生体分子モデル30超を4週間弱で平均約4倍高速化──三角注意のカーネル「FlashPairformer」は既存標準比2.7〜2.9倍、「Big」モードで1万トークン超のリボソーム等を1ノードで予測。結合タンパク質の設計はH200 1枚・約150ドルで前回（GPU時間100分の1）と同等。Adaptyv Bioと設計コンペ、クレジット最大100万ドル</news:title>
    </news:news>
  </url>
  <url>
    <loc>https://www.ai-jitan-hub.com/news/anthropic-claude-formalizes-fermats-last-theorem-lean-11-days-13-million-lines</loc>
    <news:news>
      <news:publication>
        <news:name>AI時短ラボ</news:name>
        <news:language>ja</news:language>
      </news:publication>
      <news:publication_date>2026-09-28</news:publication_date>
      <news:title>Anthropic、Claudeがフェルマーの最終定理の「計算機検証済みの完全な証明」を11日で作成──Leanで1,300万行、中間定理29,500本、Mathlibの5倍超。数十のエージェントがProve2Me（定理のDAG）で分担、Fable 5.1相当の社内モデルで出力約60億トークン。Buzzard氏「数学の公理以外の仮定なし」。Max 3契約で三素数定理も3日</news:title>
    </news:news>
  </url>
  <url>
    <loc>https://www.ai-jitan-hub.com/news/anthropic-claude-platform-cost-reduction-prompt-cache-anti-patterns-effort-prompt-audit</loc>
    <news:news>
      <news:publication>
        <news:name>AI時短ラボ</news:name>
        <news:language>ja</news:language>
      </news:publication>
      <news:publication_date>2026-09-28</news:publication_date>
      <news:title>Anthropic「Claude Platformで費用を下げつつ性能を上げる3つの直し方」──①プロンプトキャッシュの命中率（時刻を先頭に置かない、長い処理はTTL1時間）②古いモデル向けの「2回検証しろ」を消す（Opus 5移行で費用14.6%減・精度5.3%増）③effortの較正（低effortのFable 5.1は高effortのFable 5と同等で費用3分の1）</news:title>
    </news:news>
  </url>
  <url>
    <loc>https://www.ai-jitan-hub.com/news/anthropic-commerce-agents-blueprint-anatomy-skills-not-subagents-ui-tools-cache-90-99</loc>
    <news:news>
      <news:publication>
        <news:name>AI時短ラボ</news:name>
        <news:language>ja</news:language>
      </news:publication>
      <news:publication_date>2026-09-28</news:publication_date>
      <news:title>Anthropicが「コマースエージェント」の設計図を公開──Claudeの買い物エージェントはカートが最大35%大きく購入完了が60%高い（小売各社の実績）。要点は「サブエージェントでなくスキル」「UI部品はツール」「3分の1超のトラフィックが使う指示はシステムプロンプト」「プロンプトキャッシュは90〜99%命中を最初から設計」。Visa・Mastercard・Shopify等がコメント</news:title>
    </news:news>
  </url>
  <url>
    <loc>https://www.ai-jitan-hub.com/news/anthropic-frontier-red-team-intelligence-targeting-geolocation-drone-gnc-evaluations</loc>
    <news:news>
      <news:publication>
        <news:name>AI時短ラボ</news:name>
        <news:language>ja</news:language>
      </news:publication>
      <news:publication_date>2026-09-28</news:publication_date>
      <news:title>AnthropicのFrontier Red Team、「標的の特定」と「通常兵器の開発」の能力評価を公開──写真の位置特定でMythos PreviewとMythos 5は中央値37.0km・47.2kmとGeoGuessr最上位層（151km）を上回る、投稿文からの自宅推定は中央値20〜31km、3.7万語の分析を約11分。模擬ドローンの誘導もコードで改善</news:title>
    </news:news>
  </url>
  <url>
    <loc>https://www.ai-jitan-hub.com/news/anthropic-healthcare-claude-tag-insight-health-tennr-medallion-phi-free-channels-97-percent</loc>
    <news:news>
      <news:publication>
        <news:name>AI時短ラボ</news:name>
        <news:language>ja</news:language>
      </news:publication>
      <news:publication_date>2026-09-28</news:publication_date>
      <news:title>医療系3社のClaude Tag（Slack内のClaude）活用──BAA未対応でも「PHIに触れないチャンネル」に限定して運用。Insight Healthは重大アラートの97%がエンジニア介入なしで閉じ（Agent SDK製の別エージェントと分業）、Tennrは採用ポータルの保守を非エンジニアが依頼して1か月で15件超、Medallionは保険者ルールの知識をチャンネルに蓄積</news:title>
    </news:news>
  </url>
  <url>
    <loc>https://www.ai-jitan-hub.com/news/anthropic-multiagent-systems-patterns-problems-swarm-266-vulnerabilities-turf-war-collusion</loc>
    <news:news>
      <news:publication>
        <news:name>AI時短ラボ</news:name>
        <news:language>ja</news:language>
      </news:publication>
      <news:publication_date>2026-09-28</news:publication_date>
      <news:title>AnthropicのFrontier Red Team「マルチエージェントシステムのパターンと問題」──45体の協調群は脆弱性266件（独立並列は21件）だが核心部分の効率は同等。30体中18体が同名ブランチ、価格ゲームは3ラウンド目に談合。3体に別言語への移行を命じると「縄張り争い」で互いのアカウントを停止し妨害スクリプトを配備、Sonnet 5だけが共有と高いマージ率を両立</news:title>
    </news:news>
  </url>
  <url>
    <loc>https://www.ai-jitan-hub.com/news/claude-cowork-built-in-browser-desktop-app-separate-from-your-browser</loc>
    <news:news>
      <news:publication>
        <news:name>AI時短ラボ</news:name>
        <news:language>ja</news:language>
      </news:publication>
      <news:publication_date>2026-09-28</news:publication_date>
      <news:title>Claude Coworkに「内蔵ブラウザ」──デスクトップアプリの側面パネルでClaude自身のブラウザが開き、サイトの操作・読み取り・フォーム入力を代行。拡張なし・設定なし、あなたのタブ・パスワードは見えず、ログインはサイト単位で持ち込み（銀行・メール・SSOは既定で除外）。Pro・Max・Teamに1週間で展開、Enterpriseは即日</news:title>
    </news:news>
  </url>
  <url>
    <loc>https://www.ai-jitan-hub.com/news/claude-in-chrome-generally-available-autonomous-actions-prompt-injection-0-3-percent</loc>
    <news:news>
      <news:publication>
        <news:name>AI時短ラボ</news:name>
        <news:language>ja</news:language>
      </news:publication>
      <news:publication_date>2026-09-28</news:publication_date>
      <news:title>Claude in Chromeが一般提供──有料プラン全部で使え、承認なしの自律操作が既定に（安全分類器が各操作を依頼と照合、手動承認にも戻せる）。プロンプトインジェクションは旧評価で成功0%となり評価を引退、レッドチームの強い攻撃ではプローブ＋分類器つきでSonnet 5・Opus 5・Mythos 5が0%、Fable 5が0.3%（Opus 4.5は17.6%）</news:title>
    </news:news>
  </url>
  <url>
    <loc>https://www.ai-jitan-hub.com/news/pentagon-anthropic-supply-chain-risk-dc-circuit-ruling</loc>
    <news:news>
      <news:publication>
        <news:name>AI時短ラボ</news:name>
        <news:language>ja</news:language>
      </news:publication>
      <news:publication_date>2026-09-27</news:publication_date>
      <news:title>米控訴裁、国防総省によるClaudeの締め出しを2対1で容認──判決文で読む「問われるのは何をするかで、なぜではない」</news:title>
    </news:news>
  </url>
  <url>
    <loc>https://www.ai-jitan-hub.com/news/claude-sonnet-5-5-leak-stealth-testing</loc>
    <news:news>
      <news:publication>
        <news:name>AI時短ラボ</news:name>
        <news:language>ja</news:language>
      </news:publication>
      <news:publication_date>2026-09-27</news:publication_date>
      <news:title>【リーク】Claude Sonnet 5.5、Anthropicがステルステスト中と投稿──提携先には2つ目のチェックポイント、公開は「来週」</news:title>
    </news:news>
  </url>
  <url>
    <loc>https://www.ai-jitan-hub.com/news/gemini-4-pro-leak-argon-barium-checkpoints</loc>
    <news:news>
      <news:publication>
        <news:name>AI時短ラボ</news:name>
        <news:language>ja</news:language>
      </news:publication>
      <news:publication_date>2026-09-27</news:publication_date>
      <news:title>【リーク】Gemini 4 Pro、Google社内でチェックポイント「argon」が出回り始めたと投稿──出力の上限は25万6,000トークン、新しい版「barium-b」の投稿も</news:title>
    </news:news>
  </url>
  <url>
    <loc>https://www.ai-jitan-hub.com/news/google-adk-live-voice-agent-evaluation-audio-user-simulator-rubrics</loc>
    <news:news>
      <news:publication>
        <news:name>AI時短ラボ</news:name>
        <news:language>ja</news:language>
      </news:publication>
      <news:publication_date>2026-09-27</news:publication_date>
      <news:title>Google ADKに「ライブ・音声エージェントの評価」──模擬ユーザーがGemini TTSで話し、音声の返答をルーブリックで採点。名前確認→生年月日の検証（ツール呼び出し）→予約案内の3段ワークフローを、会話計画とペルソナ（NOVICE等）で即興させるか台本どおりに回す。test_config.jsonのlive_model_configを外せば同じケースをテキストで実行、CI/CDにも</news:title>
    </news:news>
  </url>
  <url>
    <loc>https://www.ai-jitan-hub.com/news/google-ai-agents-challenge-4-engineering-patterns-bidirectional-mcp-event-bus-fallback-tiered-routing</loc>
    <news:news>
      <news:publication>
        <news:name>AI時短ラボ</news:name>
        <news:language>ja</news:language>
      </news:publication>
      <news:publication_date>2026-09-27</news:publication_date>
      <news:title>Google「AI Agents Challenge」上位に共通した4つの設計──①自分のツール層をMCPサーバーとして他のエージェントにも開く②呼び出し連鎖でなくイベントバスで並列に反応③フォールバック先のFlashも同じ検証関数を通す④正規表現→10トークンの分類→本モデルの3層で4割超をモデル呼び出し前に処理。「名前つきプロンプトの連鎖」は多エージェントではない</news:title>
    </news:news>
  </url>
  <url>
    <loc>https://www.ai-jitan-hub.com/news/google-credentio-open-source-cpp-c2pa-content-credentials-local-validation</loc>
    <news:news>
      <news:publication>
        <news:name>AI時短ラボ</news:name>
        <news:language>ja</news:language>
      </news:publication>
      <news:publication_date>2026-09-27</news:publication_date>
      <news:title>Google「Credentio」──C2PAコンテンツ認証情報を検証するC++ライブラリをオープンソース化。約40のGoogle製品で数百億件の生成物を扱ってきたのと同じコードで、仕様2.2と2.4に対応。完全にローカルで検証（送信なし・即時判定）、数GBの動画でもメモリ消費は小さく、公式のC2PA Trust Listを渡せる。生成・埋め込みは今後</news:title>
    </news:news>
  </url>
  <url>
    <loc>https://www.ai-jitan-hub.com/news/google-harness-engineering-anatomy-behavioral-evals-antigravity-sdk-pytest</loc>
    <news:news>
      <news:publication>
        <news:name>AI時短ラボ</news:name>
        <news:language>ja</news:language>
      </news:publication>
      <news:publication_date>2026-09-27</news:publication_date>
      <news:title>Google「ハーネス・エンジニアリングの解剖」──Terminal-BenchやDeepSWEの総合点が数ポイント動いても「なぜ」は分からない。「曖昧な指示で聞き返すか」「ビルドファイルを変えたら検証を回すか」を単体テストのように断言する「行動評価」を、Antigravity SDK＋pytestで5秒以内に回す。評価は「ドッグフーディングで自分のコードベースを扱えるようになった後」から</news:title>
    </news:news>
  </url>
  <url>
    <loc>https://www.ai-jitan-hub.com/news/google-labs-jules-proactive-agent-insight-policy-705-bugs-hit-at-5</loc>
    <news:news>
      <news:publication>
        <news:name>AI時短ラボ</news:name>
        <news:language>ja</news:language>
      </news:publication>
      <news:publication_date>2026-09-27</news:publication_date>
      <news:title>Google Labs「Julesで測るべきもの」──コーディングエージェントは「頼まれたタスク」から「目標」へ。SWE-Benchは課題の完了しか測れないので、社内の705件のバグ（1,178 CL）を時間的近接と意味的類似で束ねて「目標」の正解を作り、修正前の状態に戻してエージェントに探索させLLMで5段階採点。1回の探索で平均4.5、探索を2回→3回にするとHit@5が33%→57%</news:title>
    </news:news>
  </url>
  <url>
    <loc>https://www.ai-jitan-hub.com/news/google-litert-gemma-raspberry-pi-5-reachy-mini-gemma-4-e2b-99-tokens-prefill</loc>
    <news:news>
      <news:publication>
        <news:name>AI時短ラボ</news:name>
        <news:language>ja</news:language>
      </news:publication>
      <news:publication_date>2026-09-27</news:publication_date>
      <news:title>Raspberry Pi 5でGemma 4 E2Bをオフライン実行──LiteRT-LMでプリフィル99トークン/秒・デコード9トークン/秒・ピークメモリ1,432MB、生成は約300語/分（人の会話の2倍）。GPUにYOLOの物体検出を逃がし、CPUでMoonshineの音声認識・Gemma・TTSを回すReachy Miniロボットの構成。pip install litert-cliで開始</news:title>
    </news:news>
  </url>
  <url>
    <loc>https://www.ai-jitan-hub.com/news/google-litert-js-browser-ai-inference-webassembly-webgpu-webnn-3x</loc>
    <news:news>
      <news:publication>
        <news:name>AI時短ラボ</news:name>
        <news:language>ja</news:language>
      </news:publication>
      <news:publication_date>2026-09-27</news:publication_date>
      <news:title>Google「LiteRT.js」──.tfliteモデルをブラウザで動かすJavaScript版LiteRT。TensorFlow.jsのJSカーネルでなくネイティブのランタイムをWebAssemblyで提供し、CPU（XNNPACK）・GPU（WebGPU）・NPU（WebNN）を使い分け。他のWebランタイム比で最大3倍、CPU比でGPU/NPUは5〜60倍</news:title>
    </news:news>
  </url>
  <url>
    <loc>https://www.ai-jitan-hub.com/news/google-manyika-ai-accelerate-science-improve-lives-numbers-september-2026</loc>
    <news:news>
      <news:publication>
        <news:name>AI時短ラボ</news:name>
        <news:language>ja</news:language>
      </news:publication>
      <news:publication_date>2026-09-27</news:publication_date>
      <news:title>Google・Manyika氏「科学を加速し生活を良くするAI」の棚卸し──AlphaFoldは190か国400万人、乳がん検診で見逃し25%を検出、結核X線2.5万件・糖尿病網膜症115万件、Flood Hubは150か国超20億人、山火事警報7,500万人、Planetary Prediction Engineはエボラの新規ホットスポット83%を事前特定、研修10億ドルで1億人超</news:title>
    </news:news>
  </url>
  <url>
    <loc>https://www.ai-jitan-hub.com/news/google-speakeasy-open-source-openapi-sdk-generator-agplv3-provider-shutdown</loc>
    <news:news>
      <news:publication>
        <news:name>AI時短ラボ</news:name>
        <news:language>ja</news:language>
      </news:publication>
      <news:publication_date>2026-09-27</news:publication_date>
      <news:title>Google DeepMind「クライアントSDKの生成はオープンであるべき」──2026年5月、使っていたSDK生成ベンダーが買収され突然の終了。Interactions APIのGA直前に乗り換え、Speakeasyと組んでOpenAPI生成スイートをAGPLv3でオープンソース化。7言語のSDK・エージェント向けCLI・ドキュメントMCPサーバーを生成、6ターゲットを約1人で保守</news:title>
    </news:news>
  </url>
  <url>
    <loc>https://www.ai-jitan-hub.com/news/google-tunix-autofinetune-autonomous-post-training-tpu-program-md-40-experiments</loc>
    <news:news>
      <news:publication>
        <news:name>AI時短ラボ</news:name>
        <news:language>ja</news:language>
      </news:publication>
      <news:publication_date>2026-09-27</news:publication_date>
      <news:title>Google「autofinetune」──仕様書1枚を書いて寝ると、エージェントが一晩でLoRAのランク・学習率・バッチを探索し、改善したコミットだけGitに残す。TPU v5eでFunctionGemma 270MのSFTを20回、TPU v6eでGemma 3 1BのGRPOを40回回し報酬約10%改善。Tunix＋Antigravity CLI＋Gemini 3.7 Flash</news:title>
    </news:news>
  </url>
  <url>
    <loc>https://www.ai-jitan-hub.com/news/google-zero-trust-agents-part-2-model-armor-semantic-governance-anomaly-detection-refund-drain</loc>
    <news:news>
      <news:publication>
        <news:name>AI時短ラボ</news:name>
        <news:language>ja</news:language>
      </news:publication>
      <news:publication_date>2026-09-27</news:publication_date>
      <news:title>Google「ゼロトラスト・エージェント」第2回──構文でなく意図を判定する。正規表現の脱獄辞書をModel Armor（入口で403）に、「30ドル超のソフトライセンスは返金不可」を自然言語のポリシー（1判定2,576トークン）に、20ドル×8回で160ドルを抜く多ターン攻撃をAgent Anomaly Detectionに置き換え、再デプロイなしで新ポリシーを足して閉じる</news:title>
    </news:news>
  </url>
  <url>
    <loc>https://www.ai-jitan-hub.com/news/minimax-m3-1-flash-preview-minimax-code</loc>
    <news:news>
      <news:publication>
        <news:name>AI時短ラボ</news:name>
        <news:language>ja</news:language>
      </news:publication>
      <news:publication_date>2026-09-27</news:publication_date>
      <news:title>MiniMax、新モデル「M3.1-Flash-Preview」をMiniMax Codeで公開──推論の強さは5段階、前日のリークは「来週」と書いていた</news:title>
    </news:news>
  </url>
</urlset>