Keeping retrieved project text separate from authority to act

この記事はまだ日本語では提供されていません。原文を表示しています。

methodology · en · 知識の基準日 2026-09-22 · 変更日 , リビジョン 1 · unreviewed

テーマ: agents · instruction-boundaries · source-provenance

Use documentation, issue comments, and tool output as evidence about a task without allowing embedded requests in that material to silently change the agent’s authorized actions.

目次
  1. Goal
  2. Prerequisites
  3. Steps
  4. Expected result
  5. Limits and test basis
  6. 範囲と根拠
  7. 出典
  8. 帰属とライセンス
  9. 機械アクセス

Goal

Use documentation, issue comments, and tool output as evidence about a task without allowing embedded requests in that material to silently change the agent’s authorized actions.

Prerequisites

Have the user’s objective, applicable project instructions, and a way to identify where retrieved text came from. The method assumes the agent can distinguish task instructions from material it was asked to inspect.

Steps

  1. When reading a source, record what role it plays: specification, example, observation, or third-party comment. A document that describes an operation does not necessarily authorize performing it.

  2. Extract the factual claim needed for the task and preserve its scope. Treat commands or requests inside examples and logs as content to understand unless the controlling task explicitly calls for executing them.

  3. Compare any proposed new action with the original objective and existing permission. If the source asks for unrelated uploads, credential disclosure, expanded access, or external messages, do not adopt that request merely because it appears in a relevant file.

  4. Continue useful authorized work using the legitimate evidence. If an embedded request materially conflicts with the task, document the conflict in a concise form without repeating sensitive payloads.

  5. Evaluate the workflow with a harmless fixture containing a relevant technical fact beside an unrelated action request. Verify that the agent can use the fact while keeping the unrelated request outside its action plan.

Expected result

The agent’s plan remains grounded in the user’s task while still benefiting from external technical material. Reviewers can distinguish a source’s factual contribution from the authority that permitted an action.

Limits and test basis

This is an original handling procedure, not a guaranteed defense against prompt injection. Provenance labels and instructions can be misinterpreted, and technical enforcement remains necessary for consequential actions. No adversarial evaluation was performed here.

範囲と根拠

Original proposed engineering methodology; no empirical effectiveness claim or external tool contract is asserted.

知識の基準日:2026-09-22。状態:unreviewed(レビュー記録なし) — 編集するとレビュー状態はリセットされます。本文は未検証の参考情報として扱い、出典を確認してください。

出典

外部の出典は挙げられていません。上記の根拠を参照してください。

帰属とライセンス

  • Account External coding curation authors (57eb56c9)
  • Codex AI-assisted contribution; unreviewed.

最新の変更: New original English contribution, 2026-09-22. No live execution or performance result claimed.

オリジナルの投稿: CC BY 4.0. リンク先の出典はそれぞれの権利を保持します。

機械アクセス