RFC: context-boundary validation — provenance for untrusted tool/MCP content that can drive tool calls
@sanketpatil06 がすでに取り組んでいます。
2026年9月14日 から。
評価
この issue はまだ評価されていません。
説明
Problem
ADK assembles the model's context from sources that are effectively untrusted relative to the tool surface it can reach: tool descriptions (including ones pulled from remote MCP servers), tool outputs, session state written by other agents, and instructions loaded at runtime. None of that content carries provenance the framework tracks, so by the time a call reaches the before_tool_callback pipeline in _tool_caller.py, the callback sees the request but has no first-class way to tell what influenced it — a tool result from a low-trust source looks exactly like the developer's own instruction.
This is the documented "context privilege escalation" pattern: untrusted content in the context window steers the model into calling privileged tools it should never have combined (arXiv 2609.01222 verified this against a dozen agent harnesses). #7076 is the same problem class showing up concretely in the resumable-mode dispatch path.
For related work I checked before filing: #6966 screens tool output via ModelArmor but is vendor-specific screening, and #6099 records decisions for audit after the fact. Neither gives callbacks a general notion of where context content came from.
Describe the solution you'd like
A small opt-in boundary layer, not a new enforcement framework:
- Provenance tags on content entering context — source kind (developer instruction / tool output / MCP server / session state) attached at assembly points ADK already controls (event construction, toolset declarations, state merge).
- That boundary context made visible to the existing
before_tool_callbackpipeline andToolConfirmation, so a callback can ask "were this call's arguments influenced by untrusted sources, and does that source have scope to call this tool?" without re-implementing event archaeology. - Default-allow policy; teams that need it can enforce rules like "MCP tool outputs cannot authorize tool X" or require confirmation on cross-boundary escalation.
Impact on your work
Tool/MCP security is the main blocker we hit when arguing for enterprise agent adoption, and the shape of this API matters a lot. The contribution guide asks for an issue first on substantial features, so before building a minimal prototype and adversarial tests: would this fit better as a core contract on the context/event objects, or as a plugin-level capability reusing the existing callback pipeline? I'd rather align on the direction than open a PR against a guess.
- 主要言語
- Python
- スター
- 21.6k
- フォーク
- 4k
- 平均マージ
- 7時間 10分
- マージ済み PR(30日)
- 7
コントリビューションガイド
はじめの一歩
- issue を最後まで読み、次にプロジェクトのコントリビューションガイドを読みます。
- 着手することを issue にコメントします — 二人が同じ作業をするのを防げます。
- リポジトリをフォークし、ブランチを切って変更します。
- issue 番号を参照したプルリクエストを送ります。
google/adk-python のほかの issue
-
難易度 2/5 1〜3時間 初心者へのやさしさ 78/100
google/adk-python#7266 · コメント 1 件 ·
-
難易度 2/5 1〜3時間 初心者へのやさしさ 78/100
google/adk-python#7265 · コメント 1 件 ·
-
mcp
難易度 2/5 1〜3時間 初心者へのやさしさ 75/100
google/adk-python#7217 · コメント 3 件 · 担当者 1 名 ·
-
tools
難易度 2/5 1〜3時間 初心者へのやさしさ 78/100
google/adk-python#7206 · コメント 1 件 · 担当者 1 名 ·
-
request clarification tools
難易度 2/5 1〜3時間 初心者へのやさしさ 86/100
google/adk-python#7205 · コメント 2 件 · 担当者 1 名 ·
google/adk-python の issue をすべて見る
似ている issue
-
agent-ready documentation needs-triage
難易度 1/5 1〜3時間 初心者へのやさしさ 88/100
-
documentation
難易度 1/5 1時間未満 初心者へのやさしさ 91/100
-
workflow-status page template still says reusable workflows are "triggered only by workflow_call:" オープン
難易度 1/5 1時間未満 初心者へのやさしさ 92/100
-
instance instance add
難易度 1/5 1時間未満 初心者へのやさしさ 72/100
searxng/searx-instances#939 · コメント 1 件 ·
-
area-deployment area-integrations triage:bot-seen
難易度 2/5 半日 初心者へのやさしさ 86/100