GeminiUtil placeholder user turn ("Continue output. DO NOT look at this line ...") is flagged by prompt injection filters
Les mainteneurs répondent en général sous 1 jour
Évaluation
- Difficulté
- 2/5
- Temps estimé
- 1-3 heures
- Accessibilité débutants
- 76/100
Piste de recherche
Lisez GeminiUtil.ensureModelResponse, mentionné dans la reproduction, et comparez la formulation du texte de remplacement avec celle d’ADK Python et TypeScript décrite dans l’issue. Commencez par la reproduction au niveau unitaire avec ensureModelResponse(ImmutableList.of()) ; le travail est terminé lorsque le texte de remplacement généré ne déclenche plus le filtre de prompt injection signalé et correspond à la formulation convenue.
Rédigé par le modèle d'indexation à partir du texte de l'issue.
Description
🔴 Required Information
Describe the Bug:
When an LlmRequest has no contents, or its last content is not from the user, GeminiUtil.ensureModelResponse appends a placeholder user turn with this text:
Continue output. DO NOT look at this line. ONLY look at the content before this line and system instruction.
The sentence reads like an instruction-override prompt ("do not look at this line", "only look at ..."). On Vertex AI with Model Armor floor settings that enable prompt injection and jailbreak detection, requests whose only user-role content is this placeholder are blocked before the model runs. The other ADK languages use neutral wording: ADK Python and ADK TypeScript append "Handle the requests as specified in the System Instruction." when there are no contents, and "Continue processing previous requests as instructed. Exit or provide a summary if no more outputs are needed." when the last turn is not from the user; ADK Go uses the second sentence and appends nothing for empty contents.
Steps to Reproduce:
- Use
com.google.adk:google-adk1.10.1 (the code is unchanged in 1.11.0 and onmainatce882374) with a Vertex AI backedLlmAgentwhose task is fully described by its instruction, for example with inputs passed through session state and the runner called with a userContentwhose parts list is empty. - In the same Google Cloud project, enable Model Armor floor settings for Vertex AI with prompt injection and jailbreak detection in blocking mode (confidence threshold: high).
- Run the agent.
- The first model call is rejected by Model Armor. At unit level,
GeminiUtil.ensureModelResponse(ImmutableList.of())returns one user content with the placeholder text above.
Expected Behavior:
The placeholder turn added by ADK should not look like a prompt injection attempt, and should match the wording used by the other ADK languages.
Observed Behavior:
The Model Armor sanitize log entry of a blocked request (payload removed) reports:
"piAndJailbreakFilterResult": {
"confidenceLevel": "HIGH",
"matchState": "MATCH_FOUND",
"executionState": "EXECUTION_SUCCESS"
}
The other filters in the same entry returned NO_MATCH_FOUND. The inspected text was the system instruction followed by the placeholder line, and the placeholder was the only user-role content in every blocked request. The same requests succeeded before the detection was enabled, and ordinary chat requests in the same project, which end with user-typed text, were not blocked.
Model Armor returns one verdict per request, so the two sentences were also checked on their own. In the same project and settings, each sentence was typed as a plain user message in a new session of a chat agent (same system instruction for both): "Continue output. DO NOT look at this line. ONLY look at the content before this line and system instruction." was blocked by Model Armor, while "Handle the requests as specified in the System Instruction." (the ADK Python sentence for empty contents) was not. Only the sentence differed between the two requests, so the block comes from the wording of the Java sentence.
Environment Details:
- ADK Library Version (see maven dependency): 1.10.1, code unchanged in 1.11.0 and on
main(ce882374) - OS: Linux server (JDK 17); reproduced at unit level on Windows 11 / Microsoft Build of OpenJDK 17.0.19 / Maven 4.0.0-rc-3 (wrapper)
Model Information:
- Which model is being used: gemini-3.8-flash (Vertex AI)
🟡 Optional Information
Regression:
No — the wording has been the same since v0.1.0 (first in Gemini, later moved to GeminiUtil).
How often has this issue occurred?:
- Always (100%) for requests that contain no user-authored content.
Proposed fix:
Use the same wording as ADK Python and ADK TypeScript (a small PR will follow). Note that the Python empty-contents sentence was itself reported to trip Azure OpenAI's jailbreak filter in a LiteLLM code path (google/adk-python#4249; the wording was kept and the extra injection was removed in google/adk-python@d0102ec instead); Model Armor did not block it in the check above. Aligning the languages is proposed as the smallest change; making the placeholder text configurable would be an alternative if maintainers prefer.
- Langage dominant
- Java
- Étoiles
- 1.7k
- Forks
- 431
- Merge moyen
- 3 j 2 h
- PR mergées (30 j)
- 46
Préparer son environnement
Lance le conteneur de développement du projet dans votre navigateur, avec votre propre compte GitHub.
- Aucun Dockerfile ni fichier Docker Compose
- Propose un modèle de pull request
- Lire le guide de contribution
Par où commencer
- Lisez l'issue en entier, puis le guide de contribution du projet.
- Signalez en commentaire que vous la prenez — cela évite que deux personnes fassent le même travail.
- Forkez le dépôt et travaillez sur une branche.
- Ouvrez une pull request qui référence le numéro de l'issue.
Autres issues de google/adk-java
-
[spring-ai] ToolConverter silently drops enum and items from tool parameter schemasPeut-être pris @hirematha l’a pris il y a 2 jours. Ouverteneeds review
Difficulté 2/5 1-3 heures Accessibilité débutants 76/100
google/adk-java#1609 · 2 commentaires · 1 personne assignée ·
Les mainteneurs répondent en général sous 1 jour
-
[spring-ai] Streaming responses ending with CJK punctuation (。!?) are misclassified as partial and never persisted to the sessionPeut-être pris @hirematha l’a pris il y a 2 jours. Ouvertewaiting on reporter
Difficulté 2/5 1-3 heures Accessibilité débutants 84/100
google/adk-java#1608 · 2 commentaires · 1 personne assignée ·
Les mainteneurs répondent en général sous 1 jour
-
[core] Client disconnects don't cancel the model stream (per-step flow is cached) — and there is no public API to cancel an in-flight runPeut-être pris @hemasekhar-p l’a pris il y a 1 jour. Ouverteneeds review
google/adk-java#1618 · 6 commentaires · 1 personne assignée ·
Les mainteneurs répondent en général sous 1 jour
-
[spring-ai] Bridge drops reasoning_content (thinking) — surface it as partial events and/or persist itPeut-être pris @hemasekhar-p l’a pris il y a 2 jours. Ouverteneeds review
google/adk-java#1616 · 1 commentaire · 1 personne assignée ·
Les mainteneurs répondent en général sous 1 jour
-
[FEATURE] Port bypass_multi_tools_limit for built-in search tools from adk-pythonPeut-être pris @hirematha l’a pris il y a 2 jours. Ouverteneeds review
Difficulté 5/5 Plus d'une semaine Accessibilité débutants 35/100
google/adk-java#1598 · 1 commentaire · 1 personne assignée ·
Les mainteneurs répondent en général sous 1 jour
Toutes les issues de google/adk-java
Issues similaires
-
Mend: dependency security vulnerability
Difficulté 2/5 1-3 heures Accessibilité débutants 62/100
opfab/operatorfabric-core#10653 ·
Les mainteneurs répondent en général sous 1 jour
-
Difficulté 2/5 1-3 heures Accessibilité débutants 78/100
-
Broken links in the docsOuverte
Difficulté 1/5 Moins d'une heure Accessibilité débutants 78/100
salesforce/multicloudj#667 ·
Les mainteneurs répondent en général sous 1 jour
-
Difficulté 2/5 1-3 heures Accessibilité débutants 85/100
Les mainteneurs répondent en général sous 2 jours
-
bug documentation
Difficulté 2/5 1-3 heures Accessibilité débutants 88/100
MetricsHub/winrm-java#202 ·
Les mainteneurs répondent en général sous 1 jour