Allow user-defined key/value tags on runs for filtering by test intent
I maintainer di solito rispondono entro 1 giorno
Nessuno ha ancora preso questa issue.
Valutazione
- Difficoltà
- 5/5
- Tempo stimato
- Più di una settimana
- Idoneità per principianti
- 35/100
- Tipo di issue
- Funzionalità
- Chiarezza
- Abbastanza chiara
- Stato di attività
- Tranquilla
- Stack tecnologico
- typescript
- Ambito
- backend-api-design, frontend
Direzione di ricerca
No files, tests, or entry points are named. Start by tracing run/request creation and the Runs table, then identify the API and persistence paths involved; done means optional tags can be created, edited, returned, displayed, filtered, and preserved across retries or reruns.
Scritto dal modello di indicizzazione a partire dal testo della issue.
Descrizione
Original author: @JaGord
Summary
Please add support for user-defined key/value tags on SCOPE runs and requests so users can filter and group runs by the reason they were created.
Why
When running benchmark matrices, the profile name and task prompt are not always enough to explain the intent of a run. Users need to mark runs with labels such as:
tag: purpose
value: test-agent-kit-branch
or:
tag: criteria_focus
value: vector-index-policy
This would make it easier to answer questions like:
- Which runs were created to test a specific Agent Kit change?
- Which runs were created to validate a specific criterion or failure mode?
- Which runs are part of a rerun batch after a harness fix?
- Which runs should be included or excluded from a comparison table?
Requested capability
Allow users to attach arbitrary key/value tags to a run or request at creation time and edit them later.
Example shape:
{
"tags": {
"purpose": "test-agent-kit-branch",
"criteria_focus": "source_excludes_embedding_from_range_index",
"batch": "rag-chat-rerun-20260626",
"include_in_totals": "true"
}
}
Expected behavior
- Tags are visible in run/request details.
- Tags can be added at request creation time.
- Tags can be edited after creation.
- Runs can be filtered by tag key and key/value pair.
- Tags are returned by the API so automation can group and summarize runs reliably.
- Reports should include tags or link back to tagged request metadata.
Use case
For Agent Kit benchmarking, I may run the same task prompt against several profiles, branches, reruns, or criteria-focused variants. A tag such as purpose=test-vector-index-guidance or batch=agent-kit-split-vs-monolith-rerun would make it much easier to filter the Runs table and avoid accidentally mixing unrelated runs in aggregate results.
Acceptance criteria
- Request/run creation accepts optional key/value tags.
- The Runs table supports filtering by tag key and tag value.
- Request/run APIs return tags.
- Tags are preserved across retries/reruns or clearly copied to replacement requests.
- Tags are visible enough that users can tell "this run was to test X" without opening external notes.
- Lingua principale
- TypeScript
- Stelle
- 7
- Fork
- 13
- Merge medio
- 3g 5h
- PR unite (30g)
- 26
Preparare l'ambiente
- Include un Dockerfile o un file Docker Compose
- Ha un modello di pull request
- Leggi la guida per i contributori
Come iniziare
- Leggi tutta la issue e poi la guida ai contributi del progetto.
- Commenta sulla issue per dire che te ne occupi tu — evita che due persone facciano lo stesso lavoro.
- Fai un fork del repository e lavora su un branch.
- Apri una pull request che faccia riferimento al numero della issue.
Altre issue di microsoft/scope
-
type: worker-update
Difficoltà 1/5 1-3 ore Idoneità per principianti 85/100
I maintainer di solito rispondono entro 1 giorno
-
type: worker-update
Difficoltà 1/5 Meno di un'ora Idoneità per principianti 88/100
I maintainer di solito rispondono entro 1 giorno
-
type: worker-update
Difficoltà 1/5 1-3 ore Idoneità per principianti 78/100
I maintainer di solito rispondono entro 1 giorno
-
Clarify that prompt features only categorize prompts, don't impact runsForse già presa @DerrickUnleashed l’ha presa 3 giorni fa. Apertaauthor: JaGord documentation good first issue UI
Difficoltà 2/5 1-3 ore Idoneità per principianti 72/100
I maintainer di solito rispondono entro 1 giorno
-
Submit Run task prompt picker should only search requirement promptsForse già presa @cedricvidal l’ha presa 69 giorni fa. Apertaauthor: cedricvidal bug portal
Difficoltà 2/5 1-3 ore Idoneità per principianti 85/100
I maintainer di solito rispondono entro 1 giorno
Tutte le issue di microsoft/scope
Issue simili
-
Difficoltà 2/5 1-3 ore Idoneità per principianti 83/100
I maintainer di solito rispondono entro 1 giorno
-
Signals (Failure Detector): a tool call and its own execution are reported as a repeated callAperta
Difficoltà 2/5 1-3 ore Idoneità per principianti 75/100
I maintainer di solito rispondono entro 1 giorno
-
Difficoltà 2/5 1-3 ore Idoneità per principianti 68/100
I maintainer di solito rispondono entro 1 giorno
-
Difficoltà 2/5 1-3 ore Idoneità per principianti 85/100
platformatic/mcp#208 ·
I maintainer di solito rispondono entro 1 giorno
-
🐛 bug
Difficoltà 2/5 1-3 ore Idoneità per principianti 66/100
margelo/react-native-vision-camera#4211 ·
I maintainer di solito rispondono entro 4 giorni