[Phase 16] S9 — Compliance evidence integrity and GDPR erasure completeness
Nessuno ha ancora preso questa issue.
Valutazione
- Difficoltà
- 4/5
- Tempo stimato
- Più di una settimana
- Idoneità per principianti
- 22/100
Direzione di ricerca
Start with _artefact_targets_for_run and the _legacy_staging_owned_by marker pattern in _purge.py, then trace how generate_training_manifest() and export_compliance_artifacts() name the bundle files. The first step is a writer-to-reader integration test that exports a real bundle and purges by run id. Done means the matching run's bundle is fully removed, a wrong run id removes nothing, and pipeline_manifest.json and audit_log.jsonl survive.
Scritto dal modello di indicizzazione a partire dal testo della issue.
Descrizione
Roadmap · Phase 16, step S9 · size L · 5 unit(s) · planned in
docs/roadmap/phase-16-trust-surface-hardening.md:553-581
Summary
The compliance surface claims a stronger evidentiary state than it holds. purge --kind artefacts cannot find the compliance bundle ForgeLM itself writes: the exporter emits fixed basenames while the resolver looks for file names that embed the run id. It fails safe (deletes nothing, emits data.erasure_failed with NoMatchingArtefacts, exits 1), so the defect is completeness, not destruction. The trainer also writes data_governance_report.json there, and the pipeline writes compliance/pipeline_manifest.json, which belongs to the whole pipeline and must not be deleted by one run's purge.
Units
| Unit | Tracked by (October 2026 review) |
|---|---|
F-W20260729-TRUST-02 |
#101 · #102 · #477 · #554 · #635 |
F-W20260729-TRUST-13, F-W20260729-TRUST-01 |
#386 · #651 |
F-W20260729-TRUST-14 |
#386 · #554 |
F-W20260729-TRUST-07 |
#386 · #679 · #680 · #681 · #715 |
Plan
- Rewrite
_artefact_targets_for_runto the allow-list + ownership-proof model the design document already specifies, mirroring the_legacy_staging_owned_bymarker pattern in the same file. - Breaking changes planned here:
AuditLogger.log_eventbecomes fail-closed;reverse-piiandpurgemove to non-zero exits on states that returned 0.
Exit criteria
- A real bundle produced by
generate_training_manifest()+export_compliance_artifacts()is fully removed by a matching run id and untouched by a wrong one, whilepipeline_manifest.jsonandaudit_log.jsonlsurvive — a writer → reader integration test, not a synthetic-filename fixture; the fabricated docstring examples in_purge.pyare deleted. - A malformed line sharing an identifier blocks the completion record.
-
reverse-piireturns non-zero whenAuditLoggerinitialisation raises. - An
OSErrorinjected separately at the manifestopen,fsyncandos.replacesteps never yields silent success. - A
required → rejected → requiredchain is served by one shared ordered-state helper across the listing and mutation surfaces.
How to deliver
- One root-cause family, committed on its own; then an Opus review round and a Sonnet review round, each with verified findings fixed and committed before the next step begins.
- A reproducer that is red before the fix and green after; the full gauntlet from
CLAUDE.mdpasses under./.venv/bin/python. - Every English documentation edit ships with its Turkish mirror in the same commit; a new or renamed audit event gets its EN and TR catalog rows in the same PR.
- A step that grows a deferred module carries its
budget_historyjustification in the same diff; a widened guard rolls out non-strict first, then flips. - Commit bodies cite
Absorbs F-W20260729-…; the CHANGELOG[Unreleased]section gets an entry only for a change a user or library consumer can observe (decision C-20).
Planned in docs/roadmap/phase-16-trust-surface-hardening.md (read at f94595f). Unit IDs (F-W20260729-…) come from the 2026-07-29/30 full-project review; the issue numbers next to them are the October 2026 review's records of the same defects.
- Lingua principale
- Python
- Stelle
- 9
- Fork
- 1
- Merge medio
- 4h 3m
- PR unite (30g)
- 3
Preparare l'ambiente
- Include un Dockerfile o un file Docker Compose
- Ha un modello di pull request
- Leggi la guida per i contributori
Come iniziare
- Leggi tutta la issue e poi la guida ai contributi del progetto.
- Commenta sulla issue per dire che te ne occupi tu — evita che due persone facciano lo stesso lavoro.
- Fai un fork del repository e lavora su un branch.
- Apri una pull request che faccia riferimento al numero della issue.
Altre issue di HodeTech/ForgeLM
-
area: dev-tooling bug severity: low source: roadmap wave: 4
Difficoltà 2/5 1-3 ore Idoneità per principianti 70/100
-
area: site bug good first issue severity: low source: review-2026-09 wave: 4
Difficoltà 2/5 1-3 ore Idoneità per principianti 66/100
-
documentation good first issue severity: medium source: review-2026-09 wave: 3
Difficoltà 2/5 1-3 ore Idoneità per principianti 75/100
-
documentation good first issue severity: low source: review-2026-09 wave: 4
Difficoltà 1/5 1-3 ore Idoneità per principianti 80/100
-
documentation severity: medium source: review-2026-09 wave: 3
Difficoltà 2/5 1-3 ore Idoneità per principianti 72/100
Tutte le issue di HodeTech/ForgeLM
Issue simili
-
Difficoltà 2/5 1-3 ore Idoneità per principianti 72/100
NousResearch/hermes-agent#136483 ·
I maintainer di solito rispondono entro 1 giorno
-
Difficoltà 1/5 Meno di un'ora Idoneità per principianti 88/100
I maintainer di solito rispondono entro 1 giorno
-
[BUG] LazyStackedTensorDictStore zeroes the last byte of a new key set on the last elementForse già presa @peterdsharpe l’ha presa oggi. Apertabug
Difficoltà 2/5 1-3 ore Idoneità per principianti 78/100
pytorch/tensordict#2307 ·
I maintainer di solito rispondono entro 1 giorno
-
Difficoltà 2/5 1-3 ore Idoneità per principianti 78/100
I maintainer di solito rispondono entro 1 giorno
-
GrokModel.generate/a_generate pass an OpenAI-style list-of-dicts to xai_sdk.chat.user(), so every call crashes with a protobuf TypeError before any network I/OForse già presa @Christian-Sidak l’ha presa oggi. Aperta
Difficoltà 2/5 1-3 ore Idoneità per principianti 70/100
confident-ai/deepeval#3436 · 1 commento ·
I maintainer di solito rispondono entro 1 giorno