aidotse/LeakPro
Standardize naming for target outputs and attack signals across MIA attacks
开放
#444 创建于 2026年8月18日
documentationenhancementgood first issuepriority - 4
仓库指标
- 星标
- (23 个星标)
- PR 合并指标
- (平均合并 68天 23小时) (30 天内合并 5 个 PR)
描述
Different MIA implementations currently use different names for similar values. Examples include:
- base: logits_target
- attack_p: attack_signal and audit_signal
- loss_trajectory: target
- LiRA and MS-LiRA: taget_model_logitsm, target_signals, shadow_models_signals, and sample_target_signals
Use the same naming pattern across all similar MIA attacks.
The names should make it clear whether a variable contains the model’s original output or a signal calculated from that output.
For example, use target_outputs and shadow_outputs or target_model_outputs and shadow_model_outputs for original model outputs.
It might be better with outputs than logits, because LeakPro also supports regression and forecasting models, which do not necessarily produce logits.