aidotse/LeakPro

Standardize naming for target outputs and attack signals across MIA attacks

开放

#444 创建于 2026年8月18日

 (1 条评论) (0 个反应) (0 位负责人)Python (26 个派生)auto 404
documentationenhancementgood first issuepriority - 4

仓库指标

星标
 (23 个星标)
PR 合并指标
 (平均合并 68天 23小时) (30 天内合并 5 个 PR)

描述

Different MIA implementations currently use different names for similar values. Examples include:

  • base: logits_target
  • attack_p: attack_signal and audit_signal
  • loss_trajectory: target
  • LiRA and MS-LiRA: taget_model_logitsm, target_signals, shadow_models_signals, and sample_target_signals

Use the same naming pattern across all similar MIA attacks. The names should make it clear whether a variable contains the model’s original output or a signal calculated from that output. For example, use target_outputs and shadow_outputs or target_model_outputs and shadow_model_outputs for original model outputs.

It might be better with outputs than logits, because LeakPro also supports regression and forecasting models, which do not necessarily produce logits.

贡献者指南