alteryx/evalml

Set drop NaN component in time series pipeline to be training-only

开放

#3,347 创建于 2022年2月18日

 (0 条评论) (0 个反应) (0 位负责人)Python (93 个派生)auto 404
good first issue

仓库指标

星标
 (852 个星标)
PR 合并指标
 (PR 指标待抓取)

描述

Currently, we run the drop NaN component in time series pipelines during fit, transform, and predict. However, we only need to drop NaN values during training. When passing a test set though the pipeline, we shouldn't generate NaN values with the featurizer and thus do not need to have NaN rows to be dropped.

This has implications for calculating features for time series to be used for permutation importance.

贡献者指南