alteryx/evalml

Set drop NaN component in time series pipeline to be training-only

Open

#3,347 opened on Feb 18, 2022

 (0 comments) (0 reactions) (0 assignees)Python (93 forks)auto 404
good first issue

Repository metrics

Stars
 (852 stars)
PR merge metrics
 (PR metrics pending)

Description

Currently, we run the drop NaN component in time series pipelines during fit, transform, and predict. However, we only need to drop NaN values during training. When passing a test set though the pipeline, we shouldn't generate NaN values with the featurizer and thus do not need to have NaN rows to be dropped.

This has implications for calculating features for time series to be used for permutation importance.

Contributor guide