Evaluation of Active Feature Acquisition Methods for Time-varying Feature Settings

📅 2023-12-03
🏛️ arXiv.org
📈 Citations: 2
✨ Influential: 0
📄 PDF
🤖 AI Summary
This paper addresses the distribution shift induced by active feature acquisition (AFA) deployment and formally introduces the “Active Feature Acquisition Performance Evaluation” (AFAPE) task—the first of its kind. Under the assumptions of no direct effect (NDE) and no unobserved confounding (NUC), we propose a semi-offline reinforcement learning framework that relaxes the conventional strong positivity assumption. Within this framework, we develop three novel estimators grounded in missing-data mechanisms: Direct Method (DM), Inverse Probability Weighting (IPW), and Doubly Robust Learning (DRL). We establish their consistency and asymptotic normality theoretically. Empirical evaluation demonstrates that our methods significantly improve estimation accuracy and robustness under time-varying feature availability, outperforming existing baselines across diverse benchmarks.
📝 Abstract
Machine learning methods often assume that input features are available at no cost. However, in domains like healthcare, where acquiring features could be expensive or harmful, it is necessary to balance a feature's acquisition cost against its predictive value. The task of training an AI agent to decide which features to acquire is called active feature acquisition (AFA). By deploying an AFA agent, we effectively alter the acquisition strategy and trigger a distribution shift. To safely deploy AFA agents under this distribution shift, we present the problem of active feature acquisition performance evaluation (AFAPE). We examine AFAPE under i) a no direct effect (NDE) assumption, stating that acquisitions do not affect the underlying feature values; and ii) a no unobserved confounding (NUC) assumption, stating that retrospective feature acquisition decisions were only based on observed features. We show that one can apply missing data methods under the NDE assumption and offline reinforcement learning under the NUC assumption. When NUC and NDE hold, we propose a novel semi-offline reinforcement learning framework. This framework requires a weaker positivity assumption and introduces three new estimators: A direct method (DM), an inverse probability weighting (IPW), and a double reinforcement learning (DRL) estimator.
Problem

Research questions and friction points this paper is trying to address.

Evaluating active feature acquisition under distribution shifts
Balancing feature cost and predictive value in healthcare
Proposing semi-offline reinforcement learning for AFA evaluation
Innovation

Methods, ideas, or system contributions that make the work stand out.

Active feature acquisition under distribution shift
Missing data methods with no direct effect
Semi-offline reinforcement learning framework
Helmholtz Munich | Technical University of Munich | Johns Hopkins University | Fraunhofer Institute for Cognitive Systems IKS
H
Henrik von Kleist
Institute of AI for Health, Helmholtz Munich - German Research Center for Environmental Health, Neuherberg, Germany; TUM School of Computation, Information and Technology, Technical University of Munich, Garching, Germany; Department of Computer Science, Johns Hopkins University Baltimore, Baltimore, MD, USA
Alireza Zamanian
Alireza Zamanian
TUM School of Computation, Information and Technology, Technical University of Munich, Garching, Germany; Fraunhofer Institute for Cognitive Systems IKS, Munich, Germany
I
I. Shpitser
Department of Computer Science, Johns Hopkins University Baltimore, Baltimore, MD, USA
N
N. Ahmidi
Institute of AI for Health, Helmholtz Munich - German Research Center for Environmental Health, Neuherberg, Germany; Department of Computer Science, Johns Hopkins University Baltimore, Baltimore, MD, USA; Fraunhofer Institute for Cognitive Systems IKS, Munich, Germany