Data Poisoning Attacks on Off-Policy Policy Evaluation Methods

Lobo, Elita; Singh, Harvineet; Petrik, Marek; Rudin, Cynthia; Lakkaraju, Himabindu

Computer Science > Machine Learning

arXiv:2404.04714 (cs)

[Submitted on 6 Apr 2024]

Title:Data Poisoning Attacks on Off-Policy Policy Evaluation Methods

Authors:Elita Lobo, Harvineet Singh, Marek Petrik, Cynthia Rudin, Himabindu Lakkaraju

View PDF HTML (experimental)

Abstract:Off-policy Evaluation (OPE) methods are a crucial tool for evaluating policies in high-stakes domains such as healthcare, where exploration is often infeasible, unethical, or expensive. However, the extent to which such methods can be trusted under adversarial threats to data quality is largely unexplored. In this work, we make the first attempt at investigating the sensitivity of OPE methods to marginal adversarial perturbations to the data. We design a generic data poisoning attack framework leveraging influence functions from robust statistics to carefully construct perturbations that maximize error in the policy value estimates. We carry out extensive experimentation with multiple healthcare and control datasets. Our results demonstrate that many existing OPE methods are highly prone to generating value estimates with large errors when subject to data poisoning attacks, even for small adversarial perturbations. These findings question the reliability of policy values derived using OPE methods and motivate the need for developing OPE methods that are statistically robust to train-time data poisoning attacks.

Comments:	Accepted at UAI 2022
Subjects:	Machine Learning (cs.LG); Artificial Intelligence (cs.AI); Cryptography and Security (cs.CR)
Cite as:	arXiv:2404.04714 [cs.LG]
	(or arXiv:2404.04714v1 [cs.LG] for this version)
	https://doi.org/10.48550/arXiv.2404.04714

Submission history

From: Elita Lobo [view email]
[v1] Sat, 6 Apr 2024 19:27:57 UTC (11,547 KB)

Computer Science > Machine Learning

Title:Data Poisoning Attacks on Off-Policy Policy Evaluation Methods

Submission history

Access Paper:

References & Citations

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators

Computer Science > Machine Learning

Title:Data Poisoning Attacks on Off-Policy Policy Evaluation Methods

Submission history

Access Paper:

References & Citations

BibTeX formatted citation

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators