Evasion Attacks Against Bayesian Predictive Models

Conference on Uncertainty in Artificial Intelligence (UAI), 2025

11 June 2025

Abstract

There is an increasing interest in analyzing the behavior of machine learning systems against adversarial attacks. However, most of the research in adversarial machine learning has focused on studying weaknesses against evasion or poisoning attacks to predictive models in classical setups, with the susceptibility of Bayesian predictive models to attacks remaining underexplored. This paper introduces a general methodology for designing optimal evasion attacks against such models. We investigate two adversarial objectives: perturbing specific point predictions and altering the entire posterior predictive distribution. For both scenarios, we propose novel gradient-based attacks and study their implementation and properties in various computational setups.

View on arXiv

Main:9 Pages

14 Figures

Bibliography:2 Pages

5 Tables

Appendix:8 Pages

Comments on this paper