☆ 4.6 Article

Systematic Poisoning Attacks on and Defenses for Machine Learning in Healthcare

IEEE JOURNAL OF BIOMEDICAL AND HEALTH INFORMATICS (2015)

期刊

IEEE JOURNAL OF BIOMEDICAL AND HEALTH INFORMATICS

卷 19, 期 6, 页码 1893-1905

出版社

IEEE-INST ELECTRICAL ELECTRONICS ENGINEERS INC

DOI: 10.1109/JBHI.2014.2344095

关键词

Healthcare; machine learning; poisoning attacks; security

类别

Computer Science, Information Systems Computer Science, Interdisciplinary Applications Mathematical & Computational Biology Medical Informatics

资金

National Science Foundation [CNS-1219570]
Direct For Computer & Info Scie & Enginr
Division Of Computer and Network Systems [1219587, 1219570] Funding Source: National Science Foundation

向作者/读者索取更多资源

Protocol

社区支持

Reagent

社区支持

摘要

Machine learning is being used in a wide range of application domains to discover patterns in large datasets. Increasingly, the results of machine learning drive critical decisions in applications related to healthcare and biomedicine. Such health-related applications are often sensitive, and thus, any security breach would be catastrophic. Naturally, the integrity of the results computed by machine learning is of great importance. Recent research has shown that some machine-learning algorithms can be compromised by augmenting their training datasets with malicious data, leading to a new class of attacks called poisoning attacks. Hindrance of a diagnosis may have life-threatening consequences and could cause distrust. On the other hand, not only may a false diagnosis prompt users to distrust the machine-learning algorithm and even abandon the entire system but also such a false positive classification may cause patient distress. In this paper, we present a systematic, algorithm-independent approach for mounting poisoning attacks across a wide range of machine-learning algorithms and healthcare datasets. The proposed attack procedure generates input data, which, when added to the training set, can either cause the results of machine learning to have targeted errors (e.g., increase the likelihood of classification into a specific class), or simply introduce arbitrary errors ( incorrect classification). These attacks may be applied to both fixed and evolving datasets. They can be applied even when only statistics of the training dataset are available or, in some cases, even without access to the training dataset, although at a lower efficacy. We establish the effectiveness of the proposed attacks using a suite of six machine-learning algorithms and five healthcare datasets. Finally, we present countermeasures against the proposed generic attacks that are based on tracking and detecting deviations in various accuracy metrics, and benchmark their effectiveness.

Systematic Poisoning Attacks on and Defenses for Machine Learning in Healthcare

期刊

IEEE JOURNAL OF BIOMEDICAL AND HEALTH INFORMATICS

出版社

IEEE-INST ELECTRICAL ELECTRONICS ENGINEERS INC

关键词

类别

资金

向作者/读者索取更多资源

Protocol

Reagent

作者

我是这篇论文的作者

评论

主要评分

次要评分

新颖性

重要性

科学严谨性

评价这篇论文

推荐

Systematic Poisoning Attacks on and Defenses for Machine Learning in Healthcare

期刊

IEEE JOURNAL OF BIOMEDICAL AND HEALTH INFORMATICS

出版社

IEEE-INST ELECTRICAL ELECTRONICS ENGINEERS INC

关键词

类别

资金

向作者/读者索取更多资源

Protocol

Reagent

作者

我是这篇论文的作者

评论

主要评分

次要评分

新颖性

重要性

科学严谨性

评价这篇论文

推荐

导出引文

分享论文