Abstract:Poisoning attacks are a primary threat to machine learning models, aiming to compromise their performance and reliability by manipulating training datasets. This paper introduces a novel attack - Outlier-Oriented Poisoning (OOP) attack, which manipulates labels of most distanced samples from the decision boundaries. The paper also investigates the adverse impact of such attacks on different machine learning algorithms within a multiclass classification scenario, analyzing their variance and correlation between different poisoning levels and performance degradation. To ascertain the severity of the OOP attack for different degrees (5% - 25%) of poisoning, we analyzed variance, accuracy, precision, recall, f1-score, and false positive rate for chosen ML <a class="link-external link-http" href="http://models.Benchmarking" rel="external noopener nofollow">this http URL</a> our OOP attack, we have analyzed key characteristics of multiclass machine learning algorithms and their sensitivity to poisoning attacks. Our experimentation used three publicly available datasets: IRIS, MNIST, and ISIC. Our analysis shows that KNN and GNB are the most affected algorithms with a decrease in accuracy of 22.81% and 56.07% while increasing false positive rate to 17.14% and 40.45% for IRIS dataset with 15% poisoning. Further, Decision Trees and Random Forest are the most resilient algorithms with the least accuracy disruption of 12.28% and 17.52% with 15% poisoning of the IRIS dataset. We have also analyzed the correlation between number of dataset classes and the performance degradation of models. Our analysis highlighted that number of classes are inversely proportional to the performance degradation, specifically the decrease in accuracy of the models, which is normalized with increasing number of classes. Further, our analysis identified that imbalanced dataset distribution can aggravate the impact of poisoning for machine learning models

Poisoning Attacks on Machine Learning Models in Cyber Systems and Mitigation Strategies

Poisoning Attacks and Data Sanitization Mitigations for Machine Learning Models in Network Intrusion Detection Systems

A Flexible Poisoning Attack Against Machine Learning.

MetaPoison: Practical General-purpose Clean-label Data Poisoning

Data Sanitization Approach to Mitigate Clean-Label Attacks Against Malware Detection Systems

Hidden Poison: Machine Unlearning Enables Camouflaged Poisoning Attacks

From Adversarial Examples to Data Poisoning Instances: Utilizing an Adversarial Attack Method to Poison a Transfer Learning Model

Stronger Data Poisoning Attacks Break Data Sanitization Defenses

How to Sift Out a Clean Data Subset in the Presence of Data Poisoning?

Data Poisoning Attacks on Regression Learning and Corresponding Defenses

Manipulating Machine Learning: Poisoning Attacks and Countermeasures for Regression Learning

Pick your Poison: Undetectability versus Robustness in Data Poisoning Attacks

Detection of Adversarial Training Examples in Poisoning Attacks through Anomaly Detection

Certified Defenses for Data Poisoning Attacks

Amplifying Membership Exposure via Data Poisoning

Poisoning Web-Scale Training Datasets is Practical

Poison Forensics: Traceback of Data Poisoning Attacks in Neural Networks

Outlier-Oriented Poisoning Attack: A Grey-box Approach to Disturb Decision Boundaries by Perturbing Outliers in Multiclass Learning

Poison as a Cure: Detecting & Neutralizing Variable-Sized Backdoor Attacks in Deep Neural Networks

Poison Attack and Defense on Deep Source Code Processing Models

Detecting Poisoning Attacks on Machine Learning in IoT Environments