Abstract:Today, the security of many domains rely on the use of Machine Learning to detect threats, identify vulnerabilities, and safeguard systems from attacks. Recently, transformer architectures have improved the state-of-the-art performance on a wide range of tasks such as malware detection and network intrusion detection. But, before abandoning current approaches to transformers, it is crucial to understand their properties and implications on cybersecurity applications. In this paper, we evaluate the robustness of transformers to adversarial samples for system defenders (i.e., resiliency to adversarial perturbations generated on different types of architectures) and their adversarial strength for system attackers (i.e., transferability of adversarial samples generated by transformers to other target models). To that effect, we first fine-tune a set of pre-trained transformer, Convolutional Neural Network (CNN), and hybrid (an ensemble of transformer and CNN) models to solve different downstream image-based tasks. Then, we use an attack algorithm to craft 19,367 adversarial examples on each model for each task. The transferability of these adversarial examples is measured by evaluating each set on other models to determine which models offer more adversarial strength, and consequently, more robustness against these attacks. We find that the adversarial examples crafted on transformers offer the highest transferability rate (i.e., 25.7% higher than the average) onto other models. Similarly, adversarial examples crafted on other models have the lowest rate of transferability (i.e., 56.7% lower than the average) onto transformers. Our work emphasizes the importance of studying transformer architectures for attacking and defending models in security domains, and suggests using them as the primary architecture in transfer attack settings.

Dual stage black-box adversarial attack against vision transformer

Protego: Detecting Adversarial Examples for Vision Transformers Via Intrinsic Capabilities

Towards transferable adversarial attacks on vision transformers for image classification

On Improving Adversarial Transferability of Vision Transformers

Attacking Transformers with Feature Diversity Adversarial Perturbation

Towards Transferable Adversarial Attacks on Image and Video Transformers

Downstream Transfer Attack: Adversarial Attacks on Downstream Models with Pre-trained Vision Transformers

Query-Efficient Hard-Label Black-Box Attack against Vision Transformers

Enhancing the robustness of vision transformer defense against adversarial attacks based on squeeze-and-excitation module

Improving transferable adversarial attack for vision transformers via global attention and local drop

Transferable Adversarial Attacks on Vision Transformers with Token Gradient Regularization

Towards Efficient Adversarial Training on Vision Transformers

Improving the Transferability of Adversarial Examples with Restructure Embedded Patches

On the Adversarial Robustness of Vision Transformers

On the Robustness of Vision Transformers to Adversarial Examples

When Adversarial Training Meets Vision Transformers: Recipes from Training to Architecture

Transferable Adversarial Attack for Both Vision Transformers and Convolutional Networks Via Momentum Integrated Gradients

Reveal of Vision Transformers Robustness against Adversarial Attacks

Improving Robustness for Vision Transformer with a Simple Dynamic Scanning Augmentation

The Efficacy of Transformer-based Adversarial Attacks in Security Domains