AI Safety in Practice: Enhancing Adversarial Robustness in Multimodal Image Captioning

Rashid, Maisha Binte; Rivas, Pablo

Computer Science > Computer Vision and Pattern Recognition

arXiv:2407.21174 (cs)

[Submitted on 30 Jul 2024]

Title:AI Safety in Practice: Enhancing Adversarial Robustness in Multimodal Image Captioning

Authors:Maisha Binte Rashid, Pablo Rivas

View PDF HTML (experimental)

Abstract:Multimodal machine learning models that combine visual and textual data are increasingly being deployed in critical applications, raising significant safety and security concerns due to their vulnerability to adversarial attacks. This paper presents an effective strategy to enhance the robustness of multimodal image captioning models against such attacks. By leveraging the Fast Gradient Sign Method (FGSM) to generate adversarial examples and incorporating adversarial training techniques, we demonstrate improved model robustness on two benchmark datasets: Flickr8k and COCO. Our findings indicate that selectively training only the text decoder of the multimodal architecture shows performance comparable to full adversarial training while offering increased computational efficiency. This targeted approach suggests a balance between robustness and training costs, facilitating the ethical deployment of multimodal AI systems across various domains.

Comments:	Accepted into KDD 2024 workshop on Ethical AI
Subjects:	Computer Vision and Pattern Recognition (cs.CV); Artificial Intelligence (cs.AI); Audio and Speech Processing (eess.AS)
ACM classes:	I.2.7
Cite as:	arXiv:2407.21174 [cs.CV]
	(or arXiv:2407.21174v1 [cs.CV] for this version)
	https://doi.org/10.48550/arXiv.2407.21174

Submission history

From: Pablo Rivas [view email]
[v1] Tue, 30 Jul 2024 20:28:31 UTC (2,349 KB)

Computer Science > Computer Vision and Pattern Recognition

Title:AI Safety in Practice: Enhancing Adversarial Robustness in Multimodal Image Captioning

Submission history

Access Paper:

References & Citations

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators

Computer Science > Computer Vision and Pattern Recognition

Title:AI Safety in Practice: Enhancing Adversarial Robustness in Multimodal Image Captioning

Submission history

Access Paper:

References & Citations

BibTeX formatted citation

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators