跳到主要导航 跳到搜索 跳到主要内容

Safety of Multimodal Large Language Models on Images and Text

  • Xin Liu
  • , Yichen Zhu
  • , Yunshi Lan*
  • , Chao Yang*
  • , Yu Qiao
  • *此作品的通讯作者
  • East China Normal University
  • Shanghai AI Laboratory
  • Midea Group

科研成果: 书/报告/会议事项章节会议稿件同行评审

摘要

Attracted by the impressive power of Multimodal Large Language Models (MLLMs), the public is increasingly utilizing them to improve the efficiency of daily work. Nonetheless, the vulnerabilities of MLLMs to unsafe instructions bring huge safety risks when these models are deployed in real-world scenarios. In this paper, we systematically survey current efforts on the evaluation, attack, and defense of MLLMs' safety on images and text. We begin with introducing the overview of MLLMs on images and text and understanding of safety, which helps researchers know the detailed scope of our survey. Then, we review the evaluation datasets and metrics for measuring the safety of MLLMs. Next, we comprehensively present attack and defense techniques related to MLLMs' safety. Finally, we analyze several unsolved issues and discuss promising research directions. The relevant papers are collected at https://github.com/isXinLiu/Awesome-MLLM-Safety.

源语言英语
主期刊名Proceedings of the 33rd International Joint Conference on Artificial Intelligence, IJCAI 2024
编辑Kate Larson
出版商International Joint Conferences on Artificial Intelligence
8151-8159
页数9
ISBN(电子版)9781956792041
出版状态已出版 - 2024
活动33rd International Joint Conference on Artificial Intelligence, IJCAI 2024 - Jeju, 韩国
期限: 3 8月 20249 8月 2024

丛书

姓名IJCAI International Joint Conference on Artificial Intelligence
ISSN(印刷版)1045-0823

会议

会议33rd International Joint Conference on Artificial Intelligence, IJCAI 2024
国家/地区韩国
Jeju
时期3/08/249/08/24

学术指纹

探究 'Safety of Multimodal Large Language Models on Images and Text' 的科研主题。它们共同构成独一无二的学术指纹。

引用此