Align Is Not Enough: Multimodal Universal Jailbreak Attack Against Multimodal Large Language Models

计算机科学 模式 脆弱性(计算) 背景(考古学) 对抗制 人工智能 计算机安全 数据科学 社会科学 生物 社会学 古生物学
作者
Youze Wang,Wenbo Hu,Yinpeng Dong,Jing Liu,Hanwang Zhang,Richang Hong
出处
期刊:IEEE Transactions on Circuits and Systems for Video Technology [Institute of Electrical and Electronics Engineers]
卷期号:35 (6): 5475-5488
标识
DOI:10.1109/tcsvt.2025.3526248
摘要

Large Language Models (LLMs) have evolved into Multimodal Large Language Models (MLLMs), significantly enhancing their capabilities by integrating visual information and other types, thus aligning more closely with the nature of human intelligence, which processes a variety of data forms beyond just text. Despite advancements, the undesirable generation of these models remains a critical concern, particularly due to vulnerabilities exposed by text-based jailbreak attacks, which have represented a significant threat by challenging existing safety protocols. Motivated by the unique security risks posed by the integration of new and old modalities for MLLMs, we propose a unified multimodal universal jailbreak attack framework that leverages iterative image-text interactions and transfer-based strategy to generate a universal adversarial suffix and image. Our work not only highlights the interaction of image-text modalities can be used as a critical vulnerability but also validates that multimodal universal jailbreak attacks can bring higher-quality undesirable generations across different MLLMs. We evaluate the undesirable context generation of MLLMs like LLaVA, Yi-VL, MiniGPT4, MiniGPT-v2, and InstructBLIP, and reveal significant multimodal safety alignment issues, highlighting the inadequacy of current safety mechanisms against sophisticated multimodal attacks. This study underscores the urgent need for robust safety measures in MLLMs, advocating for a comprehensive review and enhancement of security protocols to mitigate potential risks associated with multimodal capabilities.
最长约 10秒,即可获得该文献文件

科研通智能强力驱动
Strongly Powered by AbleSci AI
科研通是完全免费的文献互助平台,具备全网最快的应助速度,最高的求助完成率。 对每一个文献求助,科研通都将尽心尽力,给求助人一个满意的交代。
实时播报
王11发布了新的文献求助10
刚刚
科研通AI6.2的应助被Hannah采纳,获得10
1秒前
科研通AI6.2的应助被祁忆采纳,获得10
1秒前
wang发布了新的文献求助10
1秒前
桐桐的应助被老的火龙果采纳,获得10
1秒前
晓君完成签到,获得积分10
2秒前
2秒前
薯条完成签到,获得积分10
4秒前
天天快乐的应助被落寞的楼房采纳,获得10
5秒前
赘婿的应助被李鹏采纳,获得10
5秒前
coolxixi完成签到,获得积分10
5秒前
6秒前
菠萝完成签到,获得积分10
6秒前
8秒前
陈景深发布了新的文献求助10
8秒前
脑洞疼的应助被Oscillator采纳,获得10
8秒前
响响完成签到,获得积分10
9秒前
科研通AI2S的应助被GRS采纳,获得10
9秒前
9秒前
10秒前
10秒前
10秒前
11秒前
科目三的应助被下里巴人采纳,获得10
11秒前
伶俐送终完成签到,获得积分20
12秒前
cc发布了新的文献求助10
13秒前
无私纹发布了新的文献求助10
14秒前
领导范儿的应助被北风采纳,获得10
14秒前
可爱的函函的应助被GGGGA采纳,获得10
15秒前
Endeavor发布了新的文献求助10
15秒前
15秒前
害羞的可燕完成签到,获得积分10
15秒前
OriC发布了新的文献求助10
16秒前
16秒前
Amireux发布了新的文献求助10
17秒前
yang完成签到,获得积分10
17秒前
DW的应助被澜生采纳,获得10
18秒前
18秒前
18秒前
开心的路人完成签到,获得积分20
18秒前
高分求助中
(应助此贴封号)通过应助OA文献获取积分 10000
Rosenblum, Global Change Biology 800
A Silent Apostrophe:The Fayum Portraits 520
Organizational Behavior 510
Sing with Understanding: Introduction to Theology in Christian Congregational Song, 3rd ed 330
Auslegung und Untersuchung einer invers ausgelegten Beschaufelung eines einstufigen Axialverdichters mit Vorleitrad (German) 300
AI-Contracting 300
热门求助领域 (近24小时)
化学 材料科学 医学 生物 计算机科学 工程类 纳米技术 有机化学 化学工程 内科学 物理 生物化学 复合材料 催化作用 细胞生物学 人工智能 心理学 无机化学 基因 遗传学
热门帖子
关注 科研通微信公众号,转发送积分 7838473
求助须知:如何正确求助?哪些是违规求助? 9360781
关于积分的说明 20617332
捐赠科研通 7432774
什么是DOI,文献DOI怎么找? 3339100
关于科研通互助平台的介绍 2483467
邀请新用户注册赠送积分活动 2360276