计算机科学
计算机安全
软件部署
分类学(生物学)
脆弱性(计算)
欺骗攻击
稳健性(进化)
具身认知
威胁模型
组分(热力学)
生存能力
安全编码
弹性(材料科学)
脆弱性评估
概念框架
数据科学
人机交互
钥匙(锁)
感知
蜜罐
推论
作者
Wenpeng Xing,Minghao Li,Mohan Li,Meng Han
摘要
Embodied AI systems, integrating Large Vision-Language Models (LVLMs) and Large Language Models (LLMs) with physical actuators and sensors, face unique robustness and security challenges stemming from the complex interplay between perception, cognition, and actuation in real-world environments. This survey provides a systematic analysis of these vulnerabilities and associated attack surfaces. We propose a tripartite vulnerability taxonomy comprising foundational, integration, and contextual risks. Foundational vulnerabilities arise from inherent limitations in current AI architectures and training paradigms; Integration vulnerabilities emerge from the composition of cyber-physical components; And contextual vulnerabilities stem from dynamic physical environments and deployment conditions. Correspondingly, we present a comprehensive attack taxonomy that encompasses foundational attacks on LLMs/LVLMs (including logits-based, optimization-based, prompt-based, and cross-modality attacks), integration-level cybersecurity threats (such as man-in-the-middle, firmware, side-channel, and supply chain attacks), and contextual attacks (primarily sensor spoofing across multiple modalities). We further examine representative failure modes of the cognitive core, review existing evaluation methodologies and benchmarks, and synthesize a multi-layer defense framework that integrates perceptual redundancy, runtime monitoring, and hardware-enforced safety mechanisms. This work offers a unified conceptual framework and practical roadmap for designing and evaluating robust, secure Embodied AI systems in safety-critical real-world deployments.
科研通智能强力驱动
Strongly Powered by AbleSci AI