Paper deep dive
Adversarial Attacks on Vision-Language Model-Empowered Chatbots in Consumer Electronics
Yingjia Shang, Zhijun Liu, Jiawen Kang, M. Shamim Hossain, Yi Wu
Models: ChatGPT, unspecified popular VLM chatbots (3 systems)
Intelligence
Status: succeeded | Model: google/gemini-3.1-flash-lite-preview | Prompt: intel-v1 | Confidence: 93%
Last extracted: 3/11/2026, 1:07:28 AM
Summary
This paper investigates the vulnerability of Vision-Language Model (VLM)-empowered chatbots in consumer electronics to adversarial attacks. The authors demonstrate that insufficient alignment of multi-modal data allows for effective adversarial manipulation, leading to the generation of harmful content, and emphasize the need for robust security measures.
Entities (4)
Relation Signals (3)
Adversarial Attacks → targets → Vision-Language Models
confidence 95% · This paper investigates and analyzes the susceptibility of VLMs-empowered chatbots to adversarial manipulation
Adversarial Attacks → exploits → Insufficient Alignment
confidence 90% · we designed three adversarial attacks, which all exploited the insufficient alignment of VLMs on multi-modal data
Vision-Language Models → usedin → Consumer Electronics
confidence 90% · with the increasing integration of vision-language models (VLMs, a representative of AIGC) in consumer electronics
Cypher Suggestions (0)
No Cypher suggestions yet.
Abstract
Artificial Intelligence-Generated Content (AIGC) technology has revolutionized content creation, distribution, and engagement in the consumer electronics sector, propelling its applications to unprecedented heights. Within this landscape, AIGC-driven conversational agents, exemplified by renowned chatbots like ChatGPT, have gained widespread popularity globally. These advanced conversational agents are instrumental in significantly enhancing user efficiency and overall experience within consumer electronics applications. However, with the increasing integration of vision-language models (VLMs, a representative of AIGC) in consumer electronics, the vulnerability of chatbots to adversarial attacks has become a critical concern. This paper investigates and analyzes the susceptibility of VLMs-empowered chatbots to adversarial manipulation, particularly in the context of consumer electronics applications. The study employs a comprehensive approach, combining vision and language modalities, to explore potential attack vectors and vulnerabilities. Specifically, we designed three adversarial attacks, which all exploited the insufficient alignment of VLMs on multi-modal data to implement effective attacks and make chatbots output harmful content. A series of experiments demonstrate the efficacy of adversarial attacks on three popular chatbot systems, revealing vulnerabilities that may compromise the reliability and security of these systems in real-world scenarios. The findings emphasize the importance of robust defenses against adversarial attacks in VLMs-driven chatbots, urging the development of enhanced security measures to safeguard users and prevent malicious exploitation of consumer electronics. Our data and code are available at <uri xmlns:mml="http://www.w3.org/1998/Math/MathML" xmlns:xlink="http://www.w3.org/1999/xlink">https://github.com/yxc0731/VLM-Adversarial-Attacks</uri>.
Tags
Links
Full Text
833 characters extracted from source content.
Expand or collapse full text
Adversarial Attacks on Vision-Language Model-Empowered Chatbots in Consumer Electronics | IEEE Journals & Magazine | IEEE Xplore IEEE Account Change Username/Password Update Address Purchase Details Payment Options Order History View Purchased Documents Profile Information Communications Preferences Profession and Education Technical Interests Need Help? US & Canada: +1 800 678 4333 Worldwide: +1 732 981 0060 Contact & Support About IEEE Xplore Contact Us Help Accessibility Terms of Use Nondiscrimination Policy Sitemap Privacy & Opting Out of Cookies A not-for-profit organization, IEEE is the world's largest technical professional organization dedicated to advancing technology for the benefit of humanity.© Copyright 2026 IEEE - All rights reserved. Use of this web site signifies your agreement to the terms and conditions.