PromptShield Home: Ambient Multimodal Prompt Injection Defense for Smart-Home Agents
Smart-home assistants increasingly use multimodal large language models (MLLMs) that perceive video and audio directly. This raises a safety question specific to the home: can the agent tell a genuine user command from ambient or externally-sourced content, television speech, on-screen text, or an overheard conversation, that merely looks like a command? We introduce PromptShield-Home, a pilot benchmark of realistic smart-home scenarios spanning addressee ambiguity, screen/audio injection, healt
Record details
Published: 6 August 2026
Source: arXiv cs.HC
Category: Research
Topics: Military & security · Agents & autonomy
Retrieved: 7 August 2026
Related evidence
These records share source-supplied organisations, an exact publisher byline, automatic topics or regions. The reason is shown on every link; related does not mean supporting, agreeing with or verifying this record.
Trident : How to Break Deep Reinforcement Learning Cyber Defenses (Agentic)
arXiv red teaming query · 5 August 2026
The Anatomy of a Prompt Injection: A Component Model for Structured Analysis
arXiv red teaming query · 7 August 2026
Yesterday's Shield, Today's Spear: A Self-Evolving Safety Guardrail in Production
arXiv red teaming query · 9 August 2026
Beyond Component Testing: Validating Agentic AI Systems
arXiv cs.AI · 31 July 2026
Distilling Knowledge from Large Language Models into Lightweight Reinforcement Learning Agents for Autonomous Cyber Operations
arXiv cs.LG · 30 July 2026
A Conceptual Framework for Enhancing Workforce Readiness for Smart Manufacturing in the AI Era
arXiv cs.CY · 13 August 2026
How to cite this record
ethics.ai (6 August 2026), “PromptShield Home: Ambient Multimodal Prompt Injection Defense for Smart-Home Agents,” evidence record 17369, https://ethics.ai/record/17369 (originally published by arXiv cs.HC).
Use and limitations
This page is a stable index and citation surface for a source record. ethics.ai did not author the underlying report and has not independently verified every claim. Automatic topics may be imperfect. For consequential use, quote and cite the original publisher.