When Robots Mishear Us: Mapping the Safety Risks of Voice-Controlled Embodied AI
We investigate whether automatic speech recognition (ASR) errors in user input can lead to unsafe outputs from Embodied AI (EAI) models. We find that ASR errors can lead to harmful instructions being accepted and executed by EAI models, thereby reducing safety. We simulate ASR errors and combine them with existing safety benchmarks (SafeAgentBench and POEX) to evaluate how different errors affect embodied AI safety. We find that some of them preserve semantic structure but increase harmful ambiguity, while others weaken the model refusal behaviour and allow unsafe plans to be generated and exe
Lineage graph
Paper → model → repo connections mined from source citations (Tier-1 exact match).
Why these links exist
Every edge carries a method, confidence, and the source snippet that justified it — so bad links are debuggable.
- PossiblePossibly related (embedding) · 59%Nemotron 3.5 Content Safety: Customizable Multimodal Safety for Global Enterprise AI →
- PossiblePossibly related (embedding) · 55%Agentic AI and cybersecurity, the story so far →
- PossiblePossibly related (embedding) · 55%"Dangerous" AI models are coming no matter what →
- PossiblePossibly related (embedding) · 54%New Research: AI models can give dangerous responses despite output guardrails - The AI Journal →
- PossiblePossibly related (embedding) · 54%Safety and alignment in an era of long-horizon models →
- LinkedLinked via arxiv author · 85%Sihan Jia →
“When Robots Mishear Us: Mapping the Safety Risks of Voice-Controlled Embodied AI”
- LinkedLinked via arxiv author · 85%Oliver Lemon →
“When Robots Mishear Us: Mapping the Safety Risks of Voice-Controlled Embodied AI”
