Overview
How might we design voice visualizer so people can trust and act on AI output?
When to use
- Essential for voice assistants, voice-first applications, and hands-free interfaces where visual feedback enhances user understanding of voice interaction states.
When to skip
- Pure text chat with no voice path.
- Accessibility settings request reduced motion.
- Background voice where visual feedback is irrelevant.
Rules
Same animation for listening and speaking, so state is ambiguous.
Visualizer with no caption or icon for deaf or low-vision users.
Flashy motion that distracts from transcript content.
No idle state when mic is off but UI looks “live”.
Evidence
| Product | Implementation |
|---|---|
| Siri | Orb and waveform states on iPhone and HomePod screens. |
| Google Assistant | Animated dots and color shifts for listen vs respond. |
| Alexa | Light ring and screen pulse tied to audio activity. |
| ChatGPT | Voice mode orb with listening and speaking animations. |

