πŸ“° Key Takeaway

63% of Gemini users talk to the assistant directly through voice, showing voice has become the mainstream way to interact; Google also revealed Gemini generates over 150 million images daily. See the original article for details.


πŸ’¬ JudyAI Lab Take

63% of Gemini users talk to the assistant directly through voice β€” that’s the mainstream interaction mode now, not some side option tacked onto text input. That number alone is worth every AI builder’s attention.

It signals a shift in design thinking: when most users naturally reach for voice instead of typing, you can’t treat voice as a “nice-to-have” anymore β€” it’s one of the core interaction paths. At the same time, Google revealed that Gemini generates over 150 million images per day. Put those two data points together and you see user-AI interaction moving fast, from pure text conversation to a mode where multimodal input and output run in parallel. For AI builders, this means the interface assumptions baked into your product (users will type, users will read text replies) might already be outdated. Voice-first, image-first experience design is going to be the next competitive bar, not an optional add-on.

Next time you’re designing the interaction flow for an AI product, ask yourself first: if the user never typed a single word, would this feature still work?


πŸ“… Source Info


πŸ”— Further Reading