Google is taking a major step forward in AI-powered assistance with the rollout of Gemini’s real-time video and screen-reading features. A Google spokesperson confirmed to The Verge that these capabilities are now being introduced to Gemini Advanced subscribers as part of the Google One AI Premium plan.
Google has since confirmed that Gemini 3.0 will launch by the end of 2025.
The new features allow Gemini to analyze either your smartphone screen or live camera feed and provide instant responses. A Reddit user first spotted the screen-reading function on a Xiaomi phone, later sharing a demo video. Additionally, Google showcased live video interpretation in a recent demonstration, where a user asked Gemini for advice on choosing a paint color for pottery.
This advancement stems from “Project Astra,” a Google initiative first revealed nearly a year ago. The real-time AI processing sets Gemini apart from competitors, as Amazon’s Alexa Plus is still in early testing, and Apple has delayed its Siri upgrade. Even Samsung, despite having Bixby, has made Gemini the default assistant on its phones.
Google’s rapid deployment of these features highlights its lead in the AI assistant race. While rivals are still refining their offerings, Gemini is already delivering real-time visual analysis, making it a powerful tool for productivity, creativity, and everyday assistance.
As AI continues to evolve, Google’s early adoption of multimodal AI—combining text, voice, and visual inputs—positions Gemini as a frontrunner in the next generation of digital assistants.

Your First 10 AI Skills
10 practical AI skills, copy-paste prompts and a 7-day plan to start using AI with confidence.
Download the guide →
