Google is rolling out a feature that turns your phone's camera into a talking guide, describing objects, reading fine print, and narrating surroundings in real time. It arrives on Android through Gemini Live starting today, and it says less about one product update than about where phone-based AI is headed next: into the camera, as a default input.

Guided Vision works by letting users share their live camera feed with Gemini, Google's AI assistant, which then responds with spoken descriptions. Point your phone at a restaurant menu with tiny print, and it reads it aloud. Point it at a cluttered shelf, and it helps locate a specific item. Point it at an unfamiliar gadget, and it describes what it's looking at and how it might work.

The feature was built with accessibility in mind, particularly for users with low vision, and it closely mirrors a feature Apple shipped earlier this year called VoiceOver Live Recognition, available on iPhone and Vision Pro. Both companies are racing to make their AI assistants useful through the camera lens rather than just a text box, a shift that mirrors what OpenAI has also pushed with ChatGPT's vision capabilities and what Meta has baked into its camera-equipped Ray-Ban glasses.

This is not Google's first attempt at visual AI assistance. Google Lens has offered object and text recognition for years, and earlier versions of Gemini could already analyze a static photo. What's new here is the real-time, conversational layer: the AI keeps looking and talking as you move the camera, rather than requiring a snapshot and a prompt each time.

Accessibility features have a well-worn path to becoming mainstream. Predictive text, voice dictation, and screen readers all started as tools for users with disabilities before becoming default conveniences for everyone. Camera-based AI narration looks like it's following the same route, launching as an accessibility tool but built on infrastructure that will likely expand into general-purpose use, such as document scanning, product research, or on-the-spot translation.

For small business owners, the more immediate implications split into two categories: how customers might use this on your premises, and how you might use it yourself.

If customers start relying on AI camera tools to read your menus, contracts, price tags, or signage, it's worth checking whether your printed materials are actually legible and well-lit enough for a phone camera to parse. Low-contrast fonts, glossy lamination, or cramped fine print that a human squints at will likely trip up an AI reader too. This is a low-cost audit: walk your storefront or paperwork with a phone camera and see what holds up.

On the usage side, business owners who travel, source materials abroad, or deal with physical paperwork could use this for quick translation or document review. But it's worth remembering that sharing a live camera feed sends that visual data to Google's servers for processing. Before pointing a phone at a signed contract, a medical form, or anything covered by confidentiality obligations, consider whether that's a reasonable trade-off for your business and your clients.

Watch for two things in the coming months: whether Apple expands VoiceOver Live Recognition beyond its current scope to compete more directly, and whether Google eventually gates advanced Guided Vision features behind a paid Gemini subscription tier, as it has done with other AI capabilities. Also worth tracking is whether either company issues enterprise or business-specific terms for camera-sharing features, since current consumer terms may not address commercial document handling.

For now, Guided Vision is a free, opt-in feature on compatible Android devices, built primarily for accessibility but positioned to become a general-purpose tool. The practical step for business owners this week is simple: test it on your own signage and paperwork, and decide where you're comfortable letting AI read your documents, and where you're not.