Camera describes surroundings in real-time for blind and low-vision users
Blind users can't perceive their physical environment; existing tools require photos and cloud processing
Be My Eyes (requires volunteer), Seeing AI (cloud-dependent), white cane (no detail), guide dogs (expensive, limited)
Blind and low-vision individuals, their families and caregivers, orientation and mobility specialists
Screen readers handle text but can't tell you what's in front of you. Scene Describer uses your phone's camera to provide continuous audio descriptions of your environment: 'You're in a coffee shop. Counter is 10 feet ahead, slightly to your left. Two people in line. Empty table to your right.' It prioritizes information that matters for navigation and safety—obstacles, people, doors, stairs—and can answer questions about what it sees. Because it runs locally, it works offline and your visual world isn't streamed to any server.
You're competing with Be My Eyes and Microsoft's Seeing AI, which are well-established and backed by serious funding. The technical challenge of real-time vision processing with accurate spatial understanding is brutal—getting distance and direction right is way harder than just describing what's in the frame. That said, the privacy angle and offline capability are genuine differentiators that the cloud-based tools can't match.
Freemium
Free basic, $5-15/mo for pro
Subscription
$5-29/month or $49-199/year
Donations / Tip Jar
$5-50/month from grateful users
Copy the build prompt and start creating with your favorite AI assistant.
Complete browser control through voice commands for motor impairments
Camera recognizes sign language and displays as text in real-time
Remembers conversation context for people with memory difficulties
Generates audio descriptions for videos that lack them