This is a submission for the Google I/O Writing Challenge




Right now, every AI agent that tries to use a website is basically doing this:


Take a screenshot
Guess what's on screen
Click something and hope
Take another screenshot
Repeat until it works or gives up


It's the digital equivalent of reading someone's lips through a frosted glass...