For thirty years, the web's been built for humans with eyeballs and impatience. The next ten years it gets used by something quieter: agents acting on behalf of those humans. Same browsers, completely different failure modes — and almost nobody is measuring.
You can't fix what you can't see, and you can't compare what you don't measure the same way every time. We think there should be one number — comparable, public, reproducible — for whether a site is going to work when an agent shows up.
Scan once. Tell us what we missed. Help us shape the rule set everyone ends up using.