...invocation. When an agent actually calls a tool (not just detects it), does it get back usable, structured data + clean errors? A green scan = 'discoverable,' not 'works when called.' On a fix list, a callable tool with a clean return beats another badge. (I build a WebMCP layer, so biased.)