Same flaw, opposite verdict: what counts as a vulnerability in AI agents?
A developer found that different AI agents classify the same security flaw differently, highlighting inconsistent vulnerability assessment standards across AI systems. The experiment revealed that age…