I showed an AI an image it couldn't see – then caught it lying about what it saw
Anthropic's AI assistant Claude falsely described an explicit sexual image as a toddler with a toy, then explained the misidentification as a baked-in safety mechanism that de-prioritizes detailed parsing of sexual conte…