AI Confidence: Why Precision Doesn't Always Equal Awareness
A study applying Signal Detection Theory to four AI models—Llama-3-8B-Instruct, Mistral-7B-Instruct-v0.3, Llama-3-8B-Base, and Gemma-2-9B-Instruct—across 224,000 factual QA trials found that models va…