Randy Olson on X: "Question 2 from @ArtificialAnlys data: how often do top AI models get specific facts wrong without search? Even the best get about 1 in 4 answers wrong. If a fact matters, have the model look it up. Made with evident-charts, my open source agent skill for explanatory charts."
Question 2 from @ArtificialAnlys data: how often do top AI models get specific facts wrong without search? Even the best get about 1 in 4 answers wrong. If a fact matters, have the model look it up. Made with evident-charts, my open source agent skill for explanatory charts.
Question 2 from @ArtificialAnlys data: how often do top AI models get specific facts wrong without search? Even the best get about 1 in 4 answers wrong. If a fact matters, have the model look it up. Made with evident-charts, my open source agent skill for explanatory charts.
Data: AA-Omniscience, 6,000 questions about specific facts across 42 work topics, answered without search. Wrong answers as a share of answers given; declined questions don't count. Get the evident-charts agent skill here: