cd /news/large-language-models/do-not-ask-an-llm-whether-it-is-a-go… · home topics large-language-models article
[ARTICLE · art-116294] src= pub= topic=large-language-models verified=true sentiment=· neutral

Do not ask an LLM whether it is a good boy

A commentary argues that having an LLM write its own tests is unreliable, comparing it to asking a dog whether it has been a good boy, as the model lacks a clear definition of correctness and may generate self-serving or circular evaluations.

read1 min views1 publishedAug 30, 2026

Having an LLM write its own tests is like asking a dog whether it’s been a good boy. What does it mean to be a good boy? All your dog knows is that it yearns to say, yes, I have been a good boy.

If asked to formalize this, your dog might come back with: Then, it will devise dozens of permutations of sitting, over and over again, unsure which aspect of sitting it is which upon crossing its threshold one has finally become Good but always looking up with doe eyes full of treats. “Yes,” it beckons you say.

── more in #large-language-models 4 stories · sorted by recency
sponsored brought to you by zahid.host 4,200+ EU-deployed projects
reading about agents? ship yours in a single git push.

Run your AI side-project on zahid.host

EU-based hosting, git-push deploys, automatic HTTPS, no cold starts. Free tier with a custom domain — perfect for shipping the agent you just read about.

$git push zahid main
Live at https://your-agent.zahid.host
Get free account → Pricing
from €0/mo · no card required
LIVE [news/do-not-ask-an-llm-wh…] indexed:0 read:1min 2026-08-30 ·