Cheaper LLM labelling
A developer used GPT 5.6 Luna via Simon Willison's `llm` CLI tool to label git commits as "maintenance" or "new development", reporting that the model matched their own manual labels across the entire…
A developer used GPT 5.6 Luna via Simon Willison's `llm` CLI tool to label git commits as "maintenance" or "new development", reporting that the model matched their own manual labels across the entire…
Simon Willison released a new AI code comment detector built on public data, replacing a previous version trained on private data that could not be shared. The tool runs entirely in the user's browser…
A developer criticized an AI-generated code comment from Anthropic's Claude for asserting that skipping a full collection walk is 'the usual case,' a claim the model cannot verify and that misleads fu…
GLM 5.2, a new open-weights model, achieved 15% fewer achievements than Gemini 3 Flash in text adventure games, a statistically significant difference. The benchmark, costing $5.1, controlled for game…