I gave the same 18 coding tasks to three models this week: Z.ai’s brand-new GLM-5.2, OpenAI’s GPT-5.5, and the open-weight DeepSeek V4-Pro… Continue reading on Towards AI »
source & further reading
pub.towardsai.net — original article
LangGraph Workflows: Sequential & Parallel
Replace Guesswork with Statistics for Testing Your LLM Applications.
Evolution of LLMs