{"slug": "process-matters-more-than-output-for-distinguishing-humans-from-machines", "title": "Process Matters More Than Output for Distinguishing Humans from Machines", "summary": "A process-based framework called the Process Turing Test distinguished humans from AI agents with a classifier AUC of 0.88 across cognitive tasks spanning decision-making, working memory, and planning, according to an arXiv paper (2605.06524v3) submitted 7 May 2026 and last revised 28 Sep 2026 by Milena Rmus and co-authors. In a red-teaming study, broad fine-tuning on 10.7M human decisions made agents' task processes more human-like than off-the-shelf frontier agents Claude Sonnet 4.5, GPT-5, and Gemini 2.5 Pro, and task-specific process-level fine-tuning (P-SFT) improved mimicry further, though that advantage largely disappeared under cross-task transfer. The authors conclude process specification is a central bottleneck to achieving human-like cognitive processes in machines.", "body_md": "# Computer Science > Artificial Intelligence\n\n  [Submitted on 7 May 2026 (\n\n[v1](https://arxiv.org/abs/2605.06524v1)), last revised 28 Sep 2026 (this version, v3)]\n# Title:Process Matters more than Output for Distinguishing Humans from Machines\n\n[View PDF](https://arxiv.org/pdf/2605.06524)\n\n[HTML (experimental)](https://arxiv.org/html/2605.06524v3)\n\nAbstract:Reliable human-machine discrimination is becoming increasingly important as Large Language Models and autonomous agents are deployed in online settings. Existing approaches evaluate whether a system can produce responses indistinguishable from those of a human. This approach follows the focus on the output of a machine, as suggested by Alan Turing. Cognitive science provides an alternative approach: considering the process by which that behavior is produced. To evaluate whether processes can reliably distinguish humans from machines, we introduce a process-based framework, the Process Turing Test, and evaluate it across a battery of cognitive tasks spanning decision-making, working memory, and planning. These tasks, such as mental rotation and sequence prediction, yield process-level measures complementing conventional measures of overall task performance. We also include multiple CAPTCHA tasks in the battery. Across the battery, process-level features provide substantially stronger discriminative signal than performance metrics alone, reliably distinguishing humans from agents even when task performance is matched (process-based classifier AUC = 0.88). We also conducted a controlled red-teaming study comparing off-the-shelf frontier agents (Claude Sonnet 4.5, GPT-5, Gemini 2.5 Pro), Centaur (LLM fine-tuned on 10.7M human decisions), and two task-specific fine-tuning methods: action-level supervised fine-tuning (A-SFT) and process-level fine-tuning (P-SFT), which directly optimizes process features. We find that broad fine-tuning on human choices makes task processes more human-like relative to off-the-shelf frontier agents, and task-specific P-SFT further improves human-like behavioral mimicry, though this advantage largely disappears under cross-task transfer. These results highlight process specification as a central bottleneck in achieving human-like cognitive processes in machines.\n    \n\n## Submission history\n\nFrom: Milena Rmus [\n[view email](https://arxiv.org/show-email/679e3e0a/2605.06524)]\n\n**Thu, 7 May 2026 16:30:35 UTC (1,215 KB)**\n\n[\\[v1\\]](https://arxiv.org/abs/2605.06524v1)\n**Sat, 9 May 2026 12:52:35 UTC (1,215 KB)**\n\n[\\[v2\\]](https://arxiv.org/abs/2605.06524v2)\n**[v3]** Mon, 28 Sep 2026 23:35:53 UTC (1,646 KB)\n\n### References & Citations\n\nLoading...\n\n# Bibliographic and Citation Tools\n\nBibliographic Explorer \n\n*(*[What is the Explorer?](https://info.arxiv.org/labs/showcase.html#arxiv-bibliographic-explorer))\nConnected Papers \n\n*(*[What is Connected Papers?](https://www.connectedpapers.com/about))\nLitmaps \n\n*(*[What is Litmaps?](https://www.litmaps.co/))\nscite Smart Citations \n\n*(*[What are Smart Citations?](https://www.scite.ai/))\n# Code, Data and Media Associated with this Article\n\nalphaXiv \n\n*(*[What is alphaXiv?](https://alphaxiv.org/))\nCatalyzeX Code Finder for Papers \n\n*(*[What is CatalyzeX?](https://www.catalyzex.com))\nDagsHub \n\n*(*[What is DagsHub?](https://dagshub.com/))\nGotit.pub \n\n*(*[What is GotitPub?](http://gotit.pub/faq))\nHugging Face \n\n*(*[What is Huggingface?](https://huggingface.co/huggingface))\nScienceCast \n\n*(*[What is ScienceCast?](https://sciencecast.org/welcome))\n# Demos\n\n# Recommenders and Search Tools\n\nInfluence Flower \n\n*(*[What are Influence Flowers?](https://influencemap.cmlab.dev/))\nCORE Recommender \n\n*(*[What is CORE?](https://core.ac.uk/services/recommender))\n# arXivLabs: experimental projects with community collaborators\n\narXivLabs is a framework that allows collaborators to develop and share new arXiv features directly on our website.\n\nBoth individuals and organizations that work with arXivLabs have embraced and accepted our values of openness, community, excellence, and user data privacy. arXiv is committed to these values and only works with partners that adhere to them.\n\nHave an idea for a project that will add value for arXiv's community? [**Learn more about arXivLabs**](https://info.arxiv.org/labs/index.html).", "url": "https://wpnews.pro/news/process-matters-more-than-output-for-distinguishing-humans-from-machines", "canonical_source": "https://arxiv.org/abs/2605.06524", "published_at": "2026-10-03 20:16:50+00:00", "updated_at": "2026-10-03 20:36:06.609245+00:00", "lang": "en", "topics": ["artificial-intelligence", "ai-research", "ai-safety", "large-language-models", "ai-agents"], "entities": ["Process Turing Test", "Milena Rmus", "Claude Sonnet 4.5", "GPT-5", "Gemini 2.5 Pro", "Centaur", "arXiv"], "also_reported_by": [], "alternates": {"html": "https://wpnews.pro/news/process-matters-more-than-output-for-distinguishing-humans-from-machines", "markdown": "https://wpnews.pro/news/process-matters-more-than-output-for-distinguishing-humans-from-machines.md", "text": "https://wpnews.pro/news/process-matters-more-than-output-for-distinguishing-humans-from-machines.txt", "jsonld": "https://wpnews.pro/news/process-matters-more-than-output-for-distinguishing-humans-from-machines.jsonld"}}