Hadith computational science in the age of large language models: a critical narrative review A critical narrative review from arXiv (arXiv:2608.20364v1) finds uneven progress in hadith computational science, with advances in data resources, segmentation tasks, and LLM-assisted workflows, but persistent gaps in corpus breadth, benchmark comparability, and expert validation. The authors argue the field should be assessed as an evidence infrastructure problem requiring knowledge integration, provenance, and expert supervision, and propose a research agenda to strengthen methodology and utility for Islamic scholarship. arXiv:2608.20364v1 Announce Type: new Abstract: We examine how hadith computational science is being reshaped by transformer models, retrieval-grounded pipelines, and large language models LLMs . Recent reviews document growth in the literature, but they do not yet provide a critical account of which advances are methodologically robust, which remain benchmark-bound, and which unresolved problems still limit scholarly use. We address this gap through a critical narrative review that combines critique of existing reviews, paper-level appraisal of representative original studies, and synthesis of Islamic scholar and domain-expert perspectives on authenticity, authority, and responsible use. We find uneven progress. Data resources have expanded, segmentation tasks have matured, narrator and source-verification problems are better formalized, and LLM-assisted workflows now support corpus-scale enrichment, multilingual access, and grounded evaluation. At the same time, progress remains constrained by narrow corpora, weak benchmark comparability, synthetic-to-real transfer gaps, narrator identity resolution, preprocessing fragility, limited reproducibility, and sparse expert-grounded validation. We show that important gaps lie beyond dominant benchmarks: non-canonical and obscure corpora, commentary and explanatory literature, cross-source links with Qur'an and seerah, and fiqh-facing evidence support. We argue that hadith computation should be assessed less as isolated model performance than as an evidence infrastructure problem requiring knowledge integration, provenance, and expert supervision. On this basis, we define a research agenda for making the field methodologically stronger and more useful to Islamic scholarship.