Null-Model Tests of Structural Claims About the Qur'anic Text: Ring Composition and Major Derivation, with an Audit of Machine-Mined Links
Computational studies of the Qur’anic text increasingly report structural regularities, yet few test them against an explicit null model. We test two widely repeated claims under a protocol that was written and hashed before analysis, on a public root-annotated corpus (6,236 verses, 1,651 roots), with multiple-comparison control and power analysis. (1) Ring composition, operationalized lexically: mirrored verses of a surah share more roots than chance. Across all 95 surahs with at least ten verses, 5 reach raw p < 0.05 under a permutation null (4.75 expected by chance) and none survives Holm or Benjamini–Hochberg correction; the corpus-level Stouffer z is −2.12. Because mirrored pairs near the centre are adjacent verses, we add a distance-matched null; the corpus-level z is −0.21, and the single surah that passes correction does so on one shared root in one verse pair with a zero-variance null. A planted ring in which each mirrored verse carries 20% of its partner’s roots is detected in 92% of surahs, so the test is not blind to rings of that size. (2) Major derivation (al-ishtiqāq al-akbar): roots sharing two letters in the same positions, or permutations of the same three letters, are distributionally closer than random root pairs. Kinship pairs: mean cosine 0.1910 vs. 0.1883 ± 0.0019 (3,042 pairs, z = 1.40, p = 0.081; frequency-matched p = 0.117), with a minimal detectable effect of 2.5% of the null mean. Permutation pairs: 0.1848 vs. 0.1883 ± 0.0073 (208 pairs, z = −0.49, p = 0.685). Both results agree in direction and significance with an earlier measurement on the author’s own verified corpus. (3) We also report an audit of an automated pipeline that produced 331,442 “discoveries”, 201,845 links and 1,221 rules: no rule specified a falsifier, and of 146,611 candidate pairs 587 (0.40%) survived a permutation null. We release the protocol and code so that structural claims about the text can be tested before they are asserted. Bilingual edition: the English paper is followed by the full Arabic edition in the same file. الملخّص: تكثر في الدراسات الحاسوبيّة للنصّ القرآنيّ دعاوى عن انتظامات بنيويّة، وقلّما تُمتحن أمام نموذج عدم صريح. نمتحن دعويين شائعتين ببروتوكول كُتب وخُتم ببصمة رقميّة قبل التحليل، على مدوّنة عامّة مجذّرة، مع تصحيح المقارنات المتعدّدة وتحليل القوّة الإحصائيّة. الأولى نظم الحلقة بصيغته المعجميّة: أنّ الآيات المتقابلة في السورة تشترك في الجذور أكثر من الصدفة؛ فلم تصمد سورة واحدة بعد تصحيح المقارنات المتعدّدة، وحلقة مزروعة صناعيّاً تُكتشف في أكثر السور، فالاختبار ليس أعمى عن حلقات بذلك الحجم. والثانية الاشتقاق الأكبر: أنّ الجذور المشتركة في حرفين في الموضعين نفسيهما، أو تقليبات الأحرف الثلاثة نفسها، أقرب توزيعيّاً من أزواج جذور عشوائيّة؛ فلم تتجاوز الصدفة في الصورتين، والنتيجتان توافقان قياساً سابقاً للمؤلّف على مدوّنته الموثّقة. ونعرض كذلك تدقيقاً لمنظومة آليّة أنتجت مئات الآلاف من «الاكتشافات» والروابط: لم تحدّد أيّ قاعدة ما يُبطلها، ولم يصمد أمام نموذج التبديل إلّا أقلّ من نصف بالمئة من الأزواج المرشّحة. وننشر البروتوكول والشيفرة لكي تُمتحن الدعاوى البنيويّة عن النصّ قبل أن تُقال.
Authors
- Firas Assaf
Publication Details
- Journal
- Zenodo (CERN European Organization for Nuclear Research)
- Published
- 2026-09-28
- DOI
- https://doi.org/10.5281/zenodo.23010536
- Primary Topic
- Text and Document Classification Technologies
- Type
- preprint