AI-Search productization + data-quality sweep
Corpus packaged as v4.0.0: 617,912 unique graded records (19 true-duplicate narrations removed) across 128 works, with narrator records deduplicated to 27,118 (180 duplicate-PID merges; 33 homonym-risk groups held for review). All 13 tiered Azure AI Search indexes rebuilt to v4 (public / research / scholar) and served live, with canonical matn-parallels published as a dedicated clusters index (8,170 clusters, all ≥2 members). v4 is a packaging + data-quality release on the v3.176 data lineage; scholar-grade citation readiness remains in progress.
- 617,912 unique graded records (19 duplicate narrations removed); narrators deduplicated to 27,118 (180 duplicate-PID merges, 33 homonym-risk groups held)
- All 13 tiered Azure AI Search indexes rebuilt to v4 (public / research / scholar) and served live
- Canonical matn-parallels as a dedicated clusters index — 8,170 clusters, all ≥2 members
- Public Hadith Explorer: detail view with grade, attestation, and cross-collection parallels