AI-Predicted Full-Spectrum UV Correction Factors: An Analyte-Standard-Free Framework for Reaction Yield Quantification
Abstract High-throughput experimentation (HTE) accelerates molecular discovery but struggles to quantify yields without pure analyte standards. The current need to purify standards increases cost, complexity, and cycle time, limiting the scalability of new molecular discovery. Here, we introduce an integrated “detect-predict” framework that leverages artificial intelligence (AI) for analyte standard-free yield quantification. This approach hierarchically trains a prediction model of full-spectrum UV correction factor (CF) by combining quantum chemistry (QC) pre-training on 22,609 compounds for broad chemical space coverage with experimental fine-tuning on 1,845 LC-UV spectra to adapt to real instrumental conditions, thereby mitigating spectral discontinuities and data heterogeneity between computational and experimental domains. These predicted CFs enable direct quantification of analyte concentrations and reaction yields from UV absorbance data. We validated the accuracy of this framework through three independent test cases. First, concentration estimation was tested against 144 FDA compounds with known experimental concentrations. Second, yield quantification for 178 real-world reactions (spanning amide coupling, Suzuki–Miyaura coupling, SN2, and Buchwald-Hartwig transformations) was compared with calibration-curve-derived yields, achieving a mean absolute error of 4.7%. Finally, model-quantified yields were compared with isolated yields for 60 reactions as practical stress tests. The systematic positive bias aligned with the expected difference between crude analytical yield and post-purification isolated yield, enabling chemists to distinguish reaction performance from workup losses. This capability is critical for decision-making in real-world optimization. This AI-driven “detect-predict” framework offers an alternative to the traditional “detect-purify-quantify” workflow, providing a faster and more scalable foundation for HTE-accelerated molecular discovery.
Authors
- Sarah Trice
- Christopher J. Welch (ORCID: https://orcid.org/0000-0002-8899-4470)
- Jiuchuang Yuan (ORCID: https://orcid.org/0000-0001-5462-719X)
- Lin Zhang (ORCID: https://orcid.org/0000-0003-3825-4973)
- Jian Ma
- Jing Guo (ORCID: https://orcid.org/0000-0003-2106-3169)
- Yan Fan
- Mingjun Yang (ORCID: https://orcid.org/0000-0003-0928-8541)
- Shuhao Wen
- Xuekun Shi
- Guosheng Dou
- Qun Zeng
- Minjun Liu
- Liwen Fang
- Jun Yan
- Guanfeng Yang
- Lu Tan
- Peiyu Zhang
Institutions
- Chestnut Hill College (US)
- Welch Foundation (US)
- Fetal Medicine Foundation (GB)
Publication Details
- Journal
- JACS Au
- Published
- 2026-09-15
- DOI
- https://doi.org/10.1021/jacsau.6c00812
- Primary Topic
- Machine Learning in Materials Science
- Type
- article
- Field-Weighted Citation Impact
- 0.00
Funders
- XtalPi