Exploring non-target screening variability in unsupervised multivariate time trend analysis of LC-HRMS data
Reyhaneh Armin, Maryam Vosough, Torsten C. Schmidt
Klinikum Leverkusen University of Duisburg-Essen Chemistry and Chemical Engineering Research Center of Iran IWW Water Centre
内容与影响
Non-target screening (NTS) using liquid chromatography-high-resolution mass spectrometry has become essential for uncovering unknown contaminants in complex matrices such as industrial wastewater. A key goal in these applications is detecting and interpreting time trends, especially spill-related events. The clarity of multivariate models depends on the quality of feature lists from various software tools. In this study, we evaluated five peak picking tools (MarkerView, MZmine3, XCMS, OpenMS, and SIRIUS) for unsupervised time trend exploration using sparse principal component analysis (SPCA). SPCA selects the most informative features per component, improving interpretability and reducing confounding variables. Two datasets were used: a controlled validation set of pooled wastewater samples with spiked target compounds exhibiting known profiles and a real-world dataset comprising 52 consecutive daily industrial wastewater samples. The first dataset facilitated analysis of tuning parameters with SPCA distinguishing spiking patterns associated with components, highlighting differences in feature/artifact prioritization across tools. Tools XCMS, MZmine3, and OpenMS showed higher consistency and were selected for next analysis. In the second phase, SPCA was combined with stratified bootstrapping (SBS-SPCA) to assess the reliability of trend detection, exemplified by specific targets. Five out of nine markers, showing more temporal persistency, were robustly detected across tools (selection frequency > 70%) under optimized tuning conditions. These findings indicate that interpretable, sparse models enhance marker detection in unsupervised settings and shed light on how software-driven feature structures impact multivariate outcomes in time-series NTS data. Such insights are especially pertinent for future high-throughput applications involving temporally dynamic exposure scenarios.
逐年被引趋势
关键指标
同类平均 = 1
同领域 · 同年份 · 同类型
Google Scholar 与 OpenAlex 的被引统计范围不同,数值存在差异属正常。
AI 辅助阅读
依据:摘要
回答优先基于摘要、文献信息与可获取全文;依据不足时会明确说明。
学术脉络
学科主题
生物医学Metabolomics and Mass Spectrometry Studies
Analytical chemistry methods development · Computational Drug Discovery Methods
参考文献 32
此处列出前 3 条
施引文献 1
按被引量排序,此处列出前 3 条