Title: | VOCCluster: Untargeted Metabolomics Feature Clustering Approach for Clinical Breath Gas Chromatography/Mass Spectrometry Data |
Author(s): | Alkhalifah Y; Phillips I; Soltoggio A; Darnley K; Nailon WH; McLaren D; Eddleston M; Thomas CLP; Salman D; |
Address: | "Edinburgh Cancer Centre , NHS Lothian , Edinburgh EH4 2SP , U.K. Pharmacology, Toxicology and Therapeutics Unit , University of Edinburgh , Edinburgh EH8 9YL , U.K" |
DOI: | 10.1021/acs.analchem.9b03084 |
ISSN/ISBN: | 1520-6882 (Electronic) 0003-2700 (Linking) |
Abstract: | "Metabolic profiling of breath analysis involves processing, alignment, scaling, and clustering of thousands of features extracted from gas chromatography/mass spectrometry (GC/MS) data from hundreds of participants. The multistep data processing is complicated, operator error-prone, and time-consuming. Automated algorithmic clustering methods that are able to cluster features in a fast and reliable way are necessary. These accelerate metabolic profiling and discovery platforms for next-generation medical diagnostic tools. Our unsupervised clustering technique, VOCCluster, prototyped in Python, handles features of deconvolved GC/MS breath data. VOCCluster was created from a heuristic ontology based on the observation of experts undertaking data processing with a suite of software packages. VOCCluster identifies and clusters groups of volatile organic compounds (VOCs) from deconvolved GC/MS breath with similar mass spectra and retention index profiles. VOCCluster was used to cluster more than 15 000 features extracted from 74 GC/MS clinical breath samples obtained from participants with cancer before and after a radiation therapy. Results were evaluated against a panel of ground truth compounds and compared to other clustering methods (DBSCAN and OPTICS) that were used in previous metabolomics studies. VOCCluster was able to cluster those features into 1081 groups (including endogenous and exogenous compounds and instrumental artifacts) with an accuracy rate of 96% (+/-0.04 at 95% confidence interval)" |
Keywords: | Algorithms Breath Tests Cluster Analysis Gas Chromatography-Mass Spectrometry Humans *Metabolomics *Software Volatile Organic Compounds/analysis/*metabolism; |
Notes: | "MedlineAlkhalifah, Yaser Phillips, Iain Soltoggio, Andrea Darnley, Kareen Nailon, William H McLaren, Duncan Eddleston, Michael Thomas, C L Paul Salman, Dahlia eng Research Support, Non-U.S. Gov't 2019/12/04 Anal Chem. 2020 Feb 18; 92(4):2937-2945. doi: 10.1021/acs.analchem.9b03084. Epub 2020 Feb 5" |