BACKGROUND: Gas chromatography-mass spectrometry (GC-MS) is a technique frequently used in targeted and non-targeted measurements of metabolites. Most existing software tools for processing of raw instrument GC-MS data tightly integrate data processing methods with graphical user interface facilitating interactive data processing. While interactive processing remains critically important in GC-MS applications, high-throughput studies increasingly dictate the need for command line tools, suitable for scripting of high-throughput, customized processing pipelines. RESULTS: PyMS comprises a library of functions for processing of instrument GC-MS data developed in Python. PyMS currently provides a complete set of GC-MS processing functions, including reading of standard data formats (ANDI- MS/NetCDF and JCAMP-DX), noise smoothing, baseline correction, peak detection, peak deconvolution, peak integration, and peak alignment by dynamic programming. A novel common ion single quantitation algorithm allows automated, accurate quantitation of GC-MS electron impact (EI) fragmentation spectra when a large number of experiments are being analyzed. PyMS implements parallel processing for by-row and by-column data processing tasks based on Message Passing Interface (MPI), allowing processing to scale on multiple CPUs in distributed computing environments. A set of specifically designed experiments was performed in-house and used to comparatively evaluate the performance of PyMS and three widely used software packages for GC-MS data processing (AMDIS, AnalyzerPro, and XCMS). CONCLUSIONS: PyMS is a novel software package for the processing of raw GC-MS data, particularly suitable for scripting of customized processing pipelines and for data processing in batch mode. PyMS provides limited graphical capabilities and can be used both for routine data processing and interactive/exploratory data analysis. In real-life GC-MS data processing scenarios PyMS performs as well or better than leading software packages. We demonstrate data processing scenarios simple to implement in PyMS, yet difficult to achieve with many conventional GC-MS data processing software. Automated sample processing and quantitation with PyMS can provide substantial time savings compared to more traditional interactive software systems that tightly integrate data processing with the graphical user interface.
BACKGROUND: Gas chromatography-mass spectrometry (GC-MS) is a technique frequently used in targeted and non-targeted measurements of metabolites. Most existing software tools for processing of raw instrument GC-MS data tightly integrate data processing methods with graphical user interface facilitating interactive data processing. While interactive processing remains critically important in GC-MS applications, high-throughput studies increasingly dictate the need for command line tools, suitable for scripting of high-throughput, customized processing pipelines. RESULTS: PyMS comprises a library of functions for processing of instrument GC-MS data developed in Python. PyMS currently provides a complete set of GC-MS processing functions, including reading of standard data formats (ANDI- MS/NetCDF and JCAMP-DX), noise smoothing, baseline correction, peak detection, peak deconvolution, peak integration, and peak alignment by dynamic programming. A novel common ion single quantitation algorithm allows automated, accurate quantitation of GC-MS electron impact (EI) fragmentation spectra when a large number of experiments are being analyzed. PyMS implements parallel processing for by-row and by-column data processing tasks based on Message Passing Interface (MPI), allowing processing to scale on multiple CPUs in distributed computing environments. A set of specifically designed experiments was performed in-house and used to comparatively evaluate the performance of PyMS and three widely used software packages for GC-MS data processing (AMDIS, AnalyzerPro, and XCMS). CONCLUSIONS: PyMS is a novel software package for the processing of raw GC-MS data, particularly suitable for scripting of customized processing pipelines and for data processing in batch mode. PyMS provides limited graphical capabilities and can be used both for routine data processing and interactive/exploratory data analysis. In real-life GC-MS data processing scenarios PyMS performs as well or better than leading software packages. We demonstrate data processing scenarios simple to implement in PyMS, yet difficult to achieve with many conventional GC-MS data processing software. Automated sample processing and quantitation with PyMS can provide substantial time savings compared to more traditional interactive software systems that tightly integrate data processing with the graphical user interface.
Authors: Georg Weingart; Bernhard Kluger; Astrid Forneck; Rudolf Krska; Rainer Schuhmacher Journal: Phytochem Anal Date: 2011-10-18 Impact factor: 3.373
Authors: John M Halket; Daniel Waterman; Anna M Przyborowska; Raj K P Patel; Paul D Fraser; Peter M Bramley Journal: J Exp Bot Date: 2004-12-23 Impact factor: 6.992
Authors: Diana I Serrazanetti; Maurice Ndagijimana; Sylvain L Sado-Kamdem; Aldo Corsetti; Rudi F Vogel; Matthias Ehrmann; M Elisabetta Guerzoni Journal: Appl Environ Microbiol Date: 2011-02-18 Impact factor: 4.792
Authors: David S Wishart; Dan Tzur; Craig Knox; Roman Eisner; An Chi Guo; Nelson Young; Dean Cheng; Kevin Jewell; David Arndt; Summit Sawhney; Chris Fung; Lisa Nikolai; Mike Lewis; Marie-Aude Coutouly; Ian Forsythe; Peter Tang; Savita Shrivastava; Kevin Jeroncic; Paul Stothard; Godwin Amegbey; David Block; David D Hau; James Wagner; Jessica Miniaci; Melisa Clements; Mulu Gebremedhin; Natalie Guo; Ying Zhang; Gavin E Duggan; Glen D Macinnis; Alim M Weljie; Reza Dowlatabadi; Fiona Bamforth; Derrick Clive; Russ Greiner; Liang Li; Tom Marrie; Brian D Sykes; Hans J Vogel; Lori Querengesser Journal: Nucleic Acids Res Date: 2007-01 Impact factor: 16.971
Authors: David J Beale; Farhana R Pinu; Konstantinos A Kouremenos; Mahesha M Poojary; Vinod K Narayana; Berin A Boughton; Komal Kanojia; Saravanan Dayalan; Oliver A H Jones; Daniel A Dias Journal: Metabolomics Date: 2018-11-17 Impact factor: 4.290
Authors: Ivan A Titaley; O Maduka Ogba; Leah Chibwe; Eunha Hoh; Paul H-Y Cheong; Staci L Massey Simonich Journal: J Chromatogr A Date: 2018-02-07 Impact factor: 4.759
Authors: Katherine J Jeppe; Konstantinos A Kouremenos; Kallie R Townsend; Daniel F MacMahon; David Sharley; Dedreia L Tull; Ary A Hoffmann; Vincent Pettigrove; Sara M Long Journal: Metabolites Date: 2017-12-18
Authors: Christoph Steinbeck; Pablo Conesa; Kenneth Haug; Tejasvi Mahendraker; Mark Williams; Eamonn Maguire; Philippe Rocca-Serra; Susanna-Assunta Sansone; Reza M Salek; Julian L Griffin Journal: Metabolomics Date: 2012-09-25 Impact factor: 4.290
Authors: Anne Julie Overgaard; Jacquelyn M Weir; David Peter De Souza; Dedreia Tull; Claus Haase; Peter J Meikle; Flemming Pociot Journal: Metabolomics Date: 2015-11-17 Impact factor: 4.290