Pith. sign in

REVIEW 1 cited by

Metabolomics in the Cloud: Scaling Computational Tools to Big Data

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 1904.02288 v2 pith:B5MWD7TA submitted 2019-04-04 cs.DC

classification cs.DC
keywords cloudmetabolomicsplatformsdataphenomenaltoolscomputationalcomputing
verification ladder T0 review T1 audit T2 compute T3 formal
0 comments
read the original abstract

Background: Metabolomics datasets are becoming increasingly large and complex, with multiple types of algorithms and workflows needed to process and analyse the data. A cloud infrastructure with portable software tools can provide much needed resources enabling faster processing of much larger datasets than would be possible at any individual lab. The PhenoMeNal project has developed such an infrastructure, allowing users to run analyses on local or commercial cloud platforms. We have examined the computational scaling behaviour of the PhenoMeNal platform using four different implementations across 1-1000 virtual CPUs using two common metabolomics tools. Results: Our results show that data which takes up to 4 days to process on a standard desktop computer can be processed in just 10 min on the largest cluster. Improved runtimes come at the cost of decreased efficiency, with all platforms falling below 80% efficiency above approximately 1/3 of the maximum number of vCPUs. An economic analysis revealed that running on large scale cloud platforms is cost effective compared to traditional desktop systems. Conclusions: Overall, cloud implementations of PhenoMeNal show excellent scalability for standard metabolomics computing tasks on a range of platforms, making them a compelling choice for research computing in metabolomics.

Discussion (0). Continue with ORCID to comment.

Forward citations

Cited by 1 Pith paper

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score. Full citation record

  1. PyOD 2: A Python Library for Outlier Detection with LLM-powered Model Selection

    cs.LG 2024-12 conditional novelty 4.0 of 10

    PyOD 2 is a library update that unifies deep outlier detectors on PyTorch and uses GPT-4o to automatically select a model, reporting the best mean AUROC rank on 17 ADBench datasets.

Pith tools