Pith. sign in

REVIEW 1 cited by

Learn molecular representations from large-scale unlabeled molecules for drug discovery

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 2012.11175 v1 pith:2DPT4IAF submitted 2020-12-21 cs.LG q-bio.BMq-bio.QM

classification cs.LGq-bio.BMq-bio.QM
keywords moleculardiscoverydrugmoleculesmolgnetpre-trainingrepresentationsunlabeled
verification ladder T0 review T1 audit T2 compute T3 formal
0 comments
read the original abstract

How to produce expressive molecular representations is a fundamental challenge in AI-driven drug discovery. Graph neural network (GNN) has emerged as a powerful technique for modeling molecular data. However, previous supervised approaches usually suffer from the scarcity of labeled data and have poor generalization capability. Here, we proposed a novel Molecular Pre-training Graph-based deep learning framework, named MPG, that leans molecular representations from large-scale unlabeled molecules. In MPG, we proposed a powerful MolGNet model and an effective self-supervised strategy for pre-training the model at both the node and graph-level. After pre-training on 11 million unlabeled molecules, we revealed that MolGNet can capture valuable chemistry insights to produce interpretable representation. The pre-trained MolGNet can be fine-tuned with just one additional output layer to create state-of-the-art models for a wide range of drug discovery tasks, including molecular properties prediction, drug-drug interaction, and drug-target interaction, involving 13 benchmark datasets. Our work demonstrates that MPG is promising to become a novel approach in the drug discovery pipeline.

Discussion (0). Continue with ORCID to comment.

Forward citations

Cited by 1 Pith paper

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score. Full citation record

  1. Self-Supervised Learning for Graph-Structured Data in Healthcare Applications: A Comprehensive Review

    cs.LG 2024-11 conditional novelty 4.0 of 10

    This review consolidates self-supervised graph learning methods for healthcare into contrastive, generative, and predictive categories, and surveys datasets, metrics, and open challenges.

Pith tools