Pith. sign in

REVIEW

A Unified Continuous Learning Framework for Multi-modal Knowledge Discovery and Pre-training

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 2206.05555 v1 pith:PFKDAB4T submitted 2022-06-11 cs.CL cs.CV

classification cs.CLcs.CV
keywords knowledgediscoveryframeworklearningmodelmulti-modalpre-trainingcontinuous
verification ladder T0 review T1 audit T2 compute T3 formal
0 comments
read the original abstract

Multi-modal pre-training and knowledge discovery are two important research topics in multi-modal machine learning. Nevertheless, none of existing works make attempts to link knowledge discovery with knowledge guided multi-modal pre-training. In this paper, we propose to unify them into a continuous learning framework for mutual improvement. Taking the open-domain uni-modal datasets of images and texts as input, we maintain a knowledge graph as the foundation to support these two tasks. For knowledge discovery, a pre-trained model is used to identify cross-modal links on the graph. For model pre-training, the knowledge graph is used as the external knowledge to guide the model updating. These two steps are iteratively performed in our framework for continuous learning. The experimental results on MS-COCO and Flickr30K with respect to both knowledge discovery and the pre-trained model validate the effectiveness of our framework.

Discussion (0). Sign in to comment.

Pith tools