Pith. sign in

REVIEW 1 cited by

Probing Contextualized Sentence Representations with Visual Awareness

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 1911.02971 v1 pith:3KJD4DIG submitted 2019-11-07 cs.CL cs.CVcs.LG

classification cs.CLcs.CVcs.LG
keywords representationssentenceawarenesscontextualizedimageslanguagemultimodalnatural
verification ladder T0 review T1 audit T2 compute T3 formal

Signed reviews

No signed human review yet.

0 comments
read the original abstract

We present a universal framework to model contextualized sentence representations with visual awareness that is motivated to overcome the shortcomings of the multimodal parallel data with manual annotations. For each sentence, we first retrieve a diversity of images from a shared cross-modal embedding space, which is pre-trained on a large-scale of text-image pairs. Then, the texts and images are respectively encoded by transformer encoder and convolutional neural network. The two sequences of representations are further fused by a simple and effective attention layer. The architecture can be easily applied to text-only natural language processing tasks without manually annotating multimodal parallel corpora. We apply the proposed method on three tasks, including neural machine translation, natural language inference and sequence labeling and experimental results verify the effectiveness.

Discussion (0). Continue with ORCID to comment.

Forward citations

Cited by 1 Pith paper

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score. Full citation record

  1. SG-Net: Syntax-Guided Machine Reading Comprehension

    cs.CL 2019-08 conditional novelty 5.0 of 10

    Masking self-attention to syntactic ancestors and averaging it with BERT attention improves SQuAD 2.0 exact match from 84.1 to 85.1 and RACE accuracy from 72.6 to 74.2.

Pith tools