Pith. sign in

REVIEW 1 cited by

Multi-modal Retrieval of Tables and Texts Using Tri-encoder Models

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 2108.04049 v2 pith:6Z32CZAJ submitted 2021-08-09 cs.CL cs.IR

classification cs.CLcs.IR
keywords tablesquestiontexttextsmodelsmulti-modalretrievaldataset
verification ladder T0 review T1 audit T2 compute T3 formal
0 comments
read the original abstract

Open-domain extractive question answering works well on textual data by first retrieving candidate texts and then extracting the answer from those candidates. However, some questions cannot be answered by text alone but require information stored in tables. In this paper, we present an approach for retrieving both texts and tables relevant to a question by jointly encoding texts, tables and questions into a single vector space. To this end, we create a new multi-modal dataset based on text and table datasets from related work and compare the retrieval performance of different encoding schemata. We find that dense vector embeddings of transformer models outperform sparse embeddings on four out of six evaluation datasets. Comparing different dense embedding models, tri-encoders with one encoder for each question, text and table, increase retrieval performance compared to bi-encoders with one encoder for the question and one for both text and tables. We release the newly created multi-modal dataset to the community so that it can be used for training and evaluation.

Discussion (0). Continue with ORCID to comment.

Forward citations

Cited by 1 Pith paper

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score. Full citation record

  1. Benchmarking Table Comprehension In The Wild

    cs.CL 2024-12 reject novelty 5.0 of 10

    TableQuest evaluates LLMs on table tasks embedded in real 10-K reports and finds that current models handle extraction but struggle with multi-step reasoning, though the evaluation has a self-judging bias.

Pith tools