Pith. sign in

REVIEW 1 cited by

WikiTableEdit: A Benchmark for Table Editing by Natural Language Instruction

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 2403.02962 v1 pith:NKOR76BX submitted 2024-03-05 cs.AI

classification cs.AI
keywords tablesdatasetlanguageeditingwikitableeditchallengecodedata
verification ladder T0 review T1 audit T2 compute T3 formal
0 comments
read the original abstract

Tabular data, as a crucial form of data representation, exists in diverse formats on the Web. When confronted with complex and irregular tables, manual modification becomes a laborious task. This paper investigates the performance of Large Language Models (LLMs) in the context of table editing tasks. Existing research mainly focuses on regular-shaped tables, wherein instructions are used to generate code in SQL, Python, or Excel Office-script for manipulating the tables. Nevertheless, editing tables with irregular structures, particularly those containing merged cells spanning multiple rows, poses a challenge when using code. To address this, we introduce the WikiTableEdit dataset. Leveraging 26,531 tables from the WikiSQL dataset, we automatically generate natural language instructions for six distinct basic operations and the corresponding outcomes, resulting in over 200,000 instances. Subsequently, we evaluate several representative large language models on the WikiTableEdit dataset to demonstrate the challenge of this task. The dataset will be released to the community to promote related researches.

Discussion (0). Continue with ORCID to comment.

Forward citations

Cited by 1 Pith paper

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score. Full citation record

  1. MiMoTable: A Multi-scale Spreadsheet Benchmark with Meta Operations for Table Reasoning

    cs.CL 2024-12 conditional novelty 6.0 of 10

    MiMoTable is a real-world spreadsheet benchmark with 1,719 bilingual question-answer pairs and a meta-operation difficulty criterion on which the best LLM scores 77.4%.

Pith tools