Pith. sign in

Paper Citation Record · LEDGER

Benchmarking Table Comprehension In The Wild

As of 13 August 2026, this Paper Citation Record lists 28 of 28 outbound references and 2 inbound Pith citation observations for arXiv:2412.09884.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2412.09884 v1

Coverage vector

measured 28 of 28 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-11T16:40:57.798648Z

measured 30 of 30 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-13T06:32:02.005865+00:00

measured 2 of 2 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-05-21T00:25:32.898938Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-21T00:29:17.502555Z

Reference resolution

28 of 28 outbound references displayed

  • verified exact2
  • verified fuzzy12
  • unresolved14
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 85133aa0-27fa-4445-93e7-c248515112f6 · outbound

This paper cites Leave No Document Behind: Benchmarking Long-Context LLMs with Extended Multi-Doc QA, June.

Benchmarking Table Comprehension In The Wild Leave No Document Behind: Benchmarking Long-Context LLMs with Extended Multi-Doc QA, June

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T16:40:58.540397Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-11T16:40:57.649194Z digest=sha256:0b1421ed513aea1f55c856ecc20bf7a0eedb4015285f3ae78c58ff0dc2e4de44

Observation ca1f054e-8c19-4737-9139-f46249c9c19f · outbound

This paper cites FinanceBench: A New Benchmark for Financial Question Answering.

Benchmarking Table Comprehension In The Wild FinanceBench: A New Benchmark for Financial Question Answering

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-11T16:40:57.660430Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T16:40:57.660430Z digest=sha256:4c062478b3272d3b8235f1e9c9ee00fa5718f806fd5b11eb93360fb1062438a9

Observation fa496fd2-40cf-40ab-b5fb-83512b8f237e · outbound

This paper cites FinQA: A Dataset of Numerical Reasoning over Financial Data.

Benchmarking Table Comprehension In The Wild FinQA: A Dataset of Numerical Reasoning over Financial Data

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T16:40:58.520539Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-11T16:40:57.665359Z digest=sha256:49e7c73d1118287a4236036e66921e7cbe8e5e256d2ed8e489a242dd11686ef9

Observation 83fb4bd0-c950-4256-b96b-a36c4abb6c7b · outbound

This paper cites Global Table Extractor (GTE): A Framework for Joint Table Identification and Cell Structure Recognition Using Visual Context.

Benchmarking Table Comprehension In The Wild Global Table Extractor (GTE): A Framework for Joint Table Identification and Cell Structure Recognition Using Visual Context

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-11T16:40:57.669649Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T16:40:57.669649Z digest=sha256:c835ae393feff274c9af543fc7e25bee2d90adb4c85da464f489213bf9a66e2e

Observation e4739c39-6c5f-4c81-a52a-f94eab6041ce · outbound

This paper cites TableLLM: Enabling Tabular Data Manipulation by LLMs in Real Office Usage Scenarios.

Benchmarking Table Comprehension In The Wild TableLLM: Enabling Tabular Data Manipulation by LLMs in Real Office Usage Scenarios

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-11T16:40:57.675052Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T16:40:57.675052Z digest=sha256:fa56c8c7147d14030aac3eef4e68d0a7e7efbf29668d2e59102e3f648a441521

Observation 52e67c06-38b7-43f9-b26c-e58e0a0d6b98 · outbound

This paper cites Parikh, Xuezhi Wang, Sebastian Gehrmann, Manaal Faruqui, Bhuwan Dhingra, Diyi Yang, and Dipanjan Das.

Benchmarking Table Comprehension In The Wild Parikh, Xuezhi Wang, Sebastian Gehrmann, Manaal Faruqui, Bhuwan Dhingra, Diyi Yang, and Dipanjan Das

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T16:40:58.492416Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-11T16:40:57.681716Z digest=sha256:a0f257adc8acb18c70cf03dcdcf4e379fab27b7f91db2e991befa25270eb43b8

Observation 1f953bb6-89e2-4dbc-a124-4cd4c5f000ad · outbound

This paper cites FeTaQA: Free-form Table Question Answering.

Benchmarking Table Comprehension In The Wild FeTaQA: Free-form Table Question Answering

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T16:40:58.457638Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-11T16:40:57.690855Z digest=sha256:d12cddf6868e61d397ec23964243b7bc70361ab323a65366142c31e9f288d253

Observation e063154c-956b-4fa8-9fc0-4ad7819061ac · outbound

This paper cites Tables as Texts or Images: Evaluating the Table Reasoning Ability of LLMs and MLLMs.

Benchmarking Table Comprehension In The Wild Tables as Texts or Images: Evaluating the Table Reasoning Ability of LLMs and MLLMs

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T16:40:58.429729Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-11T16:40:57.696005Z digest=sha256:96c5420f7a114bf583141f630375d8c925ebefde485c37adc30700939fcc8f64

Observation 714fe418-655d-4eda-95ac-6da6c0ced6bf · outbound

This paper cites FinanceMath: Knowledge-Intensive Math Reasoning in Finance Domains, November 2023.

Benchmarking Table Comprehension In The Wild FinanceMath: Knowledge-Intensive Math Reasoning in Finance Domains, November 2023

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T16:40:58.404330Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-11T16:40:57.700639Z digest=sha256:ebad288f0a2dabc0cc08f649f6774713b454c557278345fe0763ce8b612d995c

Observation e7ed09eb-7aa0-45ca-894b-adf93ba5bad4 · outbound

This paper cites TAT-QA: A Question Answering Benchmark on a Hybrid of Tabular and Textual Content in Finance.

Benchmarking Table Comprehension In The Wild TAT-QA: A Question Answering Benchmark on a Hybrid of Tabular and Textual Content in Finance

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T16:40:58.383535Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-11T16:40:57.705130Z digest=sha256:ef6d9a88ef6aa1aee4b9fc441002794ed0fbea18e741f08cb306e6531c5df349

Observation 098070b3-34fb-49f4-a77c-531d9ea16540 · outbound

This paper cites Open-WikiTable: Dataset for Open Domain Question Answering with Complex Reasoning over Table.

Benchmarking Table Comprehension In The Wild Open-WikiTable: Dataset for Open Domain Question Answering with Complex Reasoning over Table

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-11T16:40:57.709807Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T16:40:57.709807Z digest=sha256:bc5b7a4d7a8bd065a54ba911a1dc6933d4972a0491bf60122da93ef2de4a60ea

Observation 39097fcb-8b25-450f-a257-36aee1467022 · outbound

This paper cites Multi-modal Retrieval of Tables and Texts Using Tri-encoder Models.

Benchmarking Table Comprehension In The Wild Multi-modal Retrieval of Tables and Texts Using Tri-encoder Models

Reference 12

Resolution
verified exact
local_arxiv, observed 2026-08-11T16:40:58.063081Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-11T16:40:57.715120Z digest=sha256:f21c85a36d65c42a4a5493b22c66d656a937ed57652b7e97fbc74052895b2a98

Observation d0c0d8e6-5c4e-425f-8f15-156762c383ff · outbound

This paper cites Mixed-modality Representation Learning and Pre-training for Joint Table-and-Text Retrieval in OpenQA.

Benchmarking Table Comprehension In The Wild Mixed-modality Representation Learning and Pre-training for Joint Table-and-Text Retrieval in OpenQA

Reference 13

Resolution
verified exact
local_arxiv, observed 2026-08-11T16:40:58.017444Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-11T16:40:57.719854Z digest=sha256:72c1774df87a9b144107c83b4c2887a847b33a6da6c4086ffccc01cfce0802ac

Observation ccab84bf-01ea-4802-9c13-f2b891a5dc42 · outbound

This paper cites Needle in a haystack.

Benchmarking Table Comprehension In The Wild Needle in a haystack

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T16:40:58.362686Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-11T16:40:57.726235Z digest=sha256:f8551ec1d97f72d921208fb608833c91bc721fc7d1306338dabd1ebd595f4a28

Observation b998f43f-c610-4f9a-8248-f67b2c1e7075 · outbound

This paper cites InternLM2 Technical Report.

Benchmarking Table Comprehension In The Wild InternLM2 Technical Report

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-11T16:40:57.731337Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T16:40:57.731337Z digest=sha256:aaa3035cc15d2b07ad1a72b2b064c1a3daa7fc17fed8d05f2b9b3315038cc2f5

Observation 3c09ab68-820c-4a6d-a15c-c9e155d0a9db · outbound

This paper cites Large Language Models(LLMs) on Tabular Data: Prediction, Generation, and Understanding -- A Survey.

Benchmarking Table Comprehension In The Wild Large Language Models(LLMs) on Tabular Data: Prediction, Generation, and Understanding -- A Survey

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-11T16:40:57.737721Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T16:40:57.737721Z digest=sha256:4cd061f66ee90ac35479c775d90d2e5362ebc755271a3c43c32880d163b01443

Observation 60ae4601-edf5-4e84-98e3-0c02509f87d5 · outbound

This paper cites Large Language Model for Table Processing: A Survey.

Benchmarking Table Comprehension In The Wild Large Language Model for Table Processing: A Survey

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-11T16:40:57.742900Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T16:40:57.742900Z digest=sha256:4ede54069027cf859b61f8b1f5c26cf2b8f96fc9e6c04741c3e485b359abdf51

Observation 64198580-62e7-45d7-be64-562a0ac0fc38 · outbound

This paper cites Exploring the Limits of Transfer Learning with a Unified Text-to-Text Transformer.

Benchmarking Table Comprehension In The Wild Exploring the Limits of Transfer Learning with a Unified Text-to-Text Transformer

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-11T16:40:57.750453Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T16:40:57.750453Z digest=sha256:ba58290d91e734a1b4bcb8d3c7bab6888154d4727670afbd4aca3a97a9d623ac

Observation 51c5665f-8ba9-406c-9943-92f802ae7f9a · outbound

This paper cites The FineWeb Datasets: Decanting the Web for the Finest Text Data at Scale.

Benchmarking Table Comprehension In The Wild The FineWeb Datasets: Decanting the Web for the Finest Text Data at Scale

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-11T16:40:57.755600Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T16:40:57.755600Z digest=sha256:848465741f9eca89a24148468403de5419d05b34b569850ef0c1a227f95a7077

Observation 65b5ac68-69a6-407e-b366-43bd8b3b0b90 · outbound

This paper cites MultiTabQA: Generat- ing Tabular Answers for Multi-Table Question Answering.

Benchmarking Table Comprehension In The Wild MultiTabQA: Generat- ing Tabular Answers for Multi-Table Question Answering

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T16:40:58.339603Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-11T16:40:57.761337Z digest=sha256:fb2a7db8c40a463fe7fe3b9e0c9001e0eb0d8d2fd18b9da3c1c3f2a4f515816a

Observation 3ff6452d-fa77-449c-81c0-764265e18429 · outbound

This paper cites Gonzalez, Hao Zhang, and Ion Stoica.

Benchmarking Table Comprehension In The Wild Gonzalez, Hao Zhang, and Ion Stoica

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-11T16:40:57.766151Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T16:40:57.766151Z digest=sha256:498e54ee50ffd2a2980f0bb41c2c8ce972e79fb6a9595f8638b72417388839c4

Observation ebde495f-ab60-44a0-8208-1869a8b4d20f · outbound

This paper cites Bleu: a Method for Automatic Evaluation of Machine Translation.

Benchmarking Table Comprehension In The Wild Bleu: a Method for Automatic Evaluation of Machine Translation

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T16:40:58.301939Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-11T16:40:57.774843Z digest=sha256:511d84cfed08cd59b5ae8423aa51a562060b47d32b4d4ad5dfc722ee58db553e

Observation b47d1a99-8802-4350-904a-8b7371353b7f · outbound

This paper cites ROUGE: A Package for Automatic Evaluation of Summaries.

Benchmarking Table Comprehension In The Wild ROUGE: A Package for Automatic Evaluation of Summaries

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-11T16:40:57.779249Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T16:40:57.779249Z digest=sha256:ba0e85f1171afe5522fe0a9abac793613ec8a7dcb990142ea9185908e6480860

Observation 6deefa9c-1ad0-4543-a15f-d5b0d33cde5e · outbound

This paper cites METEOR: An Automatic Metric for MT Evaluation with Improved Correlation with Human Judgments.

Benchmarking Table Comprehension In The Wild METEOR: An Automatic Metric for MT Evaluation with Improved Correlation with Human Judgments

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T16:40:58.268828Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-11T16:40:57.785799Z digest=sha256:8d44e7a298112d9a3e84069c17d329704955f4f8edd002dc6a61f09bed14be99

Observation 351daf2d-1789-4532-9155-f87c0e41db14 · outbound

This paper cites BERTScore: Evaluating Text Generation with BERT.

Benchmarking Table Comprehension In The Wild BERTScore: Evaluating Text Generation with BERT

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-11T16:40:57.792310Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T16:40:57.792310Z digest=sha256:ea0db019363ca2df1f7ed6aebbf6c1b1ae9f1b73aad93d9954d31efcbdafc5ca

Observation b043f4ae-1a8f-4c73-866d-59efea13e10c · outbound

This paper cites steps"), then provide a succinct answer (marked by.

Benchmarking Table Comprehension In The Wild steps"), then provide a succinct answer (marked by

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T16:40:58.241295Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-11T16:40:57.798648Z digest=sha256:9f44608577e385ce239e9ab0eeccfb908f6e92a53b0324e65c08c9824397c675

Observation 7e868007-b39c-4647-b7b6-2c38aee871ec · outbound

This paper cites ToTTo: A Controlled Table-To-Text Generation Dataset.

Benchmarking Table Comprehension In The Wild ToTTo: A Controlled Table-To-Text Generation Dataset

Reference 2020

Resolution
unresolved
no resolver link, observed 2026-08-11T16:40:57.686154Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T16:40:57.686154Z digest=sha256:95543d5291780576524286175a3db261f7aa010995e83261ff127b31c4ce3490

Observation 13078fb9-262c-44c3-affc-631e0890a90e · outbound

This paper cites Leave No Document Behind: Benchmarking Long-Context LLMs with Extended Multi-Doc QA.

Benchmarking Table Comprehension In The Wild Leave No Document Behind: Benchmarking Long-Context LLMs with Extended Multi-Doc QA

Reference 2024

Resolution
unresolved
no resolver link, observed 2026-08-11T16:40:57.655302Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T16:40:57.655302Z digest=sha256:fc274ef16588b239c0b3c04093bbe3c3a0c682ea3388a730dd7d572448ed7354

Pith citing papers

Observation 80c2e246-13d6-4cac-b536-0f52d36733ba · inbound

DataClawBench: An Agent Benchmark for Exploratory Real-World Financial Data Analysis cites this paper.

DataClawBench: An Agent Benchmark for Exploratory Real-World Financial Data Analysis Benchmarking Table Comprehension In The Wild

Reference 3

Resolution
metadata mismatch
arxiv_id, observed 2026-05-09T06:30:42.416195Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-05-08T18:24:15.160643Z digest=sha256:7c90516bb94d340dc1ab54f04486a0c91c50ff4bf4d78f24275ab62e8d68cc20

Observation 5548c944-463f-4cc7-ad76-478133b638f7 · inbound

DataClawBench: An Agent Benchmark for Exploratory Real-World Financial Data Analysis cites this paper.

DataClawBench: An Agent Benchmark for Exploratory Real-World Financial Data Analysis Benchmarking Table Comprehension In The Wild

Reference 4

Resolution
malformed identifier
arxiv_id, observed 2026-05-21T00:29:17.505356Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-05-21T00:25:32.898938Z digest=sha256:cd9f16e125d54377147ba918b395e389961ab80cd3bc548cf16058a12c1a67cb