Pith. sign in

Paper Citation Record · LEDGER

Towards Benchmarking Foundation Models for Tabular Data With Text

As of 7 August 2026, this Paper Citation Record lists 26 of 26 outbound references and 3 inbound Pith citation observations for arXiv:2507.07829.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2507.07829 v1

Coverage vector

measured 26 of 26 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-06T18:36:47.958225Z

measured 29 of 29 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00

measured 3 of 3 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-06-30T07:34:47.760875Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-06-30T08:04:29.090042Z

Reference resolution

26 of 26 outbound references displayed

  • verified exact0
  • verified fuzzy2
  • unresolved23
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch1

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 1c7d4019-78e6-4a09-b12b-428e12e41c91 · outbound

This paper cites write newline.

Towards Benchmarking Foundation Models for Tabular Data With Text write newline

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-06T18:36:46.629336Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T18:36:46.629336Z digest=sha256:05de3853b6ee4d27d11496b7d4585e281cdd2f2abfc9b2d56b6960b12e06a97a

Observation baf2414d-065f-4a50-a179-fcdb89fd3b6b · outbound

This paper cites OpenML Benchmarking Suites.

Towards Benchmarking Foundation Models for Tabular Data With Text OpenML Benchmarking Suites

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-06T18:36:46.741921Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T18:36:46.741921Z digest=sha256:bb0994f757447e5cfc2886b28ba22990f1047d7fecafe44e98e70a4b09aafc20

Observation 377f81f8-07a2-4e19-8c55-b165bf9a981d · outbound

This paper cites Enriching Word Vectors with Subword Information.

Towards Benchmarking Foundation Models for Tabular Data With Text Enriching Word Vectors with Subword Information

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-06T18:36:46.868838Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T18:36:46.868838Z digest=sha256:75fd4264dc743f73a98d5eab5aeb6bf4e5699dde4ef958a8cdc974586cac0b0f

Observation 483d32a2-ff1e-4cbd-a28b-b3533f925e72 · outbound

This paper cites V., Na, L., Ma, Y., Boussioux, L., Zeng, C., Soenksen, L.

Towards Benchmarking Foundation Models for Tabular Data With Text V., Na, L., Ma, Y., Boussioux, L., Zeng, C., Soenksen, L

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-06T18:36:47.052197Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T18:36:47.052197Z digest=sha256:accfcbfde943db138b5d4949da14173c7779f93c0c5de55587d0fa8c033bd2a5

Observation 877e9463-acaa-49d1-abe0-80df7238f766 · outbound

This paper cites and Guestrin, C.

Towards Benchmarking Foundation Models for Tabular Data With Text and Guestrin, C

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:36:48.592767Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-06T18:36:47.160510Z digest=sha256:c245a935894801ed0e4f272c6748cec68a41157ba9a02f1ce8b9ea9ad85f7e33

Observation da1cfbfd-2367-4aa7-b9d3-abe8e32e0193 · outbound

This paper cites AutoGluon-Tabular: Robust and Accurate AutoML for Structured Data.

Towards Benchmarking Foundation Models for Tabular Data With Text AutoGluon-Tabular: Robust and Accurate AutoML for Structured Data

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-06T18:36:47.379543Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T18:36:47.379543Z digest=sha256:01558dbc8ea85f74b60bb66c6ca51b5369ebe5fd2893589abb675df12fc9ed89

Observation fd806105-fa00-42c0-8498-182df74f00b6 · outbound

This paper cites AMLB: an AutoML Benchmark.

Towards Benchmarking Foundation Models for Tabular Data With Text AMLB: an AutoML Benchmark

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-06T18:36:47.491670Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T18:36:47.491670Z digest=sha256:a047d3d4bf057beadcbc44c677a483d0ab659c91d0e082f35ef60045a2a1ab09

Observation bacee3e0-ffd2-49d5-a053-b5645203fb63 · outbound

This paper cites L., Amaral, L.

Towards Benchmarking Foundation Models for Tabular Data With Text L., Amaral, L

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:36:48.572115Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-06T18:36:47.546239Z digest=sha256:30723324775a385eaf4af36f86962e538fe7d050c389315ae9690df595c51ea9

Observation 397e6b2c-c2bc-48c7-9a81-fad9763d348b · outbound

This paper cites Vectorizing string entries for data processing on tables: when are larger language models better?.

Towards Benchmarking Foundation Models for Tabular Data With Text Vectorizing string entries for data processing on tables: when are larger language models better?

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-06T18:36:47.624023Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T18:36:47.624023Z digest=sha256:0f4f833973730288aad316a08ffa1ff3108582abe76fe1d2a2b28653debafbda

Observation 553da31f-b73a-417e-bc29-3f3a0fffff2d · outbound

This paper cites TabLLM: Few-shot Classification of Tabular Data with Large Language Models.

Towards Benchmarking Foundation Models for Tabular Data With Text TabLLM: Few-shot Classification of Tabular Data with Large Language Models

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-06T18:36:47.741196Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T18:36:47.741196Z digest=sha256:cd50f05a11be7f97869f7ec337bee49adc6801c430e7bc5e6c420f87c5112711

Observation 2a0e2f26-28a3-4db7-99dd-1e2047a4d95b · outbound

This paper cites Machine Learning for Health symposium 2023 -- Findings track.

Towards Benchmarking Foundation Models for Tabular Data With Text Machine Learning for Health symposium 2023 -- Findings track

Reference 12

Resolution
metadata mismatch
local_arxiv, observed 2026-08-06T18:36:48.343264Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-06T18:36:47.771782Z digest=sha256:e99c1415158a7348febf00c926c8dec3eae938b3ed73e51cb23c3b969633bf4c

Observation 3d999283-7563-4363-a603-ddc2a8712cc4 · outbound

This paper cites u ller, S., Purucker, L., Krishnakumar, A., K \.

Towards Benchmarking Foundation Models for Tabular Data With Text u ller, S., Purucker, L., Krishnakumar, A., K \

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-06T18:36:47.776363Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T18:36:47.776363Z digest=sha256:48922be0ccf04c3798ad41ec21aecf725686555bf49c85022e74eb84a366cf7f

Observation 2f229c0d-7b38-45e0-871a-c7df06729364 · outbound

This paper cites CARTE: Pretraining and Transfer for Tabular Learning.

Towards Benchmarking Foundation Models for Tabular Data With Text CARTE: Pretraining and Transfer for Tabular Learning

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-06T18:36:47.781086Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T18:36:47.781086Z digest=sha256:d4eabe5b3dd187412d2faea708f0a981e516d4d535cf4b1c4fdf50e0cd63fc6f

Observation 4cfe51fd-5639-4fb0-b60c-704ef2f0020a · outbound

This paper cites LLM Embeddings for Deep Learning on Tabular Data.

Towards Benchmarking Foundation Models for Tabular Data With Text LLM Embeddings for Deep Learning on Tabular Data

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-06T18:36:47.785860Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T18:36:47.785860Z digest=sha256:aba635dc775c6868af5940cc762be6a3dafc2a37ad8d7ca5bea4fea47262bce2

Observation f2b6bb55-8deb-414c-8e46-6085c0fd301d · outbound

This paper cites TALENT: A Tabular Analytics and Learning Toolbox.

Towards Benchmarking Foundation Models for Tabular Data With Text TALENT: A Tabular Analytics and Learning Toolbox

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-06T18:36:47.790726Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T18:36:47.790726Z digest=sha256:6617722461b2f5cbcd1b4cc8b7a2139cfa3c81b309f01a9393b18e5425dfdc0c

Observation d212f918-239c-42a2-a6cb-30ab6f33bb59 · outbound

This paper cites Mug: A multimodal classification benchmark on game data with tabular, textual, and visual fields.

Towards Benchmarking Foundation Models for Tabular Data With Text Mug: A multimodal classification benchmark on game data with tabular, textual, and visual fields

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-06T18:36:47.796759Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T18:36:47.796759Z digest=sha256:aaa2c1a61f90e28c99b73a243640c418afe7fa9f42efb72bbc13dbb0e6aacd35

Observation 0e193bd7-f170-4e32-afac-bd6d0fee7595 · outbound

This paper cites an unresolved cited work.

Towards Benchmarking Foundation Models for Tabular Data With Text Unresolved cited work

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-06T18:36:47.801181Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T18:36:47.801181Z digest=sha256:f1e78a4dc0f0ebfe7f2b6a4ed94c7df9492b591358cf3720f5b9e9d86943d1a2

Observation dd1bc50d-5e66-4035-9d14-95163b290aee · outbound

This paper cites C., Golestan, K., Yu, G., Volkovs, M., and Caterini, A.

Towards Benchmarking Foundation Models for Tabular Data With Text C., Golestan, K., Yu, G., Volkovs, M., and Caterini, A

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-06T18:36:47.805612Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T18:36:47.805612Z digest=sha256:5c9ee15ec501b8cfd83b5beaa883abaf45593d2f22a2636e5f5380206eb5efec

Observation de99a1bf-5ee9-4548-b326-cb14bf0be713 · outbound

This paper cites and Ratajczak, W.

Towards Benchmarking Foundation Models for Tabular Data With Text and Ratajczak, W

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-06T18:36:47.923586Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T18:36:47.923586Z digest=sha256:86de2786cfdb60280c44d66069aafcb746201b6702ddab9224564d9bdbb09534

Observation b77ceb41-ea15-4e30-a9a0-62a0e3e34b3d · outbound

This paper cites an unresolved cited work.

Towards Benchmarking Foundation Models for Tabular Data With Text Unresolved cited work

Reference 21

Resolution
unresolved
raw_fallback, observed 2026-08-06T18:36:48.542595Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-06T18:36:47.928885Z digest=sha256:eca1da7b805fc178178c00bbaa71feec58899cbd220f22b63cb82c67d6a23074

Observation 3c40f97a-7d09-4260-87f3-95a17bc30202 · outbound

This paper cites When Do Neural Nets Outperform Boosted Trees on Tabular Data?.

Towards Benchmarking Foundation Models for Tabular Data With Text When Do Neural Nets Outperform Boosted Trees on Tabular Data?

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-06T18:36:47.933777Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T18:36:47.933777Z digest=sha256:4f1a677b139ddcdd5d852ef6a5346edf904644468997e5af845f04d3d35472ad

Observation 20f4a16d-b51d-445d-b8e4-85069cf05237 · outbound

This paper cites Benchmarking Multimodal AutoML for Tabular Data with Text Fields.

Towards Benchmarking Foundation Models for Tabular Data With Text Benchmarking Multimodal AutoML for Tabular Data with Text Fields

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-06T18:36:47.938962Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T18:36:47.938962Z digest=sha256:a1a12c42a0f6c350c51a0be82a4a38a0d887d1139a592f174a23371d74b5f4fc

Observation 68150fff-f389-4666-be6f-8c868c976331 · outbound

This paper cites JoLT: Joint Probabilistic Predictions on Tabular Data Using LLMs.

Towards Benchmarking Foundation Models for Tabular Data With Text JoLT: Joint Probabilistic Predictions on Tabular Data Using LLMs

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-06T18:36:47.943525Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T18:36:47.943525Z digest=sha256:c0f25250dd01239ccad29e16dfd48e70de7e997b730dedd6db869b9fb768033e

Observation b7cd1408-79a6-4cfc-98cf-5c62396afbdd · outbound

This paper cites AutoGluon-Multimodal (AutoMM): Supercharging Multimodal AutoML with Foundation Models.

Towards Benchmarking Foundation Models for Tabular Data With Text AutoGluon-Multimodal (AutoMM): Supercharging Multimodal AutoML with Foundation Models

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-06T18:36:47.948093Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T18:36:47.948093Z digest=sha256:5cac6aaf2b4b9cc92164093b4544b6618fedd1071e68079b238e9396b5c34db9

Observation 39807671-9e31-42b6-b726-dd81eba8cd87 · outbound

This paper cites MiniLM: Deep Self-Attention Distillation for Task-Agnostic Compression of Pre-Trained Transformers.

Towards Benchmarking Foundation Models for Tabular Data With Text MiniLM: Deep Self-Attention Distillation for Task-Agnostic Compression of Pre-Trained Transformers

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-06T18:36:47.952726Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T18:36:47.952726Z digest=sha256:ca1daf3e65f8265ef7f44951878826b9a4493df346f8da3f0267b58beb803fe0

Observation df858277-3b94-4495-b64a-1bbc205d457e · outbound

This paper cites TableLLM: Enabling Tabular Data Manipulation by LLMs in Real Office Usage Scenarios.

Towards Benchmarking Foundation Models for Tabular Data With Text TableLLM: Enabling Tabular Data Manipulation by LLMs in Real Office Usage Scenarios

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-06T18:36:47.958225Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T18:36:47.958225Z digest=sha256:1f040738bbeecfbfc481d046357d0b323aa5476ec44f850edd4613221f26ba9a

Pith citing papers

Observation 036e564d-2896-4d9d-a8c4-e74202eb4aa7 · inbound

STRABLE: Benchmarking Tabular Machine Learning with Strings cites this paper.

STRABLE: Benchmarking Tabular Machine Learning with Strings Towards Benchmarking Foundation Models for Tabular Data With Text

Reference 42

Resolution
verified exact
arxiv_id, observed 2026-05-13T05:17:18.441754Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-13T05:13:15.039160Z digest=sha256:2b036ac588837bb43f030d6095868a654fca7dd527b5e3be91aafaa23136a3fb

Observation 378dff40-5ec7-4857-906e-3d4ce8efecff · inbound

Beyond IID: How General Are Tabular Foundation Models, Really? cites this paper.

Beyond IID: How General Are Tabular Foundation Models, Really? Towards Benchmarking Foundation Models for Tabular Data With Text

Reference 16

Resolution
verified exact
arxiv_id, observed 2026-06-30T07:04:21.420092Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-06-30T06:59:14.626274Z digest=sha256:272bc5c6970cf2eab12e17671de5ca693c3eb810ad74c50701cb4f716fc4c0b5

Observation 4a372cec-af06-47c5-9513-4821a51dd7f5 · inbound

Exploring Differences Between Tabular Enterprise Data and Public Benchmarks cites this paper.

Exploring Differences Between Tabular Enterprise Data and Public Benchmarks Towards Benchmarking Foundation Models for Tabular Data With Text

Reference 10

Resolution
metadata mismatch
arxiv_id, observed 2026-06-30T08:04:29.091444Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-06-30T07:34:47.760875Z digest=sha256:106e3cf5fc47940af8a282d48a7338e2145182b8cdc937bb4b05bb6e577fbbe4