Pith. sign in

Paper Citation Record · LEDGER

Towards Benchmarking Foundation Models for Tabular Data With Text

As of 7 August 2026, this Paper Citation Record lists 26 of 26 outbound references and 3 inbound Pith citation observations for arXiv:2507.07829.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2507.07829 v1

Coverage vector

measured 26 of 26 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-06T18:36:47.958225Z

measured 29 of 29 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00

measured 3 of 3 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-06-30T07:34:47.760875Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-06-30T08:04:29.090042Z

Reference resolution

26 of 26 outbound references displayed

  • verified exact0
  • verified fuzzy2
  • unresolved23
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch1

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 1c7d4019-78e6-4a09-b12b-428e12e41c91 · outbound

This paper cites write newline.

Towards Benchmarking Foundation Models for Tabular Data With Text write newline

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-06T18:36:46.629336Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T18:36:46.629336Z digest=sha256:ca65f12f9ac132c237096757ddf36aae6075b118a656beea309648a942b0324d

Observation baf2414d-065f-4a50-a179-fcdb89fd3b6b · outbound

This paper cites OpenML Benchmarking Suites.

Towards Benchmarking Foundation Models for Tabular Data With Text OpenML Benchmarking Suites

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-06T18:36:46.741921Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T18:36:46.741921Z digest=sha256:9ab0730003a5bbbf97df8af52b6aad3c398b8c8f04e1bdcce902a4e628a09b6a

Observation 377f81f8-07a2-4e19-8c55-b165bf9a981d · outbound

This paper cites Enriching Word Vectors with Subword Information.

Towards Benchmarking Foundation Models for Tabular Data With Text Enriching Word Vectors with Subword Information

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-06T18:36:46.868838Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T18:36:46.868838Z digest=sha256:eabe3ee5b93f54a032899732be350c708b84a82d8e1d95a5f4acd8fbdb364ec7

Observation 483d32a2-ff1e-4cbd-a28b-b3533f925e72 · outbound

This paper cites V., Na, L., Ma, Y., Boussioux, L., Zeng, C., Soenksen, L.

Towards Benchmarking Foundation Models for Tabular Data With Text V., Na, L., Ma, Y., Boussioux, L., Zeng, C., Soenksen, L

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-06T18:36:47.052197Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T18:36:47.052197Z digest=sha256:2b2aeccf05c26403f3cd91b288356c42bc116d5424ad8cce7873e9cbaa32f96b

Observation 877e9463-acaa-49d1-abe0-80df7238f766 · outbound

This paper cites and Guestrin, C.

Towards Benchmarking Foundation Models for Tabular Data With Text and Guestrin, C

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:36:48.592767Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-06T18:36:47.160510Z digest=sha256:441c7c112e6429bf2d28014930551a17e36ed263bc045f18d75036bc60511995

Observation da1cfbfd-2367-4aa7-b9d3-abe8e32e0193 · outbound

This paper cites AutoGluon-Tabular: Robust and Accurate AutoML for Structured Data.

Towards Benchmarking Foundation Models for Tabular Data With Text AutoGluon-Tabular: Robust and Accurate AutoML for Structured Data

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-06T18:36:47.379543Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T18:36:47.379543Z digest=sha256:9cd433b8108eb53a4ba7d6c268209394a52f3dd96e1fae7af398e49ef4e2a6c4

Observation fd806105-fa00-42c0-8498-182df74f00b6 · outbound

This paper cites AMLB: an AutoML Benchmark.

Towards Benchmarking Foundation Models for Tabular Data With Text AMLB: an AutoML Benchmark

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-06T18:36:47.491670Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T18:36:47.491670Z digest=sha256:509ddcfbb7a8f13cfd9d6c2563d618b3e7fb7e9113a6ef9391e1701d38e50d12

Observation bacee3e0-ffd2-49d5-a053-b5645203fb63 · outbound

This paper cites L., Amaral, L.

Towards Benchmarking Foundation Models for Tabular Data With Text L., Amaral, L

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:36:48.572115Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-06T18:36:47.546239Z digest=sha256:ffec06b94c7ccea4d855bec7f149beace02a0eb42145ccb41d0f4c7d0d985125

Observation 397e6b2c-c2bc-48c7-9a81-fad9763d348b · outbound

This paper cites Vectorizing string entries for data processing on tables: when are larger language models better?.

Towards Benchmarking Foundation Models for Tabular Data With Text Vectorizing string entries for data processing on tables: when are larger language models better?

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-06T18:36:47.624023Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T18:36:47.624023Z digest=sha256:0e811d9250239d643d53c9dc8d04b66e4ffd191c0978a37008e873d58c1c2885

Observation 553da31f-b73a-417e-bc29-3f3a0fffff2d · outbound

This paper cites TabLLM: Few-shot Classification of Tabular Data with Large Language Models.

Towards Benchmarking Foundation Models for Tabular Data With Text TabLLM: Few-shot Classification of Tabular Data with Large Language Models

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-06T18:36:47.741196Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T18:36:47.741196Z digest=sha256:9c5fb46b00dcbe3d13401b7ffe8b9abae61c571fbbde2df0702f9ee45e7486ea

Observation 2a0e2f26-28a3-4db7-99dd-1e2047a4d95b · outbound

This paper cites Machine Learning for Health symposium 2023 -- Findings track.

Towards Benchmarking Foundation Models for Tabular Data With Text Machine Learning for Health symposium 2023 -- Findings track

Reference 12

Resolution
metadata mismatch
local_arxiv, observed 2026-08-06T18:36:48.343264Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-06T18:36:47.771782Z digest=sha256:ee5bc99754a1c453bc08c58f45b93ec0417f0a078648499d02a5651f98b4cacc

Observation 3d999283-7563-4363-a603-ddc2a8712cc4 · outbound

This paper cites u ller, S., Purucker, L., Krishnakumar, A., K \.

Towards Benchmarking Foundation Models for Tabular Data With Text u ller, S., Purucker, L., Krishnakumar, A., K \

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-06T18:36:47.776363Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T18:36:47.776363Z digest=sha256:427883667543c94871c86f8abaab80c6b8905a2a0c93ce93ad01853e7ef823f0

Observation 2f229c0d-7b38-45e0-871a-c7df06729364 · outbound

This paper cites CARTE: Pretraining and Transfer for Tabular Learning.

Towards Benchmarking Foundation Models for Tabular Data With Text CARTE: Pretraining and Transfer for Tabular Learning

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-06T18:36:47.781086Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T18:36:47.781086Z digest=sha256:10ced83f88227462f04ff08ebc368a05a7ecd980503d019b7c26e09bb60328b4

Observation 4cfe51fd-5639-4fb0-b60c-704ef2f0020a · outbound

This paper cites LLM Embeddings for Deep Learning on Tabular Data.

Towards Benchmarking Foundation Models for Tabular Data With Text LLM Embeddings for Deep Learning on Tabular Data

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-06T18:36:47.785860Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T18:36:47.785860Z digest=sha256:e3a46e661ceeaf7e17939c2d910defc1fb4377aff834573c3a2901605d9e75ef

Observation f2b6bb55-8deb-414c-8e46-6085c0fd301d · outbound

This paper cites TALENT: A Tabular Analytics and Learning Toolbox.

Towards Benchmarking Foundation Models for Tabular Data With Text TALENT: A Tabular Analytics and Learning Toolbox

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-06T18:36:47.790726Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T18:36:47.790726Z digest=sha256:9323d5a262db93edb96b896a27f1173f4bd8e92cba3a98d034ab1def1d9729b5

Observation d212f918-239c-42a2-a6cb-30ab6f33bb59 · outbound

This paper cites Mug: A multimodal classification benchmark on game data with tabular, textual, and visual fields.

Towards Benchmarking Foundation Models for Tabular Data With Text Mug: A multimodal classification benchmark on game data with tabular, textual, and visual fields

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-06T18:36:47.796759Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T18:36:47.796759Z digest=sha256:e61219d8b79d5386053b77038bb92adecc9f07e50c87646cdfca4385adbae033

Observation 0e193bd7-f170-4e32-afac-bd6d0fee7595 · outbound

This paper cites an unresolved cited work.

Towards Benchmarking Foundation Models for Tabular Data With Text Unresolved cited work

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-06T18:36:47.801181Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T18:36:47.801181Z digest=sha256:973bdd3b044941e06b06188538df2f56db2ed2f8ea75fd55dbd96f79da7ee916

Observation dd1bc50d-5e66-4035-9d14-95163b290aee · outbound

This paper cites C., Golestan, K., Yu, G., Volkovs, M., and Caterini, A.

Towards Benchmarking Foundation Models for Tabular Data With Text C., Golestan, K., Yu, G., Volkovs, M., and Caterini, A

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-06T18:36:47.805612Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T18:36:47.805612Z digest=sha256:929b57b2e5b002d5a1ccc380bacbb26731e21fdb0a64b6ad4f0c378402ae6779

Observation de99a1bf-5ee9-4548-b326-cb14bf0be713 · outbound

This paper cites and Ratajczak, W.

Towards Benchmarking Foundation Models for Tabular Data With Text and Ratajczak, W

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-06T18:36:47.923586Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T18:36:47.923586Z digest=sha256:dd2e25da75d845eb19cf90a4317986a2653d85a755d9803275eb257155c0e56c

Observation b77ceb41-ea15-4e30-a9a0-62a0e3e34b3d · outbound

This paper cites an unresolved cited work.

Towards Benchmarking Foundation Models for Tabular Data With Text Unresolved cited work

Reference 21

Resolution
unresolved
raw_fallback, observed 2026-08-06T18:36:48.542595Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-06T18:36:47.928885Z digest=sha256:73a5ded71a93ba367f4b76c080b94dfb6de633606d6ddbd54215b49edcef7c61

Observation 3c40f97a-7d09-4260-87f3-95a17bc30202 · outbound

This paper cites When Do Neural Nets Outperform Boosted Trees on Tabular Data?.

Towards Benchmarking Foundation Models for Tabular Data With Text When Do Neural Nets Outperform Boosted Trees on Tabular Data?

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-06T18:36:47.933777Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T18:36:47.933777Z digest=sha256:246f0686ee30ac6b749849cdee92d9f95d7283b05ee884cb4bdce4350f70aab9

Observation 20f4a16d-b51d-445d-b8e4-85069cf05237 · outbound

This paper cites Benchmarking Multimodal AutoML for Tabular Data with Text Fields.

Towards Benchmarking Foundation Models for Tabular Data With Text Benchmarking Multimodal AutoML for Tabular Data with Text Fields

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-06T18:36:47.938962Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T18:36:47.938962Z digest=sha256:0c54ecda45bda24b6a099d5b1f7b19653efde7fc097976f178a4896eef687a14

Observation 68150fff-f389-4666-be6f-8c868c976331 · outbound

This paper cites JoLT: Joint Probabilistic Predictions on Tabular Data Using LLMs.

Towards Benchmarking Foundation Models for Tabular Data With Text JoLT: Joint Probabilistic Predictions on Tabular Data Using LLMs

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-06T18:36:47.943525Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T18:36:47.943525Z digest=sha256:ad86213109945c1f97e72350f5c2abe4b39da0ff174780696b92e7982cff906c

Observation b7cd1408-79a6-4cfc-98cf-5c62396afbdd · outbound

This paper cites AutoGluon-Multimodal (AutoMM): Supercharging Multimodal AutoML with Foundation Models.

Towards Benchmarking Foundation Models for Tabular Data With Text AutoGluon-Multimodal (AutoMM): Supercharging Multimodal AutoML with Foundation Models

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-06T18:36:47.948093Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T18:36:47.948093Z digest=sha256:61a54705d7c06959b0f7f2ddbfadb767f01913b3e4c785178b7c5e1b0449ea7d

Observation 39807671-9e31-42b6-b726-dd81eba8cd87 · outbound

This paper cites MiniLM: Deep Self-Attention Distillation for Task-Agnostic Compression of Pre-Trained Transformers.

Towards Benchmarking Foundation Models for Tabular Data With Text MiniLM: Deep Self-Attention Distillation for Task-Agnostic Compression of Pre-Trained Transformers

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-06T18:36:47.952726Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T18:36:47.952726Z digest=sha256:a1ada71c4b15817bb29e70689e5ec6b1df5ee0ed807ef6a87842f7bda8261869

Observation df858277-3b94-4495-b64a-1bbc205d457e · outbound

This paper cites TableLLM: Enabling Tabular Data Manipulation by LLMs in Real Office Usage Scenarios.

Towards Benchmarking Foundation Models for Tabular Data With Text TableLLM: Enabling Tabular Data Manipulation by LLMs in Real Office Usage Scenarios

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-06T18:36:47.958225Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T18:36:47.958225Z digest=sha256:f82625edd3fa96a274f992642dde6c31de5b782ea5acc023053ffdb362c82270

Pith citing papers

Observation 036e564d-2896-4d9d-a8c4-e74202eb4aa7 · inbound

STRABLE: Benchmarking Tabular Machine Learning with Strings cites this paper.

STRABLE: Benchmarking Tabular Machine Learning with Strings Towards Benchmarking Foundation Models for Tabular Data With Text

Reference 42

Resolution
verified exact
arxiv_id, observed 2026-05-13T05:17:18.441754Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-13T05:13:15.039160Z digest=sha256:f2044cf012884c42cb494c1c205f1d1d0cf8ac50bf5ce63eb25ef8b23ea0c870

Observation 378dff40-5ec7-4857-906e-3d4ce8efecff · inbound

Beyond IID: How General Are Tabular Foundation Models, Really? cites this paper.

Beyond IID: How General Are Tabular Foundation Models, Really? Towards Benchmarking Foundation Models for Tabular Data With Text

Reference 16

Resolution
verified exact
arxiv_id, observed 2026-06-30T07:04:21.420092Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-06-30T06:59:14.626274Z digest=sha256:c6636a62f61864b07873ce6ef08d0c3a770a6b0f3df647ace98cb32597527328

Observation 4a372cec-af06-47c5-9513-4821a51dd7f5 · inbound

Exploring Differences Between Tabular Enterprise Data and Public Benchmarks cites this paper.

Exploring Differences Between Tabular Enterprise Data and Public Benchmarks Towards Benchmarking Foundation Models for Tabular Data With Text

Reference 10

Resolution
metadata mismatch
arxiv_id, observed 2026-06-30T08:04:29.091444Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-06-30T07:34:47.760875Z digest=sha256:3618cbc5b265401df81fb5d36d5deb30a6858bfa69e188f1cb0efb0188a68196