Pith. sign in

Paper Citation Record · LEDGER

Towards Benchmarking Foundation Models for Tabular Data With Text

As of 16 August 2026, this Paper Citation Record lists 26 of 26 outbound references and 3 inbound Pith citation observations for arXiv:2507.07829.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2507.07829 v1

Coverage vector

measured 26 of 26 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-06T18:36:47.958225Z

measured 29 of 29 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-15T06:32:42.880941+00:00

measured 3 of 3 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-06-30T07:34:47.760875Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-06-30T08:04:29.090042Z

Reference resolution

26 of 26 outbound references displayed

  • verified exact0
  • verified fuzzy2
  • unresolved23
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch1

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 1c7d4019-78e6-4a09-b12b-428e12e41c91 · outbound

This paper cites write newline.

Towards Benchmarking Foundation Models for Tabular Data With Text write newline

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-06T18:36:46.629336Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T18:36:46.629336Z digest=sha256:d73f9d554010b85d012d6e4819283ed262531b6eb0b4375c149bf521ca4d3aa8

Observation baf2414d-065f-4a50-a179-fcdb89fd3b6b · outbound

This paper cites OpenML Benchmarking Suites.

Towards Benchmarking Foundation Models for Tabular Data With Text OpenML Benchmarking Suites

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-06T18:36:46.741921Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T18:36:46.741921Z digest=sha256:cad16c0848505dc5d653467741ba9ed0b808dd9ba1dec6009de92a84a956a818

Observation 377f81f8-07a2-4e19-8c55-b165bf9a981d · outbound

This paper cites Enriching Word Vectors with Subword Information.

Towards Benchmarking Foundation Models for Tabular Data With Text Enriching Word Vectors with Subword Information

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-06T18:36:46.868838Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T18:36:46.868838Z digest=sha256:6c065f3f8d8d4e4ed9829907fd420449e9e426217980b95142e725e6a4125ef4

Observation 483d32a2-ff1e-4cbd-a28b-b3533f925e72 · outbound

This paper cites V., Na, L., Ma, Y., Boussioux, L., Zeng, C., Soenksen, L.

Towards Benchmarking Foundation Models for Tabular Data With Text V., Na, L., Ma, Y., Boussioux, L., Zeng, C., Soenksen, L

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-06T18:36:47.052197Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T18:36:47.052197Z digest=sha256:471dd1d7f2962e2eb8c41a28f58d8fa1cb71e9ed5eabd7981794f9daadd6b1d6

Observation 877e9463-acaa-49d1-abe0-80df7238f766 · outbound

This paper cites and Guestrin, C.

Towards Benchmarking Foundation Models for Tabular Data With Text and Guestrin, C

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:36:48.592767Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=arxiv_source observed=2026-08-06T18:36:47.160510Z digest=sha256:644479e2a20ab79d13d5631c654e2ed96354ed095e6694a476c64b2711a7b0ed

Observation da1cfbfd-2367-4aa7-b9d3-abe8e32e0193 · outbound

This paper cites AutoGluon-Tabular: Robust and Accurate AutoML for Structured Data.

Towards Benchmarking Foundation Models for Tabular Data With Text AutoGluon-Tabular: Robust and Accurate AutoML for Structured Data

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-06T18:36:47.379543Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T18:36:47.379543Z digest=sha256:5e9767e1403a3be2b158304ecb3cae36e1cefd3af3445201df19353ea9c6eaeb

Observation fd806105-fa00-42c0-8498-182df74f00b6 · outbound

This paper cites AMLB: an AutoML Benchmark.

Towards Benchmarking Foundation Models for Tabular Data With Text AMLB: an AutoML Benchmark

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-06T18:36:47.491670Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T18:36:47.491670Z digest=sha256:46ad158aba6795264c82076f70150dee64c3473bc49f2248e9cb9fc22b3928d0

Observation bacee3e0-ffd2-49d5-a053-b5645203fb63 · outbound

This paper cites L., Amaral, L.

Towards Benchmarking Foundation Models for Tabular Data With Text L., Amaral, L

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:36:48.572115Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=arxiv_source observed=2026-08-06T18:36:47.546239Z digest=sha256:1699fc2c00116859de0764c4a5293c12c83db56272a3490cb3f69ac64f17b8e3

Observation 397e6b2c-c2bc-48c7-9a81-fad9763d348b · outbound

This paper cites Vectorizing string entries for data processing on tables: when are larger language models better?.

Towards Benchmarking Foundation Models for Tabular Data With Text Vectorizing string entries for data processing on tables: when are larger language models better?

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-06T18:36:47.624023Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T18:36:47.624023Z digest=sha256:c5f9a20307ac3e813c558aba70ebab100b55d1459125fca10a318583af86ae4b

Observation 553da31f-b73a-417e-bc29-3f3a0fffff2d · outbound

This paper cites TabLLM: Few-shot Classification of Tabular Data with Large Language Models.

Towards Benchmarking Foundation Models for Tabular Data With Text TabLLM: Few-shot Classification of Tabular Data with Large Language Models

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-06T18:36:47.741196Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T18:36:47.741196Z digest=sha256:ceb82160b89b2309aa4bdf61e63ae79333fdd7dac4577966acb2f25eed12f401

Observation 2a0e2f26-28a3-4db7-99dd-1e2047a4d95b · outbound

This paper cites Machine Learning for Health symposium 2023 -- Findings track.

Towards Benchmarking Foundation Models for Tabular Data With Text Machine Learning for Health symposium 2023 -- Findings track

Reference 12

Resolution
metadata mismatch
local_arxiv, observed 2026-08-06T18:36:48.343264Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=arxiv_source observed=2026-08-06T18:36:47.771782Z digest=sha256:36d60efbe89ef5b22aa57ecb2f7fde14f41d1f8e8d6d988d75e7ee5be69b9cce

Observation 3d999283-7563-4363-a603-ddc2a8712cc4 · outbound

This paper cites u ller, S., Purucker, L., Krishnakumar, A., K \.

Towards Benchmarking Foundation Models for Tabular Data With Text u ller, S., Purucker, L., Krishnakumar, A., K \

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-06T18:36:47.776363Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T18:36:47.776363Z digest=sha256:15232a4b1acc62ef2bc581bd95d8aae50a3173e2ba01dfbb4b027545ee4082eb

Observation 2f229c0d-7b38-45e0-871a-c7df06729364 · outbound

This paper cites CARTE: Pretraining and Transfer for Tabular Learning.

Towards Benchmarking Foundation Models for Tabular Data With Text CARTE: Pretraining and Transfer for Tabular Learning

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-06T18:36:47.781086Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T18:36:47.781086Z digest=sha256:97075964565d6889c43c5df281eed41027ac20d9e72323850cf79887ee49e537

Observation 4cfe51fd-5639-4fb0-b60c-704ef2f0020a · outbound

This paper cites LLM Embeddings for Deep Learning on Tabular Data.

Towards Benchmarking Foundation Models for Tabular Data With Text LLM Embeddings for Deep Learning on Tabular Data

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-06T18:36:47.785860Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T18:36:47.785860Z digest=sha256:5acaeffca499f2d1e2e47327d5542fe18e8bf01abc2032074c0f4693d12fb798

Observation f2b6bb55-8deb-414c-8e46-6085c0fd301d · outbound

This paper cites TALENT: A Tabular Analytics and Learning Toolbox.

Towards Benchmarking Foundation Models for Tabular Data With Text TALENT: A Tabular Analytics and Learning Toolbox

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-06T18:36:47.790726Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T18:36:47.790726Z digest=sha256:aa12b3b7d35c6b4ce9340c140b9887fa345b59d38a3d8191b51585391b196cd4

Observation d212f918-239c-42a2-a6cb-30ab6f33bb59 · outbound

This paper cites Mug: A multimodal classification benchmark on game data with tabular, textual, and visual fields.

Towards Benchmarking Foundation Models for Tabular Data With Text Mug: A multimodal classification benchmark on game data with tabular, textual, and visual fields

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-06T18:36:47.796759Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T18:36:47.796759Z digest=sha256:8eb7823fe38d259149435d57a22f76b0361cc4abd8f53f4e3b66ece9aae8e011

Observation 0e193bd7-f170-4e32-afac-bd6d0fee7595 · outbound

This paper cites an unresolved cited work.

Towards Benchmarking Foundation Models for Tabular Data With Text Unresolved cited work

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-06T18:36:47.801181Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T18:36:47.801181Z digest=sha256:43b7028eff4c81441562d9c55a34591cae7024485d5d53bd5ee88b1914b9ee8b

Observation dd1bc50d-5e66-4035-9d14-95163b290aee · outbound

This paper cites C., Golestan, K., Yu, G., Volkovs, M., and Caterini, A.

Towards Benchmarking Foundation Models for Tabular Data With Text C., Golestan, K., Yu, G., Volkovs, M., and Caterini, A

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-06T18:36:47.805612Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T18:36:47.805612Z digest=sha256:4167ddfd011357d87ff9116353954b387f56e7577348415694dc4bc19fed7e49

Observation de99a1bf-5ee9-4548-b326-cb14bf0be713 · outbound

This paper cites and Ratajczak, W.

Towards Benchmarking Foundation Models for Tabular Data With Text and Ratajczak, W

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-06T18:36:47.923586Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T18:36:47.923586Z digest=sha256:ba6bc5989b44f00e3750db1fdd5e672be4e473acfa072a8db2d4b237f9f4a61c

Observation b77ceb41-ea15-4e30-a9a0-62a0e3e34b3d · outbound

This paper cites an unresolved cited work.

Towards Benchmarking Foundation Models for Tabular Data With Text Unresolved cited work

Reference 21

Resolution
unresolved
raw_fallback, observed 2026-08-06T18:36:48.542595Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=arxiv_source observed=2026-08-06T18:36:47.928885Z digest=sha256:25e8e7dfb0a580f5106ea5c0cfc0481b8abc2019ac6692440c34b67a58973d45

Observation 3c40f97a-7d09-4260-87f3-95a17bc30202 · outbound

This paper cites When Do Neural Nets Outperform Boosted Trees on Tabular Data?.

Towards Benchmarking Foundation Models for Tabular Data With Text When Do Neural Nets Outperform Boosted Trees on Tabular Data?

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-06T18:36:47.933777Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T18:36:47.933777Z digest=sha256:299d5938546f891cabe58473421004b955e9eb16c32b502b0c726adc0db6f2ce

Observation 20f4a16d-b51d-445d-b8e4-85069cf05237 · outbound

This paper cites Benchmarking Multimodal AutoML for Tabular Data with Text Fields.

Towards Benchmarking Foundation Models for Tabular Data With Text Benchmarking Multimodal AutoML for Tabular Data with Text Fields

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-06T18:36:47.938962Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T18:36:47.938962Z digest=sha256:dc6a1d50dfbe6cf90f2a92721049b303ccfe5761023107a5c443b6cfd3fd71fa

Observation 68150fff-f389-4666-be6f-8c868c976331 · outbound

This paper cites JoLT: Joint Probabilistic Predictions on Tabular Data Using LLMs.

Towards Benchmarking Foundation Models for Tabular Data With Text JoLT: Joint Probabilistic Predictions on Tabular Data Using LLMs

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-06T18:36:47.943525Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T18:36:47.943525Z digest=sha256:036cee2e22537a058c5b74756f762e5fde3c0313c4f5a8672c37fe99bcf09b89

Observation b7cd1408-79a6-4cfc-98cf-5c62396afbdd · outbound

This paper cites AutoGluon-Multimodal (AutoMM): Supercharging Multimodal AutoML with Foundation Models.

Towards Benchmarking Foundation Models for Tabular Data With Text AutoGluon-Multimodal (AutoMM): Supercharging Multimodal AutoML with Foundation Models

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-06T18:36:47.948093Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T18:36:47.948093Z digest=sha256:5325dd2cdb910654afe91fa9b23cd7881e7f98fe073b46f1829eaeaa9fbd605d

Observation 39807671-9e31-42b6-b726-dd81eba8cd87 · outbound

This paper cites MiniLM: Deep Self-Attention Distillation for Task-Agnostic Compression of Pre-Trained Transformers.

Towards Benchmarking Foundation Models for Tabular Data With Text MiniLM: Deep Self-Attention Distillation for Task-Agnostic Compression of Pre-Trained Transformers

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-06T18:36:47.952726Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T18:36:47.952726Z digest=sha256:676e7fe52d12347a93305af623e56636c420fe8d8e1f16e6aedad1f776caf0cb

Observation df858277-3b94-4495-b64a-1bbc205d457e · outbound

This paper cites TableLLM: Enabling Tabular Data Manipulation by LLMs in Real Office Usage Scenarios.

Towards Benchmarking Foundation Models for Tabular Data With Text TableLLM: Enabling Tabular Data Manipulation by LLMs in Real Office Usage Scenarios

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-06T18:36:47.958225Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T18:36:47.958225Z digest=sha256:94904c3114b6c2499b846b41491254f3dc56c0253be4bb91d707c31f252e4801

Pith citing papers

Observation 036e564d-2896-4d9d-a8c4-e74202eb4aa7 · inbound

STRABLE: Benchmarking Tabular Machine Learning with Strings cites this paper.

STRABLE: Benchmarking Tabular Machine Learning with Strings Towards Benchmarking Foundation Models for Tabular Data With Text

Reference 42

Resolution
verified exact
arxiv_id, observed 2026-05-13T05:17:18.441754Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-05-13T05:13:15.039160Z digest=sha256:a53b99c932fd3248752406d9cff4c75ad3baf1e8ba182b9c69e340d4b4c8aa80

Observation 378dff40-5ec7-4857-906e-3d4ce8efecff · inbound

Beyond IID: How General Are Tabular Foundation Models, Really? cites this paper.

Beyond IID: How General Are Tabular Foundation Models, Really? Towards Benchmarking Foundation Models for Tabular Data With Text

Reference 16

Resolution
verified exact
arxiv_id, observed 2026-06-30T07:04:21.420092Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-06-30T06:59:14.626274Z digest=sha256:7c8685cb632a8d49eb60b80b7f31b94a0a36586a3d09a32bcbfcde872d49f37f

Observation 4a372cec-af06-47c5-9513-4821a51dd7f5 · inbound

Exploring Differences Between Tabular Enterprise Data and Public Benchmarks cites this paper.

Exploring Differences Between Tabular Enterprise Data and Public Benchmarks Towards Benchmarking Foundation Models for Tabular Data With Text

Reference 10

Resolution
metadata mismatch
arxiv_id, observed 2026-06-30T08:04:29.091444Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=arxiv_source observed=2026-06-30T07:34:47.760875Z digest=sha256:b9523c602ecf2ab2a74705b9a4c92d5a09c80255b9890ee547a4a17617c5c913