Pith. sign in

Paper Citation Record · LEDGER

Decoupling Skeleton and Flesh: Efficient Multimodal Table Reasoning with Disentangled Alignment and Structure-aware Guidance

As of 8 August 2026, this Paper Citation Record lists 40 of 40 outbound references and 1 inbound Pith citation observation for arXiv:2602.03491.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2602.03491 v2

Coverage vector

measured 40 of 40 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-03T05:02:45.938777Z

measured 41 of 41 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00

measured 1 of 1 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-05-10T05:10:45.144421Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-10T09:38:43.272359Z

Reference resolution

40 of 40 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved40
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation a96168ff-0864-409e-9274-5187f1021601 · outbound

This paper cites write newline.

Decoupling Skeleton and Flesh: Efficient Multimodal Table Reasoning with Disentangled Alignment and Structure-aware Guidance write newline

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-03T05:02:40.374433Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-03T05:02:40.374433Z digest=sha256:0f98429cb89358d6edff3fdb466fdae293a4ccf7443d87e431a78b7730b5eaad

Observation 3c386893-da49-4d43-a0ac-b47dd6f68971 · outbound

This paper cites P ub H ealth T ab: A public health table-based dataset for evidence-based fact checking.

Decoupling Skeleton and Flesh: Efficient Multimodal Table Reasoning with Disentangled Alignment and Structure-aware Guidance P ub H ealth T ab: A public health table-based dataset for evidence-based fact checking

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-03T05:02:40.444113Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-03T05:02:40.444113Z digest=sha256:ed25ee3eb68bf89fc36619ed25524f726e8ddebd2f0c9c22b2644f7248620f70

Observation 366fc016-2765-49be-b27e-805256faa7e8 · outbound

This paper cites Flamingo: a visual language model for few-shot learning.

Decoupling Skeleton and Flesh: Efficient Multimodal Table Reasoning with Disentangled Alignment and Structure-aware Guidance Flamingo: a visual language model for few-shot learning

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-03T05:02:40.527472Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-03T05:02:40.527472Z digest=sha256:c213038e4c63abe4c97a04c0db597e0bc863618b51cf6c31eb7eded59fb09249

Observation eb666695-a8e2-49f6-a2db-7e3bfd543aac · outbound

This paper cites Qwen3-VL Technical Report.

Decoupling Skeleton and Flesh: Efficient Multimodal Table Reasoning with Disentangled Alignment and Structure-aware Guidance Qwen3-VL Technical Report

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-03T05:02:40.578669Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-03T05:02:40.578669Z digest=sha256:81018bfeb24828bc48f5c1711a9f8ee8578b48b5cd4de092ab36e1438c8da7ce

Observation 0560699f-674c-4eb8-bf64-da53fc79cdbc · outbound

This paper cites Qwen2.5-VL Technical Report.

Decoupling Skeleton and Flesh: Efficient Multimodal Table Reasoning with Disentangled Alignment and Structure-aware Guidance Qwen2.5-VL Technical Report

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-03T05:02:40.827814Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-03T05:02:40.827814Z digest=sha256:9e4cbce2d7d9c6a326296e7dba60e03ff18f6066458e13852cc739bb75308c5d

Observation 5c8a4522-171e-4b93-97b2-05af100dbf80 · outbound

This paper cites Deep neural networks and tabular data: A survey.

Decoupling Skeleton and Flesh: Efficient Multimodal Table Reasoning with Disentangled Alignment and Structure-aware Guidance Deep neural networks and tabular data: A survey

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-03T05:02:40.921830Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-03T05:02:40.921830Z digest=sha256:96951c668a9ea94027a0f775ec6a01340d047d77fbd08141ee217fb4263f4807

Observation 63e546db-d8d2-4c1a-ae08-08c15baf63d9 · outbound

This paper cites an unresolved cited work.

Decoupling Skeleton and Flesh: Efficient Multimodal Table Reasoning with Disentangled Alignment and Structure-aware Guidance Unresolved cited work

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-03T05:02:40.984984Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-03T05:02:40.984984Z digest=sha256:2e97e90b8bf480a6c174d10ea042cc1451de8b1249a483bb11d354b7a442eb7a

Observation 74e7b04e-88ee-4298-bf87-5c31cb86346c · outbound

This paper cites an unresolved cited work.

Decoupling Skeleton and Flesh: Efficient Multimodal Table Reasoning with Disentangled Alignment and Structure-aware Guidance Unresolved cited work

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-03T05:02:41.081979Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-03T05:02:41.081979Z digest=sha256:b8a869c1d4fc9d051c582b7c1fb4500f1ee735612a781b097ff59e6f312bc100

Observation 189f5409-6c27-49fd-8262-49e08b6c6b5d · outbound

This paper cites Vision-language models can self-improve reasoning via reflection.

Decoupling Skeleton and Flesh: Efficient Multimodal Table Reasoning with Disentangled Alignment and Structure-aware Guidance Vision-language models can self-improve reasoning via reflection

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-03T05:02:41.153759Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-03T05:02:41.153759Z digest=sha256:11ae92931aa8b2c08d5d4d2b60e7b4e60567c4e464fb5cf26703d94f3232e8db

Observation 3618a62e-e520-4a98-87e4-6dc1c00d95ef · outbound

This paper cites H i T ab: A hierarchical table dataset for question answering and natural language generation.

Decoupling Skeleton and Flesh: Efficient Multimodal Table Reasoning with Disentangled Alignment and Structure-aware Guidance H i T ab: A hierarchical table dataset for question answering and natural language generation

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-03T05:02:41.315577Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-03T05:02:41.315577Z digest=sha256:1c9e0a35ba2f6a0cf355ebce4db04055fd423ca9d413bc6fa014b8c98b011b75

Observation 278ea329-4517-4677-ab09-dc1d8db7b25d · outbound

This paper cites R., Lu, Y., Yang, J., Roth, D., Florencio, D., and Zhang, C.

Decoupling Skeleton and Flesh: Efficient Multimodal Table Reasoning with Disentangled Alignment and Structure-aware Guidance R., Lu, Y., Yang, J., Roth, D., Florencio, D., and Zhang, C

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-03T05:02:41.438800Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-03T05:02:41.438800Z digest=sha256:bea1264a952e95767b0da8ebb7cd629d2807cf569d5532ca89d9b6784f43a845

Observation 7b69bafd-3088-4011-a0c5-0878cf80f3ac · outbound

This paper cites Hallusionbench: an advanced diagnostic suite for entangled language hallucination and visual illusion in large vision-language models.

Decoupling Skeleton and Flesh: Efficient Multimodal Table Reasoning with Disentangled Alignment and Structure-aware Guidance Hallusionbench: an advanced diagnostic suite for entangled language hallucination and visual illusion in large vision-language models

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-03T05:02:41.541256Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-03T05:02:41.541256Z digest=sha256:9cc8d9c652ab212975b21b5e48bbdef934228b3c6cfcc7d265f0fc2b645ca48e

Observation d398df8f-c3c9-460b-b66d-9bcc584ca12c · outbound

This paper cites INFOTABS : Inference on tables as semi-structured data.

Decoupling Skeleton and Flesh: Efficient Multimodal Table Reasoning with Disentangled Alignment and Structure-aware Guidance INFOTABS : Inference on tables as semi-structured data

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-03T05:02:41.654242Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-03T05:02:41.654242Z digest=sha256:700e6063ad26b279df812da55896518c6b504c7f976170bd8756b9e58df5248a

Observation f1984095-b767-4a48-b931-e49c4a03aeea · outbound

This paper cites K., M \"u ller, T., Piccinno, F., and Eisenschlos, J.

Decoupling Skeleton and Flesh: Efficient Multimodal Table Reasoning with Disentangled Alignment and Structure-aware Guidance K., M \"u ller, T., Piccinno, F., and Eisenschlos, J

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-03T05:02:41.867758Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-03T05:02:41.867758Z digest=sha256:da335da1b0ee8ec77db68d65a254d40deac86a29ab99d7d04326d87c67b637fa

Observation c839312f-92f2-4d96-ba62-0662538747ec · outbound

This paper cites J., yelong shen, Wallis, P., Allen-Zhu, Z., Li, Y., Wang, S., Wang, L., and Chen, W.

Decoupling Skeleton and Flesh: Efficient Multimodal Table Reasoning with Disentangled Alignment and Structure-aware Guidance J., yelong shen, Wallis, P., Allen-Zhu, Z., Li, Y., Wang, S., Wang, L., and Chen, W

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-03T05:02:42.032337Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-03T05:02:42.032337Z digest=sha256:e282cbcfb2ddb98547ad78e126ce6a8b084d522b5cb4484378a196fb80ae80b0

Observation d17f7b36-2fff-43af-a08a-6dc4f2bdfbdc · outbound

This paper cites TABBIE : Pretrained representations of tabular data.

Decoupling Skeleton and Flesh: Efficient Multimodal Table Reasoning with Disentangled Alignment and Structure-aware Guidance TABBIE : Pretrained representations of tabular data

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-03T05:02:42.194034Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-03T05:02:42.194034Z digest=sha256:b4cdb1391c99069b51aafa0a56b5bfecc9e60b187786c65d7f7ef4f31da857c3

Observation cf35cd21-9e5c-45b2-bf9a-4871a7eb2d0d · outbound

This paper cites TabMCQ: A Dataset of General Knowledge Tables and Multiple-choice Questions.

Decoupling Skeleton and Flesh: Efficient Multimodal Table Reasoning with Disentangled Alignment and Structure-aware Guidance TabMCQ: A Dataset of General Knowledge Tables and Multiple-choice Questions

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-03T05:02:42.380233Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-03T05:02:42.380233Z digest=sha256:e11a6671eefee3bd4ef469c741d2573560b52dc35938ea33ef2d78d4542ff4f3

Observation 9efd305c-9966-4c44-8704-9ab1cdce17f9 · outbound

This paper cites Multimodal Tabular Reasoning with Privileged Structured Information.

Decoupling Skeleton and Flesh: Efficient Multimodal Table Reasoning with Disentangled Alignment and Structure-aware Guidance Multimodal Tabular Reasoning with Privileged Structured Information

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-03T05:02:42.510199Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-03T05:02:42.510199Z digest=sha256:b2db2c9d5eb5cf761abf90c04a2648d64234b89cce14362509f53cdce6099a3a

Observation 5b6bb6cf-e4e5-45d6-a049-4017ebd954bd · outbound

This paper cites Can GRPO boost complex multimodal table understanding? In Proceedings of the 2025 Conference on Empirical Methods in Natural Language Processing, pp.\ 12642--12655, 2025.

Decoupling Skeleton and Flesh: Efficient Multimodal Table Reasoning with Disentangled Alignment and Structure-aware Guidance Can GRPO boost complex multimodal table understanding? In Proceedings of the 2025 Conference on Empirical Methods in Natural Language Processing, pp.\ 12642--12655, 2025

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-03T05:02:42.664690Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-03T05:02:42.664690Z digest=sha256:96e7047dc575750aa4543bd6ab1fca902c039c1bcf6ca1d9a15fbe7d1fc4c70d

Observation 54f68201-6f0d-434a-8c5b-c55468a9c11d · outbound

This paper cites Ait-qa: Question answering dataset over complex tables in the airline industry.

Decoupling Skeleton and Flesh: Efficient Multimodal Table Reasoning with Disentangled Alignment and Structure-aware Guidance Ait-qa: Question answering dataset over complex tables in the airline industry

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-03T05:02:42.839670Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-03T05:02:42.839670Z digest=sha256:7dd50d4bce757178949d625adca99dcaead885242f4222c44bf9c3b892201c25

Observation 29a8e89d-d813-4c87-b71c-77e14a6386ff · outbound

This paper cites H., Gonzalez, J., Zhang, H., and Stoica, I.

Decoupling Skeleton and Flesh: Efficient Multimodal Table Reasoning with Disentangled Alignment and Structure-aware Guidance H., Gonzalez, J., Zhang, H., and Stoica, I

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-03T05:02:43.095142Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-03T05:02:43.095142Z digest=sha256:a4bf532ee60120dffa27e2f6c746a554c466761717631e32f7def7de8a210a9d

Observation fd5983d2-6b31-4ef5-9e72-064afefc68b1 · outbound

This paper cites Multimodal A r X iv: A dataset for improving scientific comprehension of large vision-language models.

Decoupling Skeleton and Flesh: Efficient Multimodal Table Reasoning with Disentangled Alignment and Structure-aware Guidance Multimodal A r X iv: A dataset for improving scientific comprehension of large vision-language models

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-03T05:02:43.298270Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-03T05:02:43.298270Z digest=sha256:a52148fe316f563b84649c89fb7e85d83167bbfbd0e4243fd596b77bc6ef84fd

Observation bc47a30c-2318-4b94-8cc1-e63b557e98ae · outbound

This paper cites an unresolved cited work.

Decoupling Skeleton and Flesh: Efficient Multimodal Table Reasoning with Disentangled Alignment and Structure-aware Guidance Unresolved cited work

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-03T05:02:43.510822Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-03T05:02:43.510822Z digest=sha256:f6e00e4b91851e51a60ff3aa69a87cf019aec47eda871672aa808d766eb65152

Observation b4c77cf9-7e8c-4bda-a361-625ca7d21fb8 · outbound

This paper cites an unresolved cited work.

Decoupling Skeleton and Flesh: Efficient Multimodal Table Reasoning with Disentangled Alignment and Structure-aware Guidance Unresolved cited work

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-03T05:02:43.695962Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-03T05:02:43.695962Z digest=sha256:d4780710ca7d5c8895f8de946e68adb809e90d4e865e001f31509e9680b7002e

Observation c25d1b99-9ee3-4bd2-b01b-85a1fcc8a7eb · outbound

This paper cites TAPEX : Table pre-training via learning a neural SQL executor.

Decoupling Skeleton and Flesh: Efficient Multimodal Table Reasoning with Disentangled Alignment and Structure-aware Guidance TAPEX : Table pre-training via learning a neural SQL executor

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-03T05:02:43.854060Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-03T05:02:43.854060Z digest=sha256:ecc0257e9ed1fc68817bbc82a04267b95dc48026a53ff962ec0f3e624277ef7c

Observation 3e461837-2994-4956-9a81-fd75eb3f9343 · outbound

This paper cites Hippo: Enhancing the table understanding capability of large language models through hybrid-modal preference optimization.

Decoupling Skeleton and Flesh: Efficient Multimodal Table Reasoning with Disentangled Alignment and Structure-aware Guidance Hippo: Enhancing the table understanding capability of large language models through hybrid-modal preference optimization

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-03T05:02:43.977441Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-03T05:02:43.977441Z digest=sha256:a0768fbc5f1d84d988d3edb73520cdf03d908c4f32d5c15a90fd8e60b6689ad3

Observation b21b6af7-f9d7-4bee-b368-1fa2222760e5 · outbound

This paper cites Learn to explain: Multimodal reasoning via thought chains for science question answering.

Decoupling Skeleton and Flesh: Efficient Multimodal Table Reasoning with Disentangled Alignment and Structure-aware Guidance Learn to explain: Multimodal reasoning via thought chains for science question answering

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-03T05:02:44.086960Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-03T05:02:44.086960Z digest=sha256:e290596b74ad25513fa01dd3d8de64b03f0a9661cbb198f36610bc73d8a2509e

Observation 0678052f-0e0d-4a32-b5a1-6e36f0d1389d · outbound

This paper cites Large language model for table processing: A survey.

Decoupling Skeleton and Flesh: Efficient Multimodal Table Reasoning with Disentangled Alignment and Structure-aware Guidance Large language model for table processing: A survey

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-03T05:02:44.217419Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-03T05:02:44.217419Z digest=sha256:629812084e3344b5fd0f2e12289ca27d30e720c2b7b0c34256e26b363a44b89d

Observation e5797a1b-d794-4373-8b26-97d8dba2a7cc · outbound

This paper cites and Liang, P.

Decoupling Skeleton and Flesh: Efficient Multimodal Table Reasoning with Disentangled Alignment and Structure-aware Guidance and Liang, P

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-03T05:02:44.389351Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-03T05:02:44.389351Z digest=sha256:188b4e23307a3f49c92bf94e1415f901e1bf28f648759ea5cb17a2617fca22c6

Observation bc72ec60-401b-4613-8a68-14a8bc439497 · outbound

This paper cites Towards vqa models that can read.

Decoupling Skeleton and Flesh: Efficient Multimodal Table Reasoning with Disentangled Alignment and Structure-aware Guidance Towards vqa models that can read

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-03T05:02:44.648919Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-03T05:02:44.648919Z digest=sha256:5fb07d73f2f567c5f59b20690485e37f6a9ab91db03eee6aef613459e0ac60c8

Observation 5ea8cc7e-59b2-42b5-8024-1c5c11293d03 · outbound

This paper cites Table meets llm: Can large language models understand structured table data? a benchmark and empirical study.

Decoupling Skeleton and Flesh: Efficient Multimodal Table Reasoning with Disentangled Alignment and Structure-aware Guidance Table meets llm: Can large language models understand structured table data? a benchmark and empirical study

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-03T05:02:44.879105Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-03T05:02:44.879105Z digest=sha256:6ee6c62b987be4c69fd9ad7092c613db4b118649165f8a6dc4d3f22d3ed07dc3

Observation 7c435001-45bf-4377-873d-c99e8d02a151 · outbound

This paper cites Gemma 3 Technical Report.

Decoupling Skeleton and Flesh: Efficient Multimodal Table Reasoning with Disentangled Alignment and Structure-aware Guidance Gemma 3 Technical Report

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-03T05:02:45.006686Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-03T05:02:45.006686Z digest=sha256:a4879389529d5b84cb921fa18491ac59604222962c88bf5b8209ac68d122bf53

Observation 77fd1f6a-69ac-40ff-ab25-a8e9ed3142d7 · outbound

This paper cites MMTABREAL: Real-World Benchmark for Multimodal Table Understanding.

Decoupling Skeleton and Flesh: Efficient Multimodal Table Reasoning with Disentangled Alignment and Structure-aware Guidance MMTABREAL: Real-World Benchmark for Multimodal Table Understanding

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-03T05:02:45.118121Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-03T05:02:45.118121Z digest=sha256:1fbe32e1a6e78944bbdeaead6bf3fea0a18a8e7a52d738e1f20f2cb6f89ca116

Observation 97f06c00-4c13-4968-bd7f-ad4a48ed7872 · outbound

This paper cites The all-seeing project v2: Towards general relation comprehension of the open world.

Decoupling Skeleton and Flesh: Efficient Multimodal Table Reasoning with Disentangled Alignment and Structure-aware Guidance The all-seeing project v2: Towards general relation comprehension of the open world

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-03T05:02:45.216036Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-03T05:02:45.216036Z digest=sha256:9b1bc44cf3b86ad7226440d4d49947260e9bc6a81af9aaa23a55e7ca946f1003

Observation 47e93e13-c652-4a0d-8ce1-d36f7ae29aeb · outbound

This paper cites Tuta: Tree-based transformers for generally structured table pre-training.

Decoupling Skeleton and Flesh: Efficient Multimodal Table Reasoning with Disentangled Alignment and Structure-aware Guidance Tuta: Tree-based transformers for generally structured table pre-training

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-03T05:02:45.295647Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-03T05:02:45.295647Z digest=sha256:5b7c814e02b75fcddfc1ebc6f595d7547c126d990279b4113901917fe4ca99d0

Observation a333b059-4bd0-4215-a16a-de5250f31047 · outbound

This paper cites Tabpedia: Towards comprehensive visual table understanding with concept synergy.

Decoupling Skeleton and Flesh: Efficient Multimodal Table Reasoning with Disentangled Alignment and Structure-aware Guidance Tabpedia: Towards comprehensive visual table understanding with concept synergy

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-03T05:02:45.384421Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-03T05:02:45.384421Z digest=sha256:d97bd9a42e7b51333859caa89b410dce6dbfa6768eb2a3975de90ee8cb415f9b

Observation a1709a1a-663b-417b-a496-325521b893ac · outbound

This paper cites Multimodal table understanding.

Decoupling Skeleton and Flesh: Efficient Multimodal Table Reasoning with Disentangled Alignment and Structure-aware Guidance Multimodal table understanding

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-03T05:02:45.570311Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-03T05:02:45.570311Z digest=sha256:ad3abdb721568b5719e40171497e1841cddb4c278c181cd99761df3b461696ea

Observation c6b29b89-e504-4f46-b6b8-e4cfab3b727c · outbound

This paper cites Syntab-llava: Enhancing multimodal table understanding with decoupled synthesis.

Decoupling Skeleton and Flesh: Efficient Multimodal Table Reasoning with Disentangled Alignment and Structure-aware Guidance Syntab-llava: Enhancing multimodal table understanding with decoupled synthesis

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-03T05:02:45.663169Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-03T05:02:45.663169Z digest=sha256:8bb66d0e923a91b2f12c45ae0a56d8f78d0260e332588ed6e0d9bea50e9f01de

Observation 9375986b-bb5d-496b-877a-9538da6d435b · outbound

This paper cites TAT - QA : A question answering benchmark on a hybrid of tabular and textual content in finance.

Decoupling Skeleton and Flesh: Efficient Multimodal Table Reasoning with Disentangled Alignment and Structure-aware Guidance TAT - QA : A question answering benchmark on a hybrid of tabular and textual content in finance

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-03T05:02:45.805430Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-03T05:02:45.805430Z digest=sha256:b2f8df35152eb9c1393230deb353b28275053b22667ced75efe56e2a0184e118

Observation c7303acf-20b0-44d4-9af7-c6b1a28f91a7 · outbound

This paper cites Benchmarking and improving large vision-language models for fundamental visual graph understanding and reasoning.

Decoupling Skeleton and Flesh: Efficient Multimodal Table Reasoning with Disentangled Alignment and Structure-aware Guidance Benchmarking and improving large vision-language models for fundamental visual graph understanding and reasoning

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-03T05:02:45.938777Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-03T05:02:45.938777Z digest=sha256:da87dfa5b5b5ba317f98bd323c76bbd1283294b06e8a5729f605ab06eab536ca

Pith citing papers

Observation f3904501-2dca-452b-a671-d1ec34ac2a69 · inbound

Mitigating Multimodal Hallucination via Phase-wise Self-reward cites this paper.

Mitigating Multimodal Hallucination via Phase-wise Self-reward Decoupling Skeleton and Flesh: Efficient Multimodal Table Reasoning with Disentangled Alignment and Structure-aware Guidance

Reference 65

Resolution
verified exact
arxiv_id, observed 2026-05-28T03:04:45.483203Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-10T05:10:45.144421Z digest=sha256:bbe4346bd960d3b591bc5f3d1980c7966af7420f43f156c0064b323491fc5549