Pith. sign in

Paper Citation Record · LEDGER

CLEVR-Math: A Dataset for Compositional Language, Visual and Mathematical Reasoning

As of 9 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 28 inbound Pith citation observations for arXiv:2208.05358.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2208.05358 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 28 of 28 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00

measured 28 of 28 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T20:06:36.633214Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-01T23:26:23.166439Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 0c50f2c5-2587-463f-9c0e-b6553ccf1129 · inbound

How Far Are We to GPT-4V? Closing the Gap to Commercial Multimodal Models with Open-Source Suites cites this paper.

How Far Are We to GPT-4V? Closing the Gap to Commercial Multimodal Models with Open-Source Suites CLEVR-Math: A Dataset for Compositional Language, Visual and Mathematical Reasoning

Reference 58

Resolution
verified exact
arxiv_id, observed 2026-05-12T20:58:59.072590Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-12T20:58:58.849040Z digest=sha256:ffd6e37888f1026ea1e6bbebc9531c4ea663165fd7ee97d382d886e41f1d41b3

Observation 0aabe920-1575-433f-9466-556cfbcaf193 · inbound

InternLM-XComposer-2.5: A Versatile Large Vision Language Model Supporting Long-Contextual Input and Output cites this paper.

InternLM-XComposer-2.5: A Versatile Large Vision Language Model Supporting Long-Contextual Input and Output CLEVR-Math: A Dataset for Compositional Language, Visual and Mathematical Reasoning

Reference 81

Resolution
verified exact
arxiv_id, observed 2026-05-17T10:46:28.709257Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-17T10:46:28.447347Z digest=sha256:a9e1796babc61231f92740924879128e09ad749177c04ec96120f39c114afaf2

Observation c8a137e7-1bcb-43f7-9200-52495be5a626 · inbound

Enhancing the Reasoning Ability of Multimodal Large Language Models via Mixed Preference Optimization cites this paper.

Enhancing the Reasoning Ability of Multimodal Large Language Models via Mixed Preference Optimization CLEVR-Math: A Dataset for Compositional Language, Visual and Mathematical Reasoning

Reference 52

Resolution
verified exact
arxiv_id, observed 2026-05-16T09:16:17.274108Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-16T09:16:17.150383Z digest=sha256:516297f8d32eea5fc8a9ed9cfdc07d3aaa2a0ec856ece413c5017a8900b6af5f

Observation 41b5613c-370e-4814-8adf-55f401ee457f · inbound

Expanding Performance Boundaries of Open-Source Multimodal Models with Model, Data, and Test-Time Scaling cites this paper.

Expanding Performance Boundaries of Open-Source Multimodal Models with Model, Data, and Test-Time Scaling CLEVR-Math: A Dataset for Compositional Language, Visual and Mathematical Reasoning

Reference 144

Resolution
verified exact
arxiv_id, observed 2026-05-10T13:23:58.300967Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-10T13:23:57.588851Z digest=sha256:16557a0420f29e7dfaad513395d099c5e97e178a0f7a4dceaeffd395840be992

Observation a2cff039-6b07-4082-a7f2-b173d14f87da · inbound

TaskGalaxy: Scaling Multi-modal Instruction Fine-tuning with Tens of Thousands Vision Task Types cites this paper.

TaskGalaxy: Scaling Multi-modal Instruction Fine-tuning with Tens of Thousands Vision Task Types CLEVR-Math: A Dataset for Compositional Language, Visual and Mathematical Reasoning

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-07T20:06:36.633214Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T20:06:36.633214Z digest=sha256:37215c3e4e19232cde2796badea7a043cbedba27b24deeaba1b4128faab906e9

Observation 77e8777d-4cd4-4455-a64f-3b5e89ca1698 · inbound

LMM-R1: Empowering 3B LMMs with Strong Reasoning Abilities Through Two-Stage Rule-Based RL cites this paper.

LMM-R1: Empowering 3B LMMs with Strong Reasoning Abilities Through Two-Stage Rule-Based RL CLEVR-Math: A Dataset for Compositional Language, Visual and Mathematical Reasoning

Reference 38

Resolution
verified exact
arxiv_id, observed 2026-05-16T15:15:46.403002Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-16T15:15:46.255296Z digest=sha256:d6809d256f2604ab439a05e9dc757cfe20809e3d3a1ffff9e13ef45381f27ea2

Observation 027b7536-7492-4f42-9d0e-d2ceb2612068 · inbound

InternVL3: Exploring Advanced Training and Test-Time Recipes for Open-Source Multimodal Models cites this paper.

InternVL3: Exploring Advanced Training and Test-Time Recipes for Open-Source Multimodal Models CLEVR-Math: A Dataset for Compositional Language, Visual and Mathematical Reasoning

Reference 72

Resolution
verified exact
arxiv_id, observed 2026-05-10T13:41:08.372728Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-10T13:41:07.991012Z digest=sha256:39c98ada8e31c976dd674940c250f18871a977c1545b24ee50b6e115efc88b77

Observation b417db3f-86af-4854-94a3-596d2ebdd857 · inbound

Advancing Multimodal Reasoning via Reinforcement Learning with Cold Start cites this paper.

Advancing Multimodal Reasoning via Reinforcement Learning with Cold Start CLEVR-Math: A Dataset for Compositional Language, Visual and Mathematical Reasoning

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-07T13:12:54.565717Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:12:54.565717Z digest=sha256:bf00a6a9e9854518c60d12d0483b307f8183dfca679df2ec3472df078c202e3a

Observation 5f02627a-643d-4853-a0f2-9e91805af65d · inbound

Vision-EKIPL: External Knowledge-Infused Policy Learning for Visual Reasoning cites this paper.

Vision-EKIPL: External Knowledge-Infused Policy Learning for Visual Reasoning CLEVR-Math: A Dataset for Compositional Language, Visual and Mathematical Reasoning

Reference 38

Resolution
verified exact
arxiv_id, observed 2026-05-19T10:37:15.011268Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-19T10:34:48.849524Z digest=sha256:8aab9ff53855fcc24eae562cc63150791523a58b5c70999aa718b2cbfff9b567

Observation eabeb259-7914-4223-8ee5-9a69dae88dbc · inbound

WeThink: Toward General-purpose Vision-Language Reasoning via Reinforcement Learning cites this paper.

WeThink: Toward General-purpose Vision-Language Reasoning via Reinforcement Learning CLEVR-Math: A Dataset for Compositional Language, Visual and Mathematical Reasoning

Reference 79

Resolution
unresolved
no resolver link, observed 2026-08-07T05:27:05.200325Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:27:05.200325Z digest=sha256:da9114c2dd680d8026d4e2d1dc30c456c40cd6156dab5ae52e3e68a0422641fd

Observation 04c94780-f903-44a1-811f-6c40ec8063e4 · inbound

FinLMM-R1: Enhancing Financial Reasoning in LMM through Scalable Data and Reward Design cites this paper.

FinLMM-R1: Enhancing Financial Reasoning in LMM through Scalable Data and Reward Design CLEVR-Math: A Dataset for Compositional Language, Visual and Mathematical Reasoning

Reference 85

Resolution
unresolved
no resolver link, observed 2026-08-07T00:40:47.123875Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:40:47.123875Z digest=sha256:54ead1d99aecae6560ce1d9c822c9ea8c85e475f30a524edc48ad49d5631833c

Observation 65f50570-1ed8-4434-a038-0a613d7da31b · inbound

Mono-InternVL-1.5: Towards Cheaper and Faster Monolithic Multimodal Large Language Models cites this paper.

Mono-InternVL-1.5: Towards Cheaper and Faster Monolithic Multimodal Large Language Models CLEVR-Math: A Dataset for Compositional Language, Visual and Mathematical Reasoning

Reference 84

Resolution
unresolved
no resolver link, observed 2026-08-06T16:50:04.072385Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:50:04.072385Z digest=sha256:0666b9cdff71acfafc140a3dda4b76cdd1c1d37d99f1e3e904eb2b1e394529f6

Observation a272faa2-9efb-4ed2-af7d-868140e90a10 · inbound

VL-Cogito: Progressive Curriculum Reinforcement Learning for Advanced Multimodal Reasoning cites this paper.

VL-Cogito: Progressive Curriculum Reinforcement Learning for Advanced Multimodal Reasoning CLEVR-Math: A Dataset for Compositional Language, Visual and Mathematical Reasoning

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-06T11:35:13.544048Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T11:35:13.544048Z digest=sha256:a10a1fe7468dd0b391374e28ca4a0db0b41fa533e5e99a6612b665698381d9c6

Observation c31eb727-db27-4436-a80c-b3ffc4035120 · inbound

LaRe: Latent Refocusing for Multimodal Reasoning cites this paper.

LaRe: Latent Refocusing for Multimodal Reasoning CLEVR-Math: A Dataset for Compositional Language, Visual and Mathematical Reasoning

Reference 104

Resolution
unresolved
no resolver link, observed 2026-08-04T00:15:22.380554Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T00:15:22.380554Z digest=sha256:5c09bf9b0de8c1d0ad4ef8a32ceeedaba82e588142ec8eb550e1fccaff9c289e

Observation 487b835a-0e14-4146-bd88-5a68ae8ab72b · inbound

NVIDIA Nemotron 3: Efficient and Open Intelligence cites this paper.

NVIDIA Nemotron 3: Efficient and Open Intelligence CLEVR-Math: A Dataset for Compositional Language, Visual and Mathematical Reasoning

Reference 77

Resolution
metadata mismatch
arxiv_id, observed 2026-05-18T01:40:42.484808Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-05-18T01:40:42.190369Z digest=sha256:1cd5c755b36db6e05efc15c4a5489c37ca4bb235f01fd98e3e0079755bb7c2a9

Observation 0d714d0f-e193-44c1-8414-099d916ecd0d · inbound

Spectral Imbalance Causes Forgetting in Low-Rank Continual Adaptation cites this paper.

Spectral Imbalance Causes Forgetting in Low-Rank Continual Adaptation CLEVR-Math: A Dataset for Compositional Language, Visual and Mathematical Reasoning

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-03T06:04:27.013621Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T06:04:27.013621Z digest=sha256:ecaf40f5f475d0765d4ad9ec7448e892720c11ea12f7db0d9b845c7880bcb17a

Observation 73bdd87a-c9a2-4c61-9686-c49291556a89 · inbound

PaLMR: Towards Faithful Visual Reasoning via Multimodal Process Alignment cites this paper.

PaLMR: Towards Faithful Visual Reasoning via Multimodal Process Alignment CLEVR-Math: A Dataset for Compositional Language, Visual and Mathematical Reasoning

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-02T19:57:31.647125Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T19:57:31.647125Z digest=sha256:bd028b1557f1a760aa014895debf1435562344b5dce237f551154384fafdd235

Observation fef882a6-be9c-4836-b062-1b43150410d4 · inbound

Open, Reliable, and Collective: A Community-Driven Framework for Tool-Using AI Agents cites this paper.

Open, Reliable, and Collective: A Community-Driven Framework for Tool-Using AI Agents CLEVR-Math: A Dataset for Compositional Language, Visual and Mathematical Reasoning

Reference 27

Resolution
unresolved
no resolver link, observed 2026-07-14T19:59:59.030585Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-07-14T19:59:59.030585Z digest=sha256:d40502e15cb7fe250ac50d7433890922d873fad6804ad531d0b3d3002702b1b5

Observation 800f413e-37e1-4a21-94e6-ebfd3de0d1fe · inbound

Quantifying and Understanding Uncertainty in Large Reasoning Models cites this paper.

Quantifying and Understanding Uncertainty in Large Reasoning Models CLEVR-Math: A Dataset for Compositional Language, Visual and Mathematical Reasoning

Reference 4

Resolution
metadata mismatch
arxiv_id, observed 2026-05-10T13:55:28.972190Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-10T13:53:00.861503Z digest=sha256:24760b1a3ba597fb2724fea19c858d1e9017613ae8b5f1b682a459a35d9262ce

Observation c6f7b924-964b-4b6d-928c-f54dc5d0333c · inbound

MAny: Merge Anything for Multimodal Continual Instruction Tuning cites this paper.

MAny: Merge Anything for Multimodal Continual Instruction Tuning CLEVR-Math: A Dataset for Compositional Language, Visual and Mathematical Reasoning

Reference 8

Resolution
verified exact
arxiv_id, observed 2026-05-10T14:00:29.531333Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-10T13:55:28.165204Z digest=sha256:e81133f8e1cd13e3a0a96c656d23669acc0c55bad4ee1128631a8a7168354b13

Observation 8b39bcc0-dd14-4972-8212-afb74da35ff3 · inbound

Targeted Exploration via Unified Entropy Control for Reinforcement Learning cites this paper.

Targeted Exploration via Unified Entropy Control for Reinforcement Learning CLEVR-Math: A Dataset for Compositional Language, Visual and Mathematical Reasoning

Reference 8

Resolution
metadata mismatch
arxiv_id, observed 2026-05-10T10:49:56.082079Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-10T10:48:27.733823Z digest=sha256:236797e557ec71599790a53676b795cde2340799851e6aecd123b0d96664e8ce

Observation 936ecdef-7328-4b2c-9786-d53c4df5de49 · inbound

Dynamic Cross-Modal Prompt Generation for Multimodal Continual Instruction Tuning cites this paper.

Dynamic Cross-Modal Prompt Generation for Multimodal Continual Instruction Tuning CLEVR-Math: A Dataset for Compositional Language, Visual and Mathematical Reasoning

Reference 22

Resolution
verified exact
arxiv_id, observed 2026-05-12T07:21:25.170678Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-12T03:27:45.784361Z digest=sha256:983a124c36ed9844d3a82519bf0ac5dee1d226586944d8b5282f39aefbe9e5c5

Observation 0adaa04e-1617-40fd-8114-4190be07999d · inbound

Octopus: History-Free Gradient Orthogonalization for Continual Learning in Multimodal Large Language Models cites this paper.

Octopus: History-Free Gradient Orthogonalization for Continual Learning in Multimodal Large Language Models CLEVR-Math: A Dataset for Compositional Language, Visual and Mathematical Reasoning

Reference 38

Resolution
verified exact
arxiv_id, observed 2026-07-01T14:35:46.670607Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-06-30T20:56:38.324342Z digest=sha256:eab4f4f9e97c9a6a5f6b3d34e5582be6c3b4062fbdf20cdc1e5261671d81185c

Observation 8161aa63-f862-49da-ae67-ed2dca173ef3 · inbound

From Seeing to Thinking: Decoupling Perception and Reasoning Improves Post-Training of Vision-Language Models cites this paper.

From Seeing to Thinking: Decoupling Perception and Reasoning Improves Post-Training of Vision-Language Models CLEVR-Math: A Dataset for Compositional Language, Visual and Mathematical Reasoning

Reference 15

Resolution
verified exact
arxiv_id, observed 2026-05-20T05:13:21.543836Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-20T05:13:03.237427Z digest=sha256:5964fe9329b60a19c2fc30f44cc1fa5196bb3342dcc843d9b312540912c54474

Observation 8a963dff-6db5-47e0-92e0-6f16744b0366 · inbound

Prism: A Plug-in Reproducible Infrastructure for Scalable Multimodal Continual Instruction Tuning cites this paper.

Prism: A Plug-in Reproducible Infrastructure for Scalable Multimodal Continual Instruction Tuning CLEVR-Math: A Dataset for Compositional Language, Visual and Mathematical Reasoning

Reference 3

Resolution
metadata mismatch
arxiv_id, observed 2026-06-29T23:14:01.558823Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-06-29T23:10:19.308836Z digest=sha256:3285c9c41295ad24f88f4b26f1eb15b8ad4e8715587d0854c18b3d29af403498

Observation 6b520ef2-1013-44ed-b828-f8091164f932 · inbound

CRAM: Centroid-Routing and Adaptive MoE for Multimodal Continual Instruction Tuning cites this paper.

CRAM: Centroid-Routing and Adaptive MoE for Multimodal Continual Instruction Tuning CLEVR-Math: A Dataset for Compositional Language, Visual and Mathematical Reasoning

Reference 55

Resolution
metadata mismatch
arxiv_id, observed 2026-07-01T23:26:23.168349Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-06-28T14:18:11.452378Z digest=sha256:1c813eac8535146316e0f665a98e570277fe2a79f1ab5f77a1675054ffa13b03

Observation ddee5902-6d8d-421a-a467-3c2e4bc7c263 · inbound

ProtoAda: Prototype-Guided Adaptive Adapter Expansion and Geometric Consolidation for Multimodal Continual Instruction Tuning cites this paper.

ProtoAda: Prototype-Guided Adaptive Adapter Expansion and Geometric Consolidation for Multimodal Continual Instruction Tuning CLEVR-Math: A Dataset for Compositional Language, Visual and Mathematical Reasoning

Reference 55

Resolution
metadata mismatch
arxiv_id, observed 2026-07-01T22:36:17.984812Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-06-28T15:09:00.320198Z digest=sha256:ebf5337408b57c56a7da78ffb0dbc2d777fbeb2d7e08ce83e5d775b15f17dc34

Observation bdfa2952-9e42-41bc-beb7-1fb5663bdd04 · inbound

ParVL: Parallel Scaling and Expandable Compute Allocation for Multimodal LLMs cites this paper.

ParVL: Parallel Scaling and Expandable Compute Allocation for Multimodal LLMs CLEVR-Math: A Dataset for Compositional Language, Visual and Mathematical Reasoning

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-05T04:16:08.055656Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T04:16:08.055656Z digest=sha256:1072b115cca8ba6941dc04d128ef30e41e2d0fa5869617241e9539da6b451ead