Pith. sign in

Paper Citation Record · LEDGER

Task-conditioned probing of instruction-tuned multimodal LLMs: Region-specific brain alignment patterns under naturalistic stimuli

As of 20 August 2026, this Paper Citation Record lists 22 of 22 outbound references and 1 inbound Pith citation observation for arXiv:2506.08277.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2506.08277 v3

Coverage vector

measured 22 of 22 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-05-22T00:02:41.893373Z

measured 23 of 23 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-19T06:32:44.657259+00:00

measured 1 of 1 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-06-29T12:53:45.827516Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-06-29T13:03:26.839969Z

Reference resolution

22 of 22 outbound references displayed

  • verified exact7
  • verified fuzzy11
  • unresolved2
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch2

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 9247353d-2d6d-46c8-a949-2b8c31632d98 · outbound

This paper cites Phi-3 Technical Report: A Highly Capable Language Model Locally on Your Phone.

Task-conditioned probing of instruction-tuned multimodal LLMs: Region-specific brain alignment patterns under naturalistic stimuli Phi-3 Technical Report: A Highly Capable Language Model Locally on Your Phone

Reference 1

Resolution
verified exact
local_arxiv, observed 2026-05-22T00:04:27.054934Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-05-22T00:02:41.893373Z digest=sha256:801f465f0d7c01e6b9ab59cd40032f11519aca803a46f9374d4eedf954c1430f

Observation e4d9f9e8-464f-419b-9dff-4098bbb04d87 · outbound

This paper cites Mae-ast: Masked autoencoding audio spectrogram transformer.

Task-conditioned probing of instruction-tuned multimodal LLMs: Region-specific brain alignment patterns under naturalistic stimuli Mae-ast: Masked autoencoding audio spectrogram transformer

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-05-22T00:04:27.287209Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-05-22T00:02:41.893373Z digest=sha256:d4b30a74778603775900ff0dbb1d35958a2f45a0b803ab2ee561b93894dde054

Observation 9e0954fe-bac3-4ba0-b162-7f8603e32b93 · outbound

This paper cites Qwen2-Audio Technical Report.

Task-conditioned probing of instruction-tuned multimodal LLMs: Region-specific brain alignment patterns under naturalistic stimuli Qwen2-Audio Technical Report

Reference 3

Resolution
metadata mismatch
local_arxiv, observed 2026-05-22T00:04:27.052288Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-05-22T00:02:41.893373Z digest=sha256:991d4f40b718dfd0b9437099671aeef1cfbfb54e020b9780e719dd12865aec83

Observation efdf35dd-cb56-4ba0-8021-379b9f1dcf0e · outbound

This paper cites What can 1.8 billion regressions tell us about the pressures shaping high-level visual representation in brains and machines? bioRxiv, pp.

Task-conditioned probing of instruction-tuned multimodal LLMs: Region-specific brain alignment patterns under naturalistic stimuli What can 1.8 billion regressions tell us about the pressures shaping high-level visual representation in brains and machines? bioRxiv, pp

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-05-22T00:04:27.285189Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-05-22T00:02:41.893373Z digest=sha256:8a025fb604a4937217f3c9ef1d1480fe11d23a7b22da0e2a00597b79b387950c

Observation 9715bfa2-67b0-41ad-b544-735da853fc9f · outbound

This paper cites Visual representations in the human brain are aligned with large language models.

Task-conditioned probing of instruction-tuned multimodal LLMs: Region-specific brain alignment patterns under naturalistic stimuli Visual representations in the human brain are aligned with large language models

Reference 5

Resolution
metadata mismatch
arxiv_id, observed 2026-05-22T00:04:27.048641Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-05-22T00:02:41.893373Z digest=sha256:02fc1722ccb23f7a69020b9c1769afdae2dcf9ee35eb92c7845003155c35ccb5

Observation 76930e47-ee7f-4ed6-b513-d193e6c655cc · outbound

This paper cites Vision-Language Integration in Multimodal Video Transformers (Partially) Aligns with the Brain.

Task-conditioned probing of instruction-tuned multimodal LLMs: Region-specific brain alignment patterns under naturalistic stimuli Vision-Language Integration in Multimodal Video Transformers (Partially) Aligns with the Brain

Reference 6

Resolution
verified exact
arxiv_id, observed 2026-05-22T00:04:27.047326Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-05-22T00:02:41.893373Z digest=sha256:2810f881523e9e1649da9aea7365db6db74446acb1f7237104296dd9e7da29f0

Observation 0f735a8a-5188-41ae-932e-3d575835f6fc · outbound

This paper cites Mistral 7B.

Task-conditioned probing of instruction-tuned multimodal LLMs: Region-specific brain alignment patterns under naturalistic stimuli Mistral 7B

Reference 7

Resolution
verified exact
local_arxiv, observed 2026-05-22T00:04:27.059751Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-05-22T00:02:41.893373Z digest=sha256:a88ca69a6c5f0fb39a298ce6ef941598d057bd9a62fb7c285a54da8ebde6354f

Observation 37320b77-74e9-4433-921e-1f4164d9d5ac · outbound

This paper cites VideoChat-R1: Enhancing Spatio-Temporal Perception via Reinforcement Fine-Tuning.

Task-conditioned probing of instruction-tuned multimodal LLMs: Region-specific brain alignment patterns under naturalistic stimuli VideoChat-R1: Enhancing Spatio-Temporal Perception via Reinforcement Fine-Tuning

Reference 8

Resolution
verified exact
local_arxiv, observed 2026-05-22T00:04:27.053881Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-05-22T00:02:41.893373Z digest=sha256:b621e956f397925d1ba105912a063cfbcf6f6ff53be0b4f1a531d4b4e20dcb74

Observation 3b23392d-797b-4374-89ff-f3631fdbd9e1 · outbound

This paper cites Video-llava: Learning united visual representation by alignment before projection.

Task-conditioned probing of instruction-tuned multimodal LLMs: Region-specific brain alignment patterns under naturalistic stimuli Video-llava: Learning united visual representation by alignment before projection

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-05-22T00:04:27.274146Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-05-22T00:02:41.893373Z digest=sha256:7fafc3e0d318d42c7b5297fbf9d3d4137ccf576f5cee5c73e3412ccb5e199122

Observation 625aa75e-907d-4bf0-8787-59144ae213a6 · outbound

This paper cites Video-chatgpt: Towards detailed video understanding via large vision and language models.

Task-conditioned probing of instruction-tuned multimodal LLMs: Region-specific brain alignment patterns under naturalistic stimuli Video-chatgpt: Towards detailed video understanding via large vision and language models

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-05-22T00:04:27.279461Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-05-22T00:02:41.893373Z digest=sha256:14d6229effbe3d9b3265b7adf69e8a844386421ff9b021f5449932ba9146c88b

Observation 7b3b12e9-57e4-4798-bb9d-5fb156e41aa9 · outbound

This paper cites Yuko Nakagi, Takuya Matsuyama, Naoko Koide-Majima, Hiroto Yamaguchi, Rieko Kubo, Shinji Nishimoto, and Yu Takagi.

Task-conditioned probing of instruction-tuned multimodal LLMs: Region-specific brain alignment patterns under naturalistic stimuli Yuko Nakagi, Takuya Matsuyama, Naoko Koide-Majima, Hiroto Yamaguchi, Rieko Kubo, Shinji Nishimoto, and Yu Takagi

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-05-22T00:04:27.275120Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-05-22T00:02:41.893373Z digest=sha256:dc645788f6f8c0286f61f3f829816897e3cc46c4a3a61d5bf250c07f933ae1c8

Observation c8c34f79-e349-404d-9333-6913a3b9e154 · outbound

This paper cites The cost of compression: Investigating the impact of compression on parametric knowledge in language models.

Task-conditioned probing of instruction-tuned multimodal LLMs: Region-specific brain alignment patterns under naturalistic stimuli The cost of compression: Investigating the impact of compression on parametric knowledge in language models

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-05-22T00:04:27.283002Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-05-22T00:02:41.893373Z digest=sha256:7c3d8c502db66864f2620987a32098dedd7f03d11fd638cdd39c6784705c0713

Observation 4da21162-b1b2-4b45-a792-31febf25b320 · outbound

This paper cites URL https://aclanthology.

Task-conditioned probing of instruction-tuned multimodal LLMs: Region-specific brain alignment patterns under naturalistic stimuli URL https://aclanthology

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-05-22T00:04:27.277456Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-05-22T00:02:41.893373Z digest=sha256:53d300e2c82b67391463321ac018ee733e7d8c2599202e1b12c41f344e41c333

Observation 63b56686-57f7-4834-a598-187e1d725da5 · outbound

This paper cites Tuning in to neural encoding: Linking human brain and artificial supervised representations of language.

Task-conditioned probing of instruction-tuned multimodal LLMs: Region-specific brain alignment patterns under naturalistic stimuli Tuning in to neural encoding: Linking human brain and artificial supervised representations of language

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-05-22T00:04:27.283506Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-05-22T00:02:41.893373Z digest=sha256:80e408f2ca3557b8816aaa1363ce4af3ae23e1a9d3b7f780f03f1f9514c9a2b5

Observation 88ad10ad-238a-4de3-b50c-c1ce2d88cae3 · outbound

This paper cites Llama 2: Open Foundation and Fine-Tuned Chat Models.

Task-conditioned probing of instruction-tuned multimodal LLMs: Region-specific brain alignment patterns under naturalistic stimuli Llama 2: Open Foundation and Fine-Tuned Chat Models

Reference 15

Resolution
verified exact
local_arxiv, observed 2026-05-22T00:04:27.056811Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-05-22T00:02:41.893373Z digest=sha256:e2f04a5184b737f44c080f6ba02a58cf66e27e7f3a1cc3c744d42a5c7eed51f9

Observation b53ecedc-025a-4479-ab5e-9cc7d79a8560 · outbound

This paper cites BrainWavLM: Fine-tuning Speech Representations with Brain Responses to Language.

Task-conditioned probing of instruction-tuned multimodal LLMs: Region-specific brain alignment patterns under naturalistic stimuli BrainWavLM: Fine-tuning Speech Representations with Brain Responses to Language

Reference 16

Resolution
verified exact
arxiv_id, observed 2026-05-22T00:04:27.036797Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-05-22T00:02:41.893373Z digest=sha256:e2258ecede621c5ba6e675a1b91e18f88f7fadf5e0fb162ccf9579a4ce1aeb59

Observation 8d32732a-66aa-4ffe-88db-a2de1fe8eab3 · outbound

This paper cites Qwen2-VL: Enhancing Vision-Language Model's Perception of the World at Any Resolution.

Task-conditioned probing of instruction-tuned multimodal LLMs: Region-specific brain alignment patterns under naturalistic stimuli Qwen2-VL: Enhancing Vision-Language Model's Perception of the World at Any Resolution

Reference 17

Resolution
verified exact
local_arxiv, observed 2026-05-22T00:04:27.032285Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-05-22T00:02:41.893373Z digest=sha256:f88e09ae163e252516087de704e89c96481edb6d054479f4e33c234ed1492588

Observation f055e0a8-b3d8-45bb-bcc6-077daa71a230 · outbound

This paper cites an unresolved cited work.

Task-conditioned probing of instruction-tuned multimodal LLMs: Region-specific brain alignment patterns under naturalistic stimuli Unresolved cited work

Reference 18

Resolution
unresolved
raw_fallback, observed 2026-05-22T00:04:27.265269Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-05-22T00:02:41.893373Z digest=sha256:026897a5ff519d2d0bae8e1f1820a298b76626ba7b44c85ebd65eff40958357b

Observation 90792d82-4934-43df-a3a7-db0831f052ac · outbound

This paper cites an unresolved cited work.

Task-conditioned probing of instruction-tuned multimodal LLMs: Region-specific brain alignment patterns under naturalistic stimuli Unresolved cited work

Reference 19

Resolution
unresolved
raw_fallback, observed 2026-05-22T00:04:27.272127Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-05-22T00:02:41.893373Z digest=sha256:9aeccfde2a0656047b6ac96f724282c7da58cc91c7035e41246c24ee24186018

Observation 456d5385-2b71-4c7c-a6f4-ec843d0f4ea4 · outbound

This paper cites The Wolf of Wall Street.

Task-conditioned probing of instruction-tuned multimodal LLMs: Region-specific brain alignment patterns under naturalistic stimuli The Wolf of Wall Street

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-05-22T00:04:27.263053Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-05-22T00:02:41.893373Z digest=sha256:f97f11081d7cf21addb5187e081904fab2e307587773bbeb2f056b49e1defaba

Observation 41a82bbb-fceb-431d-a08c-7b38b63d1e85 · outbound

This paper cites Goodfellas.

Task-conditioned probing of instruction-tuned multimodal LLMs: Region-specific brain alignment patterns under naturalistic stimuli Goodfellas

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-05-22T00:04:27.267166Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-05-22T00:02:41.893373Z digest=sha256:c4939229a24f9404be02bf503b1946efebe9676786861297a573a8ae591708ca

Observation 580d4fca-ed72-4077-8784-bbefbe91d4ac · outbound

This paper cites The Hangover.

Task-conditioned probing of instruction-tuned multimodal LLMs: Region-specific brain alignment patterns under naturalistic stimuli The Hangover

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-05-22T00:04:27.278693Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-05-22T00:02:41.893373Z digest=sha256:44463fa80f79fda34820f92e81a60749d0b1cc78b490f449bf7165c2d1a364a1

Pith citing papers

Observation 9b99c8da-1be8-41d9-b1ae-56b31d0727f8 · inbound

Do VLMs Align Better with Humans than LLMs during Natural Reading? cites this paper.

Do VLMs Align Better with Humans than LLMs during Natural Reading? Task-conditioned probing of instruction-tuned multimodal LLMs: Region-specific brain alignment patterns under naturalistic stimuli

Reference 3

Resolution
verified exact
local_arxiv, observed 2026-06-29T13:03:26.841301Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-06-29T12:53:45.827516Z digest=sha256:aed05b4e9804c4497968bb6f5b8d498f7813f86f5ac6f6af25fb9d1c4c679c0b