Pith. sign in

Paper Citation Record · LEDGER

Can We Talk Models Into Seeing the World Differently?

As of 11 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 7 inbound Pith citation observations for arXiv:2403.09193.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2403.09193 v2

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 7 of 7 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-10T06:31:04.303077+00:00

measured 7 of 7 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-09T15:04:35.363001Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-07-10T15:47:23.264103Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 89178649-e182-40a8-8d17-59889033142a · inbound

The in-context inductive biases of vision-language models differ across modalities cites this paper.

The in-context inductive biases of vision-language models differ across modalities Can We Talk Models Into Seeing the World Differently?

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-09T15:04:35.363001Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T15:04:35.363001Z digest=sha256:15a93c2046f39e1886daf4e77f51bace4055e08e4197f044b39f70ba3eaed93d

Observation 63cf50e5-78c6-4775-abac-8578483dc93d · inbound

TextureSAM: Towards a Texture Aware Foundation Model for Segmentation cites this paper.

TextureSAM: Towards a Texture Aware Foundation Model for Segmentation Can We Talk Models Into Seeing the World Differently?

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-07T15:02:52.741800Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:02:52.741800Z digest=sha256:e805c4f2e285f99dc8b2e9c7bbad6c755b43867f27855015ee73379a9beba285

Observation 6b402099-349e-4aec-86e3-ab2dafd7eae3 · inbound

AVA-Bench: Atomic Visual Ability Benchmark for Vision Foundation Models cites this paper.

AVA-Bench: Atomic Visual Ability Benchmark for Vision Foundation Models Can We Talk Models Into Seeing the World Differently?

Reference 23

Resolution
verified exact
arxiv_id, observed 2026-05-19T11:13:02.855394Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-19T11:12:41.130806Z digest=sha256:925a66bb807775aa2e59963c6fc9b2cbca73e755e82ff92b5fb4c4e647647985

Observation 9372506d-8553-4a17-9035-c33ebab0294b · inbound

Vision-Language Models display a strong gender bias cites this paper.

Vision-Language Models display a strong gender bias Can We Talk Models Into Seeing the World Differently?

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-05T20:05:43.038189Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T20:05:43.038189Z digest=sha256:7c52cc31856a4beac0a4406e19d1453cfed2493d1671de191b0976b69f4fedd6

Observation e812453a-ca43-4b2b-966c-9ab04c510e5a · inbound

Visual Persuasion: What Influences Decisions of Vision-Language Models? cites this paper.

Visual Persuasion: What Influences Decisions of Vision-Language Models? Can We Talk Models Into Seeing the World Differently?

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-02T23:00:25.759765Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-02T23:00:25.759765Z digest=sha256:b3b90550725fa4be96ac155cfae57af381f89e9f4700ae343d169ba24555669d

Observation ae529ab1-c035-4f5d-8e97-3f47ee63a695 · inbound

TraversalBench: Challenging Paths to Follow for Vision Language Models cites this paper.

TraversalBench: Challenging Paths to Follow for Vision Language Models Can We Talk Models Into Seeing the World Differently?

Reference 12

Resolution
verified exact
arxiv_id, observed 2026-05-11T10:36:02.748890Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-10T15:25:38.215013Z digest=sha256:35305510d16d8e9ebdd049f9f1ab5398ddb791d32ce12264b5f859f09c4b038b

Observation 827cb1d1-8327-4b24-8f5c-a32066a9bc2c · inbound

Multimodal Unlearning Across Vision, Language, Video, and Audio: Survey of Methods, Datasets, and Benchmarks cites this paper.

Multimodal Unlearning Across Vision, Language, Video, and Audio: Survey of Methods, Datasets, and Benchmarks Can We Talk Models Into Seeing the World Differently?

Reference 34

Resolution
metadata mismatch
local_arxiv, observed 2026-07-10T15:47:23.265281Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-07-10T15:38:58.361411Z digest=sha256:fc128622111ab4d04749cbc9f99e0cffddeded78179208a6456dfdda868ec310