Pith. sign in

Paper Citation Record · LEDGER

Language Models are General-Purpose Interfaces

As of 10 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 11 inbound Pith citation observations for arXiv:2206.06336.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2206.06336 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 11 of 11 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-10T06:31:04.303077+00:00

measured 11 of 11 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-08T04:54:02.052557Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-02T17:37:14.469995Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 7f0117a5-d294-417d-9ee2-1c781bf81b32 · inbound

Language Is Not All You Need: Aligning Perception with Language Models cites this paper.

Language Is Not All You Need: Aligning Perception with Language Models Language Models are General-Purpose Interfaces

Reference 10

Resolution
metadata mismatch
arxiv_id, observed 2026-05-15T18:32:22.931608Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-15T18:32:22.813668Z digest=sha256:6102b4c46891362d888bccf4b4a6d83d785b90633eb5a8b0944bd434b1a0a27a

Observation 74c66787-ce03-4df3-a046-8fa008fb051d · inbound

PaLM-E: An Embodied Multimodal Language Model cites this paper.

PaLM-E: An Embodied Multimodal Language Model Language Models are General-Purpose Interfaces

Reference 14

Resolution
metadata mismatch
arxiv_id, observed 2026-05-10T22:29:29.866806Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-10T22:29:29.631351Z digest=sha256:6073058bd3ff3789dbfc7ca040e968435ecf0efe1a8dc9bb5ce06b7eb4196ea6

Observation 8fd8ce0d-3f69-4b9c-a8c3-d48db2a045db · inbound

Kosmos-2: Grounding Multimodal Large Language Models to the World cites this paper.

Kosmos-2: Grounding Multimodal Large Language Models to the World Language Models are General-Purpose Interfaces

Reference 7

Resolution
metadata mismatch
arxiv_id, observed 2026-05-12T05:19:48.024121Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-12T05:19:47.907355Z digest=sha256:c3fad303de5c069c255ab68264154ed9f0723807a2856f47c313adfc2a593cf4

Observation 869fdac3-ab0a-4af8-aa10-def62fcf82e0 · inbound

Retentive Network: A Successor to Transformer for Large Language Models cites this paper.

Retentive Network: A Successor to Transformer for Large Language Models Language Models are General-Purpose Interfaces

Reference 10

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T20:29:59.857404Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-11T20:29:59.633357Z digest=sha256:5c15f5b9096e0156a65f8a42fd4bd604e54a9d3ad2a4ce353bf9068b21f1747e

Observation 64c34698-a391-4c70-9808-812e419acc76 · inbound

Large Language Models: A Survey cites this paper.

Large Language Models: A Survey Language Models are General-Purpose Interfaces

Reference 104

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T15:22:55.451779Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-11T15:22:54.023279Z digest=sha256:dd5de9c9c470177f1fa0374840f7847d033d1abd6be548607264f314fefc09c3

Observation 3db6993e-b988-44c7-9056-ed08e8d0bf0b · inbound

Revisiting 3D LLM Benchmarks: Are We Really Testing 3D Capabilities? cites this paper.

Revisiting 3D LLM Benchmarks: Are We Really Testing 3D Capabilities? Language Models are General-Purpose Interfaces

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-08T04:54:02.052557Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T04:54:02.052557Z digest=sha256:52d2ab1f2c12956cfcdb44b5667f480590af41de52064172ac144909ea414893

Observation 0e64d91c-0a70-4fc8-86fc-cc36ad3ab751 · inbound

Manager: Aggregating Insights from Unimodal Experts in Two-Tower VLMs and MLLMs cites this paper.

Manager: Aggregating Insights from Unimodal Experts in Two-Tower VLMs and MLLMs Language Models are General-Purpose Interfaces

Reference 101

Resolution
unresolved
no resolver link, observed 2026-08-07T04:08:48.279700Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:08:48.279700Z digest=sha256:e466c65e19f8a69c5685a90b547ba0b13e4d51fb1e919912d86a63854cf92345

Observation 0a76cdf3-950e-41c2-86d4-2c52527f4881 · inbound

Stable Diffusion Models are Secretly Good at Visual In-Context Learning cites this paper.

Stable Diffusion Models are Secretly Good at Visual In-Context Learning Language Models are General-Purpose Interfaces

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-05T20:45:08.328312Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T20:45:08.328312Z digest=sha256:0aa9f0e9fd3d89c2a1ece1b4e869f08a413cf24d62724ee1c886033c93db0ea8

Observation 85adbb4a-5260-44cd-968b-954f85aeb191 · inbound

RELO: Reinforcement Learning to Localize for Visual Object Tracking cites this paper.

RELO: Reinforcement Learning to Localize for Visual Object Tracking Language Models are General-Purpose Interfaces

Reference 18

Resolution
verified exact
arxiv_id, observed 2026-05-11T03:40:54.682789Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-05-11T02:21:49.402350Z digest=sha256:6e28b7a8584909d5769dcf0b79f5180616bc69d4f013c512065a275d88cc2949

Observation 677c17db-97b9-476d-8ef1-f2d1fcc4d199 · inbound

RELO: Reinforcement Learning to Localize for Visual Object Tracking cites this paper.

RELO: Reinforcement Learning to Localize for Visual Object Tracking Language Models are General-Purpose Interfaces

Reference 19

Resolution
verified exact
arxiv_id, observed 2026-05-20T23:19:13.931606Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-05-20T23:16:16.524148Z digest=sha256:67b405bb15a3aa216cdd26e9bca37c99b7c55ec85cb0bbf299139094ab09e14f

Observation e5a49191-5838-44fc-995c-fe99317acbf6 · inbound

The Last Visible Pixel: Probing Fine-Scale Perception in Vision-Language Models cites this paper.

The Last Visible Pixel: Probing Fine-Scale Perception in Vision-Language Models Language Models are General-Purpose Interfaces

Reference 24

Resolution
metadata mismatch
arxiv_id, observed 2026-07-02T17:37:14.471612Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-06-27T21:58:53.702009Z digest=sha256:ac7837885ab92bbae8186783861abb6dfdc7c4112e5017582f74b8bcb15ba6f9