Pith. sign in

Paper Citation Record · LEDGER

MUTEX: Learning Unified Policies from Multimodal Task Specifications

As of 9 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 5 inbound Pith citation observations for arXiv:2309.14320.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2309.14320 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 5 of 5 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00

measured 5 of 5 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-08T15:01:38.393426Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-06-29T08:13:14.633660Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 268623c9-800f-4a71-bfb0-6b09ad209d51 · inbound

Agent AI: Surveying the Horizons of Multimodal Interaction cites this paper.

Agent AI: Surveying the Horizons of Multimodal Interaction MUTEX: Learning Unified Policies from Multimodal Task Specifications

Reference 133

Resolution
verified exact
arxiv_id, observed 2026-05-18T14:25:59.407241Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-05-18T14:25:58.876978Z digest=sha256:4989d5170ac0ca660c93fce9f458b64eda4595e707d2d2d513066cb29a7b3468

Observation a63fec7a-da80-4305-abd7-d572bb5a9d0d · inbound

Adaptive Perception for Unified Visual Multi-modal Object Tracking cites this paper.

Adaptive Perception for Unified Visual Multi-modal Object Tracking MUTEX: Learning Unified Policies from Multimodal Task Specifications

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-08T15:01:38.393426Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T15:01:38.393426Z digest=sha256:acddc198466ba5af0f3585ecaf654c58d86601b96992920657a07eb6ac291827

Observation a67e63f2-50c3-4244-b6a0-ffe57fa8d91f · inbound

CARMA: Context-Aware Situational Grounding of Human-Robot Group Interactions by Combining Vision-Language Models with Object and Action Recognition cites this paper.

CARMA: Context-Aware Situational Grounding of Human-Robot Group Interactions by Combining Vision-Language Models with Object and Action Recognition MUTEX: Learning Unified Policies from Multimodal Task Specifications

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-06T22:53:19.742757Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T22:53:19.742757Z digest=sha256:d8d8a8049597035a2bdda46a4605d509bac373dede44ca1f177e30aa5769b260

Observation 2b7b8524-d70a-4a48-a3d3-5bf4ee6a3056 · inbound

RotVLA: Rotational Latent Action for Vision-Language-Action Model cites this paper.

RotVLA: Rotational Latent Action for Vision-Language-Action Model MUTEX: Learning Unified Policies from Multimodal Task Specifications

Reference 69

Resolution
verified exact
arxiv_id, observed 2026-05-14T17:49:23.054060Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-14T17:48:06.734816Z digest=sha256:5ea0a4ab5644fe6fc4026865ff77042612d280a18d4d5c5bbfde9e424e557aa5

Observation 10114916-1580-4ca7-95f4-2556eeccf972 · inbound

VisualThink-VLA: Visual Intermediate Reasoning for Effective and Low-Latency Vision-Language-Action Policies cites this paper.

VisualThink-VLA: Visual Intermediate Reasoning for Effective and Low-Latency Vision-Language-Action Policies MUTEX: Learning Unified Policies from Multimodal Task Specifications

Reference 50

Resolution
verified exact
arxiv_id, observed 2026-06-29T08:13:14.635232Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-06-29T08:12:24.023638Z digest=sha256:b0333724db30e9212410c0c9740b698d565db95a2f01b26afcaea50f1bf4ed0f