Pith. sign in

Paper Citation Record · LEDGER

From Multimodal LLMs to Generalist Embodied Agents: Methods and Lessons

As of 9 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 7 inbound Pith citation observations for arXiv:2412.08442.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2412.08442 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 7 of 7 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00

measured 7 of 7 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T13:53:54.825537Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-04T04:39:33.914861Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation a427c21b-8a0f-4bcb-9ca1-c34cf5463855 · inbound

$\pi_{0.5}$: a Vision-Language-Action Model with Open-World Generalization cites this paper.

$\pi_{0.5}$: a Vision-Language-Action Model with Open-World Generalization From Multimodal LLMs to Generalist Embodied Agents: Methods and Lessons

Reference 75

Resolution
verified exact
arxiv_id, observed 2026-05-22T18:05:00.858804Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-22T18:02:23.305313Z digest=sha256:b9d1c97111cd55c63e0d6f6155ac9911eb4e91075c5e69c0f3af2fac61429c5e

Observation a672600d-eea8-4f3d-a679-ce150f1dfb45 · inbound

ManiTaskGen: A Comprehensive Task Generator for Benchmarking and Improving Vision-Language Agents on Embodied Decision-Making cites this paper.

ManiTaskGen: A Comprehensive Task Generator for Benchmarking and Improving Vision-Language Agents on Embodied Decision-Making From Multimodal LLMs to Generalist Embodied Agents: Methods and Lessons

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-07T13:53:54.825537Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:53:54.825537Z digest=sha256:07754b2d032975081382d2fb1478386998f42d9c98039c6548da580004a2a3fd

Observation 42cb4d42-bbdd-4218-9e25-dc0bd74d1b1d · inbound

Knowledge Insulating Vision-Language-Action Models: Train Fast, Run Fast, Generalize Better cites this paper.

Knowledge Insulating Vision-Language-Action Models: Train Fast, Run Fast, Generalize Better From Multimodal LLMs to Generalist Embodied Agents: Methods and Lessons

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-07T12:45:23.965626Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:45:23.965626Z digest=sha256:dd79cfb1e10a07a40de0ded1643339eaf5b7d493cc2d1598a4fa0b2ebb66b3d5

Observation d64e2629-ea18-4e73-ac59-64eba5abd7ba · inbound

LLM-Enhanced Rapid-Reflex Async-Reflect Embodied Agent for Real-Time Decision-Making in Dynamically Changing Environments cites this paper.

LLM-Enhanced Rapid-Reflex Async-Reflect Embodied Agent for Real-Time Decision-Making in Dynamically Changing Environments From Multimodal LLMs to Generalist Embodied Agents: Methods and Lessons

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-07T05:43:14.217713Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:43:14.217713Z digest=sha256:a409fd4e14e1dc38eb13bd25a856c75a916fea3f123686b77c292da040f81818

Observation 90df1a9d-3276-4332-8d9f-e90baab03630 · inbound

ThinkAct: Vision-Language-Action Reasoning via Reinforced Visual Latent Planning cites this paper.

ThinkAct: Vision-Language-Action Reasoning via Reinforced Visual Latent Planning From Multimodal LLMs to Generalist Embodied Agents: Methods and Lessons

Reference 44

Resolution
verified exact
arxiv_id, observed 2026-05-19T03:22:01.009006Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-19T03:18:14.655384Z digest=sha256:8d45b041ec60d28d031552431492c71111d40ae758786d94420ca0aba535ec32

Observation 99799768-9768-4243-a64d-8d2c35cd9273 · inbound

General Covariant Action Modeling: Constructing Generalized Manifolds via Spatio-Temporal Decoupling cites this paper.

General Covariant Action Modeling: Constructing Generalized Manifolds via Spatio-Temporal Decoupling From Multimodal LLMs to Generalist Embodied Agents: Methods and Lessons

Reference 154

Resolution
metadata mismatch
arxiv_id, observed 2026-06-29T13:33:27.837709Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-06-29T13:33:03.368006Z digest=sha256:55001d3339c419b1adee05a210d924876e9ffdf453527444676833d30d10ee6b

Observation 3f4b65fa-8d09-4ca2-ae5b-cec0a03d3eca · inbound

Vesta: A Generalist Embodied Reasoning Model cites this paper.

Vesta: A Generalist Embodied Reasoning Model From Multimodal LLMs to Generalist Embodied Agents: Methods and Lessons

Reference 106

Resolution
verified exact
arxiv_id, observed 2026-07-04T04:39:33.921356Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-06-26T16:55:12.518255Z digest=sha256:af76e249e1c9796beb8948d480518c2d6fa0dcf74e9183776e59423437eb2e2a