Pith. sign in

Paper Citation Record · LEDGER

EmbodiedEval: Evaluate Multimodal LLMs as Embodied Agents

As of 8 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 22 inbound Pith citation observations for arXiv:2501.11858.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2501.11858 v2

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 22 of 22 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00

measured 22 of 22 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T13:53:56.823209Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-03T11:58:05.933512Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 68090646-e963-4644-ba35-4303bdae37c2 · inbound

From LLM Reasoning to Autonomous AI Agents: A Comprehensive Review cites this paper.

From LLM Reasoning to Autonomous AI Agents: A Comprehensive Review EmbodiedEval: Evaluate Multimodal LLMs as Embodied Agents

Reference 108

Resolution
verified exact
arxiv_id, observed 2026-05-15T02:57:38.429736Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-15T02:57:37.873567Z digest=sha256:ade6c7d9d6977a889cf006411b60199af7c2eeb911044712c0d4be07a8ea0600

Observation 19189071-c053-470b-8c64-f823e3ed0860 · inbound

ManiTaskGen: A Comprehensive Task Generator for Benchmarking and Improving Vision-Language Agents on Embodied Decision-Making cites this paper.

ManiTaskGen: A Comprehensive Task Generator for Benchmarking and Improving Vision-Language Agents on Embodied Decision-Making EmbodiedEval: Evaluate Multimodal LLMs as Embodied Agents

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-07T13:53:56.823209Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:53:56.823209Z digest=sha256:8515c2ca9812fee9e917c2586159798cea1db75dbf01a98ed2ee6475b7ec65ee

Observation 8b7d6512-f21c-434b-a8e0-1e0f7890470d · inbound

Benchmarking LLM-Assisted Blue Teaming via Standardized Threat Hunting cites this paper.

Benchmarking LLM-Assisted Blue Teaming via Standardized Threat Hunting EmbodiedEval: Evaluate Multimodal LLMs as Embodied Agents

Reference 2013

Resolution
unresolved
no resolver link, observed 2026-08-04T14:42:56.148795Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T14:42:56.148795Z digest=sha256:aad816ca614b56e89173203a7f8b66591842d1863c4c88971d5eda8e7154f8c9

Observation da282c5d-c3c2-455b-b77a-f42e637e0ae8 · inbound

Beyond Description: Cognitively Benchmarking Fine-Grained Action for Embodied Agents cites this paper.

Beyond Description: Cognitively Benchmarking Fine-Grained Action for Embodied Agents EmbodiedEval: Evaluate Multimodal LLMs as Embodied Agents

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-03T20:42:53.353480Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T20:42:53.353480Z digest=sha256:7accfb8dfd62ab73bd1429ec7cf757959360b418ad612cfcfc65e9e4e28b0427

Observation 90256d3e-1d31-42d6-ba2f-99dc485fc10a · inbound

Unified Embodied VLM Reasoning with Robotic Action via Autoregressive Discretized Pre-training cites this paper.

Unified Embodied VLM Reasoning with Robotic Action via Autoregressive Discretized Pre-training EmbodiedEval: Evaluate Multimodal LLMs as Embodied Agents

Reference 53

Resolution
unresolved
no resolver link, observed 2026-08-03T13:29:52.850495Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T13:29:52.850495Z digest=sha256:f4cc034d555cba847c54ee60cc6679e0e058db50a199db0c6951d27b01ad3e76

Observation 13637143-bc52-4438-9363-14becb285f5a · inbound

ST-BiBench: Benchmarking Multi-Stream Multimodal Coordination in Bimanual Embodied Tasks for MLLMs cites this paper.

ST-BiBench: Benchmarking Multi-Stream Multimodal Coordination in Bimanual Embodied Tasks for MLLMs EmbodiedEval: Evaluate Multimodal LLMs as Embodied Agents

Reference 31

Resolution
verified exact
arxiv_id, observed 2026-05-16T06:00:40.491132Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-16T06:00:02.043029Z digest=sha256:63f6d3486ea4473a831cc8e7a7af9f3e58230abccc23c3bdcdeb0698ddc2be4d

Observation fdce8250-ad50-4153-8923-823f3152cbf8 · inbound

DeepImageSearch: Benchmarking Multimodal Agents for Context-Aware Image Retrieval in Visual Histories cites this paper.

DeepImageSearch: Benchmarking Multimodal Agents for Context-Aware Image Retrieval in Visual Histories EmbodiedEval: Evaluate Multimodal LLMs as Embodied Agents

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-03T01:02:32.859526Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-03T01:02:32.859526Z digest=sha256:25208c99895d2a304bb31078d885bccff326e98675730c50d17263e677cb8d98

Observation 2f62933b-8799-421a-b14b-3b9db7eaa32c · inbound

RoboAgent: Chaining Basic Capabilities for Embodied Task Planning cites this paper.

RoboAgent: Chaining Basic Capabilities for Embodied Task Planning EmbodiedEval: Evaluate Multimodal LLMs as Embodied Agents

Reference 15

Resolution
verified exact
arxiv_id, observed 2026-05-11T05:15:57.017245Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-10T18:15:08.727921Z digest=sha256:45056342dcb2c01e38f5eb6b747366fe431a6a3ca48913a619946b94cc03ff1c

Observation 388f7469-35bc-45a7-a051-1dd64227fa35 · inbound

ESCAPE: Episodic Spatial Memory and Adaptive Execution Policy for Long-Horizon Mobile Manipulation cites this paper.

ESCAPE: Episodic Spatial Memory and Adaptive Execution Policy for Long-Horizon Mobile Manipulation EmbodiedEval: Evaluate Multimodal LLMs as Embodied Agents

Reference 6

Resolution
metadata mismatch
arxiv_id, observed 2026-05-10T13:10:26.310069Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-10T13:09:32.132811Z digest=sha256:e17f52206bdc7b3dd1d9b00f37a19a3db6477447c81e877ec00de81a9c3b082c

Observation b2d0b7e8-0728-4d47-ab7e-ed7da9cde4fd · inbound

MirrorBench: Evaluating Self-centric Intelligence in MLLMs by Introducing a Mirror cites this paper.

MirrorBench: Evaluating Self-centric Intelligence in MLLMs by Introducing a Mirror EmbodiedEval: Evaluate Multimodal LLMs as Embodied Agents

Reference 6

Resolution
verified exact
arxiv_id, observed 2026-05-10T10:55:03.980749Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-10T10:53:48.374637Z digest=sha256:2759690889577c8fc5b68dabc8b64b912b6dc9405df0c35f5ff1ca564f323688

Observation f83d05e5-5694-4e22-8abf-5854c1cf3237 · inbound

Breaking the Secret: Economic Interventions for Combating Collusion in Embodied Multi-Agent Systems cites this paper.

Breaking the Secret: Economic Interventions for Combating Collusion in Embodied Multi-Agent Systems EmbodiedEval: Evaluate Multimodal LLMs as Embodied Agents

Reference 6

Resolution
verified exact
arxiv_id, observed 2026-05-11T21:16:16.795898Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-08T06:13:17.633146Z digest=sha256:554fc42e76c2c8544c3623721b548f2a2228b93df4f8c2177174ae985238363b

Observation dc998bc8-fcc9-4559-9657-1d3d8b5460bb · inbound

Done, But Not Sure: Disentangling World Completion from Self-Termination in Embodied Agents cites this paper.

Done, But Not Sure: Disentangling World Completion from Self-Termination in Embodied Agents EmbodiedEval: Evaluate Multimodal LLMs as Embodied Agents

Reference 9

Resolution
verified exact
arxiv_id, observed 2026-05-12T08:01:29.145656Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-12T01:23:37.488121Z digest=sha256:7afc2f4d19185ef26f946da98ff29851982e1983dbaf26716bb4ed203d5cc6d7

Observation 33bb212e-c793-4683-a422-180639b260a9 · inbound

Done, But Not Sure: Disentangling World Completion from Self-Termination in Embodied Agents cites this paper.

Done, But Not Sure: Disentangling World Completion from Self-Termination in Embodied Agents EmbodiedEval: Evaluate Multimodal LLMs as Embodied Agents

Reference 9

Resolution
verified exact
arxiv_id, observed 2026-05-13T01:07:00.445056Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-13T01:03:30.995808Z digest=sha256:4a41b8bf1ddc16d948193b6d03504e2e9f0a946c9edafe82952e271cb0f4abf2

Observation 517ac0eb-c664-403a-8529-abc06c7bcb23 · inbound

Done, But Not Sure: Disentangling World Completion from Self-Termination in Embodied Agents cites this paper.

Done, But Not Sure: Disentangling World Completion from Self-Termination in Embodied Agents EmbodiedEval: Evaluate Multimodal LLMs as Embodied Agents

Reference 9

Resolution
verified exact
arxiv_id, observed 2026-05-15T06:29:49.819712Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-15T06:25:15.510083Z digest=sha256:ee32b460362b691d3ffff1b7ab9e3532034801e22b072f508b532fc8b22f7834

Observation 38244576-1541-469b-8da7-5704ece71b13 · inbound

ESI-Bench: Towards Embodied Spatial Intelligence that Closes the Perception-Action Loop cites this paper.

ESI-Bench: Towards Embodied Spatial Intelligence that Closes the Perception-Action Loop EmbodiedEval: Evaluate Multimodal LLMs as Embodied Agents

Reference 3

Resolution
metadata mismatch
arxiv_id, observed 2026-05-20T10:53:13.219228Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-20T10:52:22.778489Z digest=sha256:2e49ba4655db56be93cb22221f0a2054ccc53e8120d76734be65883d36b1a66a

Observation 324a0c64-dd7e-445c-bf57-38a614a91319 · inbound

ESI-Bench: Towards Embodied Spatial Intelligence that Closes the Perception-Action Loop cites this paper.

ESI-Bench: Towards Embodied Spatial Intelligence that Closes the Perception-Action Loop EmbodiedEval: Evaluate Multimodal LLMs as Embodied Agents

Reference 3

Resolution
metadata mismatch
arxiv_id, observed 2026-07-01T15:05:47.177627Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-06-30T18:25:17.831116Z digest=sha256:4f90129eb67d4aaaa41bee507eba5b14f447d9abdef45fe2ad14443bdd24a1a1

Observation 394b8485-04cd-47da-878f-2004801b393e · inbound

Artifact-Bench: Evaluating MLLMs on Detecting and Assessing the Artifacts of AI-Generated Videos cites this paper.

Artifact-Bench: Evaluating MLLMs on Detecting and Assessing the Artifacts of AI-Generated Videos EmbodiedEval: Evaluate Multimodal LLMs as Embodied Agents

Reference 6

Resolution
verified exact
arxiv_id, observed 2026-05-20T10:28:12.023952Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-20T10:26:30.661042Z digest=sha256:4bfe2e8fdc590645de887e4b10f4593fdb6c7c7def54059314d7a47379777913

Observation e3188389-ed05-4e30-9ad2-439ac55b4e17 · inbound

IntentionNav: A Benchmark for Intent-Driven Object Navigation from Implicit Human Instruction cites this paper.

IntentionNav: A Benchmark for Intent-Driven Object Navigation from Implicit Human Instruction EmbodiedEval: Evaluate Multimodal LLMs as Embodied Agents

Reference 9

Resolution
verified exact
arxiv_id, observed 2026-05-25T05:05:22.364600Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-25T05:04:39.348534Z digest=sha256:ce67b6e537e86c3043ef499efc49d257f83eb3e4ff7127a131db6034e611b63b

Observation 4cc676b1-2f72-46a1-9140-b6a3be5cac7e · inbound

SpatialWorld: Benchmarking Interactive Spatial Reasoning of Multimodal Agents in Real-World Tasks cites this paper.

SpatialWorld: Benchmarking Interactive Spatial Reasoning of Multimodal Agents in Real-World Tasks EmbodiedEval: Evaluate Multimodal LLMs as Embodied Agents

Reference 12

Resolution
verified exact
arxiv_id, observed 2026-07-03T01:27:30.581129Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-06-27T16:35:14.099586Z digest=sha256:153d56b384d56b7c2a474a948b48dc575eba9118bf507032e23f691ad369c885

Observation 266d5c4d-8a9e-4ab6-8c3c-5cdc71a9456e · inbound

A Practical Recipe Towards Improving Sim-and-Real Correlation for VLA Evaluation cites this paper.

A Practical Recipe Towards Improving Sim-and-Real Correlation for VLA Evaluation EmbodiedEval: Evaluate Multimodal LLMs as Embodied Agents

Reference 34

Resolution
verified exact
arxiv_id, observed 2026-07-03T05:47:41.102639Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-06-27T13:06:35.282837Z digest=sha256:a2dd2d7bd153697ca85f6ec878a78f9718a562286519401331402510e722065e

Observation 669ad96f-b1bf-4618-b052-6fc112548da0 · inbound

Intelligent Automation for Embodied Benchmark Construction: Pipelines, Embodiments, Simulators, and Trends cites this paper.

Intelligent Automation for Embodied Benchmark Construction: Pipelines, Embodiments, Simulators, and Trends EmbodiedEval: Evaluate Multimodal LLMs as Embodied Agents

Reference 42

Resolution
metadata mismatch
arxiv_id, observed 2026-07-03T11:58:05.935537Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-06-27T09:17:50.473747Z digest=sha256:ae6565dff8a0c88a26fa278d7210436cd5c7985ccb4bd84a5b181155c27b72a8

Observation 1d6f3c31-61af-4bd1-b0dd-d73ac9868442 · inbound

AgentOdyssey: Open-Ended Long-Horizon Text Game Generation for Test-Time Continual Learning Agents cites this paper.

AgentOdyssey: Open-Ended Long-Horizon Text Game Generation for Test-Time Continual Learning Agents EmbodiedEval: Evaluate Multimodal LLMs as Embodied Agents

Reference 9

Resolution
verified exact
arxiv_id, observed 2026-07-01T19:46:11.471278Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-06-28T21:59:25.449570Z digest=sha256:90046cc5d8a36754edb5e4e55e34c45b7ecc6885d814dd9c0dc4cacb950922bf