Pith. sign in

Paper Citation Record · LEDGER

EARBench: Towards Evaluating Physical Risk Awareness for Task Planning of Foundation Model-based Embodied AI Agents

As of 9 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 16 inbound Pith citation observations for arXiv:2408.04449.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2408.04449 v5

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 16 of 16 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00

measured 16 of 16 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T14:13:46.416331Z

measured 1 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

1
arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation e447815a-8dc2-4951-bb7f-5fb5c1d6b12a · inbound

HomeBench: Evaluating LLMs in Smart Homes with Valid and Invalid Instructions Across Single and Multiple Devices cites this paper.

HomeBench: Evaluating LLMs in Smart Homes with Valid and Invalid Instructions Across Single and Multiple Devices EARBench: Towards Evaluating Physical Risk Awareness for Task Planning of Foundation Model-based Embodied AI Agents

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-07T14:13:46.416331Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:13:46.416331Z digest=sha256:8b4bc42ba5143f04707e38c93f6edadbd83d295f76d990e239c00ad0bcfb516c

Observation 295af0db-ac06-4283-ba5a-64e714c496a5 · inbound

Towards provable probabilistic safety for scalable embodied AI systems cites this paper.

Towards provable probabilistic safety for scalable embodied AI systems EARBench: Towards Evaluating Physical Risk Awareness for Task Planning of Foundation Model-based Embodied AI Agents

Reference 10

Resolution
verified exact
arxiv_id, observed 2026-05-19T11:02:15.172947Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-19T11:00:27.799347Z digest=sha256:5e43eb775bad0b70ae655a01390c12ab018f0063784a3a59703aa881231091dc

Observation cd1bc8e5-5227-499d-ad41-af5077ea53a9 · inbound

The Scales of Justitia: A Comprehensive Survey on Safety Evaluation of LLMs cites this paper.

The Scales of Justitia: A Comprehensive Survey on Safety Evaluation of LLMs EARBench: Towards Evaluating Physical Risk Awareness for Task Planning of Foundation Model-based Embodied AI Agents

Reference 213

Resolution
unresolved
no resolver link, observed 2026-08-07T10:17:27.249700Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T10:17:27.249700Z digest=sha256:8715d83d290ef77ee82eb35ec5611951800a33c4bc52a05490734bd5832ad015

Observation e222ca2d-8c43-4ac4-afb3-9d293f8a9275 · inbound

Alignment and Safety in Large Language Models: Safety Mechanisms, Training Paradigms, and Emerging Challenges cites this paper.

Alignment and Safety in Large Language Models: Safety Mechanisms, Training Paradigms, and Emerging Challenges EARBench: Towards Evaluating Physical Risk Awareness for Task Planning of Foundation Model-based Embodied AI Agents

Reference 102

Resolution
unresolved
no resolver link, observed 2026-08-06T14:13:06.306571Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T14:13:06.306571Z digest=sha256:47f80734d06a35eedb2793be92cc4de58a25cc5288034f0f8a324b484a37c122

Observation fa6c37ac-b5e9-4ac8-a4d1-9812ca80ae24 · inbound

Context-Aware Risk Estimation in Home Environments: A Probabilistic Framework for Service Robots cites this paper.

Context-Aware Risk Estimation in Home Environments: A Probabilistic Framework for Service Robots EARBench: Towards Evaluating Physical Risk Awareness for Task Planning of Foundation Model-based Embodied AI Agents

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-05T15:33:21.281847Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T15:33:21.281847Z digest=sha256:ff26013a4473a8b7861ba1083c3ff6708c87634517e43dbe6603332d6fa39c69

Observation 2d6efce0-341c-4477-8467-016f933cf6c3 · inbound

RoboInspector: Unveiling the Unreliability of Policy Code for LLM-enabled Robotic Manipulation cites this paper.

RoboInspector: Unveiling the Unreliability of Policy Code for LLM-enabled Robotic Manipulation EARBench: Towards Evaluating Physical Risk Awareness for Task Planning of Foundation Model-based Embodied AI Agents

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-05T14:20:42.354783Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T14:20:42.354783Z digest=sha256:2c507811fbe193d48a1010f03ba2b791e8ec8d99454400e445d55dbe4d15e8ec

Observation 41ca27e6-ebc0-4c5c-8308-5842b9be456d · inbound

SENTINEL: A Multi-Level Formal Framework for Safety Evaluation of Foundation Model-based Embodied Agents cites this paper.

SENTINEL: A Multi-Level Formal Framework for Safety Evaluation of Foundation Model-based Embodied Agents EARBench: Towards Evaluating Physical Risk Awareness for Task Planning of Foundation Model-based Embodied AI Agents

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-04T09:52:21.230370Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T09:52:21.230370Z digest=sha256:c6ab46e785421d81585999526f453a56e0ddbf70017e621833addbb9fc97ef46

Observation d4f65807-a00e-44eb-8939-7fe00f1d669e · inbound

Harnessing Embodied Agents: Runtime Governance for Policy-Constrained Execution cites this paper.

Harnessing Embodied Agents: Runtime Governance for Policy-Constrained Execution EARBench: Towards Evaluating Physical Risk Awareness for Task Planning of Foundation Model-based Embodied AI Agents

Reference 24

Resolution
verified exact
arxiv_id, observed 2026-05-11T05:31:01.026413Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-10T18:05:01.925496Z digest=sha256:3640b3c638626fd0eb1a1b824118bdd3e11b37fada24aa8f33a5a53600956954

Observation 39f30135-dd5d-46c9-bc56-6d695cd35b32 · inbound

Harnessing Embodied Agents: Runtime Governance for Policy-Constrained Execution cites this paper.

Harnessing Embodied Agents: Runtime Governance for Policy-Constrained Execution EARBench: Towards Evaluating Physical Risk Awareness for Task Planning of Foundation Model-based Embodied AI Agents

Reference 24

Resolution
verified exact
arxiv_id, observed 2026-05-22T10:41:25.118516Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-22T10:40:28.374841Z digest=sha256:3bf59389f3e3187409afc9471ac8c7e9859bc46df17b415db1ce93f5124b05b6

Observation 993aa933-7740-4c44-bb57-51a801b63845 · inbound

SafetyALFRED: Evaluating Safety-Conscious Planning of Multimodal Large Language Models cites this paper.

SafetyALFRED: Evaluating Safety-Conscious Planning of Multimodal Large Language Models EARBench: Towards Evaluating Physical Risk Awareness for Task Planning of Foundation Model-based Embodied AI Agents

Reference 29

Resolution
metadata mismatch
arxiv_id, observed 2026-05-10T02:22:20.637557Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-05-10T02:21:29.463149Z digest=sha256:424899900c16f6bee1e394e22883c118699ef3b867216b6342ed4a9c411f5e6e

Observation 9083376f-cb8e-4bda-aae9-9311ad42a8ce · inbound

Benchmarking the Safety of Large Language Models for Robotic Health Attendant Control cites this paper.

Benchmarking the Safety of Large Language Models for Robotic Health Attendant Control EARBench: Towards Evaluating Physical Risk Awareness for Task Planning of Foundation Model-based Embodied AI Agents

Reference 17

Resolution
verified exact
arxiv_id, observed 2026-05-12T09:26:26.102085Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-07T10:54:56.913970Z digest=sha256:a0921f91f36d1a1374d245d0a84691ca77f5cacc84ce29fdb9a440441d4572e7

Observation 6c05f84d-ce2e-42fc-81b9-c90e15b12328 · inbound

Benchmarking the Safety of Large Language Models for Robotic Health Attendant Control cites this paper.

Benchmarking the Safety of Large Language Models for Robotic Health Attendant Control EARBench: Towards Evaluating Physical Risk Awareness for Task Planning of Foundation Model-based Embodied AI Agents

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-02T15:19:16.522130Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T15:19:16.522130Z digest=sha256:116d56c73b3a67c5868ea06b9ef929199ebc14e7a4901589a7f0d46472e97296

Observation be5ddc33-3030-49f7-8170-2df6b098c705 · inbound

RoboJailBench: Benchmarking Adversarial Attacks and Defenses in Embodied Robotic Agents cites this paper.

RoboJailBench: Benchmarking Adversarial Attacks and Defenses in Embodied Robotic Agents EARBench: Towards Evaluating Physical Risk Awareness for Task Planning of Foundation Model-based Embodied AI Agents

Reference 30

Resolution
malformed identifier
arxiv_id, observed 2026-05-20T05:08:05.135158Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-20T05:05:42.451746Z digest=sha256:f5d3e60c7a7c202af71aa3984fe77bc2ff9644a038165c4e0f50995ef4924f16

Observation 2f0844be-ec1f-41e1-8545-c38d0af6d040 · inbound

Toward Secure LLM Agents: Threat Surfaces, Attacks, Defenses, and Evaluation cites this paper.

Toward Secure LLM Agents: Threat Surfaces, Attacks, Defenses, and Evaluation EARBench: Towards Evaluating Physical Risk Awareness for Task Planning of Foundation Model-based Embodied AI Agents

Reference 265

Resolution
verified exact
arxiv_id, observed 2026-06-27T13:10:57.002106Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-06-27T12:55:22.831264Z digest=sha256:e8ba8b37d0b6be91232d7f94b4c2f0a750b277ebcd9b10e523d7877f58ef8237

Observation 02dd3a31-fca3-4fbf-8a7b-fb2e2aa900c6 · inbound

When Words Are Safe But Actions Kill: Probing Physical Danger Beyond Text Safety in Hidden-State Risk Space cites this paper.

When Words Are Safe But Actions Kill: Probing Physical Danger Beyond Text Safety in Hidden-State Risk Space EARBench: Towards Evaluating Physical Risk Awareness for Task Planning of Foundation Model-based Embodied AI Agents

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-01T23:52:53.743894Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T23:52:53.743894Z digest=sha256:38bf340ee500c51bef2b6060b1d8b4132348c9704df7d246b32d83814a495bbe

Observation 4f902688-6fe3-4256-b920-2ce9191f31b9 · inbound

Self-Evolving Just-In-Time Memory for Proactive Embodied Safety cites this paper.

Self-Evolving Just-In-Time Memory for Proactive Embodied Safety EARBench: Towards Evaluating Physical Risk Awareness for Task Planning of Foundation Model-based Embodied AI Agents

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-02T09:54:52.545077Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T09:54:52.545077Z digest=sha256:6c6096e7a5e31b04daf502259c5fc0cf2903c8573925917ac6db36bb92031766