Pith. sign in

Paper Citation Record · LEDGER

SafeAgent: Safeguarding LLM Agents via an Automated Risk Simulator

As of 16 August 2026, this Paper Citation Record lists 61 of 61 outbound references and 2 inbound Pith citation observations for arXiv:2505.17735.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2505.17735 v2

Coverage vector

measured 61 of 61 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T14:44:51.276710Z

measured 63 of 63 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-16T06:30:59.297886+00:00

measured 2 of 2 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-06-26T14:00:04.412804Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-04T06:59:37.648031Z

Reference resolution

61 of 61 outbound references displayed

  • verified exact1
  • verified fuzzy24
  • unresolved35
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch1

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation a8196e9e-67c4-40d3-8ac9-eacd51b1f54e · outbound

This paper cites GPT-4 Technical Report.

SafeAgent: Safeguarding LLM Agents via an Automated Risk Simulator GPT-4 Technical Report

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-07T14:44:44.453213Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:44:44.453213Z digest=sha256:fba3061fcad79c1bda87e195c3a2466e78b3523ba25f5d68264cfbe547ef694b

Observation cf6f3a6e-49c8-4e4a-bd5c-fd179d253a0f · outbound

This paper cites AgentHarm: A Benchmark for Measuring Harmfulness of LLM Agents.

SafeAgent: Safeguarding LLM Agents via an Automated Risk Simulator AgentHarm: A Benchmark for Measuring Harmfulness of LLM Agents

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-07T14:44:44.534682Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:44:44.534682Z digest=sha256:d58b79b404019b127b85a47b58c2e25327661b8afa42520639659e9b4628b384

Observation 42aed6b1-8561-4508-9c5d-0e0437887d17 · outbound

This paper cites The claude 3 model family: Opus, sonnet, haiku.

SafeAgent: Safeguarding LLM Agents via an Automated Risk Simulator The claude 3 model family: Opus, sonnet, haiku

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:44:58.565414Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-07T14:44:44.613658Z digest=sha256:475b168400c5ab95e99288faf28128f4dbcec78a6d0950c657a51ca4b0f4669f

Observation e9457d8b-902d-4b40-a21a-5809c592e65d · outbound

This paper cites Training a Helpful and Harmless Assistant with Reinforcement Learning from Human Feedback.

SafeAgent: Safeguarding LLM Agents via an Automated Risk Simulator Training a Helpful and Harmless Assistant with Reinforcement Learning from Human Feedback

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-07T14:44:44.695587Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:44:44.695587Z digest=sha256:e870b4ea4506fa2b2848e90c54d9c7d0110764880c3a0ecca4cf9e626ea09bb6

Observation 70853b67-7090-4199-aeef-021737cd8755 · outbound

This paper cites How (un)ethical are instruction-centric responses of LLMs? Unveiling the vulnerabilities of safety guardrails to harmful queries.

SafeAgent: Safeguarding LLM Agents via an Automated Risk Simulator How (un)ethical are instruction-centric responses of LLMs? Unveiling the vulnerabilities of safety guardrails to harmful queries

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-07T14:44:44.798595Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:44:44.798595Z digest=sha256:85689ff091ee3861fa26b4222cdaf7f94c82ace6c49c3f7903979f53c461b3b5

Observation 1dfc7352-7413-4e63-8e6b-a2ed011f0706 · outbound

This paper cites Assertiveness-based agent communication for a personalized medicine on medical imaging diagnosis.

SafeAgent: Safeguarding LLM Agents via an Automated Risk Simulator Assertiveness-based agent communication for a personalized medicine on medical imaging diagnosis

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:44:58.295419Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-07T14:44:44.888093Z digest=sha256:e3d4709971d27b179a8248a40e03931e2e3e62211860a8612cc67abdb4f9c3d4

Observation 0bd44a55-90f6-4db0-8a57-d90f4b2cfaa9 · outbound

This paper cites Defending Against Alignment-Breaking Attacks via Robustly Aligned LLM.

SafeAgent: Safeguarding LLM Agents via an Automated Risk Simulator Defending Against Alignment-Breaking Attacks via Robustly Aligned LLM

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-07T14:44:44.993336Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:44:44.993336Z digest=sha256:40073e1e6f3e22af392b1e71183fa38249dc1a35b8473f0bbd304fb26fbd8b47

Observation 6ffaacc0-5cc5-4a61-a24c-775cf41feb40 · outbound

This paper cites AgentDojo: A Dynamic Environment to Evaluate Prompt Injection Attacks and Defenses for LLM Agents.

SafeAgent: Safeguarding LLM Agents via an Automated Risk Simulator AgentDojo: A Dynamic Environment to Evaluate Prompt Injection Attacks and Defenses for LLM Agents

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-07T14:44:45.087796Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:44:45.087796Z digest=sha256:66348bccb773f0c8b700b9ddfad86e539a6a1c8bb9e5895e4382f02c15f57b59

Observation 8d3cf8a2-cf46-4668-a7ad-7faffbbb2780 · outbound

This paper cites The Llama 3 Herd of Models.

SafeAgent: Safeguarding LLM Agents via an Automated Risk Simulator The Llama 3 Herd of Models

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-07T14:44:45.213435Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:44:45.213435Z digest=sha256:79072afc0f1bae48cd7d9547fd56ec8c7d2a01a9f0f2e20161354a44c643bc7e

Observation b1036e4b-c310-48f9-8090-5b59d2352ef9 · outbound

This paper cites A Human-Centered Risk Evaluation of Biometric Systems Using Conjoint Analysis.

SafeAgent: Safeguarding LLM Agents via an Automated Risk Simulator A Human-Centered Risk Evaluation of Biometric Systems Using Conjoint Analysis

Reference 10

Resolution
metadata mismatch
local_arxiv, observed 2026-08-07T14:44:51.755916Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-07T14:44:45.325701Z digest=sha256:0cc54d220d86b0999f2fa85ceaa78792747fe420ddc65b6938f49112b3ec1f45

Observation b646852c-473d-45ff-b6f7-40d83f7b1a48 · outbound

This paper cites ChatGLM: A Family of Large Language Models from GLM-130B to GLM-4 All Tools.

SafeAgent: Safeguarding LLM Agents via an Automated Risk Simulator ChatGLM: A Family of Large Language Models from GLM-130B to GLM-4 All Tools

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-07T14:44:45.432347Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:44:45.432347Z digest=sha256:be7206f62661397fe837b4773964f60cf1ce87bf7f0dd8d718a20931a5ca70b0

Observation b7a13df9-7fd1-4c37-b624-8b6cc034e9fa · outbound

This paper cites ToRA: A tool-integrated reasoning agent for mathematical problem solving.

SafeAgent: Safeguarding LLM Agents via an Automated Risk Simulator ToRA: A tool-integrated reasoning agent for mathematical problem solving

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:44:58.028665Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-07T14:44:45.555426Z digest=sha256:72bc0508f480911a3112cf0be09dfb8b4ece42ef000b738e866d24a59ef36714

Observation 5490dc32-d303-461f-b2f6-d7a4d65b3852 · outbound

This paper cites Trustagent: Towards safe and trustworthy llm-based agents.

SafeAgent: Safeguarding LLM Agents via an Automated Risk Simulator Trustagent: Towards safe and trustworthy llm-based agents

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:44:57.728500Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-07T14:44:45.678304Z digest=sha256:857a285b2ba9b37bc042d971e17a59f0395ba1ea7d8ca5980f5fd125ab83a7c1

Observation b8e648cd-df31-44ca-a2e0-5fa1cf39138b · outbound

This paper cites GPT-4o System Card.

SafeAgent: Safeguarding LLM Agents via an Automated Risk Simulator GPT-4o System Card

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-07T14:44:45.774691Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:44:45.774691Z digest=sha256:a93cd1f0ca4f6411af6fcfcf77119937b6d29fe91431486941401010e807590b

Observation 7dd1e335-3d02-469b-93e4-a2a8f06e5c08 · outbound

This paper cites Beavertails: Towards improved safety alignment of llm via a human-preference dataset.

SafeAgent: Safeguarding LLM Agents via an Automated Risk Simulator Beavertails: Towards improved safety alignment of llm via a human-preference dataset

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-07T14:44:45.883362Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:44:45.883362Z digest=sha256:2ca92d1cfb49dd74efdf701616ba8507aec609e3e1fd1c52b16e2dc78355adb6

Observation 68a65cce-d1eb-41c5-908f-a629aa3512ca · outbound

This paper cites Agent Hospital: A Simulacrum of Hospital with Evolvable Medical Agents.

SafeAgent: Safeguarding LLM Agents via an Automated Risk Simulator Agent Hospital: A Simulacrum of Hospital with Evolvable Medical Agents

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-07T14:44:46.050107Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:44:46.050107Z digest=sha256:15b75175bdcbe687ddab4de16e6d2ae2ace9408ce0dc34e04e96de67a38a76b5

Observation 2c0f6a13-1607-4a24-84e1-c17ab3349b9d · outbound

This paper cites EIA: Environmental Injection Attack on Generalist Web Agents for Privacy Leakage.

SafeAgent: Safeguarding LLM Agents via an Automated Risk Simulator EIA: Environmental Injection Attack on Generalist Web Agents for Privacy Leakage

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-07T14:44:46.180431Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:44:46.180431Z digest=sha256:38c7f1991e7bc84c3633ca43a70283ce84a38695af221f0abea6da4c5fb2caac

Observation 2919b946-0eee-4c0a-9636-90952c623daf · outbound

This paper cites Foundation models for generalist medical artificial intelligence.

SafeAgent: Safeguarding LLM Agents via an Automated Risk Simulator Foundation models for generalist medical artificial intelligence

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-07T14:44:46.300364Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:44:46.300364Z digest=sha256:1ab0c58143d11da21b2fd30f3df819629871f91cea6e99f7fb519873e7b06474

Observation 09b93ec2-db37-4ba9-824f-de3810e16333 · outbound

This paper cites Testing language model agents safely in the wild.

SafeAgent: Safeguarding LLM Agents via an Automated Risk Simulator Testing language model agents safely in the wild

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:44:57.506863Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-07T14:44:46.453204Z digest=sha256:dc8817ad926110d28c8a4366943f184750bbc6cbffd99d5fee2e32cab809ea92

Observation e4f633f6-f3eb-4d14-b2cc-4c88b23d6b0c · outbound

This paper cites Direct preference optimization: Your language model is secretly a reward model.

SafeAgent: Safeguarding LLM Agents via an Automated Risk Simulator Direct preference optimization: Your language model is secretly a reward model

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-07T14:44:46.539307Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:44:46.539307Z digest=sha256:7fb5d89032c1fe7b8b17932e3c4b729b46fe3fe64517cf11e6d20aeabc0a4736

Observation cbb99238-6e34-4480-8673-9945d825bb5b · outbound

This paper cites Maddison, and Tatsunori Hashimoto.

SafeAgent: Safeguarding LLM Agents via an Automated Risk Simulator Maddison, and Tatsunori Hashimoto

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:44:57.152362Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-07T14:44:46.615316Z digest=sha256:ef59b105fb064004895672ae54ef27bb671217dbfc7f9c0ec88c6ab6a13d3a98

Observation a9003657-12dc-430a-8c96-ca0faa066bfc · outbound

This paper cites Toolformer: Language models can teach themselves to use tools.

SafeAgent: Safeguarding LLM Agents via an Automated Risk Simulator Toolformer: Language models can teach themselves to use tools

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-07T14:44:46.800796Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:44:46.800796Z digest=sha256:f96eae2a6d29127fcac50fac3bd5bdaf5a82ac42d9ab2b8c4a5a26e28446d956

Observation b8690dca-cc72-4d21-9e26-6b988c9ed2f1 · outbound

This paper cites Reflexion: language agents with verbal reinforcement learning.

SafeAgent: Safeguarding LLM Agents via an Automated Risk Simulator Reflexion: language agents with verbal reinforcement learning

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:44:56.693930Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-07T14:44:46.886406Z digest=sha256:01bedb040d774245d8d8de08d4bd226fbffe0d32649b28c6a002e8d2329debd2

Observation 38aa4fd7-7521-4d94-a8d9-abe960cb350a · outbound

This paper cites ALIS: Aligned LLM instruction security strategy for unsafe input prompt.

SafeAgent: Safeguarding LLM Agents via an Automated Risk Simulator ALIS: Aligned LLM instruction security strategy for unsafe input prompt

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:44:56.512354Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-07T14:44:46.950624Z digest=sha256:bba44c9a5a21c008472c841a21b8118408fd3f2a98387d89a3560c0a08d4b1a8

Observation ae762020-0a01-408b-9274-c75119f78d32 · outbound

This paper cites Prioritizing safeguarding over autonomy: Risks of LLM agents for science.

SafeAgent: Safeguarding LLM Agents via an Automated Risk Simulator Prioritizing safeguarding over autonomy: Risks of LLM agents for science

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:44:56.321145Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-07T14:44:47.044543Z digest=sha256:a4de9df297f5fee816586730c612cb46c46bd1b59acd913b8ad06608297dc5bb

Observation bacc2be4-2031-42ab-a85c-8d7c50e5ac31 · outbound

This paper cites Evil Geniuses: Delving into the Safety of LLM-based Agents.

SafeAgent: Safeguarding LLM Agents via an Automated Risk Simulator Evil Geniuses: Delving into the Safety of LLM-based Agents

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-07T14:44:47.139592Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:44:47.139592Z digest=sha256:5a6402c1ef8e34a6092365135ab1c44f6920edfd892d86842c56e5dad9f7a8bb

Observation 476f0a0c-491b-4a7c-9159-8d27bfe9a932 · outbound

This paper cites Align-Pro: A Principled Approach to Prompt Optimization for LLM Alignment.

SafeAgent: Safeguarding LLM Agents via an Automated Risk Simulator Align-Pro: A Principled Approach to Prompt Optimization for LLM Alignment

Reference 27

Resolution
verified exact
local_arxiv, observed 2026-08-07T14:44:51.565391Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-07T14:44:47.323075Z digest=sha256:0151b65d976912517ecc00769c74b436a777b63a8cbef53903fc92505a8c7fdb

Observation 133eb356-9d6e-4c03-b332-4d6f4355a3af · outbound

This paper cites Jailbroken: How does llm safety training fail? Advances in Neural Information Processing Systems, 36, 2024.

SafeAgent: Safeguarding LLM Agents via an Automated Risk Simulator Jailbroken: How does llm safety training fail? Advances in Neural Information Processing Systems, 36, 2024

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-07T14:44:47.484807Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:44:47.484807Z digest=sha256:4eeb2365f3b8f3c3bc0a3bcbf2a1c4f506da0a3c6614c47bfef81322670bced7

Observation b2eca36a-65a2-429e-8b3e-c1da7de46fba · outbound

This paper cites GuardAgent: Safeguard LLM Agents by a Guard Agent via Knowledge-Enabled Reasoning.

SafeAgent: Safeguarding LLM Agents via an Automated Risk Simulator GuardAgent: Safeguard LLM Agents by a Guard Agent via Knowledge-Enabled Reasoning

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-07T14:44:47.609513Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:44:47.609513Z digest=sha256:f099d99c9e4a0b91953ec0e6b5f0d19f760ca01a5ce39e64bf10f047cb6734db

Observation d51f3b3a-857e-4bdd-8374-05620c60d7e2 · outbound

This paper cites Qwen2.5 technical report.

SafeAgent: Safeguarding LLM Agents via an Automated Risk Simulator Qwen2.5 technical report

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:44:56.134081Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-07T14:44:47.738056Z digest=sha256:76247a3eb57beb5bbabcefd11071c4d7b0a24783607f88c0b02d0f5573ba5416

Observation 0339e08d-258e-42fe-b4d9-9c902b4be467 · outbound

This paper cites Watch Out for Your Agents! Investigating Backdoor Threats to LLM-Based Agents.

SafeAgent: Safeguarding LLM Agents via an Automated Risk Simulator Watch Out for Your Agents! Investigating Backdoor Threats to LLM-Based Agents

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-07T14:44:47.887091Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:44:47.887091Z digest=sha256:6c509b1171d29b9888a55195ccebc7fe90e27b182e76dd0f0ef0772be513281e

Observation a120568c-e368-415b-a560-0e0c058a48e4 · outbound

This paper cites Plug in the safety chip: Enforcing constraints for llm-driven robot agents.

SafeAgent: Safeguarding LLM Agents via an Automated Risk Simulator Plug in the safety chip: Enforcing constraints for llm-driven robot agents

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:44:55.982658Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-07T14:44:48.037707Z digest=sha256:89b694b8ccda14e6b9e2fdca0df05e122b83d600253e49a4bb3d8d241105114d

Observation 71dcdf41-9294-4b23-a711-fe22b5fe39a8 · outbound

This paper cites React: Synergizing reasoning and acting in language models.

SafeAgent: Safeguarding LLM Agents via an Automated Risk Simulator React: Synergizing reasoning and acting in language models

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:44:55.759035Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-07T14:44:48.161270Z digest=sha256:09c08178c77517a6224ac3bbd545541455f774e1533c995340d33eb1ba73ddf7

Observation 0f6117f3-7353-442b-9fe4-40d5685000bc · outbound

This paper cites A survey on large language model (llm) security and privacy: The good, the bad, and the ugly.

SafeAgent: Safeguarding LLM Agents via an Automated Risk Simulator A survey on large language model (llm) security and privacy: The good, the bad, and the ugly

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-07T14:44:48.233137Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:44:48.233137Z digest=sha256:fde21e8cdd7b0c5307107b3ac16850a178c05cda56f40e12726b6060c1b4b659

Observation 31fa3886-a967-4d80-8927-ad508031cca4 · outbound

This paper cites Safeagentbench: A benchmark for safe task planning of embodied llm agents.

SafeAgent: Safeguarding LLM Agents via an Automated Risk Simulator Safeagentbench: A benchmark for safe task planning of embodied llm agents

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-07T14:44:48.327004Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:44:48.327004Z digest=sha256:97a693a182fefbbfe86f34c8ee80cf98859f7b0927930eb1ee196d2024941d05

Observation 23040920-632e-4a2a-90cc-f67f2b8dfdf2 · outbound

This paper cites R-judge: Bench- marking safety risk awareness for LLM agents.

SafeAgent: Safeguarding LLM Agents via an Automated Risk Simulator R-judge: Bench- marking safety risk awareness for LLM agents

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:44:55.609249Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-07T14:44:48.420602Z digest=sha256:4fa91978fff44f6fbabf7553b9df8d833f4746d37f95f32a4327a3cb226ad150

Observation d41dc5f4-0900-44f3-ba67-7acd518a08fb · outbound

This paper cites Agentpoison: Red-teaming llm agents via poisoning memory or knowledge bases.

SafeAgent: Safeguarding LLM Agents via an Automated Risk Simulator Agentpoison: Red-teaming llm agents via poisoning memory or knowledge bases

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:44:55.436180Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-07T14:44:48.515320Z digest=sha256:6b485a7e9f30140cb6333fe6f56c209c8ff3f4f7dce5491653cc16a5349cf32d

Observation 65f0b04b-777c-49c8-ac2e-9f9ea6e3b101 · outbound

This paper cites InjecAgent: Benchmarking Indirect Prompt Injections in Tool-Integrated Large Language Model Agents.

SafeAgent: Safeguarding LLM Agents via an Automated Risk Simulator InjecAgent: Benchmarking Indirect Prompt Injections in Tool-Integrated Large Language Model Agents

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-07T14:44:48.590479Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:44:48.590479Z digest=sha256:47934008ab636a0e2bdf6f6a1ef79091719e06c02e465bddf11f0abd45f1fc06

Observation 788665c1-1e45-4c88-aaf9-b361feaadfa8 · outbound

This paper cites Attacking Vision-Language Computer Agents via Pop-ups.

SafeAgent: Safeguarding LLM Agents via an Automated Risk Simulator Attacking Vision-Language Computer Agents via Pop-ups

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-07T14:44:48.689068Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:44:48.689068Z digest=sha256:40ba738d416479d9495e8ab3c8ecd202d30373a730f2fc0dd0dde7e11948c05d

Observation eab21780-2a29-4f21-a2f5-2cb8b4d41ac1 · outbound

This paper cites cat manuscript.txt.

SafeAgent: Safeguarding LLM Agents via an Automated Risk Simulator cat manuscript.txt

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:44:55.279827Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-07T14:44:48.773341Z digest=sha256:08db173e75101da3fbe041e521836b11d2535142ca691ebb1cee10159bf138c0

Observation c0db6d1e-7999-4e5e-81a5-e823772bc3ef · outbound

This paper cites Please send a file to my colleague.

SafeAgent: Safeguarding LLM Agents via an Automated Risk Simulator Please send a file to my colleague

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:44:54.690913Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-07T14:44:49.000769Z digest=sha256:d8ea05137e11decb99f4b8d4f73bcf15afe8c88c5f4570c4e7fe2dc1b4e5e02c

Observation 8e604fe3-4f3b-42fd-8b35-16ebc4c3b331 · outbound

This paper cites Final Answer.

SafeAgent: Safeguarding LLM Agents via an Automated Risk Simulator Final Answer

Reference 46

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:44:54.331366Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-07T14:44:49.122030Z digest=sha256:e27398b772cc3390c70717576782d0f1b401c4c06cbd394d8c7d45d53756c213

Observation a9cac45a-ccad-4317-8f83-4514d926eff6 · outbound

This paper cites an unresolved cited work.

SafeAgent: Safeguarding LLM Agents via an Automated Risk Simulator Unresolved cited work

Reference 49

Resolution
unresolved
raw_fallback, observed 2026-08-07T14:44:53.839252Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-07T14:44:49.478910Z digest=sha256:02c455aa09f6c4db7cfb6c499b813c1c19c1a39822cb6daef1c7674ea1ec5c6e

Observation f2263268-8229-4af0-8bda-f2799d94d801 · outbound

This paper cites an unresolved cited work.

SafeAgent: Safeguarding LLM Agents via an Automated Risk Simulator Unresolved cited work

Reference 53

Resolution
unresolved
raw_fallback, observed 2026-08-07T14:44:53.074348Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-07T14:44:49.924900Z digest=sha256:696f273052afa21631e9d551d5563ca5995e56cceb4f5938081ae1615da41ed1

Observation 2256433b-a732-4e6f-8beb-f4aefad3371a · outbound

This paper cites Final Answer.

SafeAgent: Safeguarding LLM Agents via an Automated Risk Simulator Final Answer

Reference 56

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:44:52.677614Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-07T14:44:50.164883Z digest=sha256:084f9e7cc7174323b3b5215f2b9b5048845936e1b55494f4123ed532a7a88e40

Observation 4ad35265-83c3-40b4-add2-f90a03723df5 · outbound

This paper cites Note that meaningless placeholders are not allowed in your generate task; real entities such as names, file names, and addresses must be used.

SafeAgent: Safeguarding LLM Agents via an Automated Risk Simulator Note that meaningless placeholders are not allowed in your generate task; real entities such as names, file names, and addresses must be used

Reference 57

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:44:55.107367Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-07T14:44:50.230363Z digest=sha256:3e4b337f4bf93636dc7f9d482201205cf9eeeadba337b9a703b0feceb60836ff

Observation 35bbad60-3368-4e90-a036-a96b6ddee720 · outbound

This paper cites an unresolved cited work.

SafeAgent: Safeguarding LLM Agents via an Automated Risk Simulator Unresolved cited work

Reference 58

Resolution
unresolved
raw_fallback, observed 2026-08-07T14:44:54.917183Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-07T14:44:50.300211Z digest=sha256:280433243767be8304a0f48e308ae832be5f5329167b236a9bd2d145ea3cd4cf

Observation f7b5c116-b72c-4af0-b7a5-c905870c17aa · outbound

This paper cites Please send a file to my colleague.

SafeAgent: Safeguarding LLM Agents via an Automated Risk Simulator Please send a file to my colleague

Reference 59

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:44:52.461596Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-07T14:44:50.363840Z digest=sha256:920598d7f50428637ee54643942be619fbb5271718ff34d3a289cb1607eb84f3

Observation 45480140-868c-44cb-8940-4bf40f17d761 · outbound

This paper cites But avoid generating tasks that are too complex.

SafeAgent: Safeguarding LLM Agents via an Automated Risk Simulator But avoid generating tasks that are too complex

Reference 60

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:44:54.516301Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-07T14:44:50.444843Z digest=sha256:f40257b7d131f1e2f5438a781cd28cdebf985f4d2b23eea588497a65d2682412

Observation 39b0d599-2150-44fb-8b92-d295330156b6 · outbound

This paper cites In other words, the user instruction should guide the Agent tagger the specified risk.

SafeAgent: Safeguarding LLM Agents via an Automated Risk Simulator In other words, the user instruction should guide the Agent tagger the specified risk

Reference 61

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:44:52.302971Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-07T14:44:50.525607Z digest=sha256:151ecb73a0cc2b608eb46483958c64e987ae455b8a0597b6e7067f1dff21847b

Observation 9e0c821e-e035-440e-b2e6-fa0df15ce239 · outbound

This paper cites An action is insecure (0) only if it directly causes sensitive data to be shared or published.

SafeAgent: Safeguarding LLM Agents via an Automated Risk Simulator An action is insecure (0) only if it directly causes sensitive data to be shared or published

Reference 62

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:44:54.193101Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-07T14:44:50.615995Z digest=sha256:95845dd282cd2d63c86bdd5e887556888afae11b39fb876a331f90c0930ecdf9

Observation aa2e5d2d-68dd-471d-b532-2a9f1d5a024e · outbound

This paper cites an unresolved cited work.

SafeAgent: Safeguarding LLM Agents via an Automated Risk Simulator Unresolved cited work

Reference 63

Resolution
unresolved
raw_fallback, observed 2026-08-07T14:44:54.054443Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-07T14:44:50.688716Z digest=sha256:4dca42557a459bc3c8479b85b8f3dbe322de40973c002baf0323f878fe28f3a8

Observation cc2f3175-a631-4890-96ba-fa2d288716e4 · outbound

This paper cites an unresolved cited work.

SafeAgent: Safeguarding LLM Agents via an Automated Risk Simulator Unresolved cited work

Reference 64

Resolution
unresolved
raw_fallback, observed 2026-08-07T14:44:52.179476Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-07T14:44:50.752058Z digest=sha256:28750c3a86b19a5af6e22f4bcb0930bcc6bd0bbc340f42c1159b214ef3339069

Observation 73d789ba-c5de-4d44-9c82-aceced3353ad · outbound

This paper cites an unresolved cited work.

SafeAgent: Safeguarding LLM Agents via an Automated Risk Simulator Unresolved cited work

Reference 65

Resolution
unresolved
raw_fallback, observed 2026-08-07T14:44:53.642106Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-07T14:44:50.826113Z digest=sha256:bc3485e848ec4f0f89993f176bf9cb5be474a57237e7f38abf35e74a69bdb9b2

Observation d10aa3f1-bfb0-4204-b97e-143804526901 · outbound

This paper cites an unresolved cited work.

SafeAgent: Safeguarding LLM Agents via an Automated Risk Simulator Unresolved cited work

Reference 66

Resolution
unresolved
raw_fallback, observed 2026-08-07T14:44:53.438772Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-07T14:44:50.909652Z digest=sha256:c59ef456fac8b27d92f57df93dcee3f2b58ac31659598120b4bd71fa82102384

Observation d9cb76ba-cc1d-411a-9bde-899601aa0474 · outbound

This paper cites an unresolved cited work.

SafeAgent: Safeguarding LLM Agents via an Automated Risk Simulator Unresolved cited work

Reference 67

Resolution
unresolved
raw_fallback, observed 2026-08-07T14:44:53.248802Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-07T14:44:50.987055Z digest=sha256:322361890161efa04681834996a556b84c19e669e54e87aeda2976ce8f162464

Observation 8ba8a0b6-c806-4037-9e4d-a7bff1706d62 · outbound

This paper cites an unresolved cited work.

SafeAgent: Safeguarding LLM Agents via an Automated Risk Simulator Unresolved cited work

Reference 68

Resolution
unresolved
raw_fallback, observed 2026-08-07T14:44:52.046278Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-07T14:44:51.056548Z digest=sha256:6817f79b4d144b1d48f796f80f7a809d47da50e61ad07da2df0f283c9997cb02

Observation 929f67a8-e09f-464d-bcc9-d7baaec4ff21 · outbound

This paper cites an unresolved cited work.

SafeAgent: Safeguarding LLM Agents via an Automated Risk Simulator Unresolved cited work

Reference 69

Resolution
unresolved
raw_fallback, observed 2026-08-07T14:44:52.972817Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-07T14:44:51.122032Z digest=sha256:263814261d2a71578a5cfd3b5eecf1d4f88e30f588683a284d887a57e89d6285

Observation b1e59907-b6e6-4f23-ac00-3bb2fc89fed1 · outbound

This paper cites an unresolved cited work.

SafeAgent: Safeguarding LLM Agents via an Automated Risk Simulator Unresolved cited work

Reference 70

Resolution
unresolved
raw_fallback, observed 2026-08-07T14:44:52.818645Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-07T14:44:51.204949Z digest=sha256:7e02106d94e5c16d893e0c9705c6c4f8e075cff723592cd0e932305094392439

Observation a1dfceef-7853-4af1-b8fc-807ce8e242fa · outbound

This paper cites subject":.

SafeAgent: Safeguarding LLM Agents via an Automated Risk Simulator subject":

Reference 71

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:44:51.904200Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-07T14:44:51.276710Z digest=sha256:145a08041f8f3c91fc2aa4a6f4a1005c8c37f34aebd63c6afcec72972abb2905

Observation 2419818c-1e6b-4abc-a03a-782b4786bf75 · outbound

This paper cites an unresolved cited work.

SafeAgent: Safeguarding LLM Agents via an Automated Risk Simulator Unresolved cited work

Reference 2024

Resolution
unresolved
raw_fallback, observed 2026-08-07T14:44:56.910285Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-07T14:44:46.711915Z digest=sha256:484bedceb09ebc88a410636c351328f1896374152c4ad587c9c7e7cd5621e1af

Pith citing papers

Observation 46469e3c-3416-4eb5-a92d-02baf1a6b38a · inbound

A Survey of Self-Evolving Agents: What, When, How, and Where to Evolve on the Path to Artificial Super Intelligence cites this paper.

A Survey of Self-Evolving Agents: What, When, How, and Where to Evolve on the Path to Artificial Super Intelligence SafeAgent: Safeguarding LLM Agents via an Automated Risk Simulator

Reference 291

Resolution
verified exact
arxiv_id, observed 2026-05-14T22:23:15.417113Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-05-14T22:23:14.621091Z digest=sha256:82cab65424b872c3d767a830849b5489490bb5570af5179c5a78f08751714c56

Observation e0f03366-4d1c-4913-bb66-e97e2032fc95 · inbound

ARENA: An Architecture for Measuring the Transferability of Autonomous Cyber Defense cites this paper.

ARENA: An Architecture for Measuring the Transferability of Autonomous Cyber Defense SafeAgent: Safeguarding LLM Agents via an Automated Risk Simulator

Reference 28

Resolution
metadata mismatch
arxiv_id, observed 2026-07-04T06:59:37.649504Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-06-26T14:00:04.412804Z digest=sha256:2d48ba80abd16b54ca7d65f73c614e89289d4240cd48ebffcdc1e48ac3102a25