Pith. sign in

Paper Citation Record · LEDGER

The Limits of Predicting Agents from Behaviour

As of 23 August 2026, this Paper Citation Record lists 20 of 20 outbound references and 1 inbound Pith citation observation for arXiv:2506.02923.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2506.02923 v1

Coverage vector

measured 20 of 20 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T11:25:27.476820Z

measured 21 of 21 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-22T06:32:14.747728+00:00

measured 1 of 1 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-06-27T10:12:19.552548Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-03T10:17:57.278786Z

Reference resolution

20 of 20 outbound references displayed

  • verified exact1
  • verified fuzzy6
  • unresolved13
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 7950fef2-220c-46ef-b56c-baca23b22dfb · outbound

This paper cites assumption-free.

The Limits of Predicting Agents from Behaviour assumption-free

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:25:28.677826Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-07T11:25:27.400658Z digest=sha256:8180082697eec03bfbaf1cf67cfc533fdb16b8cdbebbd183c71cf7fd4c9c3428

Observation 62022ce5-b7dd-458f-885e-073ae0566f57 · outbound

This paper cites Superintelligent Agents Pose Catastrophic Risks: Can Scientist AI Offer a Safer Path?.

The Limits of Predicting Agents from Behaviour Superintelligent Agents Pose Catastrophic Risks: Can Scientist AI Offer a Safer Path?

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-07T11:25:27.095118Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:25:27.095118Z digest=sha256:3f7ccc0becffd9d878ddbea3d48587dcf1aa10f672cbc8f1bd347c9c3d6b56f8

Observation 2638d28b-679f-41d6-8d07-0b41a562c222 · outbound

This paper cites Does ChatGPT Have a Mind?.

The Limits of Predicting Agents from Behaviour Does ChatGPT Have a Mind?

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-07T11:25:27.213047Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:25:27.213047Z digest=sha256:994bdefb02176a411057f9e233a6695a220ec8bb89f76dc6462e86fe137cc496

Observation 2b3eb09e-176d-4597-af94-b601178faaa1 · outbound

This paper cites Language Models Represent Space and Time.

The Limits of Predicting Agents from Behaviour Language Models Represent Space and Time

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-07T11:25:27.217058Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:25:27.217058Z digest=sha256:4a7967b5b41b11a470a498725c5f938a1a2e6dd05cbf7776cd3fff8e7198822f

Observation 1ed8737b-4b4e-4a77-88f8-27b58863fe40 · outbound

This paper cites Emergent World Representations: Exploring a Sequence Model Trained on a Synthetic Task.

The Limits of Predicting Agents from Behaviour Emergent World Representations: Exploring a Sequence Model Trained on a Synthetic Task

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-07T11:25:27.253885Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:25:27.253885Z digest=sha256:e2947655b6a296e11374e734d70968dc3383157136c4f0f51b526a169384523e

Observation 9d97dfa0-8ecb-42d8-ac3d-f806e2f31504 · outbound

This paper cites Robust agents learn causal world models.

The Limits of Predicting Agents from Behaviour Robust agents learn causal world models

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-07T11:25:27.272210Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:25:27.272210Z digest=sha256:d5d5f9ee10e7cde456e7918ddc2d470fd5ba3876acedf26593979b2098eaeba3

Observation 70e9df7c-176a-40f9-9fba-98c2a26d905f · outbound

This paper cites Bounds on the conditional and average treatment effect with unobserved confounding factors.

The Limits of Predicting Agents from Behaviour Bounds on the conditional and average treatment effect with unobserved confounding factors

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-07T11:25:27.359955Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:25:27.359955Z digest=sha256:88f03915632d7ee2a5d71d5c1a38a747d5d4bd9f7234e9d2b2d834028b25ab2d

Observation 3870e3e7-28a4-42d7-a433-b05dc9e897ae · outbound

This paper cites Were I to intervene in the environment, what action do you believe is optimal?.

The Limits of Predicting Agents from Behaviour Were I to intervene in the environment, what action do you believe is optimal?

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:25:28.527939Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-07T11:25:27.414471Z digest=sha256:d7d1a899be6edec6474789d1c6a6c19485c38177b1c7791ad7b2135d92eb02ab

Observation a8974cc0-590c-4fa8-9617-de36a99e1a0b · outbound

This paper cites 𝐴 is a difference of two terms written𝐴(𝒓)=𝐴 1(𝒓)−𝐴 2(𝒓).

The Limits of Predicting Agents from Behaviour 𝐴 is a difference of two terms written𝐴(𝒓)=𝐴 1(𝒓)−𝐴 2(𝒓)

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:25:28.339013Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-07T11:25:27.427661Z digest=sha256:d03a56e4b0bf9ed1e19ae7bf80546607abfcc9dbcc886c41af968221041186c0

Observation b63fbb39-5b9a-465d-af40-336d39e097fd · outbound

This paper cites The nature of the modification is unknown but we are told that after modification, the expected probability of𝑪 is given by𝑃𝜎,𝑑(𝑪), assumed to be known and internalised by the A.

The Limits of Predicting Agents from Behaviour The nature of the modification is unknown but we are told that after modification, the expected probability of𝑪 is given by𝑃𝜎,𝑑(𝑪), assumed to be known and internalised by the A

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:25:28.205541Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-07T11:25:27.436669Z digest=sha256:ce4d38992a7a5ba0b2f6939d388256353e03c8149bab935cc76377b581a6b585

Observation 645d3b13-869c-4516-ade9-0137cb168b1f · outbound

This paper cites good” or “beneficial.

The Limits of Predicting Agents from Behaviour good” or “beneficial

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:25:27.983506Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-07T11:25:27.476820Z digest=sha256:61d31100311c129cf7b1f6f108f8e3f3bf3de86a125e7eef09c15391e3349ce8

Observation 2ca736b1-553e-40ca-b613-b5481fe69588 · outbound

This paper cites Towards Resolving Unidentifiability in Inverse Reinforcement Learning.

The Limits of Predicting Agents from Behaviour Towards Resolving Unidentifiability in Inverse Reinforcement Learning

Reference 1967

Resolution
unresolved
no resolver link, observed 2026-08-07T11:25:27.084038Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:25:27.084038Z digest=sha256:ff8d4daa089e95933f64674e28671b477bd41ef901f35fc4a58f02f5a22c6f46

Observation 41c195e1-68e0-4163-b20d-f9b5f395fb77 · outbound

This paper cites an unresolved cited work.

The Limits of Predicting Agents from Behaviour Unresolved cited work

Reference 1972

Resolution
unresolved
raw_fallback, observed 2026-08-07T11:25:28.846022Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-07T11:25:27.314884Z digest=sha256:53fb00c534de11211e007fdd1f688717cf632f204b768a5a090ba1091f8998ef

Observation bf64d97c-13b1-43df-979f-acdf4f64c7c6 · outbound

This paper cites Towards Guaranteed Safe AI: A Framework for Ensuring Robust and Reliable AI Systems.

The Limits of Predicting Agents from Behaviour Towards Guaranteed Safe AI: A Framework for Ensuring Robust and Reliable AI Systems

Reference 1996

Resolution
unresolved
no resolver link, observed 2026-08-07T11:25:27.207549Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:25:27.207549Z digest=sha256:a909a4f7dfe26968d9652a3e93794ab6163431e13dd6b4e21c9aa4b579404689

Observation 6133a42e-5572-4e9d-aeb2-e72494a0ea22 · outbound

This paper cites Risks from Learned Optimization in Advanced Machine Learning Systems.

The Limits of Predicting Agents from Behaviour Risks from Learned Optimization in Advanced Machine Learning Systems

Reference 2008

Resolution
unresolved
no resolver link, observed 2026-08-07T11:25:27.238331Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:25:27.238331Z digest=sha256:ad458ff503ffc4c93a53973753b5d45d1f7e1c330805226895d9f43f8ce1dcda

Observation 12220a8d-aaf4-4f1a-a636-0d8397a58051 · outbound

This paper cites Preference elicitation and inverse reinforcement learning.

The Limits of Predicting Agents from Behaviour Preference elicitation and inverse reinforcement learning

Reference 2010

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:25:29.038826Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-07T11:25:27.294578Z digest=sha256:5b38c47b4367f7ab1a12a246b9bdd9aef749f016703a893275ab342370e49046

Observation dd62d694-1fe9-4ecd-a76a-240c481816de · outbound

This paper cites Partial Counterfactual Identification from Observational and Experimental Data.

The Limits of Predicting Agents from Behaviour Partial Counterfactual Identification from Observational and Experimental Data

Reference 2021

Resolution
verified exact
local_arxiv, observed 2026-08-07T11:25:27.666163Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-07T11:25:27.384918Z digest=sha256:34b21e938436331febdebc0a07733526aabf3e66f5047d94ed43767aea677df9

Observation 6ddb955a-133d-4239-b92a-3487ea86450d · outbound

This paper cites Evaluating the World Model Implicit in a Generative Model.

The Limits of Predicting Agents from Behaviour Evaluating the World Model Implicit in a Generative Model

Reference 2022

Resolution
unresolved
no resolver link, observed 2026-08-07T11:25:27.336907Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:25:27.336907Z digest=sha256:21155c0acb5b5c1b28b9a7617ecb4d76d5436a974f6b561478968f681d2c7e9a

Observation 821071bc-3983-4688-a749-f6c0789e8f8c · outbound

This paper cites Subjective Causality.

The Limits of Predicting Agents from Behaviour Subjective Causality

Reference 2023

Resolution
unresolved
no resolver link, observed 2026-08-07T11:25:27.223716Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:25:27.223716Z digest=sha256:af1a917a68dfe255d04f6747653331844336eaa06b3e3f5968f15ced5171ee5f

Observation b2ea569c-aed3-423c-b25d-2431cbad024b · outbound

This paper cites Can a Bayesian Oracle Prevent Harm from an Agent?.

The Limits of Predicting Agents from Behaviour Can a Bayesian Oracle Prevent Harm from an Agent?

Reference 2024

Resolution
unresolved
no resolver link, observed 2026-08-07T11:25:27.089122Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:25:27.089122Z digest=sha256:a9e229542f735ee788c417486d59f52582211aa9972af1b91a36aceb7448c5a2

Pith citing papers

Observation dda7c65b-5681-463d-9868-0b2b5ddbf6cd · inbound

The Impossibility of Eliciting Latent Knowledge cites this paper.

The Impossibility of Eliciting Latent Knowledge The Limits of Predicting Agents from Behaviour

Reference 1

Resolution
verified exact
arxiv_id, observed 2026-07-03T10:17:57.280043Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-06-27T10:12:19.552548Z digest=sha256:6d7568119f426ed49c53e3ca254cd3eaa9f7ec942022cc9eecec26932ac77465