Pith. sign in

Paper Citation Record · LEDGER

Automating and Scaling Behavioral Scientific Research on AI Agents

As of 16 August 2026, this Paper Citation Record lists 79 of 79 outbound references and 0 inbound Pith citation observations for arXiv:2608.10030.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2608.10030 v1

Coverage vector

measured 79 of 79 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-14T04:24:51.705273Z

measured 79 of 79 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-16T06:30:59.297886+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

79 of 79 outbound references displayed

  • verified exact8
  • verified fuzzy16
  • unresolved50
  • parse uncertain0
  • malformed identifier3
  • metadata mismatch2

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 806f8c5d-9b70-4f4b-88fb-a8da2abe3c45 · outbound

This paper cites When Persuasion Overrides Truth in Multi-Agent LLM Debates: Introducing a Confidence-Weighted Persuasion Override Rate (CW-POR).

Automating and Scaling Behavioral Scientific Research on AI Agents When Persuasion Overrides Truth in Multi-Agent LLM Debates: Introducing a Confidence-Weighted Persuasion Override Rate (CW-POR)

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-14T04:24:50.620020Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-14T04:24:50.620020Z digest=sha256:1d20f782627e3032bfc9819c75b25c0efef69129a1403d68bbdf97c99965f1f1

Observation 3f91f124-6274-4282-9777-b4d640c82977 · outbound

This paper cites Playing repeated games with large language models.Nature Human Behaviour, 9 (7):1380–1390, 2025.

Automating and Scaling Behavioral Scientific Research on AI Agents Playing repeated games with large language models.Nature Human Behaviour, 9 (7):1380–1390, 2025

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-14T04:24:50.630310Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-14T04:24:50.630310Z digest=sha256:991c077420c6e95f3b2729d67cd1ba07b1feb66d88f0990dd8277534898b5cae

Observation 2bdf8d5d-94ec-4eeb-94f9-f3ded4844c98 · outbound

This paper cites Information Discernment in Large Language Models.

Automating and Scaling Behavioral Scientific Research on AI Agents Information Discernment in Large Language Models

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-14T04:24:50.650726Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-14T04:24:50.650726Z digest=sha256:4c5af93eef9fd6e5bba1bccb99a50f79e7f8fcc0975ca4f1344591f98c91a6b9

Observation 1383f290-c18b-4775-a234-a5de68d7762c · outbound

This paper cites Jagadish, Or Duek, Ilan Harpaz-Rotem, Marie-Christine Khorsandian, Achim Burrer, Erich Seifritz, Philipp Homan, Eric Schulz, and Tobias R.

Automating and Scaling Behavioral Scientific Research on AI Agents Jagadish, Or Duek, Ilan Harpaz-Rotem, Marie-Christine Khorsandian, Achim Burrer, Erich Seifritz, Philipp Homan, Eric Schulz, and Tobias R

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-14T04:24:50.662840Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-14T04:24:50.662840Z digest=sha256:280cdcd3cd1c9abec2c3f2154fc5ac6f67fdfc711722e5dee516cd568650dbc5

Observation b9f7c56a-1de2-41dd-a790-44c62cd96c29 · outbound

This paper cites Inducing state anxiety in llm agents reproduces human-like biases in consumer decision-making.npj Artificial Intelli- gence, 2(1):55, 2026.

Automating and Scaling Behavioral Scientific Research on AI Agents Inducing state anxiety in llm agents reproduces human-like biases in consumer decision-making.npj Artificial Intelli- gence, 2(1):55, 2026

Reference 5

Resolution
verified exact
doi, observed 2026-08-14T04:24:52.079152Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-14T04:24:50.681803Z digest=sha256:a189c4deb86ea02374795b176de0ab3498af95d5bbd9c86680ce9a2f26062600

Observation 589efa11-535a-46c5-ab18-e7595566622e · outbound

This paper cites Modelling monotonic effects of ordinal predictors in bayesian regression models.British Journal of Mathematical and Statistical Psychology, 73(3):420–451, 2020.

Automating and Scaling Behavioral Scientific Research on AI Agents Modelling monotonic effects of ordinal predictors in bayesian regression models.British Journal of Mathematical and Statistical Psychology, 73(3):420–451, 2020

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-14T04:24:50.707456Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-14T04:24:50.707456Z digest=sha256:f1b1d3544f1505cb8e6ab508061cfa2338273a37468195777d0f7827c0bedf22

Observation 398791e0-1613-4354-a699-f0a353134c28 · outbound

This paper cites I want to break free! persuasion and anti-social behavior of LLMs in multi-agent settings with social hierarchy.Transactions on Machine Learning Research, 2025.

Automating and Scaling Behavioral Scientific Research on AI Agents I want to break free! persuasion and anti-social behavior of LLMs in multi-agent settings with social hierarchy.Transactions on Machine Learning Research, 2025

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T04:24:57.217227Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-14T04:24:50.734948Z digest=sha256:62543511f52b253cbcc46c54f876c56a570937bad49f26dfac3b4b9715333aa5

Observation 7f969a5d-d7cb-425c-9ec6-a42a81c81952 · outbound

This paper cites ELEPHANT: Measuring and understanding social sycophancy in LLMs.

Automating and Scaling Behavioral Scientific Research on AI Agents ELEPHANT: Measuring and understanding social sycophancy in LLMs

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-14T04:24:50.768438Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-14T04:24:50.768438Z digest=sha256:fcc8dfc3e45a5297e96d4327f37b8bc7dd868adf1b25352679cc40cfbb3c02e5

Observation 6d2f408b-7813-4c93-ac74-1f1b2ae51bc2 · outbound

This paper cites A framework for studying AI agent behavior: Evidence from consumer choice experiments.

Automating and Scaling Behavioral Scientific Research on AI Agents A framework for studying AI agent behavior: Evidence from consumer choice experiments

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T04:24:57.194436Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-14T04:24:50.775423Z digest=sha256:de6b7be663fa12dec79c8d99183959e85de47e8d0f62e746cd792668df1b1d9d

Observation 3090842a-7c45-4a9b-8bcd-2eaff2d0ef48 · outbound

This paper cites Herd Behavior: Investigating Peer Influence in LLM-based Multi-Agent Systems.

Automating and Scaling Behavioral Scientific Research on AI Agents Herd Behavior: Investigating Peer Influence in LLM-based Multi-Agent Systems

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-14T04:24:50.787214Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-14T04:24:50.787214Z digest=sha256:86fc972e9547f36f8529c11bb4019fd8eb294a051eb8b5b42bb840721e9bdf02

Observation 2e2a8155-da1a-41b1-aaa3-78681e11ed4b · outbound

This paper cites Estimating the reproducibility of psychological science.Science, 349(6251):aac4716, 2015.

Automating and Scaling Behavioral Scientific Research on AI Agents Estimating the reproducibility of psychological science.Science, 349(6251):aac4716, 2015

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-14T04:24:50.808228Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-14T04:24:50.808228Z digest=sha256:80d99350ee3bf258e07af8cbf6b720e6cc623dba5ee058906f60113de7c8739d

Observation 4d37b749-d358-4223-a03d-1799a59a109e · outbound

This paper cites GameBench: Evaluating Strategic Reasoning Abilities of LLM Agents.

Automating and Scaling Behavioral Scientific Research on AI Agents GameBench: Evaluating Strategic Reasoning Abilities of LLM Agents

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-14T04:24:50.818635Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-14T04:24:50.818635Z digest=sha256:bf2a603e06e93ed06fa275cddb81a7ba852b2f1bba2c8a6dc1bde212f765957e

Observation ba42dbdc-7d2f-433e-a75c-477158e5e5c1 · outbound

This paper cites an unresolved cited work.

Automating and Scaling Behavioral Scientific Research on AI Agents Unresolved cited work

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-14T04:24:50.833112Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-14T04:24:50.833112Z digest=sha256:466cd71af9f2b34a27ecde5e3a8b03b20ce5f5251a388edd89376745605a5c2d

Observation 520220c4-ba3c-430b-86f5-4a0174d83abf · outbound

This paper cites AI on my shoulder: Supporting emotional labor in front-office roles with an LLM-based empathetic coworker.

Automating and Scaling Behavioral Scientific Research on AI Agents AI on my shoulder: Supporting emotional labor in front-office roles with an LLM-based empathetic coworker

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-14T04:24:50.838455Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-14T04:24:50.838455Z digest=sha256:7e45b0ce51d509dd8b77f31685e767ff4cdc3924f17c9aed27c8a6237ee49f44

Observation 9d8698d2-df69-4860-a308-978e30e9d8c4 · outbound

This paper cites MAEBE: Multi-Agent Emergent Behavior Framework.

Automating and Scaling Behavioral Scientific Research on AI Agents MAEBE: Multi-Agent Emergent Behavior Framework

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-14T04:24:50.855799Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-14T04:24:50.855799Z digest=sha256:a2df22d8bef6b591b3f8261dfb5f948e643dbb2c092e225aeda360577d40f1d6

Observation 9080aba5-86f2-4623-889e-8048cb46cf05 · outbound

This paper cites an unresolved cited work.

Automating and Scaling Behavioral Scientific Research on AI Agents Unresolved cited work

Reference 17

Resolution
unresolved
raw_fallback, observed 2026-08-14T04:24:57.152833Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-14T04:24:50.888062Z digest=sha256:1c06f734fe37a3ae180ed4974c77fff03819a559d0fb008178de632483393cd2

Observation de6249ec-9ec2-42e1-bc4b-11492f39cec6 · outbound

This paper cites Alignment faking in large language models.

Automating and Scaling Behavioral Scientific Research on AI Agents Alignment faking in large language models

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-14T04:24:50.897678Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-14T04:24:50.897678Z digest=sha256:e12e3ef604f8c771c08e701094c49f722cee2069e27b88e141198033523bdf43

Observation d3c4c112-9685-4687-a7e7-8b228e29ef31 · outbound

This paper cites Bowman, and Sara Price.

Automating and Scaling Behavioral Scientific Research on AI Agents Bowman, and Sara Price

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T04:24:57.074507Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-14T04:24:50.909446Z digest=sha256:cdc218ac472fc2480b0565176fcd4b749582cb91f0774f98a2d188854d2a3b6c

Observation ec3f4417-30c3-42eb-aa3c-4100b5de8991 · outbound

This paper cites Deceptionbench: A comprehensive benchmark for AI deception behaviors in real-world scenarios.

Automating and Scaling Behavioral Scientific Research on AI Agents Deceptionbench: A comprehensive benchmark for AI deception behaviors in real-world scenarios

Reference 20

Resolution
verified exact
doi, observed 2026-08-14T04:24:52.001349Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-14T04:24:50.918468Z digest=sha256:98689bf687e59d192ee559a2f5de173b5e37889bccb5aa10e69033ae3beb6c27

Observation d6087938-5963-4976-b564-e5554fb0b791 · outbound

This paper cites Per- sonaLLM: Investigating the ability of large language models to express personality traits.

Automating and Scaling Behavioral Scientific Research on AI Agents Per- sonaLLM: Investigating the ability of large language models to express personality traits

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-14T04:24:50.935705Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-14T04:24:50.935705Z digest=sha256:6754d6f2b2e953eec66bfa3ad2a842b49bae594d0a125d3845d146e05a0650ff

Observation 7d6b2803-1854-4015-9d15-164d444ec423 · outbound

This paper cites FollowBench: A multi-level fine-grained constraints following benchmark for large language models.

Automating and Scaling Behavioral Scientific Research on AI Agents FollowBench: A multi-level fine-grained constraints following benchmark for large language models

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T04:24:57.004879Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-14T04:24:50.960570Z digest=sha256:26dc04d55638b0d99b10650059a4c2cfbd3cd1b36b896c27b75cc8e3623bdf0c

Observation 22d42a15-f3ac-4ab0-b402-1240c5d36cfb · outbound

This paper cites Can large language models be good emotional supporter? miti- gating preference bias on emotional support conversation.

Automating and Scaling Behavioral Scientific Research on AI Agents Can large language models be good emotional supporter? miti- gating preference bias on emotional support conversation

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-14T04:24:50.970341Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-14T04:24:50.970341Z digest=sha256:dcea970a21933ba5b3a4b05661978f7a2262e837ea340d32444f204b093554d2

Observation cb015e79-0394-480b-820b-ca9b3ed51240 · outbound

This paper cites Toward a science of ai agent societies.

Automating and Scaling Behavioral Scientific Research on AI Agents Toward a science of ai agent societies

Reference 24

Resolution
metadata mismatch
raw_fallback, observed 2026-08-14T04:24:53.734753Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-14T04:24:50.981263Z digest=sha256:9a3f0254bb8262bde1843b9b744df54f38adb122809b38f1514118c82d2e0135

Observation 2aea2229-0122-44f5-9154-07e1783ab6ae · outbound

This paper cites Aerobat code and data repository.

Automating and Scaling Behavioral Scientific Research on AI Agents Aerobat code and data repository

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T04:24:56.924832Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-14T04:24:51.002229Z digest=sha256:1af667cc915b71c046519264c71c158ebeedd00c46450322236ec2b8c0fe4ffa

Observation 4f1c2a1f-4b68-45c4-9e1b-907c4ffd9409 · outbound

This paper cites Emergence of psychopathological computations in large language models.arXiv preprint arXiv:2504.08016, 2025.

Automating and Scaling Behavioral Scientific Research on AI Agents Emergence of psychopathological computations in large language models.arXiv preprint arXiv:2504.08016, 2025

Reference 26

Resolution
verified exact
raw_fallback, observed 2026-08-14T04:24:53.514750Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-14T04:24:51.017300Z digest=sha256:f671bbc046e806a503759cf273476025a1f5aceec06db8b80adf289dfe50ccd0

Observation 0758d6c4-d89c-4260-98c9-197426603430 · outbound

This paper cites Large Language Models Produce Responses Perceived to be Empathic.

Automating and Scaling Behavioral Scientific Research on AI Agents Large Language Models Produce Responses Perceived to be Empathic

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-14T04:24:51.023322Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-14T04:24:51.023322Z digest=sha256:af68e4587385198581c5c69abece43a6ab1aaccda4b4a5ed0b0c236dd32c82aa

Observation 3910385c-8171-4aa2-8c0f-38df0153502a · outbound

This paper cites AdaPlanBench: Evaluating Adaptive Planning in Large Language Model Agents under World and User Constraints.

Automating and Scaling Behavioral Scientific Research on AI Agents AdaPlanBench: Evaluating Adaptive Planning in Large Language Model Agents under World and User Constraints

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-14T04:24:51.036759Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-14T04:24:51.036759Z digest=sha256:e32348c4e4308963e632b7aaff92b201e9f8fabcbfda314d4e33f7865f511c80

Observation 2761bdd5-9235-4f6f-b9aa-bcc4aa24a158 · outbound

This paper cites Strategic behavior of large language models and the role of game structure versus contextual framing.Scientific Reports, 14(1):18490, 2024.

Automating and Scaling Behavioral Scientific Research on AI Agents Strategic behavior of large language models and the role of game structure versus contextual framing.Scientific Reports, 14(1):18490, 2024

Reference 29

Resolution
malformed identifier
raw_fallback, observed 2026-08-14T04:24:56.836787Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-14T04:24:51.043242Z digest=sha256:6fea135739898ad1911c9a89b95bf3ba503be7396c5ca577531d638ab96c1c24

Observation de1900e2-e63a-4c5f-9140-222913df7cff · outbound

This paper cites Agentic misalignment: How LLMs could be insider threats.arXiv preprint arXiv:2510.05179, 2025.

Automating and Scaling Behavioral Scientific Research on AI Agents Agentic misalignment: How LLMs could be insider threats.arXiv preprint arXiv:2510.05179, 2025

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-14T04:24:51.055176Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-14T04:24:51.055176Z digest=sha256:af29b8f254148024ef549585ae8cd3c290a63ef853905f9433e3f3e7db555efc

Observation 7a8ddb43-761a-469a-a98d-fac0259739d5 · outbound

This paper cites Automated Social Science: Language Models as Scientist and Subjects.

Automating and Scaling Behavioral Scientific Research on AI Agents Automated Social Science: Language Models as Scientist and Subjects

Reference 32

Resolution
metadata mismatch
local_arxiv, observed 2026-08-14T04:24:53.121067Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-14T04:24:51.083853Z digest=sha256:8b14798104a668a7dc659f5d955dfc02987831b22b80107b36748ffe240e586a

Observation aa27b351-58c1-40e6-933a-7a38a1cbc1f4 · outbound

This paper cites ALYMPICS: LLM agents meet game theory.

Automating and Scaling Behavioral Scientific Research on AI Agents ALYMPICS: LLM agents meet game theory

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T04:24:56.802910Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-14T04:24:51.094814Z digest=sha256:fb34cf6098220b9f0608d6e8229ebdface186c3c3495b7451d85aedfa464dacc

Observation 16488ed7-a5e9-45d3-9285-93e436a9ce87 · outbound

This paper cites Frontier Models are Capable of In-context Scheming.

Automating and Scaling Behavioral Scientific Research on AI Agents Frontier Models are Capable of In-context Scheming

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-14T04:24:51.169617Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-14T04:24:51.169617Z digest=sha256:df022d9ebfab813cf06dfad3aafaa9b34dcafaf7849b9cb780d3d205969b9412

Observation ca4a715c-b8a7-4d7d-b8a6-dab97ba3b988 · outbound

This paper cites Learn- ing when to plan: Efficiently allocating test-time compute for LLM agents.arXiv preprint arXiv:2509.03581, 2025.

Automating and Scaling Behavioral Scientific Research on AI Agents Learn- ing when to plan: Efficiently allocating test-time compute for LLM agents.arXiv preprint arXiv:2509.03581, 2025

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-14T04:24:51.184477Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-14T04:24:51.184477Z digest=sha256:43a4285d324adca7aa9ab6537fa922ffa3a8326b2b73324273c74a2bd850326b

Observation de270470-c86f-4b6b-907b-2c126c39c5a9 · outbound

This paper cites Do Large Language Model Agents Exhibit a Survival Instinct? An Empirical Study in a Sugarscape-Style Simulation.

Automating and Scaling Behavioral Scientific Research on AI Agents Do Large Language Model Agents Exhibit a Survival Instinct? An Empirical Study in a Sugarscape-Style Simulation

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-14T04:24:51.134751Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-14T04:24:51.134751Z digest=sha256:0593d49dad8ff29d58cab522561a736ab4fc73c39ca0f1b50a801d0f174d3a90

Observation 1357b927-ecfa-4fdd-83bc-7deb79d0dd03 · outbound

This paper cites O’Brien, Carrie Jun Cai, Meredith Ringel Morris, Percy Liang, and Michael S.

Automating and Scaling Behavioral Scientific Research on AI Agents O’Brien, Carrie Jun Cai, Meredith Ringel Morris, Percy Liang, and Michael S

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-14T04:24:51.210759Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-14T04:24:51.210759Z digest=sha256:a737736ef2d165842bc38488d62fbc8ce0845bc449cfce80ddecafbb0ed3f8f1

Observation f990cd49-3043-48d9-94a1-2479d0ba398c · outbound

This paper cites LLM Agents Grounded in Self-Reports Enable General-Purpose Simulation of Individuals.

Automating and Scaling Behavioral Scientific Research on AI Agents LLM Agents Grounded in Self-Reports Enable General-Purpose Simulation of Individuals

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-14T04:24:51.223542Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-14T04:24:51.223542Z digest=sha256:a3d547466ccf1466d288be04d6c28a76110794e6bebf3426e9a5ae4942a69f28

Observation 5688f647-3349-4f16-97d9-bcc826e0b5c7 · outbound

This paper cites Do the rewards justify the means? Mea- suring trade-offs between rewards and ethical behavior in the MACHIA VELLI benchmark.

Automating and Scaling Behavioral Scientific Research on AI Agents Do the rewards justify the means? Mea- suring trade-offs between rewards and ethical behavior in the MACHIA VELLI benchmark

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T04:24:56.752116Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-14T04:24:51.197682Z digest=sha256:6429940cd3e37e242564429d4996f921aacad4a778b74c744ce7cdbbff42c56a

Observation 8ed6ebf5-41f9-4048-a640-6c2fc72ffb50 · outbound

This paper cites Psychological predicates.

Automating and Scaling Behavioral Scientific Research on AI Agents Psychological predicates

Reference 41

Resolution
verified exact
doi, observed 2026-08-14T04:24:51.945629Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-14T04:24:51.246184Z digest=sha256:81a65fed68b7f2f6feef10a5b303d95e075c38c0184bb1cb593c6c403dbe568f

Observation f629dd28-7fca-4bb6-bfeb-e50c5be9f953 · outbound

This paper cites AGENTIF: Benchmarking Instruction Following of Large Language Models in Agentic Scenarios.

Automating and Scaling Behavioral Scientific Research on AI Agents AGENTIF: Benchmarking Instruction Following of Large Language Models in Agentic Scenarios

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-14T04:24:51.262784Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-14T04:24:51.262784Z digest=sha256:bb0a6a47772ec2c65507968707518e062d80fdcf146604bce445b040329503ba

Observation 1894fb4f-512f-44da-af36-38e3a6ff7b6b · outbound

This paper cites AgentSociety: Large-Scale Simulation of LLM-Driven Generative Agents Advances Understanding of Human Behaviors and Society.

Automating and Scaling Behavioral Scientific Research on AI Agents AgentSociety: Large-Scale Simulation of LLM-Driven Generative Agents Advances Understanding of Human Behaviors and Society

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-14T04:24:51.231665Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-14T04:24:51.231665Z digest=sha256:cc258fa338f14fc4e3a97d040eab39ce8aea8875d1bcb6bf958eec8cdde5c9ca

Observation eea0eac9-a4f7-4426-afe8-dd43e0374962 · outbound

This paper cites Learning to Make Friends: Coaching LLM Agents toward Emergent Social Ties.

Automating and Scaling Behavioral Scientific Research on AI Agents Learning to Make Friends: Coaching LLM Agents toward Emergent Social Ties

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-14T04:24:51.289596Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-14T04:24:51.289596Z digest=sha256:9eb1e882f3ba0527c76b78b02938b92af744b92a695acdf34b04e563ed9c01a5

Observation 09261913-47c0-455a-b080-d88f07b02ffe · outbound

This paper cites Bowman, Newton Cheng, Esin Durmus, Zac Hatfield-Dodds, Scott R.

Automating and Scaling Behavioral Scientific Research on AI Agents Bowman, Newton Cheng, Esin Durmus, Zac Hatfield-Dodds, Scott R

Reference 45

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T04:24:56.673074Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-14T04:24:51.299621Z digest=sha256:e726f8de10d03a5ab0719603eea1013c40a323b208250ef3c40ba3c78dbe0bca

Observation eea89723-4058-4099-b1b3-ab29b49337ff · outbound

This paper cites Escalation risks from language models in military and diplomatic decision-making.

Automating and Scaling Behavioral Scientific Research on AI Agents Escalation risks from language models in military and diplomatic decision-making

Reference 46

Resolution
verified exact
arxiv_id_nonexistent, observed 2026-08-14T04:24:52.706008Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-14T04:24:51.275476Z digest=sha256:31cef40f0dd2f6888a7fec4047e89bd2313530e65ce9473945481ddccf78e39e

Observation 4f0ce941-dceb-4b20-a3aa-9568008f1205 · outbound

This paper cites LLMs can’t handle peer pressure: Crumbling under multi-agent social interactions.arXiv preprint arXiv:2508.18321, 2025.

Automating and Scaling Behavioral Scientific Research on AI Agents LLMs can’t handle peer pressure: Crumbling under multi-agent social interactions.arXiv preprint arXiv:2508.18321, 2025

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-14T04:24:51.313872Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-14T04:24:51.313872Z digest=sha256:e78cd3faec0eeb1167bbc516cbddf18d06f3eadcb00a77436c8c547cd688dde6

Observation f1187513-89ec-4a60-ba0a-cd987907d995 · outbound

This paper cites AI-Researcher: Autonomous sci- entific innovation.

Automating and Scaling Behavioral Scientific Research on AI Agents AI-Researcher: Autonomous sci- entific innovation

Reference 48

Resolution
verified exact
doi, observed 2026-08-14T04:24:51.913415Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-14T04:24:51.327559Z digest=sha256:d73b21c808d6318a47a6d53d646c001250e432738bbe1c02f65021f69f1ba0ce

Observation 630ac671-dc36-4080-9dd3-665ffad888b9 · outbound

This paper cites Reflexion: Language agents with verbal reinforcement learning.

Automating and Scaling Behavioral Scientific Research on AI Agents Reflexion: Language agents with verbal reinforcement learning

Reference 49

Resolution
malformed identifier
raw_fallback, observed 2026-08-14T04:24:56.565613Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-14T04:24:51.305846Z digest=sha256:79f90cc23a27ebcc8e86377c59e2af231dbb01f4e26e8a2bbafd9688b0d7c465

Observation 707a48fa-9508-4b3a-a68e-0db47594eca5 · outbound

This paper cites The Instruction Hierarchy: Training LLMs to Prioritize Privileged Instructions.

Automating and Scaling Behavioral Scientific Research on AI Agents The Instruction Hierarchy: Training LLMs to Prioritize Privileged Instructions

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-14T04:24:51.367266Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-14T04:24:51.367266Z digest=sha256:f7888643cbdc929e00074bb4cb26c5453db6453d868f5d73a0c298d6895d05e7

Observation 7a088685-c2c3-4080-8e39-4c744a1ed8ce · outbound

This paper cites AgenticEval: Toward Agentic and Self-Evolving Safety Evaluation of Large Language Models.

Automating and Scaling Behavioral Scientific Research on AI Agents AgenticEval: Toward Agentic and Self-Evolving Safety Evaluation of Large Language Models

Reference 51

Resolution
verified exact
local_arxiv, observed 2026-08-14T04:24:52.456823Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-14T04:24:51.374296Z digest=sha256:07ad77d46bf647f52e0b620ddc8394409b129aa2ead254429afde40a6d3e6c85

Observation 8ac5d6e4-2adc-4270-8de5-5ea26314b28c · outbound

This paper cites PlanBench: An extensible benchmark for evaluating large language models on planning and reasoning about change.

Automating and Scaling Behavioral Scientific Research on AI Agents PlanBench: An extensible benchmark for evaluating large language models on planning and reasoning about change

Reference 52

Resolution
verified exact
doi, observed 2026-08-14T04:24:51.865352Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-14T04:24:51.358957Z digest=sha256:16aa581ddef58c2a6f27eed97e65e2faab15d453c3b2a72d6f80913ca6904955

Observation de57500b-2f34-4463-bfa5-c90649a8dfc2 · outbound

This paper cites ReAct: Synergizing reasoning and acting in language models.

Automating and Scaling Behavioral Scientific Research on AI Agents ReAct: Synergizing reasoning and acting in language models

Reference 53

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T04:24:56.514806Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-14T04:24:51.414542Z digest=sha256:6159321fbf60a062a8c769a849e1d230b91f38178d3eaa6d9b564075d6586f23

Observation 279df0e6-a3fb-4d35-9dc8-da62dba5db4e · outbound

This paper cites Position: Llms can’t jump.

Automating and Scaling Behavioral Scientific Research on AI Agents Position: Llms can’t jump

Reference 54

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T04:24:56.397207Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-14T04:24:51.428119Z digest=sha256:ca3772ee5230de282c4e84eb926cb68cec7728bf34929fdd624e953bf8c807e8

Observation 10c38f9d-123e-47e1-854d-4b1e0ea730f0 · outbound

This paper cites Nuclear deployed!: Analyzing catastrophic risks in decision-making of autonomous LLM agents.

Automating and Scaling Behavioral Scientific Research on AI Agents Nuclear deployed!: Analyzing catastrophic risks in decision-making of autonomous LLM agents

Reference 55

Resolution
unresolved
no resolver link, observed 2026-08-14T04:24:51.389322Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-14T04:24:51.389322Z digest=sha256:18c09cb557eeed1141d02b1fdd8fcdc7abcc40b70b038cef2ed3ad941bf76be2

Observation 96d0a874-9369-49dd-9215-285e988602e1 · outbound

This paper cites SocioVerse: A World Model for Social Simulation Powered by LLM Agents and A Pool of 10 Million Real-World Users.

Automating and Scaling Behavioral Scientific Research on AI Agents SocioVerse: A World Model for Social Simulation Powered by LLM Agents and A Pool of 10 Million Real-World Users

Reference 56

Resolution
unresolved
no resolver link, observed 2026-08-14T04:24:51.439987Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-14T04:24:51.439987Z digest=sha256:3117d82c42b8333ba3623e478a011dff4b739368c653ae3822568d85bf79ca55

Observation 58af1bbd-c493-466b-a407-cbf27754ac13 · outbound

This paper cites CompeteAI: Understanding the competition dynamics in large language model-based agents.

Automating and Scaling Behavioral Scientific Research on AI Agents CompeteAI: Understanding the competition dynamics in large language model-based agents

Reference 57

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T04:24:56.283155Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-14T04:24:51.447353Z digest=sha256:304645530bb3f64ee96230440db969c8e401543e2ca241b769d9647abf803870

Observation f8935cfe-df32-4625-9ddf-21621dd3225c · outbound

This paper cites Dive into the agent matrix: A realistic evaluation of self-replication risk in LLM agents.arXiv preprint arXiv:2509.25302, 2025.

Automating and Scaling Behavioral Scientific Research on AI Agents Dive into the agent matrix: A realistic evaluation of self-replication risk in LLM agents.arXiv preprint arXiv:2509.25302, 2025

Reference 58

Resolution
unresolved
no resolver link, observed 2026-08-14T04:24:51.433413Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-14T04:24:51.433413Z digest=sha256:3ab5e619199308c88d1131ac236f61f4fcc97025673183d933611c3f6066acde

Observation d47b118f-08bc-483e-aae7-c6973334a775 · outbound

This paper cites Navigating the grey area: How expres- sions of uncertainty and overconfidence affect language models.

Automating and Scaling Behavioral Scientific Research on AI Agents Navigating the grey area: How expres- sions of uncertainty and overconfidence affect language models

Reference 59

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T04:24:56.104750Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-14T04:24:51.461235Z digest=sha256:c8dafb309754412696e8676d91e65cd5a4249270a575a765fcb861465bfee013

Observation c5faa021-d339-42ea-944e-9c2c204c5ae5 · outbound

This paper cites SOTOPIA: Interactive evaluation for social intelligence in language agents.

Automating and Scaling Behavioral Scientific Research on AI Agents SOTOPIA: Interactive evaluation for social intelligence in language agents

Reference 60

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T04:24:56.038303Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-14T04:24:51.490605Z digest=sha256:71a53b3160a1dbba4cf3f7084dc6bc2be09ab8592fb8e980511a656b28e0ac1c

Observation 7f6bfdfb-3093-4560-a4f0-5e6c5d31f381 · outbound

This paper cites ALI- Agent: Assessing LLMs’ alignment with human values via agent-based evaluation.

Automating and Scaling Behavioral Scientific Research on AI Agents ALI- Agent: Assessing LLMs’ alignment with human values via agent-based evaluation

Reference 61

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T04:24:56.204748Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-14T04:24:51.455534Z digest=sha256:047ffde5cef3e46a66577f1373e1c94d6ad58fe789924f6d69cf373ca263d5c7

Observation 2ad78825-ebb6-4224-a4d9-95249781abbf · outbound

This paper cites positive.

Automating and Scaling Behavioral Scientific Research on AI Agents positive

Reference 65

Resolution
unresolved
no resolver link, observed 2026-08-14T04:24:51.498858Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-14T04:24:51.498858Z digest=sha256:ffb3825426b7234014bb36da8ccba84d01c283fd67252ef88c099e8e86308602

Observation 0cb612de-3ee3-43a6-ac86-f65e1c69f969 · outbound

This paper cites an unresolved cited work.

Automating and Scaling Behavioral Scientific Research on AI Agents Unresolved cited work

Reference 66

Resolution
unresolved
raw_fallback, observed 2026-08-14T04:24:55.998608Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-14T04:24:51.505375Z digest=sha256:8bc05e7c3faa064292b64b5d436f26326cf8385c5a1b3620ee0cb6e67724fa3a

Observation 63178901-1947-41f3-afa7-219c02df16b6 · outbound

This paper cites an unresolved cited work.

Automating and Scaling Behavioral Scientific Research on AI Agents Unresolved cited work

Reference 67

Resolution
unresolved
raw_fallback, observed 2026-08-14T04:24:55.957704Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-14T04:24:51.527886Z digest=sha256:2aa7097745795dcfe8bf4b594333d7b10bd31c9c5bbfe98a26a536d8105f0e08

Observation 79b11712-6082-4da8-a248-79072d28f17f · outbound

This paper cites an unresolved cited work.

Automating and Scaling Behavioral Scientific Research on AI Agents Unresolved cited work

Reference 68

Resolution
unresolved
raw_fallback, observed 2026-08-14T04:24:55.845708Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-14T04:24:51.539063Z digest=sha256:476b36cfdecb6c2e07bc98cedb8f369ebf821f75bcb75b6ee1f4bd993a587f4b

Observation a98993bd-d96d-4aee-a24d-61b8281d49a2 · outbound

This paper cites None" - task 2.exemption: if no meaningful interactions exist, write.

Automating and Scaling Behavioral Scientific Research on AI Agents None" - task 2.exemption: if no meaningful interactions exist, write

Reference 69

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T04:24:55.774750Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-14T04:24:51.551241Z digest=sha256:7cbe190663df9a4fc8d9f7784b1d7b74e4d324b2dd8ba1b9844bf8ced31d9a8c

Observation 796143db-41f4-4357-87d3-3541732f19a9 · outbound

This paper cites objective.

Automating and Scaling Behavioral Scientific Research on AI Agents objective

Reference 70

Resolution
malformed identifier
raw_fallback, observed 2026-08-14T04:24:55.724756Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-14T04:24:51.562766Z digest=sha256:212ad1f9e363b0d1f948fd4a9b6d8fe6bc06d8920ca2ee0c6cf5a6aa0e231031

Observation b31cc0aa-60b4-4907-882d-9e08e796def3 · outbound

This paper cites an unresolved cited work.

Automating and Scaling Behavioral Scientific Research on AI Agents Unresolved cited work

Reference 71

Resolution
unresolved
raw_fallback, observed 2026-08-14T04:24:55.587374Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-14T04:24:51.576142Z digest=sha256:dca05ee1a3a14d3d12d943164f6328053ce568529b48bf788a1a212fd290a8f2

Observation e737a7e3-1fc8-4595-b7f7-2cffc0d210f4 · outbound

This paper cites an unresolved cited work.

Automating and Scaling Behavioral Scientific Research on AI Agents Unresolved cited work

Reference 72

Resolution
unresolved
raw_fallback, observed 2026-08-14T04:24:55.456910Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-14T04:24:51.587965Z digest=sha256:b84a0426f87e43befa802eecaadf26a16e6153f56913f432764b4b15ba2461c2

Observation 83c7e6af-a54f-4f71-a285-5d6b661e6c5b · outbound

This paper cites an unresolved cited work.

Automating and Scaling Behavioral Scientific Research on AI Agents Unresolved cited work

Reference 73

Resolution
unresolved
raw_fallback, observed 2026-08-14T04:24:55.420261Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-14T04:24:51.593204Z digest=sha256:ac1311eb449e756ced64aa34dd93df462a1bbc601902c9289e1aae3eee8f7c70

Observation 57a3bb62-e5a9-4c94-a41a-86fea870700e · outbound

This paper cites an unresolved cited work.

Automating and Scaling Behavioral Scientific Research on AI Agents Unresolved cited work

Reference 74

Resolution
unresolved
raw_fallback, observed 2026-08-14T04:24:55.356918Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-14T04:24:51.602457Z digest=sha256:52d081723529a7b3773020eab7c45349d2594747a65d79743d97f95f4dc9bd1e

Observation 63a2ead3-1e7a-484e-a3ad-a607ea74e511 · outbound

This paper cites an unresolved cited work.

Automating and Scaling Behavioral Scientific Research on AI Agents Unresolved cited work

Reference 75

Resolution
unresolved
raw_fallback, observed 2026-08-14T04:24:55.263910Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-14T04:24:51.608105Z digest=sha256:aecbdd18a1fb41c9ca1bc1ddefe31d11ccf3960a3bdaa91ff982114250f9173e

Observation ab3f72f3-6805-4505-93b0-3100bda10edc · outbound

This paper cites an unresolved cited work.

Automating and Scaling Behavioral Scientific Research on AI Agents Unresolved cited work

Reference 76

Resolution
unresolved
raw_fallback, observed 2026-08-14T04:24:55.164740Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-14T04:24:51.613122Z digest=sha256:ba6a8a484555ef247cb24f89ce753724cc68dfcfdf991817992bd9d00327b76b

Observation 55a66085-e2bc-485c-97c4-c102dfa0f578 · outbound

This paper cites an unresolved cited work.

Automating and Scaling Behavioral Scientific Research on AI Agents Unresolved cited work

Reference 77

Resolution
unresolved
raw_fallback, observed 2026-08-14T04:24:55.003909Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-14T04:24:51.628160Z digest=sha256:d5407e281f742e05d6ef54414f295a86e0914c0d7346c960d32e5d7c9eecc390

Observation 2489b948-bde9-433f-97e1-c3443642ead3 · outbound

This paper cites an unresolved cited work.

Automating and Scaling Behavioral Scientific Research on AI Agents Unresolved cited work

Reference 78

Resolution
unresolved
raw_fallback, observed 2026-08-14T04:24:54.914743Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-14T04:24:51.641847Z digest=sha256:b061af7145b1be779ea0d21d2a0d2a11cfc43e72d147f1a896039c89e642c156

Observation 79d9f9ad-0c2a-4617-8e7f-773e8b4a5347 · outbound

This paper cites an unresolved cited work.

Automating and Scaling Behavioral Scientific Research on AI Agents Unresolved cited work

Reference 79

Resolution
unresolved
raw_fallback, observed 2026-08-14T04:24:54.832253Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-14T04:24:51.654911Z digest=sha256:acd20266a68c130bbe83d38112047d5e1b0ba26ca1aaf03e5762ede2807bbceb

Observation 036d3d6e-2688-46fc-85ea-2036f33388d3 · outbound

This paper cites an unresolved cited work.

Automating and Scaling Behavioral Scientific Research on AI Agents Unresolved cited work

Reference 80

Resolution
unresolved
raw_fallback, observed 2026-08-14T04:24:54.738686Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-14T04:24:51.671098Z digest=sha256:d9e00b5a55e1dfa329cb32be17322d6284c0db3636f8b57fa6189c61f19feca0

Observation a45e26c0-2160-40b6-8593-24de377c6026 · outbound

This paper cites an unresolved cited work.

Automating and Scaling Behavioral Scientific Research on AI Agents Unresolved cited work

Reference 81

Resolution
unresolved
raw_fallback, observed 2026-08-14T04:24:54.669376Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-14T04:24:51.684880Z digest=sha256:6e323ab623e55a5dd72d0b3176f5a46ba7fb8a332c9daeb76f78012dcb0bead8

Observation 54e24820-b0fa-41aa-939a-398b2940c9c4 · outbound

This paper cites In each subplot, its x-axis and y-axis ticks denote distinct evidence classes defined in its rubric yrubric.

Automating and Scaling Behavioral Scientific Research on AI Agents In each subplot, its x-axis and y-axis ticks denote distinct evidence classes defined in its rubric yrubric

Reference 82

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T04:24:54.586598Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-14T04:24:51.705273Z digest=sha256:0d11781fc633befdb344a45a3460668ac1b1134a2f4f81ce0201ce7ee087c872

Observation e0f9a378-7f09-4c59-83eb-1cb5eb113234 · outbound

This paper cites URL https://aclanthology.org/2023.

Automating and Scaling Behavioral Scientific Research on AI Agents URL https://aclanthology.org/2023

Reference 2023

Resolution
unresolved
no resolver link, observed 2026-08-14T04:24:51.480581Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-14T04:24:51.480581Z digest=sha256:bbdc17a5a04f94a515f666a86f263a57e1b7653d3d22a23bcf0d8387b3f52575

Observation 8e12f268-f866-4c19-8d1a-7c81374ad9fa · outbound

This paper cites an unresolved cited work.

Automating and Scaling Behavioral Scientific Research on AI Agents Unresolved cited work

Reference 2025

Resolution
unresolved
no resolver link, observed 2026-08-14T04:24:50.879313Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-14T04:24:50.879313Z digest=sha256:a62cbe78f566644915893d919e27f2baca3386478211f32cf9f789e461d2f095

Observation de35ca0b-7e3e-46df-966c-24d164b8c7e3 · outbound

This paper cites Who Endorsed It? Measuring Authority Bias Across Expertise Levels in Language Models.

Automating and Scaling Behavioral Scientific Research on AI Agents Who Endorsed It? Measuring Authority Bias Across Expertise Levels in Language Models

Reference 2026

Resolution
unresolved
no resolver link, observed 2026-08-14T04:24:51.073476Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-14T04:24:51.073476Z digest=sha256:da5f252f96e0e169f170959dd3ace3553a5243f5094f047197bccdca9ec901b1

Pith citing papers

No inbound Pith citation observations are available.