Pith. sign in

Paper Citation Record · LEDGER

SOTOPIA: Interactive Evaluation for Social Intelligence in Language Agents

As of 14 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 56 inbound Pith citation observations for arXiv:2310.11667.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2310.11667 v2

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 56 of 56 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-14T06:32:32.682623+00:00

measured 56 of 56 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-14T04:19:42.764152Z

measured 1 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-08-05T02:28:24.338817Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

3
pith, observed 2026-08-05T02:28:24.338817Z

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation fa0df7c5-9bd2-40e8-b042-954ab1a0f55f · inbound

Prompt Infection: LLM-to-LLM Prompt Injection within Multi-Agent Systems cites this paper.

Prompt Infection: LLM-to-LLM Prompt Injection within Multi-Agent Systems SOTOPIA: Interactive Evaluation for Social Intelligence in Language Agents

Reference 97

Resolution
verified exact
arxiv_id, observed 2026-05-15T19:32:19.870652Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=arxiv_source observed=2026-05-15T19:32:19.405615Z digest=sha256:fb49370de2853aabf5558e0f888cda4ff33f0df8349e66516920c08974351945

Observation 624c6659-d8d1-4b86-81cd-658e3f51be17 · inbound

A Survey on LLM-as-a-Judge cites this paper.

A Survey on LLM-as-a-Judge SOTOPIA: Interactive Evaluation for Social Intelligence in Language Agents

Reference 228

Resolution
verified exact
arxiv_id, observed 2026-05-23T17:35:44.038191Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-05-23T17:33:13.394338Z digest=sha256:aa3e1b0b8b2c8013cbc1797cd77ee51a748a74c29c2b68ff52524b4c4915cd9e

Observation 4256ae2e-cdb7-4ecb-98d2-a71daca0dd2f · inbound

From Individual to Society: A Survey on Social Simulation Driven by Large Language Model-based Agents cites this paper.

From Individual to Society: A Survey on Social Simulation Driven by Large Language Model-based Agents SOTOPIA: Interactive Evaluation for Social Intelligence in Language Agents

Reference 138

Resolution
unresolved
no resolver link, observed 2026-08-11T22:18:31.100824Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T22:18:31.100824Z digest=sha256:582dff4bcfea4045d542b2cab4f0c679504c0118aa513cd1be319619d16c02f2

Observation 83145d4a-e866-4445-8674-47c83576841c · inbound

BotSim: LLM-Powered Malicious Social Botnet Simulation cites this paper.

BotSim: LLM-Powered Malicious Social Botnet Simulation SOTOPIA: Interactive Evaluation for Social Intelligence in Language Agents

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-11T13:11:53.078631Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T13:11:53.078631Z digest=sha256:e9823797b70dab11ab2745cf98409fd732926e0e7082d4ced1c655a754b792e5

Observation 8b961a02-8412-4494-95b7-84599627ddaa · inbound

AgentSociety: Large-Scale Simulation of LLM-Driven Generative Agents Advances Understanding of Human Behaviors and Society cites this paper.

AgentSociety: Large-Scale Simulation of LLM-Driven Generative Agents Advances Understanding of Human Behaviors and Society SOTOPIA: Interactive Evaluation for Social Intelligence in Language Agents

Reference 116

Resolution
verified exact
arxiv_id, observed 2026-05-23T04:12:30.936693Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-05-23T04:08:36.292592Z digest=sha256:ce1ad5b977ace25663257d328585ef92ab8a8f4962623cd650f2998f241b1308

Observation 8adf801a-d4ca-48b9-b78b-968ba133e92e · inbound

Empowering LLMs in Task-Oriented Dialogues: A Domain-Independent Multi-Agent Framework and Fine-Tuning Strategy cites this paper.

Empowering LLMs in Task-Oriented Dialogues: A Domain-Independent Multi-Agent Framework and Fine-Tuning Strategy SOTOPIA: Interactive Evaluation for Social Intelligence in Language Agents

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-07T15:41:11.953688Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:41:11.953688Z digest=sha256:e33f4b60e84005006bdf402f1d71f1e4a5cf59e49fb189aa76f91b5bdf8e5a00

Observation fd96bd9b-3e57-4aa4-80ac-fe5d99d7b201 · inbound

TimeHC-RL: Temporal-aware Hierarchical Cognitive Reinforcement Learning for Enhancing LLMs' Social Intelligence cites this paper.

TimeHC-RL: Temporal-aware Hierarchical Cognitive Reinforcement Learning for Enhancing LLMs' Social Intelligence SOTOPIA: Interactive Evaluation for Social Intelligence in Language Agents

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-07T12:27:52.254318Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:27:52.254318Z digest=sha256:a8e9e0750cf3651d823ddd81db8d5f87b6e72ce61642d45fc254630648d07d44

Observation 23c052aa-133b-453d-be09-6464489841e1 · inbound

ARIA: Training Language Agents with Intention-Driven Reward Aggregation cites this paper.

ARIA: Training Language Agents with Intention-Driven Reward Aggregation SOTOPIA: Interactive Evaluation for Social Intelligence in Language Agents

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-07T12:09:01.633394Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:09:01.633394Z digest=sha256:f81f8331c2e66ad620a8f6a291d4ff6213c528e7ecd8fe1788f98d3aa21a40ec

Observation 7bbdf8d0-5212-4c78-a48b-72108fab9690 · inbound

Aligning VLM Assistants with Personalized Situated Cognition cites this paper.

Aligning VLM Assistants with Personalized Situated Cognition SOTOPIA: Interactive Evaluation for Social Intelligence in Language Agents

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-07T11:59:40.352298Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:59:40.352298Z digest=sha256:c75e432fbf1ffc083c346116a0ea37f34a68a91e56a0c89078493a8431eba51f

Observation e17006bd-a8f1-4b15-8dab-f045d180ebba · inbound

MAEBE: Multi-Agent Emergent Behavior Framework cites this paper.

MAEBE: Multi-Agent Emergent Behavior Framework SOTOPIA: Interactive Evaluation for Social Intelligence in Language Agents

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-07T11:14:47.386971Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T11:14:47.386971Z digest=sha256:43e5a2582c99570a8b005cc471a95e9ecac4703aff31815a0fb758513e877f34

Observation 88a24f31-a158-4a96-bfd8-6664721247ac · inbound

SIV-Bench: A Video Benchmark for Social Interaction Understanding and Reasoning cites this paper.

SIV-Bench: A Video Benchmark for Social Interaction Understanding and Reasoning SOTOPIA: Interactive Evaluation for Social Intelligence in Language Agents

Reference 62

Resolution
verified exact
arxiv_id, observed 2026-05-19T11:37:15.753940Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-05-19T11:36:36.687324Z digest=sha256:c60677cd83399f3200c31b0b4fd2f57c4691478cfc07db2c17a0ac6028e7dadb

Observation 807cdc1f-47b7-4cdc-9a4d-65ccf42d07b1 · inbound

The Scales of Justitia: A Comprehensive Survey on Safety Evaluation of LLMs cites this paper.

The Scales of Justitia: A Comprehensive Survey on Safety Evaluation of LLMs SOTOPIA: Interactive Evaluation for Social Intelligence in Language Agents

Reference 166

Resolution
unresolved
no resolver link, observed 2026-08-07T10:17:27.102545Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T10:17:27.102545Z digest=sha256:d00fe12e9d6df679043d4f1fbd92ccf70c7a6f2db02cef9d0402e34fbe016e3b

Observation 1bf1f8a8-b4ca-40be-8141-12d7b656c0f1 · inbound

Evolutionary Perspectives on the Evaluation of LLM-Based AI Agents: A Comprehensive Survey cites this paper.

Evolutionary Perspectives on the Evaluation of LLM-Based AI Agents: A Comprehensive Survey SOTOPIA: Interactive Evaluation for Social Intelligence in Language Agents

Reference 155

Resolution
unresolved
no resolver link, observed 2026-08-07T06:00:19.722103Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T06:00:19.722103Z digest=sha256:e32eaebb2fcfa7d58622c4f6b7b4533d30557e387f08aafecae390a6ab6ecc68

Observation af7aa177-2720-4f06-888f-ee7396bb5b47 · inbound

LIFELONG SOTOPIA: Evaluating Social Intelligence of Language Agents Over Lifelong Social Interactions cites this paper.

LIFELONG SOTOPIA: Evaluating Social Intelligence of Language Agents Over Lifelong Social Interactions SOTOPIA: Interactive Evaluation for Social Intelligence in Language Agents

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-07T00:49:29.629827Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T00:49:29.629827Z digest=sha256:3a2e19d0e3e9f1f6bb439ad2ac430c03417b14978f97fc84b5841cc85c732ddb

Observation 9edd8e6d-ef77-473d-974a-340009677af8 · inbound

MotiveBench: How Far Are We From Human-Like Motivational Reasoning in Large Language Models? cites this paper.

MotiveBench: How Far Are We From Human-Like Motivational Reasoning in Large Language Models? SOTOPIA: Interactive Evaluation for Social Intelligence in Language Agents

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-07T00:43:06.885426Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:43:06.885426Z digest=sha256:a6a0b9b55d018635e2a139e7efc66e6526a2c6a4f83d852b6dd0384b202b812a

Observation 581546fb-f6aa-406d-b764-8dc34bc9567e · inbound

Infected Smallville: How Disease Threat Shapes Sociality in LLM Agents cites this paper.

Infected Smallville: How Disease Threat Shapes Sociality in LLM Agents SOTOPIA: Interactive Evaluation for Social Intelligence in Language Agents

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-07T04:59:51.721605Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T04:59:51.721605Z digest=sha256:3cf5e76a327fd9119c855113795bc5d80aa6d40bd91e1f231982b0f3bb90b2f1

Observation bae085ed-78fd-4f07-891c-c5e2b9fdef77 · inbound

AgentGroupChat-V2: Divide-and-Conquer Is What LLM-Based Multi-Agent System Need cites this paper.

AgentGroupChat-V2: Divide-and-Conquer Is What LLM-Based Multi-Agent System Need SOTOPIA: Interactive Evaluation for Social Intelligence in Language Agents

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-06T23:59:48.116476Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:59:48.116476Z digest=sha256:9ece2bbfb74c1d75047eed3ac3d047b4377295a8d352626b295062acd1dd1323

Observation d700abbb-c455-47d6-b6b0-3d735c27d72d · inbound

Kaleidoscopic Teaming in Multi Agent Simulations cites this paper.

Kaleidoscopic Teaming in Multi Agent Simulations SOTOPIA: Interactive Evaluation for Social Intelligence in Language Agents

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-06T23:35:52.622151Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:35:52.622151Z digest=sha256:c7b42bf327d7f910495201bec01052874fa74af86c6bbd65005a46dc1a40a0df

Observation 09717eaa-c633-4529-ba60-4a80b74a1215 · inbound

Multi-Actor Generative Artificial Intelligence as a Game Engine cites this paper.

Multi-Actor Generative Artificial Intelligence as a Game Engine SOTOPIA: Interactive Evaluation for Social Intelligence in Language Agents

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-06T18:28:25.430417Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:28:25.430417Z digest=sha256:d013ce8ed6033211c0f4ed01a3ee44b3bbf268429aa5f75946cced73ce715976

Observation b633bc8e-5635-4f45-bca1-20f639157132 · inbound

ProactiveEval: A Unified Evaluation Framework for Proactive Dialogue Agents cites this paper.

ProactiveEval: A Unified Evaluation Framework for Proactive Dialogue Agents SOTOPIA: Interactive Evaluation for Social Intelligence in Language Agents

Reference 55

Resolution
unresolved
no resolver link, observed 2026-08-05T14:43:16.454680Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T14:43:16.454680Z digest=sha256:fd5c54e35f1ca15b433fed7237afe13841ca714cc51b78f4ba5d75710d9433c2

Observation ad3a86d2-7b36-4e95-a262-b2753cf25bb6 · inbound

Evalet: Evaluating Large Language Models through Functional Fragmentation cites this paper.

Evalet: Evaluating Large Language Models through Functional Fragmentation SOTOPIA: Interactive Evaluation for Social Intelligence in Language Agents

Reference 103

Resolution
metadata mismatch
arxiv_id, observed 2026-05-18T17:01:39.953124Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-05-18T16:57:25.259866Z digest=sha256:e377d744b7a40615973f653e9cc62b5f993308afdf8184687ee9c46ebeb8de7e

Observation 07ca23cc-67c5-4bcf-bf97-b093e7a865dd · inbound

DoubleAgents: Human-Agent Alignment in a Socially Embedded Workflow cites this paper.

DoubleAgents: Human-Agent Alignment in a Socially Embedded Workflow SOTOPIA: Interactive Evaluation for Social Intelligence in Language Agents

Reference 57

Resolution
verified exact
arxiv_id, observed 2026-05-18T17:11:40.270360Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-05-18T17:09:02.466082Z digest=sha256:fac16a22a9f40143aa75bb6bda6f12d1e25fb51d8caf8104a1f25149b4c3f623

Observation 33b9668c-8b8a-48e7-92aa-ba484337307d · inbound

AgentCrypt: Advancing Privacy and (Secure) Computation in AI Agent Collaboration cites this paper.

AgentCrypt: Advancing Privacy and (Secure) Computation in AI Agent Collaboration SOTOPIA: Interactive Evaluation for Social Intelligence in Language Agents

Reference 41

Resolution
verified exact
arxiv_id, observed 2026-05-16T23:58:42.686439Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=arxiv_source observed=2026-05-16T23:57:19.757902Z digest=sha256:c649b9117700429aa4ad0792cd9926929dc44c5aac161c98e008a24dbdf997ea

Observation a2c881d5-bea0-4e72-b09f-9455cb711b5b · inbound

Belief-Sim: Towards Belief-Driven Simulation of Demographic Misinformation Susceptibility cites this paper.

Belief-Sim: Towards Belief-Driven Simulation of Demographic Misinformation Susceptibility SOTOPIA: Interactive Evaluation for Social Intelligence in Language Agents

Reference 8

Resolution
malformed identifier
no resolver link, observed 2026-08-02T19:06:34.246776Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T19:06:34.246776Z digest=sha256:1ca97bf76d3e6f7f4cc34fb6a5d25947e96c81306eac42a62f7deaa1a76b6b94

Observation 5fbe1505-b2a7-4c69-a685-f65d57143f4c · inbound

OpenHospital: A Thing-in-itself Arena for Evolving and Benchmarking LLM-based Collective Intelligence cites this paper.

OpenHospital: A Thing-in-itself Arena for Evolving and Benchmarking LLM-based Collective Intelligence SOTOPIA: Interactive Evaluation for Social Intelligence in Language Agents

Reference 26

Resolution
unresolved
no resolver link, observed 2026-07-14T21:00:05.911327Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-14T21:00:05.911327Z digest=sha256:12c8c7dd1b1895b07588afb158f5aed8c99a85e571c40d3510bf81cd08d3a424

Observation 6d20427b-9a78-4efb-ac4a-e9105424920a · inbound

Sell More, Play Less: Benchmarking LLM Realistic Selling Skill cites this paper.

Sell More, Play Less: Benchmarking LLM Realistic Selling Skill SOTOPIA: Interactive Evaluation for Social Intelligence in Language Agents

Reference 51

Resolution
verified exact
arxiv_id, observed 2026-05-11T06:31:02.363636Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=arxiv_source observed=2026-05-10T17:36:10.725278Z digest=sha256:82264e0150521b363ff8fc204e129989855b6d4a0319aabcc3a5d48b5d7817de

Observation 6d65f17f-c0b6-412d-8c1e-21c22743b7ce · inbound

Imperfectly Cooperative Human-AI Interactions: Comparing the Impacts of Human and AI Attributes in Simulated and User Studies cites this paper.

Imperfectly Cooperative Human-AI Interactions: Comparing the Impacts of Human and AI Attributes in Simulated and User Studies SOTOPIA: Interactive Evaluation for Social Intelligence in Language Agents

Reference 70

Resolution
verified exact
arxiv_id, observed 2026-05-10T09:43:47.718641Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=arxiv_source observed=2026-05-10T09:39:56.132765Z digest=sha256:84be49885a0c67f8a97170d73bc3c46e1f62a7cb0ee0c33614f4364d8fd78f2e

Observation 66c273ec-9e90-4a1b-a632-da43dec9b14d · inbound

Understanding the Mechanism of Altruism in Large Language Models cites this paper.

Understanding the Mechanism of Altruism in Large Language Models SOTOPIA: Interactive Evaluation for Social Intelligence in Language Agents

Reference 254

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T13:31:02.066303Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=arxiv_source observed=2026-05-10T01:36:50.329664Z digest=sha256:215840550d107534f0fe75115d3608c22b05b2ec3e4eb23de1188a8d087e1b3a

Observation 264d9de0-884a-43f8-bef5-a53bb9a9d843 · inbound

Cooperate to Compete: Strategic Coordination in Multi-Agent Conquest cites this paper.

Cooperate to Compete: Strategic Coordination in Multi-Agent Conquest SOTOPIA: Interactive Evaluation for Social Intelligence in Language Agents

Reference 41

Resolution
verified exact
arxiv_id, observed 2026-05-11T23:36:36.058637Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=arxiv_source observed=2026-05-07T16:37:58.860183Z digest=sha256:22cb0d8060b51f8a587b128471676980e767d1fabaac8860fdc9d066db35aab1

Observation 8bcab146-a811-422e-a5a3-5890bef2125d · inbound

Agent Island: A Saturation- and Contamination-Resistant Benchmark from Multiagent Games cites this paper.

Agent Island: A Saturation- and Contamination-Resistant Benchmark from Multiagent Games SOTOPIA: Interactive Evaluation for Social Intelligence in Language Agents

Reference 20

Resolution
verified exact
arxiv_id, observed 2026-05-11T17:46:17.235903Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=arxiv_source observed=2026-05-08T17:06:32.814188Z digest=sha256:36c2f07d0eed810a4d706e103344e27a9b51c7eb83b26b51a6bf3eea68e191ad

Observation 6675a431-f82e-47f3-9676-4fc5c91b3cbb · inbound

CustomerSim: Benchmarking and Aligning Multimodal Language Models as Retail User Simulators cites this paper.

CustomerSim: Benchmarking and Aligning Multimodal Language Models as Retail User Simulators SOTOPIA: Interactive Evaluation for Social Intelligence in Language Agents

Reference 46

Resolution
verified exact
arxiv_id, observed 2026-05-12T08:41:24.072457Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-05-12T00:51:57.796883Z digest=sha256:ef238af48314015b3818f05e6ca6ab1bbfbc2a864bdcf8dc40cf63f6ac115fc5

Observation 79780cc2-ae8a-4f27-a6eb-d83ca9324c00 · inbound

CustomerSim: Benchmarking and Aligning Multimodal Language Models as Retail User Simulators cites this paper.

CustomerSim: Benchmarking and Aligning Multimodal Language Models as Retail User Simulators SOTOPIA: Interactive Evaluation for Social Intelligence in Language Agents

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-02T14:37:24.618276Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T14:37:24.618276Z digest=sha256:d50c9e56f4508bdaceff712a2ff0db1eecb96391dd7afa9ba5e9a34d260bca61

Observation a154e0f4-c7cd-4c3e-ba0d-f66eeb3e4f77 · inbound

ProactBench: Beyond What The User Asked For cites this paper.

ProactBench: Beyond What The User Asked For SOTOPIA: Interactive Evaluation for Social Intelligence in Language Agents

Reference 158

Resolution
verified exact
arxiv_id, observed 2026-05-12T02:16:16.012831Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=arxiv_source observed=2026-05-12T02:14:01.145443Z digest=sha256:89d10f1a1330b93d1c709845678cfe17b1249ed06f3e7d607794bb9e3a5a59ff

Observation 2daea6fa-770a-4114-a3ef-09a62f1868d9 · inbound

LLM Jaggedness Unlocks Scientific Creativity cites this paper.

LLM Jaggedness Unlocks Scientific Creativity SOTOPIA: Interactive Evaluation for Social Intelligence in Language Agents

Reference 17

Resolution
verified exact
arxiv_id, observed 2026-05-12T06:36:27.715342Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-05-12T04:06:57.197356Z digest=sha256:1ddbedff77af3622bcaa924490f6cfb26a9ef2f5eaf7b301892059c3be43a886

Observation f5ccb9e6-a1d0-4ded-b75d-d67ae19eb1df · inbound

LLM Jaggedness Unlocks Scientific Creativity cites this paper.

LLM Jaggedness Unlocks Scientific Creativity SOTOPIA: Interactive Evaluation for Social Intelligence in Language Agents

Reference 17

Resolution
verified exact
arxiv_id, observed 2026-05-21T08:19:52.490571Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-05-21T08:18:38.493557Z digest=sha256:9328b50fb6224cd651688f6cd9c0de33b147df0a56e3977c1f2be6fd3c6fe5c6

Observation a6741f12-72b5-44ef-96c5-f3d652dcef7f · inbound

Beyond Individual Mimicry: Constructing Human-Like Social network with Graph-Augmented LLM Agents cites this paper.

Beyond Individual Mimicry: Constructing Human-Like Social network with Graph-Augmented LLM Agents SOTOPIA: Interactive Evaluation for Social Intelligence in Language Agents

Reference 33

Resolution
metadata mismatch
arxiv_id, observed 2026-05-14T21:19:27.901802Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=arxiv_source observed=2026-05-14T21:19:05.453523Z digest=sha256:55cbf7855f94bfaad557e94cd590e1581ad4a1902ddf4a354671c37976c6178b

Observation c02ffb49-d792-4301-a5fd-655ca4a10490 · inbound

Can LLMs Think Like Consumers? Benchmarking Crowd-Level Reaction Reconstruction with ConsumerSimBench cites this paper.

Can LLMs Think Like Consumers? Benchmarking Crowd-Level Reaction Reconstruction with ConsumerSimBench SOTOPIA: Interactive Evaluation for Social Intelligence in Language Agents

Reference 38

Resolution
verified exact
arxiv_id, observed 2026-05-20T15:33:25.492532Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-05-20T15:31:25.079191Z digest=sha256:63eac841a3ffc1b815dfd4f3bc129e5db9a005f5048bc90ee26bda311d18cc91

Observation b7bf9072-b979-4211-a34e-13c79de91ef6 · inbound

Do LLM Agents Mirror Socio-Cognitive Effects in Power-Asymmetric Conversations? cites this paper.

Do LLM Agents Mirror Socio-Cognitive Effects in Power-Asymmetric Conversations? SOTOPIA: Interactive Evaluation for Social Intelligence in Language Agents

Reference 2

Resolution
metadata mismatch
arxiv_id, observed 2026-05-20T12:23:16.947332Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=arxiv_source observed=2026-05-20T12:20:24.984624Z digest=sha256:ce501b0a4d032632584990c6632b8aeb3b9169fdf4bbbf481cd337b3597aabaa

Observation 9ec357cd-8a27-46a2-9999-0e824042bce1 · inbound

Do LLM Agents Mirror Socio-Cognitive Effects in Power-Asymmetric Conversations? cites this paper.

Do LLM Agents Mirror Socio-Cognitive Effects in Power-Asymmetric Conversations? SOTOPIA: Interactive Evaluation for Social Intelligence in Language Agents

Reference 2

Resolution
metadata mismatch
arxiv_id, observed 2026-05-21T07:34:02.681492Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=arxiv_source observed=2026-05-21T07:32:10.115134Z digest=sha256:d8f46df4bbf2cd87ace7066172190a6b99936e02a24fe56a8daf8914b656b8df

Observation 26f62344-4b8d-4a69-b357-6ef63e79b412 · inbound

Boiling the Frog: A Multi-Turn Benchmark for Agentic Safety cites this paper.

Boiling the Frog: A Multi-Turn Benchmark for Agentic Safety SOTOPIA: Interactive Evaluation for Social Intelligence in Language Agents

Reference 98

Resolution
verified exact
arxiv_id, observed 2026-05-22T05:51:07.776602Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-05-22T05:50:28.114140Z digest=sha256:53ad7e11276dfb562f750d7a3848aa41afffd5a401f846fbe0cf159049cd965d

Observation cb6f09bf-7ec8-45d2-bf23-3819443ee787 · inbound

Boiling the Frog: A Multi-Turn Benchmark for Agentic Safety cites this paper.

Boiling the Frog: A Multi-Turn Benchmark for Agentic Safety SOTOPIA: Interactive Evaluation for Social Intelligence in Language Agents

Reference 98

Resolution
verified exact
arxiv_id, observed 2026-05-25T06:06:43.133940Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-05-25T06:05:27.736494Z digest=sha256:04e59ca9d7d3af1a8e117e32153a9c563aad85f04d2cf17511ef0105c3496e9c

Observation a33cb92a-6d13-4cdb-989b-fdae10fba4f0 · inbound

Distilling Game Code World Model Generation into Lightweight Large Language Models cites this paper.

Distilling Game Code World Model Generation into Lightweight Large Language Models SOTOPIA: Interactive Evaluation for Social Intelligence in Language Agents

Reference 40

Resolution
malformed identifier
arxiv_id, observed 2026-06-30T14:04:44.669446Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-06-30T13:58:37.956333Z digest=sha256:b3c04c191a723891828dcd08436bf27c1fac2f22db0beb277de9c97b2ce6dba9

Observation 37e2fa4a-a3f5-485e-bc46-307ff531f485 · inbound

CRPO: Character-centric Group Relative Policy Optimization for Role-aware Reasoning in Role-playing Agents cites this paper.

CRPO: Character-centric Group Relative Policy Optimization for Role-aware Reasoning in Role-playing Agents SOTOPIA: Interactive Evaluation for Social Intelligence in Language Agents

Reference 35

Resolution
verified exact
arxiv_id, observed 2026-06-29T21:53:59.299763Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-06-29T21:50:04.277333Z digest=sha256:13b18e0042f94feba2e9c315484929cb4223685a12a781bcab0d9d38bfe34477

Observation dfc5bccf-0713-4cfb-bdea-96bda3f6724c · inbound

Got a Secret? LLM Agents Can't Keep It: Evaluating Privacy in Multi-Agent Systems cites this paper.

Got a Secret? LLM Agents Can't Keep It: Evaluating Privacy in Multi-Agent Systems SOTOPIA: Interactive Evaluation for Social Intelligence in Language Agents

Reference 46

Resolution
metadata mismatch
arxiv_id, observed 2026-06-29T16:43:40.531528Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-06-29T16:38:21.049712Z digest=sha256:de1593e04c3da7641be23a03d69e27758eca06c7c882b7398db2796525387628

Observation 8ca7c88a-47bc-4592-a1cd-96d42e7e98b9 · inbound

MINDGAMES: A Live Arena for Evaluating Social and Strategic Reasoning in Multi-Agent LLMs cites this paper.

MINDGAMES: A Live Arena for Evaluating Social and Strategic Reasoning in Multi-Agent LLMs SOTOPIA: Interactive Evaluation for Social Intelligence in Language Agents

Reference 22

Resolution
verified exact
arxiv_id, observed 2026-06-29T07:23:13.223995Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-06-29T07:15:27.939886Z digest=sha256:f2ac3cacee0ad745c6abad3d3a5b0dd39fe2adb76ebfc88b5e2bbb229e5dfdd3

Observation 0fb19b8f-a38f-4f58-97fb-ef99afa835a8 · inbound

Resonant Minds: Closed-Loop Social Avatars with Theory of Mind cites this paper.

Resonant Minds: Closed-Loop Social Avatars with Theory of Mind SOTOPIA: Interactive Evaluation for Social Intelligence in Language Agents

Reference 57

Resolution
metadata mismatch
arxiv_id, observed 2026-07-02T12:36:56.224215Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-06-28T02:04:39.753443Z digest=sha256:6fca5564c3e036eac9d16cff26b6f0d6ed1afd340b1619970d17dd757201d1c9

Observation 0058c49c-2947-4f15-8409-20b28af87b40 · inbound

Toward Temporal Realism in City-Scale Crisis Response Simulation using LLM Agents cites this paper.

Toward Temporal Realism in City-Scale Crisis Response Simulation using LLM Agents SOTOPIA: Interactive Evaluation for Social Intelligence in Language Agents

Reference 36

Resolution
verified exact
arxiv_id, observed 2026-07-04T05:49:38.347257Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-06-26T15:16:46.281166Z digest=sha256:68a6efc1a0f5c95402dd842f74cfbd9ab47a03c793a3970f0ff61663b20a3030

Observation fbe1802b-caa6-484a-b894-f7fed37c2c71 · inbound

Social World Model for Lifelong Social Intelligence cites this paper.

Social World Model for Lifelong Social Intelligence SOTOPIA: Interactive Evaluation for Social Intelligence in Language Agents

Reference 1

Resolution
verified exact
arxiv_id, observed 2026-07-04T06:29:37.143740Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=arxiv_source observed=2026-06-26T14:29:11.743627Z digest=sha256:b23f561b2213904008183b1b4920941171b2e22bc8710fa7da240c1674b2755b

Observation c6c416aa-6e68-49bc-b5be-58470f996ba4 · inbound

Theory of Mind and Persuasion Beyond Conversation: Assessing the Capacity of LLMs to Induce Belief States via Planning and Action cites this paper.

Theory of Mind and Persuasion Beyond Conversation: Assessing the Capacity of LLMs to Induce Belief States via Planning and Action SOTOPIA: Interactive Evaluation for Social Intelligence in Language Agents

Reference 16

Resolution
metadata mismatch
arxiv_id, observed 2026-07-01T10:15:44.668838Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=arxiv_source observed=2026-07-01T05:40:54.002702Z digest=sha256:da32e3e6a888c6659a9d4014a5f32502e1a0c933745093c8dd1129565d54a27b

Observation 0f3c489d-ea83-4b7e-b4de-7384c499019e · inbound

LLM Agents for Deliberative Collaboration: A Study on Joint Decision Making Under Partial Observability cites this paper.

LLM Agents for Deliberative Collaboration: A Study on Joint Decision Making Under Partial Observability SOTOPIA: Interactive Evaluation for Social Intelligence in Language Agents

Reference 30

Resolution
metadata mismatch
local_arxiv, observed 2026-07-08T15:05:03.522465Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=arxiv_source observed=2026-07-08T15:03:14.483228Z digest=sha256:f74f3fa0967d11aa798e7707ec24a861bceaece98a90db41136628b7d9f7fb01

Observation 26b477fa-748e-461d-8f24-9beb88ae7d91 · inbound

Exposure is not manifestation: measurement target and output resolution jointly determine which behavioural-faithfulness evaluator wins cites this paper.

Exposure is not manifestation: measurement target and output resolution jointly determine which behavioural-faithfulness evaluator wins SOTOPIA: Interactive Evaluation for Social Intelligence in Language Agents

Reference 64

Resolution
unresolved
no resolver link, observed 2026-08-02T07:44:09.802630Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T07:44:09.802630Z digest=sha256:bbf5966ed978942f37042182012caf18f7759a81a728af01b01af8c717f870da

Observation 592a4709-0c57-4783-8f17-564a3787c917 · inbound

Reproducing human biases in route choice using large language models: Toward scalable behavioral modeling cites this paper.

Reproducing human biases in route choice using large language models: Toward scalable behavioral modeling SOTOPIA: Interactive Evaluation for Social Intelligence in Language Agents

Reference 57

Resolution
unresolved
no resolver link, observed 2026-07-14T04:15:33.637340Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-07-14T04:15:33.637340Z digest=sha256:94a63c559c06b2655d2e55e1754a687807e56e46731556c155389ace61e15a7d

Observation 80e8b286-9390-47e2-b01d-3e6ab3c3a26c · inbound

COSI-Lab: Conference Living Lab for Modeling Multi-Perspective Multimodal Social Intention cites this paper.

COSI-Lab: Conference Living Lab for Modeling Multi-Perspective Multimodal Social Intention SOTOPIA: Interactive Evaluation for Social Intelligence in Language Agents

Reference 2024

Resolution
malformed identifier
no resolver link, observed 2026-08-03T00:52:08.455524Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T00:52:08.455524Z digest=sha256:55617f194e801c2502b635248164d0e1e8accb9dbc783c3d31065202e25902d4

Observation 8165469f-0a38-4d57-b00c-5c60b6b73de1 · inbound

Learning from Environmental Feedback: Credit Assignment across Multiple Timescales for Agentic Reinforcement Learning cites this paper.

Learning from Environmental Feedback: Credit Assignment across Multiple Timescales for Agentic Reinforcement Learning SOTOPIA: Interactive Evaluation for Social Intelligence in Language Agents

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-12T00:20:31.002411Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T00:20:31.002411Z digest=sha256:559dfaf64f9e406f03b8efab7ac09d8102c25efb91c521e57fce6fea93763fed

Observation 06c4d98c-e1d8-46ae-bfb5-b3da4e03bc83 · inbound

CARD: Controlled Agentic Reddit Discussions for Credit Card Simulation cites this paper.

CARD: Controlled Agentic Reddit Discussions for Credit Card Simulation SOTOPIA: Interactive Evaluation for Social Intelligence in Language Agents

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-11T10:47:58.973533Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T10:47:58.973533Z digest=sha256:c72540ae543189569bede46ec7092553077c877668aae367d02387957795b518

Observation 25822fbe-3dfd-4aa2-825e-415fc8b745d2 · inbound

CARD: Controlled Agentic Reddit Discussions for Credit Card Simulation cites this paper.

CARD: Controlled Agentic Reddit Discussions for Credit Card Simulation SOTOPIA: Interactive Evaluation for Social Intelligence in Language Agents

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-14T04:19:42.764152Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-14T04:19:42.764152Z digest=sha256:c9bf0b6f0024c119a881c5faaec527347e1f1478277ade7ac7554d566573caf5