Pith. sign in

Paper Citation Record · LEDGER

Werewolf Arena: A Case Study in LLM Evaluation via Social Deduction

As of 9 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 20 inbound Pith citation observations for arXiv:2407.13943.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2407.13943 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 20 of 20 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00

measured 20 of 20 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-08T21:58:39.391646Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-04T00:49:18.497365Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation a0595463-d9c1-42c6-8e6d-42b7648a6a6f · inbound

Learning Strategic Language Agents in the Werewolf Game with Iterative Latent Space Policy Optimization cites this paper.

Learning Strategic Language Agents in the Werewolf Game with Iterative Latent Space Policy Optimization Werewolf Arena: A Case Study in LLM Evaluation via Social Deduction

Reference 2022

Resolution
unresolved
no resolver link, observed 2026-08-08T21:58:39.391646Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T21:58:39.391646Z digest=sha256:55ff62a64571c9829b810c30320b9ec57f7a18dcaf671bbeb82e05c13b9a996c

Observation b354bf91-0a7a-4a52-beff-d7b843a97177 · inbound

Game Theory Meets Large Language Models: A Systematic Survey with Taxonomy and New Frontiers cites this paper.

Game Theory Meets Large Language Models: A Systematic Survey with Taxonomy and New Frontiers Werewolf Arena: A Case Study in LLM Evaluation via Social Deduction

Reference 66

Resolution
unresolved
no resolver link, observed 2026-08-07T22:50:32.949249Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T22:50:32.949249Z digest=sha256:c3ba519edec5125bca4ae67dbf3c4680d167e67a696b6b4b216c3e42e42c1916

Observation cc9f7743-f87c-4dd4-92b1-1cfc14760251 · inbound

Agents Require Metacognitive and Strategic Reasoning to Succeed in the Coming Labor Markets cites this paper.

Agents Require Metacognitive and Strategic Reasoning to Succeed in the Coming Labor Markets Werewolf Arena: A Case Study in LLM Evaluation via Social Deduction

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-07T14:02:22.412122Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:02:22.412122Z digest=sha256:f5fad131830780f935955ad1fcd88254ff98f5a940025a081ee75d4f636cbb5a

Observation 6df2bdba-089d-4aee-bafa-a3f3e4da6171 · inbound

SocialMaze: A Benchmark for Evaluating Social Reasoning in Large Language Models cites this paper.

SocialMaze: A Benchmark for Evaluating Social Reasoning in Large Language Models Werewolf Arena: A Case Study in LLM Evaluation via Social Deduction

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-07T12:45:15.845155Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:45:15.845155Z digest=sha256:1d4a9ba8ed6f28c7b8aa940e46eb9062247b88d54a69a375b61da86400226366

Observation 724b08bb-f7b8-44c3-a214-03db4ff78204 · inbound

TextAtari: 100K Frames Game Playing with Language Agents cites this paper.

TextAtari: 100K Frames Game Playing with Language Agents Werewolf Arena: A Case Study in LLM Evaluation via Social Deduction

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-07T10:51:57.923142Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T10:51:57.923142Z digest=sha256:e6dc7d17ea2f7c6d098c63a14982b3843c43f9c1eb150851667bdfd1f61cd75e

Observation 3be38aa0-8869-492d-812c-2af4af6b75b7 · inbound

Strategy Adaptation in Large Language Model Werewolf Agents cites this paper.

Strategy Adaptation in Large Language Model Werewolf Agents Werewolf Arena: A Case Study in LLM Evaluation via Social Deduction

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-06T16:42:43.264423Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T16:42:43.264423Z digest=sha256:2a1ae6ac1cfb79c06199e66faa700e3bd1d812bb7a5c434c178c15457cb3631b

Observation 7eeaeb2a-9457-43d6-85e2-6642c708532a · inbound

Deceive, Detect, and Disclose: Large Language Models Play Mini-Mafia cites this paper.

Deceive, Detect, and Disclose: Large Language Models Play Mini-Mafia Werewolf Arena: A Case Study in LLM Evaluation via Social Deduction

Reference 2

Resolution
verified exact
arxiv_id, observed 2026-05-18T13:26:24.820227Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-05-18T13:25:17.313704Z digest=sha256:b28c76af41a1e7fd3d187cc4bf6fbcffd3c7b6f594bfd9d30130c8a5de5f48f2

Observation 14bb5245-6fcd-4294-b7f8-0e5807abde3d · inbound

AIT Academy: Cultivating the Complete Agent with a Confucian Three-Domain Curriculum cites this paper.

AIT Academy: Cultivating the Complete Agent with a Confucian Three-Domain Curriculum Werewolf Arena: A Case Study in LLM Evaluation via Social Deduction

Reference 23

Resolution
verified exact
arxiv_id, observed 2026-05-10T09:43:49.911579Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-10T05:08:54.560648Z digest=sha256:791373069d59f6c4f84bacaabc35ab54d5c25954e66300d9ea899dfb1a14d75f

Observation 41bb56e2-2d5c-40f1-b85f-5f907eb5614c · inbound

Towards Generalist Game Players: An Investigation of Foundation Models in the Game Multiverse cites this paper.

Towards Generalist Game Players: An Investigation of Foundation Models in the Game Multiverse Werewolf Arena: A Case Study in LLM Evaluation via Social Deduction

Reference 9

Resolution
verified exact
arxiv_id, observed 2026-05-12T03:26:19.123141Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-12T03:25:24.844859Z digest=sha256:084a65bab7e402a8308978a8a267bc7f94e4000230e5a1e46077fcb86d645677

Observation d4372387-bfcd-4f7f-a16a-7ad8176b37e7 · inbound

Towards Generalist Game Players: An Investigation of Foundation Models in the Game Multiverse cites this paper.

Towards Generalist Game Players: An Investigation of Foundation Models in the Game Multiverse Werewolf Arena: A Case Study in LLM Evaluation via Social Deduction

Reference 9

Resolution
verified exact
arxiv_id, observed 2026-05-13T06:47:26.904795Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-13T06:44:28.552513Z digest=sha256:8680c33652f940223cf100b31c3f0fc015244bb92fec28469c50eca06dc6ab0f

Observation 9d9b41be-bcdb-4364-a6c3-024f314de6cf · inbound

MINDGAMES: A Live Arena for Evaluating Social and Strategic Reasoning in Multi-Agent LLMs cites this paper.

MINDGAMES: A Live Arena for Evaluating Social and Strategic Reasoning in Multi-Agent LLMs Werewolf Arena: A Case Study in LLM Evaluation via Social Deduction

Reference 28

Resolution
verified exact
arxiv_id, observed 2026-06-29T07:23:13.222505Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-06-29T07:15:27.939886Z digest=sha256:c9b06778eb5fb8bb7ef270a5a8118ddfbdbe1bb729bdd72ccd53a30bc33de35c

Observation 45e7a251-a33f-4ef9-82d8-9dec8bf25823 · inbound

RogueAI: A Reverse Turing Test for Detecting Licensed AI Deception in Dialogue cites this paper.

RogueAI: A Reverse Turing Test for Detecting Licensed AI Deception in Dialogue Werewolf Arena: A Case Study in LLM Evaluation via Social Deduction

Reference 9

Resolution
verified exact
arxiv_id, observed 2026-07-03T15:18:33.728065Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-06-27T06:34:39.457798Z digest=sha256:4b466c334157ea17c33093c43b542bff66713fc7e159e1e35700667dcfee3492

Observation 62b9d302-6e6f-4bc0-a0b9-8a6ae95a92a3 · inbound

Enhancing Decision-Making with Large Language Models through Multi-Agent Fictitious Play cites this paper.

Enhancing Decision-Making with Large Language Models through Multi-Agent Fictitious Play Werewolf Arena: A Case Study in LLM Evaluation via Social Deduction

Reference 3

Resolution
verified exact
arxiv_id, observed 2026-07-04T00:49:18.499911Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-06-26T20:57:49.840546Z digest=sha256:314064c43c1b93793980ba2e77dcccf5779323066d12ee46cc52afe54db23a93

Observation 4c68c23d-7cae-4249-944a-47e68d77ebdc · inbound

Thinking Out Loud: Real-Time Deception Monitoring in Asymmetric LLM Negotiations cites this paper.

Thinking Out Loud: Real-Time Deception Monitoring in Asymmetric LLM Negotiations Werewolf Arena: A Case Study in LLM Evaluation via Social Deduction

Reference 9

Resolution
verified exact
arxiv_id, observed 2026-07-01T08:45:35.459536Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-07-01T07:11:25.568064Z digest=sha256:93e423d13aaf2610a70b25de09c2d5550d7480e93ea2b7784880d333a6c3e2cf

Observation 5b6be818-7b78-48d2-aa8f-de906eedda25 · inbound

Theory of Mind and Persuasion Beyond Conversation: Assessing the Capacity of LLMs to Induce Belief States via Planning and Action cites this paper.

Theory of Mind and Persuasion Beyond Conversation: Assessing the Capacity of LLMs to Induce Belief States via Planning and Action Werewolf Arena: A Case Study in LLM Evaluation via Social Deduction

Reference 85

Resolution
metadata mismatch
arxiv_id, observed 2026-07-01T10:15:44.661711Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-07-01T05:40:54.002702Z digest=sha256:d0221ea3e40c79b32db066538c34dd4b9407323e536433de53108b2fcc69a91a

Observation fd43ebee-2b79-4cf7-a707-d86854e5e49e · inbound

MafiaScope: Non-Invasive, Time-Resolved Belief Probing for LLM Agents in Social Deduction Games cites this paper.

MafiaScope: Non-Invasive, Time-Resolved Belief Probing for LLM Agents in Social Deduction Games Werewolf Arena: A Case Study in LLM Evaluation via Social Deduction

Reference 4

Resolution
unresolved
no resolver link, observed 2026-07-14T10:15:59.479435Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-07-14T10:15:59.479435Z digest=sha256:3e775e321c6e062d760e32620a3bbf0613a0d4faa8b997aa30411760bff79466

Observation 3b95b283-e15e-4d31-aedd-e5b1279b29c4 · inbound

MafiaScope: Non-Invasive, Time-Resolved Belief Probing for LLM Agents in Social Deduction Games cites this paper.

MafiaScope: Non-Invasive, Time-Resolved Belief Probing for LLM Agents in Social Deduction Games Werewolf Arena: A Case Study in LLM Evaluation via Social Deduction

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-02T07:14:59.839689Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-02T07:14:59.839689Z digest=sha256:88ff1f9c62bcb61ca13e35050409402648bca2bd3a56e58e763cdf768fd0ad85

Observation d3fc3623-1d5b-4819-9f6d-f4875f209f35 · inbound

Auditing Belief-Conditioned LLM Agents in Hidden-Information Social Deduction Games cites this paper.

Auditing Belief-Conditioned LLM Agents in Hidden-Information Social Deduction Games Werewolf Arena: A Case Study in LLM Evaluation via Social Deduction

Reference 3

Resolution
unresolved
no resolver link, observed 2026-07-14T09:04:16.361753Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-14T09:04:16.361753Z digest=sha256:433ffdda0f858f9913ec9f423626aaabb4341f8f8392e7855ff20fdff27664c4

Observation bcfcad8a-1b0e-47b7-9701-120e26b7325d · inbound

Cumulative suspicion and absorption dynamics in an agent-based Mafia game cites this paper.

Cumulative suspicion and absorption dynamics in an agent-based Mafia game Werewolf Arena: A Case Study in LLM Evaluation via Social Deduction

Reference 23

Resolution
unresolved
no resolver link, observed 2026-07-31T08:47:10.407663Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T08:47:10.407663Z digest=sha256:f9e635902ef468714272e93a0b85434a3dbae8c7cd50dddc9ee054f4a367e216

Observation 4d19343b-bada-4ef0-933b-0773c7bfddbb · inbound

Even More Deception: Objective Misalignment in Mixed-Motive LLM Multi-Agent Systems cites this paper.

Even More Deception: Objective Misalignment in Mixed-Motive LLM Multi-Agent Systems Werewolf Arena: A Case Study in LLM Evaluation via Social Deduction

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-01T00:51:15.781283Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T00:51:15.781283Z digest=sha256:988d545586e41d1889ff4dc4b5532b71c1ad8717ce1d7db69704cdc12d143661