Pith. sign in

Paper Citation Record · LEDGER

DVM: Towards Controllable LLM Agents in Social Deduction Games

As of 11 August 2026, this Paper Citation Record lists 31 of 31 outbound references and 1 inbound Pith citation observation for arXiv:2501.06695.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2501.06695 v1

Coverage vector

measured 31 of 31 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-10T20:57:45.050254Z

measured 32 of 32 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-11T06:34:44.6726+00:00

measured 1 of 1 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T14:55:26.052077Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-08-07T14:55:29.503801Z

Reference resolution

31 of 31 outbound references displayed

  • verified exact1
  • verified fuzzy12
  • unresolved18
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation e9e7b2ee-ba4d-4d42-8951-9420490a445b · outbound

This paper cites Language mod- els are few-shot learners,.

DVM: Towards Controllable LLM Agents in Social Deduction Games Language mod- els are few-shot learners,

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-10T20:57:44.874657Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T20:57:44.874657Z digest=sha256:10e39601ef3d9f982a2d24c7283c7fab9777a4a8cd2547ca6518ae007560c9f1

Observation 8413c0c2-eece-48df-926e-c959947c4a4f · outbound

This paper cites Palm: Scal- ing language modeling with pathways,.

DVM: Towards Controllable LLM Agents in Social Deduction Games Palm: Scal- ing language modeling with pathways,

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-10T20:57:44.880827Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T20:57:44.880827Z digest=sha256:ae5df67bc7db8d6fa4ac86b7cf358b4062929a54844d9e8948b9a0af2414f931

Observation e78d7e85-c327-418d-bf48-75405cdca201 · outbound

This paper cites LLaMA: Open and Efficient Foundation Language Models.

DVM: Towards Controllable LLM Agents in Social Deduction Games LLaMA: Open and Efficient Foundation Language Models

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-10T20:57:44.886449Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T20:57:44.886449Z digest=sha256:00202c3f2062a465e8ab7400aa7a129cbb7f08b40aa7bbd89e366475bd9393b0

Observation b2271959-9051-465e-8db1-21fa5fc2bce3 · outbound

This paper cites Determinants of LLM-assisted Decision-Making.

DVM: Towards Controllable LLM Agents in Social Deduction Games Determinants of LLM-assisted Decision-Making

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-10T20:57:44.892856Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T20:57:44.892856Z digest=sha256:4862282bf2927f751af92ad29566bf79dee6f9e74efbc5ea51f393794b0129bb

Observation 5e4931df-1ff8-4d5a-a747-b86dd9c2e3c8 · outbound

This paper cites A Survey on Large Language Model-Based Game Agents.

DVM: Towards Controllable LLM Agents in Social Deduction Games A Survey on Large Language Model-Based Game Agents

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-10T20:57:44.899515Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T20:57:44.899515Z digest=sha256:1838c3831d22b9d283ee08290615b5b68c1805c53879d38a1301925f7951d0d3

Observation 4dd40147-d179-4987-9d2a-8ddf3e38bae0 · outbound

This paper cites Cradle: Empowering Foundation Agents Towards General Computer Control.

DVM: Towards Controllable LLM Agents in Social Deduction Games Cradle: Empowering Foundation Agents Towards General Computer Control

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-10T20:57:44.906970Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T20:57:44.906970Z digest=sha256:23b8b656558c87353afa5322b3399c0e9d228beda020ff2ac96dbc323dc770cd

Observation 4fc8d56f-fcc1-47b6-bb8a-29d2789067a3 · outbound

This paper cites Mp5: A multi-modal open-ended embodied system in minecraft via active perception,.

DVM: Towards Controllable LLM Agents in Social Deduction Games Mp5: A multi-modal open-ended embodied system in minecraft via active perception,

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T20:57:45.863866Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-10T20:57:44.913162Z digest=sha256:2b824ded9097816e4dad0976225e9857f0294329f4c6a2f7f0dc4cca207f467a

Observation b602c7e7-e2c3-4960-8e8a-54089ba6219d · outbound

This paper cites Describe, explain, plan and select: Interactive planning with LLMs enables open-world multi-task agents,.

DVM: Towards Controllable LLM Agents in Social Deduction Games Describe, explain, plan and select: Interactive planning with LLMs enables open-world multi-task agents,

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T20:57:45.848296Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-10T20:57:44.918848Z digest=sha256:e5f3aa98c1d5ceb724132ed9f85deca7431eaad7c62edab85013693044700aab

Observation 468961c1-04f2-4b22-9764-f4b2965dc247 · outbound

This paper cites V oyager: An open-ended embodied agent with large language models,.

DVM: Towards Controllable LLM Agents in Social Deduction Games V oyager: An open-ended embodied agent with large language models,

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T20:57:45.832282Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-10T20:57:44.925820Z digest=sha256:916e029c85e4f813ec6abe2e538529d1386c601e4c313e143450eb4c087ce335

Observation adbc158d-911a-4905-956d-a8385a4530bd · outbound

This paper cites Baba is AI: Break the rules to beat the benchmark,.

DVM: Towards Controllable LLM Agents in Social Deduction Games Baba is AI: Break the rules to beat the benchmark,

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T20:57:45.812978Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-10T20:57:44.931633Z digest=sha256:bbe060821735fdee57d59532375f62a5b0f368b2885af6d0ef680a90d88fc290

Observation 2542daa2-a17f-4239-ad82-648ad2e98045 · outbound

This paper cites Avalonbench: Evaluating LLMs playing the game of avalon,.

DVM: Towards Controllable LLM Agents in Social Deduction Games Avalonbench: Evaluating LLMs playing the game of avalon,

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T20:57:45.790897Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-10T20:57:44.936862Z digest=sha256:b48264e90c17d7ba53b72fda8a3713164bfde697e21812f7f06fa660e135dede

Observation bdfee0b1-6063-4ec6-8d1e-867b33df333a · outbound

This paper cites Playing repeated games with Large Language Models.

DVM: Towards Controllable LLM Agents in Social Deduction Games Playing repeated games with Large Language Models

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-10T20:57:44.942547Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T20:57:44.942547Z digest=sha256:a1bc4c231cf6edd35c4ab7ad773c3feb1bf65122dd061f845bd39b7cefe7dbef

Observation 471ef80c-5718-4c39-9a72-c271c48bdd59 · outbound

This paper cites AMONGAGENTS: Evaluating Large Language Models in the Interactive Text-Based Social Deduction Game.

DVM: Towards Controllable LLM Agents in Social Deduction Games AMONGAGENTS: Evaluating Large Language Models in the Interactive Text-Based Social Deduction Game

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-10T20:57:44.947467Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T20:57:44.947467Z digest=sha256:121a9e6c9925974a1915333834719f3286b314cb3ba31ffb3e7bd32a16ec5e6c

Observation 4c1c9fc7-11c1-4678-bb3c-db38da5a3084 · outbound

This paper cites Microscopic analysis on llm players via social deduction game,.

DVM: Towards Controllable LLM Agents in Social Deduction Games Microscopic analysis on llm players via social deduction game,

Reference 14

Resolution
verified exact
raw_fallback, observed 2026-08-10T20:57:45.512341Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-10T20:57:44.955412Z digest=sha256:89b257215d6f85dbcae7bab34688b4d4cc17d15c10cdd6b109ca1c8a119b1b3c

Observation b8595c91-19a3-45c1-a304-ae3ce4fcc31e · outbound

This paper cites Language Agents with Reinforcement Learning for Strategic Play in the Werewolf Game.

DVM: Towards Controllable LLM Agents in Social Deduction Games Language Agents with Reinforcement Learning for Strategic Play in the Werewolf Game

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-10T20:57:44.960655Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T20:57:44.960655Z digest=sha256:21a90aa9060fc6f1015eb8fb51bcedf412c0d5d9f218749c8c6dbf5b1f0d2516

Observation e4cde7d7-45d9-414c-bda1-ae517cba6b25 · outbound

This paper cites Enhance Reasoning for Large Language Models in the Game Werewolf.

DVM: Towards Controllable LLM Agents in Social Deduction Games Enhance Reasoning for Large Language Models in the Game Werewolf

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-10T20:57:44.966522Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T20:57:44.966522Z digest=sha256:686ac30e8baa9d78e9e26dd10ccdda0fec808ffa877f9372c5264d46a612781e

Observation 3dde51aa-4792-4a1a-9e10-865396d5ce65 · outbound

This paper cites Emergent password signalling in the game of werewolf,.

DVM: Towards Controllable LLM Agents in Social Deduction Games Emergent password signalling in the game of werewolf,

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T20:57:45.768458Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-10T20:57:44.973586Z digest=sha256:1ce7bc4b32becb173ac713d3342aaa723b0b579ca42ac556553793f5481b69f5

Observation 3c87d585-8ddc-4bae-896a-8631fb0c9ece · outbound

This paper cites Exploring Large Language Models for Communication Games: An Empirical Study on Werewolf.

DVM: Towards Controllable LLM Agents in Social Deduction Games Exploring Large Language Models for Communication Games: An Empirical Study on Werewolf

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-10T20:57:44.978500Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T20:57:44.978500Z digest=sha256:a505dc83c20071eb1c6bc42f04333fc19ec17d982f50e0026e5a40a8f70fa0de

Observation a2374452-d6c4-4c02-abb8-f6d5c736b536 · outbound

This paper cites Werewolf among us: Multimodal resources for modeling persuasion behaviors in social deduction games,.

DVM: Towards Controllable LLM Agents in Social Deduction Games Werewolf among us: Multimodal resources for modeling persuasion behaviors in social deduction games,

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T20:57:45.749434Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-10T20:57:44.983541Z digest=sha256:4c6471e1d33e8d6c7b9d49f2911ea964293f922fb627546bf53c03affebe5a1a

Observation 0145ff2d-b4d6-4c4c-940b-5ee3ae778a33 · outbound

This paper cites Werewolf-xl: A database for identifying spontaneous affect in large competitive group interactions,.

DVM: Towards Controllable LLM Agents in Social Deduction Games Werewolf-xl: A database for identifying spontaneous affect in large competitive group interactions,

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-10T20:57:44.988346Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T20:57:44.988346Z digest=sha256:02c4bed2b5415b1ad045228f939576856f3b2b3011469dd9fe22e4079b0c4096

Observation 6826a58a-1c42-481c-bcf4-30fd5095764d · outbound

This paper cites Proximal Policy Optimization Algorithms.

DVM: Towards Controllable LLM Agents in Social Deduction Games Proximal Policy Optimization Algorithms

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-10T20:57:44.994010Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T20:57:44.994010Z digest=sha256:9b125f415399031de8df0f3530e4150e229176a0661df5f94e29fa77f1fa2728

Observation 1c7ab567-64dc-4dce-8b9c-c543d0eb0dc8 · outbound

This paper cites Proximal policy optimization with mixed distributed training,.

DVM: Towards Controllable LLM Agents in Social Deduction Games Proximal policy optimization with mixed distributed training,

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-10T20:57:45.000134Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T20:57:45.000134Z digest=sha256:fc43d2517050ea2525144f0fd097d27561bce6b3c83f1cf65575d54bb88b18f0

Observation 04894800-9558-4724-bd1d-cad8b2fefcfa · outbound

This paper cites Proximal policy optimization via enhanced exploration efficiency,.

DVM: Towards Controllable LLM Agents in Social Deduction Games Proximal policy optimization via enhanced exploration efficiency,

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T20:57:45.731947Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-10T20:57:45.005937Z digest=sha256:c0652a0f87b91003dfd0da27abf7d6f6541e5b41f8f35bc60de0d3f8598220ff

Observation 2d81d20a-d9d1-4d13-822e-030374a7855a · outbound

This paper cites Secrets of RLHF in Large Language Models Part I: PPO.

DVM: Towards Controllable LLM Agents in Social Deduction Games Secrets of RLHF in Large Language Models Part I: PPO

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-10T20:57:45.011114Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T20:57:45.011114Z digest=sha256:4be9842dd500647f956a83913d9c368bf913f4b90803f98d4839a64d68da057a

Observation 8f74dfe4-ddea-4706-b0a4-731da9010679 · outbound

This paper cites Is DPO Superior to PPO for LLM Alignment? A Comprehensive Study.

DVM: Towards Controllable LLM Agents in Social Deduction Games Is DPO Superior to PPO for LLM Alignment? A Comprehensive Study

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-10T20:57:45.016486Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T20:57:45.016486Z digest=sha256:ce08f69c0433ba475ad4ebfaab529b31a7e213988e5b2a1db0fd5ed03e8891e6

Observation aa235c4c-6abd-42c2-a020-5a79bc58a8f6 · outbound

This paper cites A theoretical analysis of optimistic proximal policy optimization in linear markov decision processes,.

DVM: Towards Controllable LLM Agents in Social Deduction Games A theoretical analysis of optimistic proximal policy optimization in linear markov decision processes,

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T20:57:45.714650Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-10T20:57:45.022320Z digest=sha256:f437248ef67c4cc050a19ed9a3725e6807c305ed19ac8c8664d860b6aec99f8f

Observation dc4e5840-d085-4ffb-b73f-74cfe5ec6a6d · outbound

This paper cites Finding friend and foe in multi-agent games,.

DVM: Towards Controllable LLM Agents in Social Deduction Games Finding friend and foe in multi-agent games,

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T20:57:45.695214Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-10T20:57:45.027627Z digest=sha256:0b9d4de65395968642ca4e08244c38a1645df7072637c7b03385fca6ccb664ee

Observation da40be48-1122-468d-8c48-e5c5ecde9d40 · outbound

This paper cites Chatglm: A family of large language models from glm-130b to glm-4 all tools,.

DVM: Towards Controllable LLM Agents in Social Deduction Games Chatglm: A family of large language models from glm-130b to glm-4 all tools,

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-10T20:57:45.033170Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T20:57:45.033170Z digest=sha256:57c22183b30da1c537b7673a29a536ffb720db0e0945c0c578c718f43d00bb74

Observation 56f6ffd7-09c7-4db5-9c00-a15a3890d821 · outbound

This paper cites Direct preference optimization: Your language model is secretly a reward model,.

DVM: Towards Controllable LLM Agents in Social Deduction Games Direct preference optimization: Your language model is secretly a reward model,

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T20:57:45.668877Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-10T20:57:45.038646Z digest=sha256:0a0719c16ca3215f00a66054c5f190145290c48967d5eb136c9e9ebbe20eca8f

Observation 8b717e92-4a91-440e-9aa1-22a4870e1697 · outbound

This paper cites ReAct: Synergizing Reasoning and Acting in Language Models.

DVM: Towards Controllable LLM Agents in Social Deduction Games ReAct: Synergizing Reasoning and Acting in Language Models

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-10T20:57:45.044229Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T20:57:45.044229Z digest=sha256:bcf9509337f7229db2a7207688746ccae8ceb6dda8a52d6a42afd051d9478739

Observation 6d15f9e8-5e92-496b-a182-f7ab1b5c780a · outbound

This paper cites Least-to-most prompting enables complex reasoning in large language models,.

DVM: Towards Controllable LLM Agents in Social Deduction Games Least-to-most prompting enables complex reasoning in large language models,

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T20:57:45.647655Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-10T20:57:45.050254Z digest=sha256:7c43e9310aacf0da32f6c4d7295591e1fe993a277b4c1d91771bf77ffc828f04

Pith citing papers

Observation 90c2a39f-f7bc-4831-9126-30e24ae2aedc · inbound

Cracking Aegis: An Adversarial LLM-based Game for Raising Awareness of Vulnerabilities in Privacy Protection cites this paper.

Cracking Aegis: An Adversarial LLM-based Game for Raising Awareness of Vulnerabilities in Privacy Protection DVM: Towards Controllable LLM Agents in Social Deduction Games

Reference 149

Resolution
verified exact
local_arxiv, observed 2026-08-07T14:55:29.578410Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-07T14:55:26.052077Z digest=sha256:668655d1951921c01a357fdd21b610215ec7f0b8ca2a1d6b016fc1471786cd53