Pith. sign in

REVIEW 2 cited by

LLM-driven Imitation of Subrational Behavior : Illusion or Reality?

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 2402.08755 v1 pith:2GBDGBOG submitted 2024-02-13 cs.AI econ.GNq-fin.EC

classification cs.AIecon.GNq-fin.EC
keywords humanllmsframeworksubrationalhumansmodelsabilityagents
verification ladder T0 review T1 audit T2 compute T3 formal
0 comments
read the original abstract

Modeling subrational agents, such as humans or economic households, is inherently challenging due to the difficulty in calibrating reinforcement learning models or collecting data that involves human subjects. Existing work highlights the ability of Large Language Models (LLMs) to address complex reasoning tasks and mimic human communication, while simulation using LLMs as agents shows emergent social behaviors, potentially improving our comprehension of human conduct. In this paper, we propose to investigate the use of LLMs to generate synthetic human demonstrations, which are then used to learn subrational agent policies though Imitation Learning. We make an assumption that LLMs can be used as implicit computational models of humans, and propose a framework to use synthetic demonstrations derived from LLMs to model subrational behaviors that are characteristic of humans (e.g., myopic behavior or preference for risk aversion). We experimentally evaluate the ability of our framework to model sub-rationality through four simple scenarios, including the well-researched ultimatum game and marshmallow experiment. To gain confidence in our framework, we are able to replicate well-established findings from prior human studies associated with the above scenarios. We conclude by discussing the potential benefits, challenges and limitations of our framework.

Discussion (0). Sign in to comment.

Forward citations

Cited by 2 Pith papers

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score. Full citation record

  1. The Effect of State Representation on LLM Agent Behavior in Dynamic Routing Games

    cs.AI 2025-06 conditional novelty 6.0 of 10

    In a repeated Braess routing game, LLM agents given summarized, regret-based, and own-action-only state representations converge closer to Nash equilibrium and behave more stably than agents given full chat transcript...

  2. Securing Agentic AI: Threat Modeling and Risk Analysis for Network Monitoring Agentic AI System

    cs.CR 2025-08 conditional novelty 3.0 of 10

    An LLM network-monitoring agent experienced nearly doubled telemetry delays under replayed DoS traffic, and edited memory files led it to choose longer, heavier packet captures, in a two-case test of the MAESTRO threa...

Pith tools