Pith. sign in

Paper Citation Record · LEDGER

MotiveBench: How Far Are We From Human-Like Motivational Reasoning in Large Language Models?

As of 17 August 2026, this Paper Citation Record lists 39 of 39 outbound references and 1 inbound Pith citation observation for arXiv:2506.13065.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2506.13065 v1

Coverage vector

measured 39 of 39 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T00:43:08.528481Z

measured 40 of 40 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-16T06:30:59.297886+00:00

measured 1 of 1 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-05-20T15:31:25.079191Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-20T15:33:25.520674Z

Reference resolution

39 of 39 outbound references displayed

  • verified exact0
  • verified fuzzy19
  • unresolved19
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch1

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 35508956-b02a-4755-8f58-3dc3c487d173 · outbound

This paper cites an unresolved cited work.

MotiveBench: How Far Are We From Human-Like Motivational Reasoning in Large Language Models? Unresolved cited work

Reference 1

Resolution
unresolved
raw_fallback, observed 2026-08-07T00:43:12.930373Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-07T00:43:06.990283Z digest=sha256:606586438dbd21e0b67c881cfde3093eba6ba2a6a6943f580851c3b14e98afe5

Observation 3c17f80e-9116-49fd-a4f7-4b8307fc33f9 · outbound

This paper cites A Sequence-to-Sequence Model for User Simulation in Spoken Dialogue Systems.

MotiveBench: How Far Are We From Human-Like Motivational Reasoning in Large Language Models? A Sequence-to-Sequence Model for User Simulation in Spoken Dialogue Systems

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-07T00:43:06.413338Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:43:06.413338Z digest=sha256:04a08266d57340de418f3f845938558bf1e8e1ac86126af30592f921ea96425f

Observation e1dfb8d4-a5d3-4b33-8f76-5f98ad7c42d0 · outbound

This paper cites Emergent autonomous scientific research capabilities of large language models.

MotiveBench: How Far Are We From Human-Like Motivational Reasoning in Large Language Models? Emergent autonomous scientific research capabilities of large language models

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-07T00:43:06.473711Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:43:06.473711Z digest=sha256:b23377d6018a9c054c66b84c147f1b119d3bb31d3afb6a29f20b41342d5e693f

Observation f5353aec-fa41-471f-b8fa-55b6bc307c1f · outbound

This paper cites The distractors must be related to certain parts of the information in the question.

MotiveBench: How Far Are We From Human-Like Motivational Reasoning in Large Language Models? The distractors must be related to certain parts of the information in the question

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:43:11.685098Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-07T00:43:07.590113Z digest=sha256:b1295325fed2fe1d688880f6494d9b5ed7402b2be2f54530907cd6739a4edb77

Observation 20551643-e8ac-48af-b96b-677e5edc69af · outbound

This paper cites This is necessary to ensure each question is challenging.

MotiveBench: How Far Are We From Human-Like Motivational Reasoning in Large Language Models? This is necessary to ensure each question is challenging

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:43:11.421325Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-07T00:43:07.642886Z digest=sha256:4c965596193542581e2914782595fc5aa0f43eb9cccf4c5a0acc77c492d762eb

Observation 1a539ba5-65b4-49da-bbb9-ce6a994edf43 · outbound

This paper cites If they do not, suggest adding the relevant distracting information or modifying the options.

MotiveBench: How Far Are We From Human-Like Motivational Reasoning in Large Language Models? If they do not, suggest adding the relevant distracting information or modifying the options

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:43:10.499118Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-07T00:43:07.930913Z digest=sha256:abfbc2d27ccfbe4cd11f52a72549b2def542a1d21c48098dae8e34f65ffb01c2

Observation ff39faf5-77e3-4a69-892b-c93f7ddf8d9d · outbound

This paper cites an unresolved cited work.

MotiveBench: How Far Are We From Human-Like Motivational Reasoning in Large Language Models? Unresolved cited work

Reference 7

Resolution
unresolved
raw_fallback, observed 2026-08-07T00:43:10.321988Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-07T00:43:07.968105Z digest=sha256:e6f72a2eaf1475bf0e9d0efc02af8495da818d5d44bb41075aac1908cc0c07f1

Observation 3420088b-b267-4d3a-9d72-8a86c395c3ce · outbound

This paper cites Please provide specific modification suggestions for the question set and give your feedback to the question author in a reasonable tone.

MotiveBench: How Far Are We From Human-Like Motivational Reasoning in Large Language Models? Please provide specific modification suggestions for the question set and give your feedback to the question author in a reasonable tone

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:43:10.128969Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-07T00:43:08.055680Z digest=sha256:d58a272bb739e86f4b643a78f6ce2d1371f79c485b57cc43adc1bf08c06e4fec

Observation db7c632b-8ec6-4419-b7cb-b1f5fecf5a77 · outbound

This paper cites LiveBench: A Challenging, Contamination-Limited LLM Benchmark.

MotiveBench: How Far Are We From Human-Like Motivational Reasoning in Large Language Models? LiveBench: A Challenging, Contamination-Limited LLM Benchmark

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-07T00:43:06.828900Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:43:06.828900Z digest=sha256:f970b395aec8cab28f50824b06345da5552ce268f3379ca7c4e7a67ee7686a87

Observation 9edd8e6d-ef77-473d-974a-340009677af8 · outbound

This paper cites SOTOPIA: Interactive Evaluation for Social Intelligence in Language Agents.

MotiveBench: How Far Are We From Human-Like Motivational Reasoning in Large Language Models? SOTOPIA: Interactive Evaluation for Social Intelligence in Language Agents

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-07T00:43:06.885426Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:43:06.885426Z digest=sha256:86b92bed847e6a12b38ce3a0c76a603b3195b7abc90bffa084d641679cfadd18

Observation 087c0df3-dd1b-4d1a-a7db-86ede95fe0fd · outbound

This paper cites an unresolved cited work.

MotiveBench: How Far Are We From Human-Like Motivational Reasoning in Large Language Models? Unresolved cited work

Reference 11

Resolution
unresolved
raw_fallback, observed 2026-08-07T00:43:12.939017Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-07T00:43:06.934518Z digest=sha256:99df817e5a6070512dbed7ee473d08c71dff13abd200405c1d69574581c8552d

Observation 79f74805-d35e-4822-a887-d59f223b79ea · outbound

This paper cites Question: {Question_Content} Options: {Options} CoT Prompt for Evaluation The following is a {Question_Type}.

MotiveBench: How Far Are We From Human-Like Motivational Reasoning in Large Language Models? Question: {Question_Content} Options: {Options} CoT Prompt for Evaluation The following is a {Question_Type}

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:43:12.913951Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-07T00:43:07.106368Z digest=sha256:7037fd11784a6c9b7e09aa15801c2a285a19f5e3b811b4acaf692f61bb7dea8d

Observation aa1e3507-a675-4c8f-80eb-74bc52971eb4 · outbound

This paper cites an unresolved cited work.

MotiveBench: How Far Are We From Human-Like Motivational Reasoning in Large Language Models? Unresolved cited work

Reference 15

Resolution
unresolved
raw_fallback, observed 2026-08-07T00:43:12.783342Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-07T00:43:07.136703Z digest=sha256:b32ef221ff21749ed5fa926928ab016ea7b89c9d28d4789f2e737b03f7f28a42

Observation 95eb9b0c-ec08-46a3-953c-f56210fcb7e1 · outbound

This paper cites A, B, C, D, E, F.

MotiveBench: How Far Are We From Human-Like Motivational Reasoning in Large Language Models? A, B, C, D, E, F

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:43:12.921343Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-07T00:43:07.196600Z digest=sha256:e3489dbb1bb70b404ccc0709988a1ce0f288b35c75e0233ab6fc0bfbea9d3d26

Observation c710a47d-e5b7-4dbd-8c20-ed454e2711f1 · outbound

This paper cites an unresolved cited work.

MotiveBench: How Far Are We From Human-Like Motivational Reasoning in Large Language Models? Unresolved cited work

Reference 17

Resolution
unresolved
raw_fallback, observed 2026-08-07T00:43:12.722206Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-07T00:43:07.249369Z digest=sha256:e69b0d43fd542d82a6c1f7ee6cf5652b8741a5c54998b696bf25dba92a2e500a

Observation 6cd9dd59-f3c1-4ea8-8cd4-21fe05c3a290 · outbound

This paper cites The question should not contain any direct description related to the predicted motivation.

MotiveBench: How Far Are We From Human-Like Motivational Reasoning in Large Language Models? The question should not contain any direct description related to the predicted motivation

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:43:12.556860Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-07T00:43:07.285303Z digest=sha256:4f2867231f09c3b5187ba78ac2e6324f7054a754075b798c7f79da92ff40f7e9

Observation 770135ea-b476-473d-9e32-c73d2db2ee5f · outbound

This paper cites The question should not contain any direct description related to the predicted behavior.

MotiveBench: How Far Are We From Human-Like Motivational Reasoning in Large Language Models? The question should not contain any direct description related to the predicted behavior

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:43:12.439760Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-07T00:43:07.340325Z digest=sha256:b298d3658873cb2f61a3f92c20851a3963da401e65f3b6a03a8697bf1fc70141

Observation f5c20ab1-dffd-4c2d-8346-44a45e97c11b · outbound

This paper cites The question should only include the complex scenario and the character’s profile.

MotiveBench: How Far Are We From Human-Like Motivational Reasoning in Large Language Models? The question should only include the complex scenario and the character’s profile

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:43:12.298433Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-07T00:43:07.393854Z digest=sha256:3c28f30fb8d956f8db060ae17c1d113af29a37624541257af5a0d6788d2b19b2

Observation 8ae918ab-fd12-4ce2-87cb-054e2557ac0c · outbound

This paper cites Please rewrite this scenario by cor- recting any logical inconsistencies, and add relevant details to make the scenario, profile, motivation, and behavior more vivid and complex.

MotiveBench: How Far Are We From Human-Like Motivational Reasoning in Large Language Models? Please rewrite this scenario by cor- recting any logical inconsistencies, and add relevant details to make the scenario, profile, motivation, and behavior more vivid and complex

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:43:12.194449Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-07T00:43:07.446648Z digest=sha256:34b07c5eccaec6b60066f43d21d714ad693126ecef00abaa4f791cf422ab8e9c

Observation 8d36fe18-465a-4713-81cc-9dd2da8644ef · outbound

This paper cites However, ensure that the motivation and behavior are only related to real human needs, not to any POIs or products in the text.

MotiveBench: How Far Are We From Human-Like Motivational Reasoning in Large Language Models? However, ensure that the motivation and behavior are only related to real human needs, not to any POIs or products in the text

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:43:12.009779Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-07T00:43:07.498704Z digest=sha256:59afefcdfdf3766aa31690029dbf85d5960f58ab00d6c3f035921fb62e53c979

Observation e82ebf1b-7ea1-4f8d-a9e0-dec0bc1035e7 · outbound

This paper cites Therefore, please ensure that each question has enough rich and complex scenario and profile information to support correct reasoning.

MotiveBench: How Far Are We From Human-Like Motivational Reasoning in Large Language Models? Therefore, please ensure that each question has enough rich and complex scenario and profile information to support correct reasoning

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:43:11.843382Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-07T00:43:07.535893Z digest=sha256:65442ee78867fdcd327cac6ba4459ed6f5313e63848cf9be28cdbc42cf903204

Observation cec63912-de7a-4c5f-b789-d14f939bbb56 · outbound

This paper cites The motivation reasoning question should include additional behavioral information about the character.

MotiveBench: How Far Are We From Human-Like Motivational Reasoning in Large Language Models? The motivation reasoning question should include additional behavioral information about the character

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:43:11.284235Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-07T00:43:07.672766Z digest=sha256:36f3fdb59c4c7db5d6b379464f48206addb18605f4048a4b3c8dcd68dd982edc

Observation 69957a79-e6f8-4ad5-91d8-628f379fe3c0 · outbound

This paper cites an unresolved cited work.

MotiveBench: How Far Are We From Human-Like Motivational Reasoning in Large Language Models? Unresolved cited work

Reference 27

Resolution
unresolved
raw_fallback, observed 2026-08-07T00:43:11.126370Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-07T00:43:07.721601Z digest=sha256:79e74d686db6ba79a99ed8fc94a9dc3309e52d957c59ba3dfc19c617e1f7c925

Observation 338d7ffd-2e14-41fc-a359-b53a0bfc8773 · outbound

This paper cites If not, suggest modifications to the scenario or character profile to make the information clearer or more comprehensive.

MotiveBench: How Far Are We From Human-Like Motivational Reasoning in Large Language Models? If not, suggest modifications to the scenario or character profile to make the information clearer or more comprehensive

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:43:11.005281Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-07T00:43:07.779326Z digest=sha256:e235234aad65007e417c4752769343dd5d5bb2a44ea353eb528e14e7ea3984e1

Observation b5de10a9-74f6-4f29-bde3-08ced9ab848e · outbound

This paper cites an unresolved cited work.

MotiveBench: How Far Are We From Human-Like Motivational Reasoning in Large Language Models? Unresolved cited work

Reference 29

Resolution
unresolved
raw_fallback, observed 2026-08-07T00:43:10.837908Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-07T00:43:07.830422Z digest=sha256:dc9bbc538d023a6c77dbc4892cd04ae7929b2c878f97589bcee339a1877b1c1a

Observation 28d6b908-c370-4144-8929-7d2eb874c42c · outbound

This paper cites an unresolved cited work.

MotiveBench: How Far Are We From Human-Like Motivational Reasoning in Large Language Models? Unresolved cited work

Reference 30

Resolution
unresolved
raw_fallback, observed 2026-08-07T00:43:10.685253Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-07T00:43:07.882397Z digest=sha256:63507b0e0bfdaa105ea7303ac3a910c7ed0cce594944fad869b0dc0910906b71

Observation efa51cdc-7d40-45da-b9dd-67536a8bfa85 · outbound

This paper cites The question should not include any description related to the predicted motivation.

MotiveBench: How Far Are We From Human-Like Motivational Reasoning in Large Language Models? The question should not include any description related to the predicted motivation

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:43:09.972730Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-07T00:43:08.123433Z digest=sha256:61ca9207453be0146b6afd6d25cb4999598ec54e4cd9e65b3c3b2cbbc8fe7ccb

Observation 315c78b7-171a-40e0-8c17-0823a9400072 · outbound

This paper cites The question should not include any description related to the predicted behavior.

MotiveBench: How Far Are We From Human-Like Motivational Reasoning in Large Language Models? The question should not include any description related to the predicted behavior

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:43:09.766268Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-07T00:43:08.173303Z digest=sha256:abe6bbd9e6d517a8527dd8056f9ba0a09d10d023b0dbf306e6057d84365b96d8

Observation ef08e0bb-8316-430e-9133-a3402584a0b6 · outbound

This paper cites an unresolved cited work.

MotiveBench: How Far Are We From Human-Like Motivational Reasoning in Large Language Models? Unresolved cited work

Reference 36

Resolution
unresolved
raw_fallback, observed 2026-08-07T00:43:09.592265Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-07T00:43:08.253355Z digest=sha256:eb0fb8c0ecd0b4d37b971e0db3267a3cf6d7d17dc063d2723186d3ec62c17a09

Observation 14a6bdec-1881-4797-a14f-10c4c2a7777f · outbound

This paper cites an unresolved cited work.

MotiveBench: How Far Are We From Human-Like Motivational Reasoning in Large Language Models? Unresolved cited work

Reference 37

Resolution
unresolved
raw_fallback, observed 2026-08-07T00:43:09.356741Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-07T00:43:08.328734Z digest=sha256:4c71ee4a38c79ea69deb11e9e80a6c07c13c1dd3c06601330031a76e29f42a57

Observation fc6598fb-7e8f-4487-b224-c7e20c52c968 · outbound

This paper cites an unresolved cited work.

MotiveBench: How Far Are We From Human-Like Motivational Reasoning in Large Language Models? Unresolved cited work

Reference 38

Resolution
unresolved
raw_fallback, observed 2026-08-07T00:43:09.107701Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-07T00:43:08.405819Z digest=sha256:d94b47942321dee49c4c44af3bf3c3750e28b4348f846be5be3a6fe52faa2516

Observation f69e1631-63be-471c-8d31-56af8534c153 · outbound

This paper cites Respondents should only reason based on the question provided, without seeing any other information.

MotiveBench: How Far Are We From Human-Like Motivational Reasoning in Large Language Models? Respondents should only reason based on the question provided, without seeing any other information

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:43:08.960846Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-07T00:43:08.452146Z digest=sha256:a80977b1672fb938e951d7e6a2279e1177452a118c358ca801f6113d98fac025

Observation ab9cadef-d8bf-4422-8d6f-9030e73f35f5 · outbound

This paper cites self-promotion.

MotiveBench: How Far Are We From Human-Like Motivational Reasoning in Large Language Models? self-promotion

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:43:08.797069Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-07T00:43:08.528481Z digest=sha256:ef89be8ec2155f08fcdc5ee94157fc422832335029223279eb3338845b949167

Observation 78b84e43-6e4c-4d62-9ab0-8e01b86a96ed · outbound

This paper cites Clever Hans or Neural Theory of Mind? Stress Testing Social Reasoning in Large Language Models.

MotiveBench: How Far Are We From Human-Like Motivational Reasoning in Large Language Models? Clever Hans or Neural Theory of Mind? Stress Testing Social Reasoning in Large Language Models

Reference 2006

Resolution
unresolved
no resolver link, observed 2026-08-07T00:43:06.727480Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:43:06.727480Z digest=sha256:1ab038bbc5147638114d359ac9a57c34b43d2e65e1a9638ff41a58d3cd68c901

Observation 55be7344-65e9-466b-a782-a9139d670165 · outbound

This paper cites Neural User Simulation for Corpus-based Policy Optimisation for Spoken Dialogue Systems.

MotiveBench: How Far Are We From Human-Like Motivational Reasoning in Large Language Models? Neural User Simulation for Corpus-based Policy Optimisation for Spoken Dialogue Systems

Reference 2010

Resolution
metadata mismatch
local_arxiv, observed 2026-08-07T00:43:08.701502Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-07T00:43:06.553376Z digest=sha256:a74521433e444488a6f70be7129bb976e5bb2ab33c67094894538b03f92955ab

Observation 958c50c7-34da-4976-b1f5-a9465258e5a3 · outbound

This paper cites To CoT or not to CoT? Chain-of-thought helps mainly on math and symbolic reasoning.

MotiveBench: How Far Are We From Human-Like Motivational Reasoning in Large Language Models? To CoT or not to CoT? Chain-of-thought helps mainly on math and symbolic reasoning

Reference 2016

Resolution
unresolved
no resolver link, observed 2026-08-07T00:43:06.779316Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:43:06.779316Z digest=sha256:6ce2d10454b5e54afec95e54286f05491a55ef7310485a5d4960220f562a57d8

Observation a88f81ec-669c-4dc9-b645-24232aca3cb7 · outbound

This paper cites Towards Social AI: A Survey on Understanding Social Interactions.

MotiveBench: How Far Are We From Human-Like Motivational Reasoning in Large Language Models? Towards Social AI: A Survey on Understanding Social Interactions

Reference 2019

Resolution
unresolved
no resolver link, observed 2026-08-07T00:43:06.603376Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:43:06.603376Z digest=sha256:422ca256ec82c98d667931dba8c97eb7898ad581e6effd9caf2b3fd6c7fefef3

Observation cfb77f19-2672-4a00-926a-64ccf72b072b · outbound

This paper cites InInternational Conference on Machine Learning, pages 337–371.

MotiveBench: How Far Are We From Human-Like Motivational Reasoning in Large Language Models? InInternational Conference on Machine Learning, pages 337–371

Reference 2023

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:43:12.949017Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-07T00:43:06.349846Z digest=sha256:1d46001d5f230645f014699027a340cf55c5bd6a5c6eae3074c6d4d7a947b8bf

Observation 7bfdf243-313a-4357-8b98-54601d3885ee · outbound

This paper cites EmoBench: Evaluating the Emotional Intelligence of Large Language Models.

MotiveBench: How Far Are We From Human-Like Motivational Reasoning in Large Language Models? EmoBench: Evaluating the Emotional Intelligence of Large Language Models

Reference 2024

Resolution
unresolved
no resolver link, observed 2026-08-07T00:43:06.686448Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:43:06.686448Z digest=sha256:4c47942fc79d6d56180d8cac90d76f8afcc07e5886992a35399b5873b3cf56f2

Pith citing papers

Observation 3eb5105e-3f69-4b94-8008-38b16aec4baa · inbound

Can LLMs Think Like Consumers? Benchmarking Crowd-Level Reaction Reconstruction with ConsumerSimBench cites this paper.

Can LLMs Think Like Consumers? Benchmarking Crowd-Level Reaction Reconstruction with ConsumerSimBench MotiveBench: How Far Are We From Human-Like Motivational Reasoning in Large Language Models?

Reference 70

Resolution
verified exact
arxiv_id, observed 2026-05-20T15:33:25.522229Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-05-20T15:31:25.079191Z digest=sha256:4f6c37adb79ade09469462b6569767aac5ed438d4a6d92f08ae84ff834f18754