Pith. sign in

Paper Citation Record · LEDGER

TimeHC-RL: Temporal-aware Hierarchical Cognitive Reinforcement Learning for Enhancing LLMs' Social Intelligence

As of 19 August 2026, this Paper Citation Record lists 38 of 38 outbound references and 0 inbound Pith citation observations for arXiv:2505.24500.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2505.24500 v1

Coverage vector

measured 38 of 38 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T12:27:52.808169Z

measured 38 of 38 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-18T06:34:40.430872+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

38 of 38 outbound references displayed

  • verified exact0
  • verified fuzzy7
  • unresolved31
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 19b10a2d-5766-43e4-baf5-9eeb9d2a81c2 · outbound

This paper cites SFT or RL? An Early Investigation into Training R1-Like Reasoning Large Vision-Language Models.

TimeHC-RL: Temporal-aware Hierarchical Cognitive Reinforcement Learning for Enhancing LLMs' Social Intelligence SFT or RL? An Early Investigation into Training R1-Like Reasoning Large Vision-Language Models

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-07T12:27:48.099504Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:27:48.099504Z digest=sha256:b4c6acbc42ef63ad7b35b50e365c301ffc2dc05dae255e9d03b1863ec7384c91

Observation 308dc582-050e-4727-b741-33d9c09ef9e0 · outbound

This paper cites SFT Memorizes, RL Generalizes: A Comparative Study of Foundation Model Post-training.

TimeHC-RL: Temporal-aware Hierarchical Cognitive Reinforcement Learning for Enhancing LLMs' Social Intelligence SFT Memorizes, RL Generalizes: A Comparative Study of Foundation Model Post-training

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-07T12:27:48.257139Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:27:48.257139Z digest=sha256:6061fc66d54d7f739b155c8c509e62633b6dff710c12c4abf711b930db2315cd

Observation 39d74bd7-4134-4130-aef8-714800e24531 · outbound

This paper cites DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning.

TimeHC-RL: Temporal-aware Hierarchical Cognitive Reinforcement Learning for Enhancing LLMs' Social Intelligence DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-07T12:27:48.555039Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:27:48.555039Z digest=sha256:b24271d5895566aaa98a414a758ec6acc7a87fe9780e5054c66d7aa11d238a5b

Observation 4fb164ae-8e0c-4378-893d-f0e5e985c877 · outbound

This paper cites Open-Reasoner-Zero: An Open Source Approach to Scaling Up Reinforcement Learning on the Base Model.

TimeHC-RL: Temporal-aware Hierarchical Cognitive Reinforcement Learning for Enhancing LLMs' Social Intelligence Open-Reasoner-Zero: An Open Source Approach to Scaling Up Reinforcement Learning on the Base Model

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-07T12:27:48.797927Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:27:48.797927Z digest=sha256:ff8e71a13f4e998c0052b4fe5304095eadf29706c718aad89674a38165f1cbda

Observation 1093a78b-dfa2-4b22-995a-1f9a9b8871c8 · outbound

This paper cites QueryAgent: A Reliable and Efficient Reasoning Framework with Environmental Feedback-based Self-Correction.

TimeHC-RL: Temporal-aware Hierarchical Cognitive Reinforcement Learning for Enhancing LLMs' Social Intelligence QueryAgent: A Reliable and Efficient Reasoning Framework with Environmental Feedback-based Self-Correction

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-07T12:27:48.938782Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:27:48.938782Z digest=sha256:45aeacd8fe86773183566b22e46728c69025da01ac5036e0b971b161afce2308

Observation 6f3181d8-1738-40d3-9e36-d59e4c93e32f · outbound

This paper cites OpenAI o1 System Card.

TimeHC-RL: Temporal-aware Hierarchical Cognitive Reinforcement Learning for Enhancing LLMs' Social Intelligence OpenAI o1 System Card

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-07T12:27:49.075144Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:27:49.075144Z digest=sha256:e73cac9def7f8922e41ce4f46ee307b2425ebbac8e792186b3e45d79457d8298

Observation b7a8e223-9433-4d61-ad96-5bda5d2fa32a · outbound

This paper cites Search-R1: Training LLMs to Reason and Leverage Search Engines with Reinforcement Learning.

TimeHC-RL: Temporal-aware Hierarchical Cognitive Reinforcement Learning for Enhancing LLMs' Social Intelligence Search-R1: Training LLMs to Reason and Leverage Search Engines with Reinforcement Learning

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-07T12:27:49.228737Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:27:49.228737Z digest=sha256:c05a2f69cf7628f3f68c67d158b8586f91c85f2970856d02b3825f7b91351b93

Observation 5d54a5a6-4a1b-4af1-a6cd-1ce9beb6e125 · outbound

This paper cites Perceptions to beliefs: Exploring precursory inferences for theory of mind in large language models.

TimeHC-RL: Temporal-aware Hierarchical Cognitive Reinforcement Learning for Enhancing LLMs' Social Intelligence Perceptions to beliefs: Exploring precursory inferences for theory of mind in large language models

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:27:55.671442Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T12:27:49.386878Z digest=sha256:bfe00a54274cf3b2461318737b172b9a94013347a719f60204ab59cb7105ee37

Observation 86bff7f8-0ddb-470d-8ee2-25c119f40e4e · outbound

This paper cites s1: Simple test-time scaling.

TimeHC-RL: Temporal-aware Hierarchical Cognitive Reinforcement Learning for Enhancing LLMs' Social Intelligence s1: Simple test-time scaling

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-07T12:27:49.679902Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:27:49.679902Z digest=sha256:4df3bf27a336c9a2b87dfa7e0b4c98aee96e5cd93ef497aeb817977d39ce4eed

Observation c307fdb2-94fb-47f7-965b-c555ddcb52af · outbound

This paper cites Explore Theory of Mind: Program-guided adversarial data generation for theory of mind reasoning.

TimeHC-RL: Temporal-aware Hierarchical Cognitive Reinforcement Learning for Enhancing LLMs' Social Intelligence Explore Theory of Mind: Program-guided adversarial data generation for theory of mind reasoning

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-07T12:27:49.910942Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:27:49.910942Z digest=sha256:36eaa112bd24c029bdbcfe8e020c066b9b8beb0d348c92f9b003d027668c8218

Observation c3fdab7c-ad12-4103-b3fe-b67b516ae7e2 · outbound

This paper cites DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models.

TimeHC-RL: Temporal-aware Hierarchical Cognitive Reinforcement Learning for Enhancing LLMs' Social Intelligence DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-07T12:27:50.005114Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:27:50.005114Z digest=sha256:04c35300f22549b6ee408bdf7672b6b0e3789a628e209a8b90e08003486294f4

Observation 7eacdde7-166a-4f7c-9059-def5e592d386 · outbound

This paper cites HybridFlow: A Flexible and Efficient RLHF Framework.

TimeHC-RL: Temporal-aware Hierarchical Cognitive Reinforcement Learning for Enhancing LLMs' Social Intelligence HybridFlow: A Flexible and Efficient RLHF Framework

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-07T12:27:50.102559Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:27:50.102559Z digest=sha256:e3538bb2fdcad1bd644bac1b3efc243265a9d0fc5165925956acbb3467c6a314

Observation e809d746-7890-44c6-bfc8-5b44b0404e16 · outbound

This paper cites ToMATO: Verbalizing the Mental States of Role-Playing LLMs for Benchmarking Theory of Mind.

TimeHC-RL: Temporal-aware Hierarchical Cognitive Reinforcement Learning for Enhancing LLMs' Social Intelligence ToMATO: Verbalizing the Mental States of Role-Playing LLMs for Benchmarking Theory of Mind

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-07T12:27:50.270022Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:27:50.270022Z digest=sha256:f58df0a36a079145fa275329efb20425fd076675aa6e35257619b849d54d9501

Observation e5df7549-894f-46fc-8577-1c81973ef424 · outbound

This paper cites R1-Searcher: Incentivizing the Search Capability in LLMs via Reinforcement Learning.

TimeHC-RL: Temporal-aware Hierarchical Cognitive Reinforcement Learning for Enhancing LLMs' Social Intelligence R1-Searcher: Incentivizing the Search Capability in LLMs via Reinforcement Learning

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-07T12:27:50.567982Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:27:50.567982Z digest=sha256:c1ecc76cf13add65d2b2f55ab3c2cf16056c501bc1c1578b09c38e4284a2e379

Observation 9edca75c-ed6c-4a03-ac76-063551775524 · outbound

This paper cites Climbing the ladder of reasoning: What llms can-and still can’t-solve after sft? arXiv preprint arXiv:2504.11741,.

TimeHC-RL: Temporal-aware Hierarchical Cognitive Reinforcement Learning for Enhancing LLMs' Social Intelligence Climbing the ladder of reasoning: What llms can-and still can’t-solve after sft? arXiv preprint arXiv:2504.11741,

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-07T12:27:50.713881Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:27:50.713881Z digest=sha256:b0ff0d94faccc669a1ae4f4fd5206b7ace1a423090ece7dace35a6d63bb6f3dd

Observation e57198dc-bc3b-4e25-9d4e-e0450c9c8766 · outbound

This paper cites HelpSteer2: Open-source dataset for training top-performing reward models.

TimeHC-RL: Temporal-aware Hierarchical Cognitive Reinforcement Learning for Enhancing LLMs' Social Intelligence HelpSteer2: Open-source dataset for training top-performing reward models

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-07T12:27:50.908285Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:27:50.908285Z digest=sha256:8858aae2119718844aa6250ca365e8796ab043d8366dc888958299e0b37cdb1d

Observation 1b5f4533-fe3b-4e24-8961-ccccea2f8de8 · outbound

This paper cites SWE-RL: Advancing LLM Reasoning via Reinforcement Learning on Open Software Evolution.

TimeHC-RL: Temporal-aware Hierarchical Cognitive Reinforcement Learning for Enhancing LLMs' Social Intelligence SWE-RL: Advancing LLM Reasoning via Reinforcement Learning on Open Software Evolution

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-07T12:27:51.038674Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:27:51.038674Z digest=sha256:4c8caba192e4396f19f69a8fd4c324d7376a4ae3b70071af3dffda6a9b51bed4

Observation 7474caf0-d8fe-4815-8c1d-71db8aaaefec · outbound

This paper cites Hi-tom: A benchmark for evaluating higher-order theory of mind reasoning in large language models.

TimeHC-RL: Temporal-aware Hierarchical Cognitive Reinforcement Learning for Enhancing LLMs' Social Intelligence Hi-tom: A benchmark for evaluating higher-order theory of mind reasoning in large language models

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:27:55.090600Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T12:27:51.153523Z digest=sha256:5442c7aa0189d877d81d243bc12a3f505730cf1150c93b36cfb2024df6467b58

Observation 2c95a59f-4855-449a-a573-10761728cc9d · outbound

This paper cites GUI-R1 : A Generalist R1-Style Vision-Language Action Model For GUI Agents.

TimeHC-RL: Temporal-aware Hierarchical Cognitive Reinforcement Learning for Enhancing LLMs' Social Intelligence GUI-R1 : A Generalist R1-Style Vision-Language Action Model For GUI Agents

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-07T12:27:51.290313Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:27:51.290313Z digest=sha256:808a6e1f30b405e1ee1b500209ae60e2b181dd9beab0ebc41bdbea603b8f9788

Observation 1634f636-6489-4528-98a5-5373a857391d · outbound

This paper cites Logic-RL: Unleashing LLM Reasoning with Rule-Based Reinforcement Learning.

TimeHC-RL: Temporal-aware Hierarchical Cognitive Reinforcement Learning for Enhancing LLMs' Social Intelligence Logic-RL: Unleashing LLM Reasoning with Rule-Based Reinforcement Learning

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-07T12:27:51.421057Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:27:51.421057Z digest=sha256:da599d77585aa0c68d9f818656605211f97f5fa28a04f9db9baf9169e0f82a31

Observation 28372f1e-66ef-40a4-a453-95af9b8631b5 · outbound

This paper cites Qwen2.5 Technical Report.

TimeHC-RL: Temporal-aware Hierarchical Cognitive Reinforcement Learning for Enhancing LLMs' Social Intelligence Qwen2.5 Technical Report

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-07T12:27:51.563161Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:27:51.563161Z digest=sha256:f1a43dceb391188728ed744af1976653216605cc42051e0d28d9d895abd3cbcb

Observation e4451955-6d9b-420b-85b6-d8f8577955af · outbound

This paper cites LIMO: Less is More for Reasoning.

TimeHC-RL: Temporal-aware Hierarchical Cognitive Reinforcement Learning for Enhancing LLMs' Social Intelligence LIMO: Less is More for Reasoning

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-07T12:27:51.686739Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:27:51.686739Z digest=sha256:321cf2a941bcd5f341662ccae2f34b6f6f6f43cae634d5dbb1640a740b4a3bdb

Observation 1f316624-31ca-4996-a14e-5958a5b4f105 · outbound

This paper cites Z1: Efficient Test-time Scaling with Code.

TimeHC-RL: Temporal-aware Hierarchical Cognitive Reinforcement Learning for Enhancing LLMs' Social Intelligence Z1: Efficient Test-time Scaling with Code

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-07T12:27:51.841949Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:27:51.841949Z digest=sha256:4703d19ad1789411d709f7337a3a8a14b7182649790b5c822de30a6e4f19f498

Observation 585c6ccc-11bf-434b-b600-bd8655e19a10 · outbound

This paper cites Scaling of Search and Learning: A Roadmap to Reproduce o1 from Reinforcement Learning Perspective.

TimeHC-RL: Temporal-aware Hierarchical Cognitive Reinforcement Learning for Enhancing LLMs' Social Intelligence Scaling of Search and Learning: A Roadmap to Reproduce o1 from Reinforcement Learning Perspective

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-07T12:27:51.974700Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:27:51.974700Z digest=sha256:64804b1c5f44ab1b58cd89f903fa347e406d36e05d68ad13a0b75fb867438158

Observation f36ab083-ece6-43d4-b9c9-3f3f29937d35 · outbound

This paper cites Autotom: Automated bayesian inverse planning and model discovery for open-ended theory of mind.

TimeHC-RL: Temporal-aware Hierarchical Cognitive Reinforcement Learning for Enhancing LLMs' Social Intelligence Autotom: Automated bayesian inverse planning and model discovery for open-ended theory of mind

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-07T12:27:52.105777Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:27:52.105777Z digest=sha256:5074ef680154584524ac05a04f05dbe8d412d61b8fd07be01b9dba02b0837618

Observation fd96bd9b-3e57-4aa4-80ac-fe5d99d7b201 · outbound

This paper cites SOTOPIA: Interactive Evaluation for Social Intelligence in Language Agents.

TimeHC-RL: Temporal-aware Hierarchical Cognitive Reinforcement Learning for Enhancing LLMs' Social Intelligence SOTOPIA: Interactive Evaluation for Social Intelligence in Language Agents

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-07T12:27:52.254318Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:27:52.254318Z digest=sha256:7d09c585a68a55c5f96e3c680c9b827b661ebd632bc5063dfd80fa352a027b4b

Observation 4f807b02-017d-4bed-a1be-c32a4eab9262 · outbound

This paper cites SWEET-RL: Training Multi-Turn LLM Agents on Collaborative Reasoning Tasks.

TimeHC-RL: Temporal-aware Hierarchical Cognitive Reinforcement Learning for Enhancing LLMs' Social Intelligence SWEET-RL: Training Multi-Turn LLM Agents on Collaborative Reasoning Tasks

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-07T12:27:52.419144Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:27:52.419144Z digest=sha256:94e1d0bc9abe98e0afce5648baac29d5a2afa878a63e0a7b790ba694ce03c509

Observation fb1d23b2-1d90-41b0-b60f-9e45ef984190 · outbound

This paper cites The comparison results are shown in Table.

TimeHC-RL: Temporal-aware Hierarchical Cognitive Reinforcement Learning for Enhancing LLMs' Social Intelligence The comparison results are shown in Table

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:27:54.888072Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T12:27:52.551126Z digest=sha256:8b401c7da681710717fdf28ef8449b2b591563c611b9bb2b2bbe23e25b3d48d1

Observation 98b817c3-08ca-43c5-9bba-9f4ad29f3d62 · outbound

This paper cites What LLMs Can—and Still Can’t—Solve after SFT?.

TimeHC-RL: Temporal-aware Hierarchical Cognitive Reinforcement Learning for Enhancing LLMs' Social Intelligence What LLMs Can—and Still Can’t—Solve after SFT?

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:27:54.535643Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T12:27:52.696376Z digest=sha256:584381b0a9fd16dd41729dd6c13b8e9fa53f20b005aa944e0ec81846759c8488

Observation dd47479f-7f0e-451e-a765-43ba4198eef4 · outbound

This paper cites budget forcing.

TimeHC-RL: Temporal-aware Hierarchical Cognitive Reinforcement Learning for Enhancing LLMs' Social Intelligence budget forcing

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:27:54.170166Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T12:27:52.808169Z digest=sha256:e0d9760f27c628072328df9ff48c48f3304393965c2b7a6f05b291a4a302fde6

Observation 43b977b2-7bc4-4a82-8d8b-6f534d1c1dde · outbound

This paper cites Scaling LLM Test-Time Compute Optimally can be More Effective than Scaling Model Parameters.

TimeHC-RL: Temporal-aware Hierarchical Cognitive Reinforcement Learning for Enhancing LLMs' Social Intelligence Scaling LLM Test-Time Compute Optimally can be More Effective than Scaling Model Parameters

Reference 1996

Resolution
unresolved
no resolver link, observed 2026-08-07T12:27:50.446081Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:27:50.446081Z digest=sha256:ed94ea0846fda373f78493e05ff0d044381eff66b6576837f120a8fdf206f616

Observation d8dc670e-9460-401a-9bd4-a8a430bd1ccf · outbound

This paper cites Revisiting the evaluation of theory of mind through question answering.

TimeHC-RL: Temporal-aware Hierarchical Cognitive Reinforcement Learning for Enhancing LLMs' Social Intelligence Revisiting the evaluation of theory of mind through question answering

Reference 2011

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:27:55.491641Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T12:27:49.528985Z digest=sha256:6f76b03a12419fc958903ce9c0564003224d17028dc91498105bdff868c9d5f3

Observation c181eec8-6d4f-42e4-a4f9-56e8da5edba2 · outbound

This paper cites Video-R1: Reinforcing Video Reasoning in MLLMs.

TimeHC-RL: Temporal-aware Hierarchical Cognitive Reinforcement Learning for Enhancing LLMs' Social Intelligence Video-R1: Reinforcing Video Reasoning in MLLMs

Reference 2013

Resolution
unresolved
no resolver link, observed 2026-08-07T12:27:48.392575Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:27:48.392575Z digest=sha256:01cf875aaed60dc0b7519f2d98d0c5d53302f9e3c451ac41a2124f45adab89bf

Observation 257d3a63-41ab-44bc-afab-438bd05e1d94 · outbound

This paper cites Relation-r1: Cognitive chain-of-thought guided reinforcement learning for unified relational comprehension.

TimeHC-RL: Temporal-aware Hierarchical Cognitive Reinforcement Learning for Enhancing LLMs' Social Intelligence Relation-r1: Cognitive chain-of-thought guided reinforcement learning for unified relational comprehension

Reference 2019

Resolution
unresolved
no resolver link, observed 2026-08-07T12:27:49.594703Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:27:49.594703Z digest=sha256:4650317aa5491ebb7d189c6f3effe7daac9ff7534eb9a47dc5ec7860672df137

Observation 0f1ef408-6abe-4356-bada-9bfd8f90590e · outbound

This paper cites Social iqa: Common- sense reasoning about social interactions.

TimeHC-RL: Temporal-aware Hierarchical Cognitive Reinforcement Learning for Enhancing LLMs' Social Intelligence Social iqa: Common- sense reasoning about social interactions

Reference 2020

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:27:55.299794Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T12:27:49.824323Z digest=sha256:6e381cd0a2652ed59a47aeefa1ef3cba81e11f621b009e8e653654be46de34b8

Observation ae8e3374-ad35-4ae9-bc01-3de7c29f3cb3 · outbound

This paper cites Mutual Reasoning Makes Smaller LLMs Stronger Problem-Solvers.

TimeHC-RL: Temporal-aware Hierarchical Cognitive Reinforcement Learning for Enhancing LLMs' Social Intelligence Mutual Reasoning Makes Smaller LLMs Stronger Problem-Solvers

Reference 2023

Resolution
unresolved
no resolver link, observed 2026-08-07T12:27:49.740854Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:27:49.740854Z digest=sha256:0c6c0a7bfe2eb8e0d32190e1cb2ef1cc72afc45edad579005c2797ad77de9177

Observation f67b7fdd-adbd-4450-a8c8-0c4c883b535a · outbound

This paper cites Large Language Monkeys: Scaling Inference Compute with Repeated Sampling.

TimeHC-RL: Temporal-aware Hierarchical Cognitive Reinforcement Learning for Enhancing LLMs' Social Intelligence Large Language Monkeys: Scaling Inference Compute with Repeated Sampling

Reference 2024

Resolution
unresolved
no resolver link, observed 2026-08-07T12:27:48.037680Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:27:48.037680Z digest=sha256:81d223242b4d1014ecae1c10cb84df8dbf56170030e3e09ff83246948d8d1b85

Observation 6e6e6052-e199-4952-ad4f-924d4dfd415b · outbound

This paper cites GenCLS++: Pushing the Boundaries of Generative Classification in LLMs Through Comprehensive SFT and RL Studies Across Diverse Datasets.

TimeHC-RL: Temporal-aware Hierarchical Cognitive Reinforcement Learning for Enhancing LLMs' Social Intelligence GenCLS++: Pushing the Boundaries of Generative Classification in LLMs Through Comprehensive SFT and RL Studies Across Diverse Datasets

Reference 2025

Resolution
unresolved
no resolver link, observed 2026-08-07T12:27:48.666190Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:27:48.666190Z digest=sha256:c841eb5d7f28cd2308b7fb82f0f005a6aacf24d71f2b01ecd4e016a958192729

Pith citing papers

No inbound Pith citation observations are available.