Pith. sign in

Paper Citation Record · LEDGER

Hypothesis generation and updating in large language models

As of 18 August 2026, this Paper Citation Record lists 100 of 300 outbound references and 1 inbound Pith citation observation for arXiv:2605.05851.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2605.05851 v1

Coverage vector

measured 100 of 300 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-05-08T14:43:56.125546Z

measured 101 of 101 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-18T06:34:40.430872+00:00

measured 1 of 1 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-07-30T19:24:34.330570Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

100 of 300 outbound references displayed

  • verified exact63
  • verified fuzzy19
  • unresolved7
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch11

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 9030ad06-daeb-4e7e-a634-20715395635e · outbound

This paper cites NVIDIA Nemotron 3: Efficient and Open Intelligence.

Hypothesis generation and updating in large language models NVIDIA Nemotron 3: Efficient and Open Intelligence

Reference 1

Resolution
verified exact
arxiv_id, observed 2026-05-18T01:40:43.189893Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-05-08T14:43:56.125546Z digest=sha256:cab2747812a84f048dc5c60afec00de0d7c8b70194cda3ece523b91a7ed88b99

Observation 9b5a28c2-1800-4576-853a-0f2f120064de · outbound

This paper cites Google , author =.

Hypothesis generation and updating in large language models Google , author =

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-05-26T11:37:37.865128Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-05-08T14:43:56.125546Z digest=sha256:9fa4392fbf13518665ff28c24c1b96c322257ac21e87d7b92b44c38d55a45ab7

Observation ec88b1fe-2ef9-4b4e-9878-b90fe54d4fde · outbound

This paper cites From ai for science to agentic science: A survey on autonomous scientific discovery.

Hypothesis generation and updating in large language models From ai for science to agentic science: A survey on autonomous scientific discovery

Reference 3

Resolution
verified exact
arxiv_id, observed 2026-05-08T21:14:12.274884Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-05-08T14:43:56.125546Z digest=sha256:0ddd6fd3c979cd4c3d1e3011080ae835d9c3b666f2030e725a94c7ccac5d1147

Observation 790abe83-e7c2-4b7d-bffd-541e91b301ff · outbound

This paper cites an unresolved cited work.

Hypothesis generation and updating in large language models Unresolved cited work

Reference 4

Resolution
unresolved
raw_fallback, observed 2026-05-26T11:37:37.848455Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-05-08T14:43:56.125546Z digest=sha256:83d057692c7e44a84357798d377b05cb7dea70d22a48ab38ce494f6c0e088520

Observation d1fec973-a031-4a46-b0ed-5e094d57f379 · outbound

This paper cites Kosmos: An AI Scientist for Autonomous Discovery.

Hypothesis generation and updating in large language models Kosmos: An AI Scientist for Autonomous Discovery

Reference 5

Resolution
metadata mismatch
arxiv_id, observed 2026-05-22T08:40:55.079629Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-05-08T14:43:56.125546Z digest=sha256:44422fa4ff359e29265a864fdb39e55dc28026cd642d0d593558d791b5733fc9

Observation a2edbde2-1db6-4964-bed1-03454921ef3d · outbound

This paper cites AI-Researcher: Autonomous Scientific Innovation.

Hypothesis generation and updating in large language models AI-Researcher: Autonomous Scientific Innovation

Reference 6

Resolution
metadata mismatch
arxiv_id, observed 2026-05-08T21:14:12.390138Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-05-08T14:43:56.125546Z digest=sha256:5b512520b1d4132677a366674b7b781262574b067672e8632f665a0512a00104

Observation 3b9fb214-94fc-42ce-8882-7553b7a3ab1f · outbound

This paper cites The AI Scientist-v2: Workshop-Level Automated Scientific Discovery via Agentic Tree Search.

Hypothesis generation and updating in large language models The AI Scientist-v2: Workshop-Level Automated Scientific Discovery via Agentic Tree Search

Reference 7

Resolution
verified exact
arxiv_id, observed 2026-05-09T01:44:55.508555Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-05-08T14:43:56.125546Z digest=sha256:45152a96dff9f605c3b889e2e698b24cee64444b3ba65304563d1e99907d3873

Observation 4702acf5-838b-4c25-8899-c735d731a04e · outbound

This paper cites Towards an AI co-scientist.

Hypothesis generation and updating in large language models Towards an AI co-scientist

Reference 8

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T13:02:46.484476Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-05-08T14:43:56.125546Z digest=sha256:cfb6da7618629c7a3e24e210288cc464a2f4cbe1b03bc2959d1a5bad054e8084

Observation 565ab26d-3941-41b3-ab0c-372112869588 · outbound

This paper cites year =.

Hypothesis generation and updating in large language models year =

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-05-26T11:37:37.839265Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-05-08T14:43:56.125546Z digest=sha256:eb6376c7d064c30666f22a62855994ce53c78254113bba738bca889e34b51a42

Observation e838d177-a85e-4a0c-95b3-bb6cf9e5c228 · outbound

This paper cites Cognitive basis of language learning in infants , volume =.

Hypothesis generation and updating in large language models Cognitive basis of language learning in infants , volume =

Reference 10

Resolution
verified exact
doi, observed 2026-05-08T21:14:12.132977Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-05-08T14:43:56.125546Z digest=sha256:33a673418a9294bc2454e7e0c5e92af4a69cbedcc573373aca220f27fe5fb06c

Observation d7e2bfd7-41ca-4699-a401-61d6c0006fe3 · outbound

This paper cites Journal of Child Language , author =.

Hypothesis generation and updating in large language models Journal of Child Language , author =

Reference 11

Resolution
verified exact
doi, observed 2026-05-08T21:14:12.097757Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-05-08T14:43:56.125546Z digest=sha256:c35ca24aa164f4f792dd26e4472cfe244acdebde67c9ad9e1a34c97b004f2f44

Observation 920bce99-57ac-4a11-a7e7-52204fb82e95 · outbound

This paper cites an unresolved cited work.

Hypothesis generation and updating in large language models Unresolved cited work

Reference 12

Resolution
unresolved
raw_fallback, observed 2026-05-26T11:37:37.877646Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-05-08T14:43:56.125546Z digest=sha256:8e4674d1281546696073d6f975ecd2554358fa8bd89676c30c0f09e8fc50afe8

Observation ec7c6509-430e-4217-b21b-499937208c35 · outbound

This paper cites How LLMs Detect and Correct Their Own Errors: The Role of Internal Confidence Signals.

Hypothesis generation and updating in large language models How LLMs Detect and Correct Their Own Errors: The Role of Internal Confidence Signals

Reference 13

Resolution
verified exact
local_arxiv, observed 2026-05-08T21:14:12.052674Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-05-08T14:43:56.125546Z digest=sha256:7f4a60b97294e399dd6d1d3ce7543066f1899d3dbc04d37370a7a627d3d77636

Observation 309979ad-b93f-4dc1-b552-0bf3f3cb9cbb · outbound

This paper cites Reasoning Theater: Disentangling Model Beliefs from Chain-of-Thought.

Hypothesis generation and updating in large language models Reasoning Theater: Disentangling Model Beliefs from Chain-of-Thought

Reference 14

Resolution
verified exact
arxiv_id, observed 2026-05-29T03:04:37.872643Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-05-08T14:43:56.125546Z digest=sha256:066045ed03ea33fc61158f01835ce55b6f953313a11bbfab32b41e7e03073fd5

Observation d101b226-0181-410a-b54c-0d954c171684 · outbound

This paper cites Real-Time Progress Prediction in Reasoning Language Models.

Hypothesis generation and updating in large language models Real-Time Progress Prediction in Reasoning Language Models

Reference 15

Resolution
verified exact
arxiv_id, observed 2026-05-27T02:05:03.856858Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-05-08T14:43:56.125546Z digest=sha256:ab20bbe502cf162db8c454b3dcaf41b85c83c6a0a702d8784f23cc0f8744d572

Observation e23c75d5-2a11-46e4-85fa-f7bd3e166791 · outbound

This paper cites Calibrating Reasoning in Language Models with Internal Consistency.

Hypothesis generation and updating in large language models Calibrating Reasoning in Language Models with Internal Consistency

Reference 16

Resolution
verified exact
arxiv_id, observed 2026-05-08T21:14:12.111952Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-05-08T14:43:56.125546Z digest=sha256:fc11e8514014e784b1054b1bab2c1bf4c7854ebb0ec468b39f47a302c5cd66d7

Observation 3b89ed0e-7218-45c9-9346-85425f3c3574 · outbound

This paper cites an unresolved cited work.

Hypothesis generation and updating in large language models Unresolved cited work

Reference 17

Resolution
verified exact
doi, observed 2026-05-08T21:14:12.014581Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-05-08T14:43:56.125546Z digest=sha256:47e4eaac22907b3404c89fa5a52a0baf96fb999afaa5b1f9e8a139e271f4b8a1

Observation 796152bf-b314-4302-94e4-7aadf23dfc05 · outbound

This paper cites Reasoning with Sampling: Your Base Model is Smarter Than You Think.

Hypothesis generation and updating in large language models Reasoning with Sampling: Your Base Model is Smarter Than You Think

Reference 18

Resolution
verified exact
arxiv_id, observed 2026-05-17T03:16:05.617665Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-05-08T14:43:56.125546Z digest=sha256:e229ece0e908fd0d76e8415974d5dd5a1c5e4c0da015ece91a0dcb3d10ae02b1

Observation 7296389a-8f57-4c26-b371-89e5c905f99e · outbound

This paper cites Chen, J., Hu, S., Liu, Z., and Sun, M.

Hypothesis generation and updating in large language models Chen, J., Hu, S., Liu, Z., and Sun, M

Reference 19

Resolution
verified exact
arxiv_id, observed 2026-05-08T21:14:12.591282Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-05-08T14:43:56.125546Z digest=sha256:f9a9c01293120a570a705d578f9b3a1f7ad2229798f7162085832fed62e7e6c9

Observation aedd9192-98de-49cb-b0cc-8a072bbdbf32 · outbound

This paper cites Temporal.

Hypothesis generation and updating in large language models Temporal

Reference 20

Resolution
verified exact
arxiv_id, observed 2026-05-08T21:14:12.484443Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-05-08T14:43:56.125546Z digest=sha256:668d03f194907500937a148042dd68f7660f9b67cfcbffe85d7a1b663b4f4ce8

Observation f06e2c7e-5456-48d2-83a6-a3975ed6a966 · outbound

This paper cites LLM Reasoning as Trajectories: Step-Specific Representation Geometry and Correctness Signals.

Hypothesis generation and updating in large language models LLM Reasoning as Trajectories: Step-Specific Representation Geometry and Correctness Signals

Reference 21

Resolution
metadata mismatch
local_arxiv, observed 2026-05-08T21:14:12.466420Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-05-08T14:43:56.125546Z digest=sha256:3e75c91b7d68d10b9271c482f4f98ccf822b3c8a6af0349e7ccdd4342b726d00

Observation 738f7e20-5654-48cc-8d9e-dce89599c7aa · outbound

This paper cites Efficient PRM Training Data Synthesis via Formal Verification.

Hypothesis generation and updating in large language models Efficient PRM Training Data Synthesis via Formal Verification

Reference 22

Resolution
verified exact
local_arxiv, observed 2026-05-08T21:14:12.601324Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-05-08T14:43:56.125546Z digest=sha256:e8fd252cdd20335f4fa4b0f8adf63999eb4b3b5875392a5bae9ffba98056b3cd

Observation ce6dcc1d-ed4c-4d33-a7ef-ec9fa6b64cdf · outbound

This paper cites Truth as a.

Hypothesis generation and updating in large language models Truth as a

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-05-26T11:37:37.858080Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-05-08T14:43:56.125546Z digest=sha256:edb3d86faeedb7735870ff4edc46858e80f2efed985b322c5d60c3b6d7daf5a6

Observation ae5f2d7e-07d3-4ef5-bae1-48a8449934dd · outbound

This paper cites Closing the confidence-faithfulness gap in large language models.

Hypothesis generation and updating in large language models Closing the confidence-faithfulness gap in large language models

Reference 24

Resolution
verified exact
arxiv_id, observed 2026-05-08T21:14:12.002783Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-05-08T14:43:56.125546Z digest=sha256:65090c8b2d7b0cf02a65f0c24b2e00ae288a2a3cf24c93ca6dcb1b87034cc068

Observation cab58a57-38f9-4728-8732-bf73168dec87 · outbound

This paper cites How do LLMs Compute Verbal Confidence.

Hypothesis generation and updating in large language models How do LLMs Compute Verbal Confidence

Reference 25

Resolution
verified exact
arxiv_id, observed 2026-05-20T02:04:53.698588Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-05-08T14:43:56.125546Z digest=sha256:709cabf2030abf4af72ccba84641da0db72ebdb6b84b99b7e06164764ea37232

Observation 934401ba-8901-4f67-91ce-b32cc00281ac · outbound

This paper cites ReProbe: Efficient Test-Time Scaling of Multi-Step Reasoning by Probing Internal States of Large Language Models.

Hypothesis generation and updating in large language models ReProbe: Efficient Test-Time Scaling of Multi-Step Reasoning by Probing Internal States of Large Language Models

Reference 26

Resolution
verified exact
local_arxiv, observed 2026-05-08T21:14:12.439467Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-05-08T14:43:56.125546Z digest=sha256:4af5a48dd82ee6b35bfdab27d9c058850b4c855caf746560223b5f587e5126f5

Observation d7a8dd25-1125-4982-8dbc-6ea001b6f201 · outbound

This paper cites Do LLMs Know What They Know? Measuring Metacognitive Efficiency with Signal Detection Theory.

Hypothesis generation and updating in large language models Do LLMs Know What They Know? Measuring Metacognitive Efficiency with Signal Detection Theory

Reference 27

Resolution
verified exact
arxiv_id, observed 2026-07-15T01:20:58.272624Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-05-08T14:43:56.125546Z digest=sha256:0c3950f5565d5902cb236b6db22d820aa244f297d85aa8c8f7f588ae2559478a

Observation cec467fe-4509-43b4-956d-92da3564b497 · outbound

This paper cites Evidence for.

Hypothesis generation and updating in large language models Evidence for

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-05-26T11:37:37.864758Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-05-08T14:43:56.125546Z digest=sha256:963cee1905d5a3b49cbdabad0a47e67c05fe1ae74ca55ad668a8342b2c1dfc3e

Observation 11b3bda7-934c-40a8-944d-f32ba7120fa2 · outbound

This paper cites an unresolved cited work.

Hypothesis generation and updating in large language models Unresolved cited work

Reference 29

Resolution
unresolved
raw_fallback, observed 2026-05-26T11:37:37.888437Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-05-08T14:43:56.125546Z digest=sha256:ed8226efc3621f4caf971fe424ae58b5f41b039de167e8a965ef8194df8d6fdd

Observation 74dd7402-4515-4cd8-903a-80bffca283c7 · outbound

This paper cites Learning to.

Hypothesis generation and updating in large language models Learning to

Reference 30

Resolution
verified exact
arxiv_id, observed 2026-05-08T21:14:12.586341Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-05-08T14:43:56.125546Z digest=sha256:79213a291bc0d97fc58f474945d7303babf7d8df48819fc4e8c5cc499c6bd4c9

Observation 40059f7c-f8a4-46ab-827f-183b3fbbac62 · outbound

This paper cites More Capable, Less Cooperative? When LLMs Fail At Zero-Cost Collaboration.

Hypothesis generation and updating in large language models More Capable, Less Cooperative? When LLMs Fail At Zero-Cost Collaboration

Reference 31

Resolution
verified exact
local_arxiv, observed 2026-05-08T21:14:12.343277Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-05-08T14:43:56.125546Z digest=sha256:564454a9bd3bcd96a7375de61fd19d3ff2e02f2b1459145ec876182ef06c6be8

Observation 943be80b-38f7-4956-8235-374e2932df85 · outbound

This paper cites Lossless data compression by large models , volume =.

Hypothesis generation and updating in large language models Lossless data compression by large models , volume =

Reference 32

Resolution
verified exact
doi, observed 2026-05-08T21:14:12.334953Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-05-08T14:43:56.125546Z digest=sha256:9e731a4ebfc32e18b28fa87967e69633faf3c3fae4944aa7ee8597224f648dbb

Observation 6707a743-e3a0-430a-a3bd-1873db53943b · outbound

This paper cites an unresolved cited work.

Hypothesis generation and updating in large language models Unresolved cited work

Reference 33

Resolution
unresolved
raw_fallback, observed 2026-05-26T11:37:37.851930Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-05-08T14:43:56.125546Z digest=sha256:72294299f0b46e5a9271c3284365fa5aec82a10171511138e6454ae97ab98c4d

Observation ebee73ff-dd3a-49be-b827-49aadcdf00df · outbound

This paper cites The No Free Lunch Theorem, Kolmogorov Complexity, and the Role of Inductive Biases in Machine Learning.

Hypothesis generation and updating in large language models The No Free Lunch Theorem, Kolmogorov Complexity, and the Role of Inductive Biases in Machine Learning

Reference 34

Resolution
verified exact
arxiv_id, observed 2026-05-08T21:14:12.395122Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-05-08T14:43:56.125546Z digest=sha256:cb81bf0d288fe3caaf8a4fa015ea48972873bab40d5bb2f1a932625ea3c1770d

Observation 167d62cb-ba3c-493e-9cae-c352b6d7ea6e · outbound

This paper cites Language.

Hypothesis generation and updating in large language models Language

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-05-26T11:37:37.842558Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-05-08T14:43:56.125546Z digest=sha256:3a3e93803db98fd95c8a4b289d403ddaa9f78e087c3405821dd5f97b037bc6b7

Observation c04cc3f6-0291-4638-b927-ec9287d678c1 · outbound

This paper cites On predictability of reinforce- ment learning dynamics for large language models.CoRR, abs/2510.00553.

Hypothesis generation and updating in large language models On predictability of reinforce- ment learning dynamics for large language models.CoRR, abs/2510.00553

Reference 36

Resolution
verified exact
arxiv_id, observed 2026-05-08T21:14:12.318454Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-05-08T14:43:56.125546Z digest=sha256:925a257c792e481ee5d3dddae71214ba081779575174734fa77aa7ba17a29bd4

Observation 24b3e082-9f2c-4b90-a23d-f7b730fd2790 · outbound

This paper cites Advances in Neural Information Processing Systems , author =.

Hypothesis generation and updating in large language models Advances in Neural Information Processing Systems , author =

Reference 37

Resolution
verified exact
doi, observed 2026-05-08T21:14:11.849183Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-05-08T14:43:56.125546Z digest=sha256:bf1886b67e27bfcafb4e9718be83c1d211a1afbe7682f39e9102fcddb1689eb3

Observation 666ed916-1592-458c-ae35-78a183f763ec · outbound

This paper cites Transactions on Machine Learning Research , author =.

Hypothesis generation and updating in large language models Transactions on Machine Learning Research , author =

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-05-26T11:37:37.869270Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-05-08T14:43:56.125546Z digest=sha256:7e8106247eb59d3c8716a020babd86958f86d5b0d135c4be414c8f1a20ea571c

Observation c8c91847-9a8f-442d-8c80-6c4a83d67baa · outbound

This paper cites BIG-Bench Extra Hard.

Hypothesis generation and updating in large language models BIG-Bench Extra Hard

Reference 39

Resolution
verified exact
arxiv_id, observed 2026-05-08T21:14:12.648060Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-05-08T14:43:56.125546Z digest=sha256:d486d1fcf12596fc3949956e4cab9d20ec152872baed5e12342f67c8c6a02677

Observation df13b7df-a3a1-4a93-a922-01e6b55f9af7 · outbound

This paper cites The Lookahead Limitation: Why Multi-Operand Addition is Hard for LLMs.

Hypothesis generation and updating in large language models The Lookahead Limitation: Why Multi-Operand Addition is Hard for LLMs

Reference 40

Resolution
verified exact
arxiv_id, observed 2026-05-08T21:14:12.026036Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-05-08T14:43:56.125546Z digest=sha256:1b3927cc10fc1f1149f6640b73e87043fc93730fc6e4f0b83efa4921c467eb6d

Observation 4cedac66-791a-4cf0-aee5-09055db8150a · outbound

This paper cites CRUXEval: A Benchmark for Code Reasoning, Understanding and Execution.

Hypothesis generation and updating in large language models CRUXEval: A Benchmark for Code Reasoning, Understanding and Execution

Reference 41

Resolution
metadata mismatch
arxiv_id, observed 2026-05-14T20:57:16.443569Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-05-08T14:43:56.125546Z digest=sha256:0ba8f99d070c8d1cd137bc407ba79c931bfd0cb8aac06be6f6431059bec1cccd

Observation 46cc2235-3f58-473f-8c6d-596d72a4c114 · outbound

This paper cites an unresolved cited work.

Hypothesis generation and updating in large language models Unresolved cited work

Reference 42

Resolution
unresolved
raw_fallback, observed 2026-05-26T11:37:37.893579Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-05-08T14:43:56.125546Z digest=sha256:bff4c7055b4894eda9e9e9d15d6119acfcb4f2ca9db906eab3eb4cb335a5592e

Observation 307ff8af-470a-4658-b215-82d2cac54d82 · outbound

This paper cites and Moros-Daval, Yael and Zhang, Seraphina and Zhao, Qinlin and Huang, Yitian and Sun, Luning and Prunty, Jonathan E.

Hypothesis generation and updating in large language models and Moros-Daval, Yael and Zhang, Seraphina and Zhao, Qinlin and Huang, Yitian and Sun, Luning and Prunty, Jonathan E

Reference 43

Resolution
verified exact
doi, observed 2026-05-08T21:14:12.365090Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-05-08T14:43:56.125546Z digest=sha256:f7ee86c68b3d80868e4fa2c0bc259e208829351dcf8f70848b8bb652dbbbb3c8

Observation 3c8dc81b-6c3d-4e59-9333-7661731bf3b6 · outbound

This paper cites What and.

Hypothesis generation and updating in large language models What and

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-05-26T11:37:37.861475Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-05-08T14:43:56.125546Z digest=sha256:e9e970a3db61bdf6c1b539f90c13987546c72fac43b395259195a32cb391ccd9

Observation 10d9403e-fb7d-4245-9450-56cbdf70834f · outbound

This paper cites Revisiting the.

Hypothesis generation and updating in large language models Revisiting the

Reference 45

Resolution
verified exact
arxiv_id, observed 2026-05-08T21:14:11.905598Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-05-08T14:43:56.125546Z digest=sha256:b9320808fcf24236fffcb8a57f2c6f040dbbca68a17547c70e5ee933e8fa7065

Observation 89bbc8d3-5a58-48bd-9d2e-b9b3fd00e777 · outbound

This paper cites and Piantadosi, Steven T.

Hypothesis generation and updating in large language models and Piantadosi, Steven T

Reference 46

Resolution
verified fuzzy
raw_fallback, observed 2026-05-26T11:37:37.855949Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-05-08T14:43:56.125546Z digest=sha256:4eeac83e810ca5fd069bee7b4aab1402324369e5c0c1dc354acbcd3e82644925

Observation 62b22744-6742-4ac3-9272-812ec128ea37 · outbound

This paper cites 2023 , keywords =.

Hypothesis generation and updating in large language models 2023 , keywords =

Reference 47

Resolution
verified fuzzy
raw_fallback, observed 2026-05-26T11:37:37.862580Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-05-08T14:43:56.125546Z digest=sha256:ea113f4febdb4b6524eef0d7370762a3c2200c09d9450c5c303fa42aa57b67c0

Observation 64ac58bc-3f4a-4b14-87b4-465d3bd8781f · outbound

This paper cites Journal of Open Psychology Data , author =.

Hypothesis generation and updating in large language models Journal of Open Psychology Data , author =

Reference 48

Resolution
verified exact
doi, observed 2026-05-08T21:14:12.360252Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-05-08T14:43:56.125546Z digest=sha256:490876ba20516cbdf16622e6ee8a4de362baa649e289839b96744bca239e14a2

Observation 842efd26-b153-4b83-a3e7-5e69534230ed · outbound

This paper cites Rules and.

Hypothesis generation and updating in large language models Rules and

Reference 49

Resolution
verified fuzzy
raw_fallback, observed 2026-05-26T11:37:37.860176Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-05-08T14:43:56.125546Z digest=sha256:7d695a6511b776524e8834ad3c702ae4a7581bc9ff4c791b20574edcd29b4ae1

Observation 807a7adf-db54-4099-bb4d-ece52fbf9e0f · outbound

This paper cites month = apr, year =.

Hypothesis generation and updating in large language models month = apr, year =

Reference 50

Resolution
verified exact
arxiv_id, observed 2026-05-08T21:14:12.208720Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-05-08T14:43:56.125546Z digest=sha256:225ba25d3b148c97b2cd90d3060afc690b5945544e0df1ecfebfbad7af10eeb2

Observation b86f6e62-086c-436f-9dd1-89c6ee9f4bb5 · outbound

This paper cites Bayesian teaching enables probabilistic reasoning in large language models , volume =.

Hypothesis generation and updating in large language models Bayesian teaching enables probabilistic reasoning in large language models , volume =

Reference 51

Resolution
verified exact
doi, observed 2026-05-08T21:14:12.033039Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-05-08T14:43:56.125546Z digest=sha256:d41e23cddba2d9978d58cb55bc16e812ee515b31aec4ca2c9025bb2db4df5557

Observation 20b79598-a638-4265-baf1-645ee0c42272 · outbound

This paper cites Magpie: Alignment Data Synthesis from Scratch by Prompting Aligned LLMs with Nothing.

Hypothesis generation and updating in large language models Magpie: Alignment Data Synthesis from Scratch by Prompting Aligned LLMs with Nothing

Reference 52

Resolution
verified exact
arxiv_id, observed 2026-05-16T06:58:37.155837Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-05-08T14:43:56.125546Z digest=sha256:54202ffc41eada35f23db5af937e6fea29a5b4ac9c0c44e6826fd90a76e8e6d8

Observation 0b494f22-b586-40aa-b01b-f9a62948d817 · outbound

This paper cites On Language Models' Sensitivity to Suspicious Coincidences.

Hypothesis generation and updating in large language models On Language Models' Sensitivity to Suspicious Coincidences

Reference 53

Resolution
verified exact
arxiv_id, observed 2026-05-08T21:14:12.424085Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-05-08T14:43:56.125546Z digest=sha256:17ef57913cfa82f47ec8185e2e900d4993b20eb75d4964e2f6351d0b382881e8

Observation da5818ff-a8b9-442f-9ed5-04fd2856a259 · outbound

This paper cites an unresolved cited work.

Hypothesis generation and updating in large language models Unresolved cited work

Reference 54

Resolution
verified exact
arxiv_id, observed 2026-05-08T21:14:11.926703Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-05-08T14:43:56.125546Z digest=sha256:5ce5a0dc4d24513d3e31013cada1b1e5826485c6a2d8bc062228946ca8a7634f

Observation 606be60e-364a-4153-ad06-b5cf858f1bae · outbound

This paper cites an unresolved cited work.

Hypothesis generation and updating in large language models Unresolved cited work

Reference 55

Resolution
verified exact
arxiv_id, observed 2026-05-08T21:14:12.349649Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-05-08T14:43:56.125546Z digest=sha256:749691f0265fbdc86b6c30cbadff844b963f7287e067d3795fddf936ec7774e1

Observation a9d14390-04ab-4387-9f8e-a6b2c7d95f72 · outbound

This paper cites Revisiting Uncertainty Estimation and Calibration of Large Language Models.

Hypothesis generation and updating in large language models Revisiting Uncertainty Estimation and Calibration of Large Language Models

Reference 56

Resolution
verified exact
arxiv_id, observed 2026-05-08T21:14:12.533436Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-05-08T14:43:56.125546Z digest=sha256:32154bd5f3b31d92c9a70a6897e4d4ee0424c3ba64efbb1a39c11de273d3777a

Observation ef6e4ed6-30c8-4bb8-bfb5-d117879768b1 · outbound

This paper cites Kapoor, N.

Hypothesis generation and updating in large language models Kapoor, N

Reference 57

Resolution
verified exact
doi, observed 2026-05-08T21:14:12.614338Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-05-08T14:43:56.125546Z digest=sha256:ea99e01d9b80b4b08178a48ed779de91611289f0f3e3d459648c8e4c9891f35c

Observation d7b39099-2378-481a-85d8-d7f7fa4b4d9e · outbound

This paper cites The Illusion of Thinking: Understanding the Strengths and Limitations of Reasoning Models via the Lens of Problem Complexity.

Hypothesis generation and updating in large language models The Illusion of Thinking: Understanding the Strengths and Limitations of Reasoning Models via the Lens of Problem Complexity

Reference 58

Resolution
verified exact
arxiv_id, observed 2026-05-15T16:10:31.796621Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-05-08T14:43:56.125546Z digest=sha256:2fb0c2e52ac4cd529bd5249f1fb6b557d3efc19884612bc2613796d3355254cb

Observation bd9e3a89-5a5c-4d6a-91de-e68dbff3f2f5 · outbound

This paper cites Contextual Position Encoding: Learning to Count What's Important.

Hypothesis generation and updating in large language models Contextual Position Encoding: Learning to Count What's Important

Reference 59

Resolution
verified exact
arxiv_id, observed 2026-05-08T21:14:12.639564Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-05-08T14:43:56.125546Z digest=sha256:de037dd2aeceff04932032dcd6866831131b66b73d2119f4914feea6ea5b1920

Observation e413db63-9e03-480d-b264-6c04ac38ec80 · outbound

This paper cites arXiv preprint arXiv:2509.10739 , year=.

Hypothesis generation and updating in large language models arXiv preprint arXiv:2509.10739 , year=

Reference 60

Resolution
verified exact
arxiv_id, observed 2026-05-08T21:14:12.330921Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-05-08T14:43:56.125546Z digest=sha256:b52c2f1a4386340cf331673168dc623dafeefee65cb8bafdfb47fb3b06c67516

Observation f00305ee-4de6-4c8f-ba48-55b9b1f716c2 · outbound

This paper cites an unresolved cited work.

Hypothesis generation and updating in large language models Unresolved cited work

Reference 61

Resolution
unresolved
raw_fallback, observed 2026-05-26T11:37:37.937338Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-05-08T14:43:56.125546Z digest=sha256:831efef6ea8ebe6bcf503464bfacbd6925600b601e87f0701580082cdaa30e5f

Observation 3de428d9-d1db-4970-95c1-a49be9cfe006 · outbound

This paper cites an unresolved cited work.

Hypothesis generation and updating in large language models Unresolved cited work

Reference 62

Resolution
unresolved
raw_fallback, observed 2026-05-26T11:37:37.918363Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-05-08T14:43:56.125546Z digest=sha256:b2b1ce0e799e8f714e8a1f913e1e4e3949aad2d4c6e84932e3751da2885943d9

Observation 68d7090a-5f83-4c76-a79e-f2fb09b009f5 · outbound

This paper cites Position.

Hypothesis generation and updating in large language models Position

Reference 63

Resolution
verified fuzzy
raw_fallback, observed 2026-05-26T11:37:37.915230Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-05-08T14:43:56.125546Z digest=sha256:835e8218e4f7b4f9a7d55111040eaeaeed528f6f2c8a98a5705c6b328ef39e0c

Observation d406970f-48ef-4374-9dee-06a89e1321cb · outbound

This paper cites EleutherAI Blog , author =.

Hypothesis generation and updating in large language models EleutherAI Blog , author =

Reference 64

Resolution
verified fuzzy
raw_fallback, observed 2026-05-26T11:37:37.911880Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-05-08T14:43:56.125546Z digest=sha256:8c3262ad5d42fbca8bb3982684c9d53af3f3fc006fd33ad0aab928a784583825

Observation f4a461d7-b74d-43f9-9e19-185f33a674a1 · outbound

This paper cites Embarrassingly Simple Self-Distillation Improves Code Generation.

Hypothesis generation and updating in large language models Embarrassingly Simple Self-Distillation Improves Code Generation

Reference 65

Resolution
verified exact
arxiv_id, observed 2026-06-26T01:15:18.478768Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-05-08T14:43:56.125546Z digest=sha256:c738989d9c4e6ed839429cbfee38ed8bfe8f6e38823b7d540e2a06cac4eed798

Observation 791cac6d-ab96-4894-99a4-67548f373226 · outbound

This paper cites Learning Montezuma's Revenge from a Single Demonstration.

Hypothesis generation and updating in large language models Learning Montezuma's Revenge from a Single Demonstration

Reference 66

Resolution
verified exact
arxiv_id, observed 2026-05-08T21:14:12.576022Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-05-08T14:43:56.125546Z digest=sha256:7f7b2687e367ff70aff54e426be72f821291a08326037b58e3f8156fa051720f

Observation 63804fe2-61a3-411d-abcd-5750c045f237 · outbound

This paper cites Cooperative Inverse Reinforcement Learning.

Hypothesis generation and updating in large language models Cooperative Inverse Reinforcement Learning

Reference 67

Resolution
verified exact
arxiv_id, observed 2026-05-08T21:14:12.404414Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-05-08T14:43:56.125546Z digest=sha256:e48975cf2b9f597a30a33f8ca253ba54248618b9748d8a4951ac4ec0826d016d

Observation cfde71cf-a7da-422d-8b3d-6f14e32563c3 · outbound

This paper cites Test your best methods on our hard.

Hypothesis generation and updating in large language models Test your best methods on our hard

Reference 68

Resolution
verified fuzzy
raw_fallback, observed 2026-05-26T11:37:37.853527Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-05-08T14:43:56.125546Z digest=sha256:61955a7554da6592cd3d19fd2e59655931be4df69ad2fbb64fb85ac1c7ebf0f8

Observation 13784bc3-b8cd-4a38-916c-2ebe908fa1a7 · outbound

This paper cites SKILL0: In-Context Agentic Reinforcement Learning for Skill Internalization.

Hypothesis generation and updating in large language models SKILL0: In-Context Agentic Reinforcement Learning for Skill Internalization

Reference 69

Resolution
metadata mismatch
arxiv_id, observed 2026-05-20T00:00:37.467312Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-05-08T14:43:56.125546Z digest=sha256:62a09fcb95c5e422f8ae7b301f6fe21c9b0b52dcfea649c8b96c99c9a71027ea

Observation b9ed458f-64a0-418d-9f1f-8e7d0ffd7a44 · outbound

This paper cites Cognitive Dark Matter: Measuring What AI Misses.

Hypothesis generation and updating in large language models Cognitive Dark Matter: Measuring What AI Misses

Reference 70

Resolution
metadata mismatch
arxiv_id, observed 2026-08-03T03:15:37.660537Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-05-08T14:43:56.125546Z digest=sha256:a6ce8b66ec44747f13f06006374c48a36d49ba9f3cfc4c7ec99a5a8f5f14d4b5

Observation 8f40c6f6-7722-473c-b08e-9b4bcaf10497 · outbound

This paper cites Cognitive models and AI algorithms provide templates for designing language agents.

Hypothesis generation and updating in large language models Cognitive models and AI algorithms provide templates for designing language agents

Reference 71

Resolution
verified exact
arxiv_id, observed 2026-05-08T21:14:12.461251Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-05-08T14:43:56.125546Z digest=sha256:92a7e235d9a71b65d5f816b5e72ca5d4a5a042e82def425d147d4b654c46b7ff

Observation fcf0aac7-129a-4af7-b8c9-649661eb5185 · outbound

This paper cites year =.

Hypothesis generation and updating in large language models year =

Reference 72

Resolution
verified fuzzy
raw_fallback, observed 2026-05-26T11:37:37.905582Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-05-08T14:43:56.125546Z digest=sha256:24c1a18d3cd14af45931f55ca2dfd6d271fcfc7b660f3c3e537bfe359bab318a

Observation 0edeb54a-7728-4a43-9d32-2e43cd6f93c0 · outbound

This paper cites and Vezhnevets, Alexander Sasha and Diaz, Manfred and Agapiou, John P.

Hypothesis generation and updating in large language models and Vezhnevets, Alexander Sasha and Diaz, Manfred and Agapiou, John P

Reference 73

Resolution
verified exact
arxiv_id, observed 2026-05-08T21:14:12.456537Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-05-08T14:43:56.125546Z digest=sha256:5cd5d84c838334035cd8f8d7ca5472ef5f29886affba22c044f7f76e260de88d

Observation 15508f36-4631-4690-a8ba-d89d2e139d7c · outbound

This paper cites Superposition Yields Robust Neural Scaling.

Hypothesis generation and updating in large language models Superposition Yields Robust Neural Scaling

Reference 74

Resolution
verified exact
local_arxiv, observed 2026-05-08T21:14:11.802350Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-05-08T14:43:56.125546Z digest=sha256:54656d8da62f579df0f494fa19073bc5610cf7a852771b82c1b90db9d2dba960

Observation 5a406431-6aac-4af1-8997-0b6b8df8c5f2 · outbound

This paper cites Science , author =.

Hypothesis generation and updating in large language models Science , author =

Reference 75

Resolution
verified exact
doi, observed 2026-05-08T21:14:12.007531Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-05-08T14:43:56.125546Z digest=sha256:7ebf6df17b1470ad0e0691f539244256cd0f9d4f386b7eeaecd6430701c33fe7

Observation b9b3855e-4a3e-49a9-92a0-cf91301bb7e0 · outbound

This paper cites Representations and generalization in artificial and brain neural networks , volume =.

Hypothesis generation and updating in large language models Representations and generalization in artificial and brain neural networks , volume =

Reference 76

Resolution
verified exact
doi, observed 2026-05-08T21:14:12.630733Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-05-08T14:43:56.125546Z digest=sha256:bc261834857ef01cda8808c2f4f77b36ff110b104b53cfc77fe3e8c0f0eb0a47

Observation 9271fc38-c9f5-41b5-adff-a55eb6e689c9 · outbound

This paper cites (Joshua Brett) , year =.

Hypothesis generation and updating in large language models (Joshua Brett) , year =

Reference 77

Resolution
verified fuzzy
raw_fallback, observed 2026-05-26T11:37:37.908809Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-05-08T14:43:56.125546Z digest=sha256:ae62965e7c2b5d17233328cec58fcdcbcc10c782bbf4e32171449768aff8628f

Observation 028f480d-b607-4d22-bfc1-86512fb8de11 · outbound

This paper cites Liar, Liar Pants on Fire.

Hypothesis generation and updating in large language models Liar, Liar Pants on Fire

Reference 78

Resolution
verified exact
doi, observed 2026-05-08T21:14:12.061280Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-05-08T14:43:56.125546Z digest=sha256:5f9c70e1a54b1e0759d814084c42003b7a3210d7ad6cf7c4400da138c6781802

Observation a265d452-25a5-4ff4-94b2-e76cf0c8da30 · outbound

This paper cites Prototypical.

Hypothesis generation and updating in large language models Prototypical

Reference 79

Resolution
verified fuzzy
raw_fallback, observed 2026-05-26T11:37:37.926413Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-05-08T14:43:56.125546Z digest=sha256:db51332d0b8f98dba06786c2ee86538710d9401241560cee322007d078afae4f

Observation ab757c56-0feb-4d27-a975-13b0aab1da32 · outbound

This paper cites Gemma 3 Technical Report.

Hypothesis generation and updating in large language models Gemma 3 Technical Report

Reference 80

Resolution
metadata mismatch
arxiv_id, observed 2026-05-08T21:14:12.283722Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-05-08T14:43:56.125546Z digest=sha256:5031e6a6d67c2142a6f97824d22e984242310e5003e3c40e42931e0063f85857

Observation 89ff62ea-41cd-441c-baca-d65d5788a0c6 · outbound

This paper cites Qwen3.5.

Hypothesis generation and updating in large language models Qwen3.5

Reference 81

Resolution
verified fuzzy
raw_fallback, observed 2026-05-26T11:37:37.930444Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-05-08T14:43:56.125546Z digest=sha256:b30dfaad3fdef984f7f2d4b0aff7ef3a6fa99717b3dc7ae6bec3a183ee65c03f

Observation 3e31b0bc-65bb-4023-a22c-645cfaabc746 · outbound

This paper cites Memory & Cognition , author =.

Hypothesis generation and updating in large language models Memory & Cognition , author =

Reference 82

Resolution
verified exact
doi, observed 2026-05-08T21:14:12.565162Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-05-08T14:43:56.125546Z digest=sha256:af05c3d75d9faf66cc3da1fff2fc00c44a4ade8b3ffa2acd19dd13879a4c68f1

Observation d99e07fe-909a-46c7-967b-e4c2d5baea09 · outbound

This paper cites year =.

Hypothesis generation and updating in large language models year =

Reference 83

Resolution
verified exact
doi, observed 2026-05-08T21:14:12.418720Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-05-08T14:43:56.125546Z digest=sha256:9dcd84fb2dd3e606749cd8c3adbea6530b14662b69552ded3fc0f975d8ca2bbf

Observation d03c56c5-dc6d-4a67-95bd-d1ad0da2d0e2 · outbound

This paper cites Instruction-Following Evaluation for Large Language Models.

Hypothesis generation and updating in large language models Instruction-Following Evaluation for Large Language Models

Reference 84

Resolution
verified exact
arxiv_id, observed 2026-05-08T21:14:11.784375Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-05-08T14:43:56.125546Z digest=sha256:68f4f1d054888e147a483932fcd2ef2660673a295fbb840d0db7969d52ba9a40

Observation d68e8c28-e01c-49e6-a560-079cb7183aaf · outbound

This paper cites month = aug, year =.

Hypothesis generation and updating in large language models month = aug, year =

Reference 85

Resolution
verified fuzzy
raw_fallback, observed 2026-05-26T11:37:37.894214Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-05-08T14:43:56.125546Z digest=sha256:3a08054dabd328b6520944a8b21e8db23c050a3ad3fc3bc8482163cf73006e77

Observation d2d4ea88-ddfc-4783-b16e-5efd0f920766 · outbound

This paper cites Advances in.

Hypothesis generation and updating in large language models Advances in

Reference 86

Resolution
verified exact
doi, observed 2026-05-08T21:14:11.760282Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-05-08T14:43:56.125546Z digest=sha256:c863541cdc3a2f3fe5f20bb0f23567df2a575e463f3e8700a068e44ae86e9bb9

Observation 4c0b3152-88bb-4b2d-aaf0-1a50cacb38a3 · outbound

This paper cites Damien Ernst, Pierre Geurts, and Louis Wehenkel.

Hypothesis generation and updating in large language models Damien Ernst, Pierre Geurts, and Louis Wehenkel

Reference 87

Resolution
metadata mismatch
doi, observed 2026-05-08T21:14:12.550713Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-05-08T14:43:56.125546Z digest=sha256:caaf31b7987214fc59168be76ce55bacdde8afa432a3edb4c00a8ed502c36a3a

Observation 0cf4e194-1fb2-4dd2-ba37-ea58b500fea7 · outbound

This paper cites Proceedings of the National Academy of Sciences , author =.

Hypothesis generation and updating in large language models Proceedings of the National Academy of Sciences , author =

Reference 88

Resolution
verified exact
doi, observed 2026-05-08T21:14:12.469751Z

Source-reported events for the cited work

correction dated 2023-03-10. Source: crossref record 10.1073/pnas.2302618120->10.1073/pnas.1915984117:correction, observed 2026-07-11T03:09:38.508366+00:00. This notice travels one citation hop only.

source=arxiv_source observed=2026-05-08T14:43:56.125546Z digest=sha256:dd3cff5fad4f2b7e526bd7e25b0b4f81e34090a96f91dc1c33b957eeac43b5e7

Observation 1c9bb7b6-aaef-413f-a408-82806c467117 · outbound

This paper cites Murray, Alberto Bernacchia, Nicholas A.

Hypothesis generation and updating in large language models Murray, Alberto Bernacchia, Nicholas A

Reference 89

Resolution
metadata mismatch
doi, observed 2026-05-08T21:14:12.411540Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-05-08T14:43:56.125546Z digest=sha256:98897e801b3ded7aea02156083282c8d5b541711ecddbf1f9ea084249863ef0c

Observation bdc6dca4-7d26-43ac-9d01-c83d5d8d5841 · outbound

This paper cites The selective updating of working memory: a predictive coding account , language =.

Hypothesis generation and updating in large language models The selective updating of working memory: a predictive coding account , language =

Reference 90

Resolution
verified fuzzy
raw_fallback, observed 2026-05-26T11:37:37.897106Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-05-08T14:43:56.125546Z digest=sha256:6f317bf2b373be2aaadf345a50a95b307609a4725a4f7baa7799fa9c7708fea0

Observation 3415c68f-2f86-40b8-bd0e-c053cf7d5dd6 · outbound

This paper cites Annual Review of Psychology , author =.

Hypothesis generation and updating in large language models Annual Review of Psychology , author =

Reference 91

Resolution
verified exact
doi, observed 2026-05-08T21:14:11.774246Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-05-08T14:43:56.125546Z digest=sha256:bb61e3a32b2ae8e70d108c8bc82a679f84955437b32165b33810119dfc7af2e5

Observation 1c025f57-b5d2-4beb-a5af-89e3231be64d · outbound

This paper cites Current Opinion in Behavioral Sciences , author =.

Hypothesis generation and updating in large language models Current Opinion in Behavioral Sciences , author =

Reference 92

Resolution
verified exact
doi, observed 2026-05-08T21:14:12.223360Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-05-08T14:43:56.125546Z digest=sha256:2d3b7ee6e4bc69ab48cf31662f2d62aaa0f21ba28a3196c510d6cf2ff0a981f7

Observation bcfe6f02-e0d6-41fa-bf4b-e10fc354eba5 · outbound

This paper cites Nature Communications , author =.

Hypothesis generation and updating in large language models Nature Communications , author =

Reference 93

Resolution
verified exact
doi, observed 2026-05-08T21:14:12.493616Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-05-08T14:43:56.125546Z digest=sha256:3c7f5debf93b015ff61376b82a7ecf418812bb539e3509a1b9548748b12442be

Observation d6e5f528-aaa3-4aca-9d1e-65333a743ce3 · outbound

This paper cites Nature Communications , author =.

Hypothesis generation and updating in large language models Nature Communications , author =

Reference 94

Resolution
verified exact
doi, observed 2026-05-08T21:14:12.488519Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-05-08T14:43:56.125546Z digest=sha256:7f3be36be8687cb7004e96e1c922b4aa0cf86c6d4e9bf3aeebe738cf29563188

Observation 9bfd6ec2-0c9e-4d72-9f29-2d86181cf552 · outbound

This paper cites Trends in Cognitive Sciences , author =.

Hypothesis generation and updating in large language models Trends in Cognitive Sciences , author =

Reference 95

Resolution
verified exact
doi, observed 2026-05-08T21:14:12.188279Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-05-08T14:43:56.125546Z digest=sha256:3b40f50cc9cf4367c03062996100b551abda1f767f500ede9b7cc7db97babafc

Observation c718251a-5550-4a8c-b5ea-4a5f9a1e85fb · outbound

This paper cites Trends in Cognitive Sciences , author =.

Hypothesis generation and updating in large language models Trends in Cognitive Sciences , author =

Reference 96

Resolution
verified exact
doi, observed 2026-05-08T21:14:12.368517Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-05-08T14:43:56.125546Z digest=sha256:55fda4785207bd6b93ef939d829dd82464a73c27364a0e51780e61db423b9156

Observation 9290a1af-cb4d-4235-b91d-249679329683 · outbound

This paper cites The Journal of Neuroscience , author =.

Hypothesis generation and updating in large language models The Journal of Neuroscience , author =

Reference 97

Resolution
verified exact
arxiv_id, observed 2026-05-08T21:14:12.252349Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-05-08T14:43:56.125546Z digest=sha256:1f4074e49dd0e3d3f02e9f96b063c8efca9dee55dabdf6ad0e814ab2d9e0c0b3

Observation 0622aeeb-d196-40de-9146-28b98fab0968 · outbound

This paper cites month = oct, year =.

Hypothesis generation and updating in large language models month = oct, year =

Reference 98

Resolution
verified exact
doi, observed 2026-05-08T21:14:12.377037Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-05-08T14:43:56.125546Z digest=sha256:8d0eec0e1846659059c7d1f2a9860452a1f87d382f20d404a129fc9edaf8b5ed

Observation 49d62c70-6557-494d-a977-d4e717a0a28f · outbound

This paper cites Nature Human Behaviour , author =.

Hypothesis generation and updating in large language models Nature Human Behaviour , author =

Reference 99

Resolution
verified exact
doi, observed 2026-05-08T21:14:12.125013Z

Source-reported events for the cited work

correction dated 2020-10-09. Source: crossref record 10.1038/s41562-020-00993-7->10.1038/s41562-020-00938-0:correction, observed 2026-07-11T03:09:11.611456+00:00. This notice travels one citation hop only.

source=arxiv_source observed=2026-05-08T14:43:56.125546Z digest=sha256:2dcc7c158a031da137c5502cb26ae8f67306de27919d6526979bede5c70b29a1

Observation e73daff0-40d0-4e7c-b47e-31bfe1080c67 · outbound

This paper cites Strong inhibitory signaling underlies stable temporal dynamics and working memory in spiking neural networks.

Hypothesis generation and updating in large language models Strong inhibitory signaling underlies stable temporal dynamics and working memory in spiking neural networks

Reference 100

Resolution
metadata mismatch
doi, observed 2026-05-08T21:14:12.415186Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-05-08T14:43:56.125546Z digest=sha256:f943cc5a4bc492c939592c384e76725cf5ca54bd745c73f35c90218ee59a6577

Pith citing papers

Observation ccbae499-94fa-4b49-a548-c657576fade6 · inbound

Thinking Under Uncertainty: Evidence Use and Information-Seeking in Language Models cites this paper.

Thinking Under Uncertainty: Evidence Use and Information-Seeking in Language Models Hypothesis generation and updating in large language models

Reference 14

Resolution
unresolved
no resolver link, observed 2026-07-30T19:24:34.330570Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-30T19:24:34.330570Z digest=sha256:a0c3967182579e0e29689c07d30dab2acb77cf4baea20d03f3fdd8c2f255ffc9