Pith. sign in

Paper Citation Record · LEDGER

Guiding Reinforcement Learning Using Uncertainty-Aware Large Language Models

As of 16 August 2026, this Paper Citation Record lists 59 of 59 outbound references and 1 inbound Pith citation observation for arXiv:2411.14457.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2411.14457 v1

Coverage vector

measured 59 of 59 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-12T19:34:41.907594Z

measured 60 of 60 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-16T06:30:59.297886+00:00

measured 1 of 1 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-06T16:23:32.867331Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-08-06T16:23:34.078888Z

Reference resolution

59 of 59 outbound references displayed

  • verified exact0
  • verified fuzzy1
  • unresolved57
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch1

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 40a1cc83-c532-4c0d-be6f-d151235db974 · outbound

This paper cites an unresolved cited work.

Guiding Reinforcement Learning Using Uncertainty-Aware Large Language Models Unresolved cited work

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-12T19:34:41.646213Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T19:34:41.646213Z digest=sha256:9fb96677c77852bfcc7379cf471986dd5b11dbc0b255eaccaa7538b364ca53be

Observation 8da97f0c-7875-4c75-bdc7-92edb1df8eef · outbound

This paper cites GPT-4 Technical Report.

Guiding Reinforcement Learning Using Uncertainty-Aware Large Language Models GPT-4 Technical Report

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-12T19:34:41.651812Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T19:34:41.651812Z digest=sha256:d69bc68b45c4194eeb3b0386582ba752c3ab8476dc1eac7cf690b69f3f3d0224

Observation 30b10aa2-93e1-4bd7-9a61-d7ab88270b15 · outbound

This paper cites an unresolved cited work.

Guiding Reinforcement Learning Using Uncertainty-Aware Large Language Models Unresolved cited work

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-12T19:34:41.656837Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T19:34:41.656837Z digest=sha256:34b4823ab001e68ccf8f8caea6aa6ee4c4d8724dfd4e1c2f2ee3cdd72e2e04cf

Observation c24f2d9d-f145-49e7-85ad-bb5e2b9e6482 · outbound

This paper cites Deep Reinforcement Learning from Policy-Dependent Human Feedback.

Guiding Reinforcement Learning Using Uncertainty-Aware Large Language Models Deep Reinforcement Learning from Policy-Dependent Human Feedback

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-12T19:34:41.668581Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T19:34:41.668581Z digest=sha256:f01e047cdf6271a336924e29c54565399e96f0c2e59580d468f9876e576a26fa

Observation 3ad25830-1a58-4ea1-be2a-b87f7ac427e8 · outbound

This paper cites an unresolved cited work.

Guiding Reinforcement Learning Using Uncertainty-Aware Large Language Models Unresolved cited work

Reference 5

Resolution
unresolved
raw_fallback, observed 2026-08-12T19:34:42.574987Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-12T19:34:41.673817Z digest=sha256:00daeb434d8c0a8dac8d230840923d946c5316d07504d934766377572ab1c591

Observation 03765f5f-3fde-4ada-b1a3-18de442e58c8 · outbound

This paper cites ChatGPT is a Knowledgeable but Inexperienced Solver: An Investigation of Commonsense Problem in Large Language Models.

Guiding Reinforcement Learning Using Uncertainty-Aware Large Language Models ChatGPT is a Knowledgeable but Inexperienced Solver: An Investigation of Commonsense Problem in Large Language Models

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-12T19:34:41.678501Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T19:34:41.678501Z digest=sha256:8921e701c62f379d2ee2c29a6d24bb17bb02900e89d9fc9993c286794ad68b90

Observation 88ebe1b2-d585-4f0d-809e-c36e72a73cc9 · outbound

This paper cites Survey on Large Language Model-Enhanced Reinforcement Learning: Concept, Taxonomy, and Methods.

Guiding Reinforcement Learning Using Uncertainty-Aware Large Language Models Survey on Large Language Model-Enhanced Reinforcement Learning: Concept, Taxonomy, and Methods

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-12T19:34:41.684020Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T19:34:41.684020Z digest=sha256:0590b077291ae7e0ceaf6a0d3df1af4c2af969d07ab35dd82087d75d182a3298

Observation df3919e8-875b-4a02-ad71-e9b17194526e · outbound

This paper cites an unresolved cited work.

Guiding Reinforcement Learning Using Uncertainty-Aware Large Language Models Unresolved cited work

Reference 8

Resolution
unresolved
raw_fallback, observed 2026-08-12T19:34:42.561724Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-12T19:34:41.689116Z digest=sha256:548bab7271f3b4a72d38553b33d5225e8e07b9f8beae344c0a1a7b8f5875790b

Observation 3f8338df-cbaf-45a4-80a9-23bb5bbf69c0 · outbound

This paper cites an unresolved cited work.

Guiding Reinforcement Learning Using Uncertainty-Aware Large Language Models Unresolved cited work

Reference 9

Resolution
unresolved
raw_fallback, observed 2026-08-12T19:34:42.550058Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-12T19:34:41.694384Z digest=sha256:63c456ac563e9c117f58fe9dc11e4c5dc3a646d97b0d50979b6dc53242065f13

Observation 7fe2e24d-4ae1-4be3-83bd-24929adf7adc · outbound

This paper cites an unresolved cited work.

Guiding Reinforcement Learning Using Uncertainty-Aware Large Language Models Unresolved cited work

Reference 10

Resolution
unresolved
raw_fallback, observed 2026-08-12T19:34:42.536955Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-12T19:34:41.698793Z digest=sha256:c15530c74271ddac2e5d2ccf3bb013ce76da88abc3783457c2f349557e2bcb85

Observation 79cbcea4-a194-45c3-a996-d3f02aa5ee61 · outbound

This paper cites Accelerating Reinforcement Learning of Robotic Manipulations via Feedback from Large Language Models.

Guiding Reinforcement Learning Using Uncertainty-Aware Large Language Models Accelerating Reinforcement Learning of Robotic Manipulations via Feedback from Large Language Models

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-12T19:34:41.703104Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T19:34:41.703104Z digest=sha256:e4a9c22a17fe9ea06c29572a82a741345a57c28187e15492153b06d41892676d

Observation d1e9cfa7-eca2-4b93-ac32-dab53e149beb · outbound

This paper cites an unresolved cited work.

Guiding Reinforcement Learning Using Uncertainty-Aware Large Language Models Unresolved cited work

Reference 12

Resolution
unresolved
raw_fallback, observed 2026-08-12T19:34:42.524723Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-12T19:34:41.709035Z digest=sha256:e4abad021a0049a0191b02e87ab608918bd4b355576f5ec1b57df4ad036b2f9e

Observation d9173cde-c8c2-4051-864b-66c47ea784ac · outbound

This paper cites an unresolved cited work.

Guiding Reinforcement Learning Using Uncertainty-Aware Large Language Models Unresolved cited work

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-12T19:34:41.713473Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T19:34:41.713473Z digest=sha256:890f1ff14a62087ad6e00d18ac959f718373d09b53ff4b37094657d0d2ea88d5

Observation 289d4a4f-b3ac-4452-97ec-90a403848d7c · outbound

This paper cites an unresolved cited work.

Guiding Reinforcement Learning Using Uncertainty-Aware Large Language Models Unresolved cited work

Reference 14

Resolution
unresolved
raw_fallback, observed 2026-08-12T19:34:42.503711Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-12T19:34:41.717636Z digest=sha256:57fcb74c61317ab8c2d6d0fa0539b503481db5f539f2fec67012df569aa5af9b

Observation f9e84247-7236-49d7-bac2-ef760519605c · outbound

This paper cites an unresolved cited work.

Guiding Reinforcement Learning Using Uncertainty-Aware Large Language Models Unresolved cited work

Reference 15

Resolution
unresolved
raw_fallback, observed 2026-08-12T19:34:42.491667Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-12T19:34:41.721365Z digest=sha256:fa563a8513f66243fa882a8b3100d8ded6afceb754f64e9a41a34c3b5773f0e8

Observation 6e84c4d1-8678-4e6b-9c0b-14d0f3e2b41b · outbound

This paper cites On the Importance of Uncertainty in Decision-Making with Large Language Models.

Guiding Reinforcement Learning Using Uncertainty-Aware Large Language Models On the Importance of Uncertainty in Decision-Making with Large Language Models

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-12T19:34:41.725138Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T19:34:41.725138Z digest=sha256:9df368c9e61cab5281cee285874c944d8797a890c51cd658876c75e8a367a84d

Observation ba09c59b-fe28-4763-b7fa-b78cf39071de · outbound

This paper cites an unresolved cited work.

Guiding Reinforcement Learning Using Uncertainty-Aware Large Language Models Unresolved cited work

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-12T19:34:41.729045Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T19:34:41.729045Z digest=sha256:1213cd92447e85a2a7c8209a6244535d136ce2483c547c0f7355ce3fa52394cf

Observation bcda8ef2-ab01-4bdd-938f-a5c99fe1707d · outbound

This paper cites an unresolved cited work.

Guiding Reinforcement Learning Using Uncertainty-Aware Large Language Models Unresolved cited work

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-12T19:34:41.733423Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T19:34:41.733423Z digest=sha256:fa22667d532f10c145812850b829d11af016c45b6d039cb85543eacf7749a2b7

Observation d865dc5d-fa4e-4b15-8a75-15d2183940fd · outbound

This paper cites Guiding Reinforcement Learning Exploration Using Natural Language.

Guiding Reinforcement Learning Using Uncertainty-Aware Large Language Models Guiding Reinforcement Learning Exploration Using Natural Language

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-12T19:34:41.737300Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T19:34:41.737300Z digest=sha256:c0731bcccfc09ac52a3487e92bc3af528914d3570bb32dfc781664a4093d98a7

Observation d0b1d895-9e7d-45f7-ab4d-89c89318ae70 · outbound

This paper cites an unresolved cited work.

Guiding Reinforcement Learning Using Uncertainty-Aware Large Language Models Unresolved cited work

Reference 20

Resolution
unresolved
raw_fallback, observed 2026-08-12T19:34:42.464841Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-12T19:34:41.741844Z digest=sha256:64e19d6f327ad4ff0923a43da54242595db30ce8627e7e72f4ec87edb7d2fd51

Observation 18d5c70a-b7a0-47e1-96c4-b420f8693394 · outbound

This paper cites Look Before You Leap: An Exploratory Study of Uncertainty Measurement for Large Language Models.

Guiding Reinforcement Learning Using Uncertainty-Aware Large Language Models Look Before You Leap: An Exploratory Study of Uncertainty Measurement for Large Language Models

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-12T19:34:41.745863Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T19:34:41.745863Z digest=sha256:30ae620c9a6f40fb6e9338be3227a61b950944a1d69446b3ca62ed2fa3bde8ba

Observation 6846a880-475e-4c48-ae8c-0154cccb66ed · outbound

This paper cites an unresolved cited work.

Guiding Reinforcement Learning Using Uncertainty-Aware Large Language Models Unresolved cited work

Reference 22

Resolution
unresolved
raw_fallback, observed 2026-08-12T19:34:42.452281Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-12T19:34:41.750658Z digest=sha256:0cc0dbde2b09e9afd3cc2b57b6f828774d214b7a38ceef59e7674c675e552b33

Observation c1a16b08-a949-45d3-a727-b674fdb2c4f8 · outbound

This paper cites an unresolved cited work.

Guiding Reinforcement Learning Using Uncertainty-Aware Large Language Models Unresolved cited work

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-12T19:34:41.754398Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T19:34:41.754398Z digest=sha256:ec73ebc303a20f815d455eeb38778e844eab042600536294f6b59a70d38e5318

Observation 5ba27ffe-2e34-4edf-b431-142243c820bb · outbound

This paper cites an unresolved cited work.

Guiding Reinforcement Learning Using Uncertainty-Aware Large Language Models Unresolved cited work

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-12T19:34:41.759746Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T19:34:41.759746Z digest=sha256:f41f55fc6c90c1ae562d964bcb85dee35c9fe6cb4a1eb39e252773d57d24cf25

Observation 26868cb3-1a4b-4e66-8672-72c29b50482e · outbound

This paper cites an unresolved cited work.

Guiding Reinforcement Learning Using Uncertainty-Aware Large Language Models Unresolved cited work

Reference 25

Resolution
unresolved
raw_fallback, observed 2026-08-12T19:34:42.422572Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-12T19:34:41.763797Z digest=sha256:e07bed40c16a2b655f2930967eadf18bf86ed7e312549c90a40f8584c2bc77d6

Observation 2240a49b-a9a5-4d9a-b14f-bd44bdb4ab36 · outbound

This paper cites an unresolved cited work.

Guiding Reinforcement Learning Using Uncertainty-Aware Large Language Models Unresolved cited work

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-12T19:34:41.768289Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T19:34:41.768289Z digest=sha256:9ee488c203c3e9561e7f956c5792c29879384b63851fc9100bfaa1367d04efb3

Observation 3afa7b64-0140-447e-a6ed-6401be30f49e · outbound

This paper cites an unresolved cited work.

Guiding Reinforcement Learning Using Uncertainty-Aware Large Language Models Unresolved cited work

Reference 27

Resolution
unresolved
raw_fallback, observed 2026-08-12T19:34:42.404161Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-12T19:34:41.772296Z digest=sha256:5a8ec3512ee16fe39c0a625acb8c9736eea8ef33691a1c3b971fabc671506281

Observation 5de79e6c-e55f-45c3-b338-91f9b0f36aad · outbound

This paper cites an unresolved cited work.

Guiding Reinforcement Learning Using Uncertainty-Aware Large Language Models Unresolved cited work

Reference 28

Resolution
unresolved
raw_fallback, observed 2026-08-12T19:34:42.391247Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-12T19:34:41.776134Z digest=sha256:c72ff7f4ed98843e017fe2cb216faf531abe45541fd77d15e711a6ae1b9ae5cb

Observation d7b793d6-24aa-43de-bdc0-96e0ef1f871c · outbound

This paper cites an unresolved cited work.

Guiding Reinforcement Learning Using Uncertainty-Aware Large Language Models Unresolved cited work

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-12T19:34:41.780489Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T19:34:41.780489Z digest=sha256:3851a8bba7e3b80d447cd4bce3483a2d52b6ab1e6e95a674fbb19ab302a41bd4

Observation 58ac7936-4156-4410-9f3e-7a1a851df054 · outbound

This paper cites an unresolved cited work.

Guiding Reinforcement Learning Using Uncertainty-Aware Large Language Models Unresolved cited work

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-12T19:34:41.784982Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T19:34:41.784982Z digest=sha256:ae7924ffd98a5b2177971d88e8907aed6f75b559dd75c697274cc14dacbf657a

Observation 5dd42ddb-6798-4bf0-a126-f722e0554956 · outbound

This paper cites an unresolved cited work.

Guiding Reinforcement Learning Using Uncertainty-Aware Large Language Models Unresolved cited work

Reference 31

Resolution
unresolved
raw_fallback, observed 2026-08-12T19:34:42.364644Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-12T19:34:41.791601Z digest=sha256:4c6c0bbd8f233133d2069db692f4e8028a3bcfbddaa620519a94663326b7b2a7

Observation 54a360bb-6601-429e-9e53-eba161315547 · outbound

This paper cites an unresolved cited work.

Guiding Reinforcement Learning Using Uncertainty-Aware Large Language Models Unresolved cited work

Reference 32

Resolution
unresolved
raw_fallback, observed 2026-08-12T19:34:42.353220Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-12T19:34:41.796144Z digest=sha256:964c46ff9bfce5f1f72d1756509f05cdd389c23005762f396a3d6ffaa4ef1c80

Observation baf9c4b9-e279-4d8b-8deb-8937d0494f74 · outbound

This paper cites Learning to Model the World with Language.

Guiding Reinforcement Learning Using Uncertainty-Aware Large Language Models Learning to Model the World with Language

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-12T19:34:41.801224Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T19:34:41.801224Z digest=sha256:ff908510f833ce7a93b8ae6e6da6e8079491b96bb7f699d47a2a9a1823dc661c

Observation c41b99d3-f011-4551-be8d-af73b8160ff2 · outbound

This paper cites an unresolved cited work.

Guiding Reinforcement Learning Using Uncertainty-Aware Large Language Models Unresolved cited work

Reference 34

Resolution
unresolved
raw_fallback, observed 2026-08-12T19:34:42.340285Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-12T19:34:41.805642Z digest=sha256:6111599e7d2abdfb902531ca2a91f805d1014ab73dc9b14883aa7902a2ce0706

Observation c0b4eb37-57da-4d0e-968b-39d14e1744cd · outbound

This paper cites an unresolved cited work.

Guiding Reinforcement Learning Using Uncertainty-Aware Large Language Models Unresolved cited work

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-12T19:34:41.809944Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T19:34:41.809944Z digest=sha256:644a900055754df08a0bece30623a815b160810798c3510e076f415d14840d42

Observation 12ebd4a6-759c-464d-a62c-c2fde5d7db11 · outbound

This paper cites Prevent the Language Model from being Overconfident in Neural Machine Translation.

Guiding Reinforcement Learning Using Uncertainty-Aware Large Language Models Prevent the Language Model from being Overconfident in Neural Machine Translation

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-12T19:34:41.814602Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T19:34:41.814602Z digest=sha256:3c55e319e50093c83b944f3cb905c8ac4cac79257a55982f86628c948c382527

Observation 58aedbe4-441c-4847-b732-d533b13f5e5c · outbound

This paper cites an unresolved cited work.

Guiding Reinforcement Learning Using Uncertainty-Aware Large Language Models Unresolved cited work

Reference 37

Resolution
unresolved
raw_fallback, observed 2026-08-12T19:34:42.317587Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-12T19:34:41.819281Z digest=sha256:78acc613d567debb2ab951c2bc8a88d6741de49bff949f5aec3c332344665838

Observation 8c626b34-ed7e-4c37-a6c0-c36d5552f1f3 · outbound

This paper cites an unresolved cited work.

Guiding Reinforcement Learning Using Uncertainty-Aware Large Language Models Unresolved cited work

Reference 38

Resolution
unresolved
raw_fallback, observed 2026-08-12T19:34:42.304675Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-12T19:34:41.823164Z digest=sha256:845e8157134eee3602753088e86dd2986ba8848374eb3460ddd892d788ceee66

Observation 87c968f1-dc22-439a-aaa4-a8ee6d685df0 · outbound

This paper cites an unresolved cited work.

Guiding Reinforcement Learning Using Uncertainty-Aware Large Language Models Unresolved cited work

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-12T19:34:41.827077Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T19:34:41.827077Z digest=sha256:9222a3505433004afea0b1015167f1989286d780eb3f2d576ec48e3c7592772b

Observation 7f4040a5-788e-4548-bc9e-213d753a65c3 · outbound

This paper cites an unresolved cited work.

Guiding Reinforcement Learning Using Uncertainty-Aware Large Language Models Unresolved cited work

Reference 40

Resolution
unresolved
raw_fallback, observed 2026-08-12T19:34:42.284370Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-12T19:34:41.831346Z digest=sha256:bbb543f62d568ab3067595ba9b828aa5ce2006a36ee52a2e5d9790e3a6bd50c8

Observation 667e3f7c-33bf-4b2e-9a3c-34bf1c2f596c · outbound

This paper cites an unresolved cited work.

Guiding Reinforcement Learning Using Uncertainty-Aware Large Language Models Unresolved cited work

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-12T19:34:41.837636Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T19:34:41.837636Z digest=sha256:078da227416ccb5692f2dcda0f2909e4cfca291a903864ca2414cb686b3f80af

Observation 3d59a283-29e4-4ee2-af23-509ce85617d1 · outbound

This paper cites Select Committee on Artificial In- telligence.

Guiding Reinforcement Learning Using Uncertainty-Aware Large Language Models Select Committee on Artificial In- telligence

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T19:34:42.262933Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-12T19:34:41.846222Z digest=sha256:38080f319ccf8f619657c426fa035eea83ce597c5a3ce7c9f8842f2f301fc27c

Observation 624b186d-c132-4f7a-a82f-6c0f25de9075 · outbound

This paper cites Unleashing the Power of Pre-trained Language Models for Offline Reinforcement Learning.

Guiding Reinforcement Learning Using Uncertainty-Aware Large Language Models Unleashing the Power of Pre-trained Language Models for Offline Reinforcement Learning

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-12T19:34:41.850850Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T19:34:41.850850Z digest=sha256:b2152e38eaa64f6719f1396005a57938af2adb40ecfaf895945f4add34f9c83c

Observation 6ed4a278-315e-4f06-84c9-21e34e4fa554 · outbound

This paper cites an unresolved cited work.

Guiding Reinforcement Learning Using Uncertainty-Aware Large Language Models Unresolved cited work

Reference 44

Resolution
unresolved
raw_fallback, observed 2026-08-12T19:34:42.249991Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-12T19:34:41.855480Z digest=sha256:276bf7950eeff6dd81215e9a3644113ba7b281e3a9c1fefaa3e5063169bc2d79

Observation ec205bc3-b000-432c-b47b-8975809dd078 · outbound

This paper cites an unresolved cited work.

Guiding Reinforcement Learning Using Uncertainty-Aware Large Language Models Unresolved cited work

Reference 45

Resolution
unresolved
raw_fallback, observed 2026-08-12T19:34:42.238396Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-12T19:34:41.859102Z digest=sha256:ec3502f9f1c4a70f41942fb4fd5a13af33bebe2e67c6f5fb7aad2636d5e6a465

Observation 54eed5f6-ff51-4dc5-9657-2b3598846450 · outbound

This paper cites an unresolved cited work.

Guiding Reinforcement Learning Using Uncertainty-Aware Large Language Models Unresolved cited work

Reference 46

Resolution
unresolved
raw_fallback, observed 2026-08-12T19:34:42.225959Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-12T19:34:41.862679Z digest=sha256:5be39aa660158cf321ce0866ae2ad1dc5c3264cb90e5054c529e03de1931a154

Observation c39d75ce-2379-4faa-b0a8-ed1aad0157cb · outbound

This paper cites Understanding the Capabilities, Limitations, and Societal Impact of Large Language Models.

Guiding Reinforcement Learning Using Uncertainty-Aware Large Language Models Understanding the Capabilities, Limitations, and Societal Impact of Large Language Models

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-12T19:34:41.866456Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T19:34:41.866456Z digest=sha256:a8f7ffc28da769356653dca28be20c08e1ac4cc02cc492126a4584355d170f07

Observation 7d100d2d-0584-4d85-8dca-df27aa16ef51 · outbound

This paper cites an unresolved cited work.

Guiding Reinforcement Learning Using Uncertainty-Aware Large Language Models Unresolved cited work

Reference 48

Resolution
unresolved
raw_fallback, observed 2026-08-12T19:34:42.213393Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-12T19:34:41.870400Z digest=sha256:1913770ff16fdd96853ddb188785f5690da97a183060dd88f378a8efd644692b

Observation 22db6783-e769-4ef9-8500-2c3220abc155 · outbound

This paper cites an unresolved cited work.

Guiding Reinforcement Learning Using Uncertainty-Aware Large Language Models Unresolved cited work

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-12T19:34:41.879299Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T19:34:41.879299Z digest=sha256:ad322ba3d8f79e758bb441dd6ad0e845a4cfb016e209750d06793ba3cfdc2d78

Observation 4ffcaaa4-aa61-4074-a1b0-beb3b85fc88c · outbound

This paper cites an unresolved cited work.

Guiding Reinforcement Learning Using Uncertainty-Aware Large Language Models Unresolved cited work

Reference 50

Resolution
unresolved
raw_fallback, observed 2026-08-12T19:34:42.191345Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-12T19:34:41.883966Z digest=sha256:40fc58c1aa0601ac51e6bd21efbf5d366bc42c7b862978ebf4d6a1676b95f579

Observation a6b82b9f-4b35-4379-bf5a-90e4f4699d2b · outbound

This paper cites Larger language models do in-context learning differently.

Guiding Reinforcement Learning Using Uncertainty-Aware Large Language Models Larger language models do in-context learning differently

Reference 51

Resolution
unresolved
no resolver link, observed 2026-08-12T19:34:41.888302Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T19:34:41.888302Z digest=sha256:1e5df0bda3045f527cdec0980c1dde733dd8b2b32acc6d1232782226c28453f2

Observation 411d1e64-98e0-434a-984c-0fcfbf684518 · outbound

This paper cites an unresolved cited work.

Guiding Reinforcement Learning Using Uncertainty-Aware Large Language Models Unresolved cited work

Reference 52

Resolution
unresolved
raw_fallback, observed 2026-08-12T19:34:42.178006Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-12T19:34:41.892387Z digest=sha256:a02840b0b92f855889534349c99369094719aa560a27ec1ebdf211de05302cd0

Observation b15d3b76-c645-42e4-8848-c53fc2b36824 · outbound

This paper cites On Hallucination and Predictive Uncertainty in Conditional Language Generation.

Guiding Reinforcement Learning Using Uncertainty-Aware Large Language Models On Hallucination and Predictive Uncertainty in Conditional Language Generation

Reference 53

Resolution
unresolved
no resolver link, observed 2026-08-12T19:34:41.896085Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T19:34:41.896085Z digest=sha256:0a80c68edfb07a87bda560f8603682359869f914ba22c950c553dcc9450d076a

Observation db6d92bd-d6d4-49bb-9b9e-173ef398cb9b · outbound

This paper cites Keep CALM and Explore: Language Models for Action Generation in Text-based Games.

Guiding Reinforcement Learning Using Uncertainty-Aware Large Language Models Keep CALM and Explore: Language Models for Action Generation in Text-based Games

Reference 54

Resolution
unresolved
no resolver link, observed 2026-08-12T19:34:41.899943Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T19:34:41.899943Z digest=sha256:327f6dfab032f86c1a3dc57a8e9298ff88bb873312fbd25bbab5124e1d186b03

Observation 37e18345-1d76-44aa-84e4-ecc5ce467812 · outbound

This paper cites Learning Shaping Strategies in Human-in-the-loop Interactive Reinforcement Learning.

Guiding Reinforcement Learning Using Uncertainty-Aware Large Language Models Learning Shaping Strategies in Human-in-the-loop Interactive Reinforcement Learning

Reference 55

Resolution
unresolved
no resolver link, observed 2026-08-12T19:34:41.903610Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T19:34:41.903610Z digest=sha256:73a39cfd2db1364e481e17d9991504124a5db42e8b5016f28c504dbd3ec1fc00

Observation 4117b3c4-8fb4-4568-862e-0a6f7bbc835a · outbound

This paper cites an unresolved cited work.

Guiding Reinforcement Learning Using Uncertainty-Aware Large Language Models Unresolved cited work

Reference 56

Resolution
unresolved
raw_fallback, observed 2026-08-12T19:34:42.165813Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-12T19:34:41.907594Z digest=sha256:60ee94956df85b54f78c81f4ec095e1246283250b3d6f2a171ed161243b0e33d

Observation 68eb32ac-8a84-4447-8776-017ba70b072a · outbound

This paper cites Proximal Policy Optimization Algorithms.

Guiding Reinforcement Learning Using Uncertainty-Aware Large Language Models Proximal Policy Optimization Algorithms

Reference 2017

Resolution
unresolved
no resolver link, observed 2026-08-12T19:34:41.841695Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T19:34:41.841695Z digest=sha256:d41ab82c9510126cc9102245f55f3914e6b3541dc63c26ab799ab0604900a97b

Observation 776fb0eb-36d4-440a-83c3-30817678588b · outbound

This paper cites Influencing Reinforcement Learning through Natural Language Guidance.

Guiding Reinforcement Learning Using Uncertainty-Aware Large Language Models Influencing Reinforcement Learning through Natural Language Guidance

Reference 2021

Resolution
metadata mismatch
local_arxiv, observed 2026-08-12T19:34:41.994320Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-12T19:34:41.874350Z digest=sha256:48d3942a1dcca09fa7abb052d7ea9daa83f4d6d22b2785b7ef8bfb08b89ad3c6

Observation 009bcc72-70d8-4344-bc2c-61ffb5e3022c · outbound

This paper cites What learning algorithm is in-context learning? Investigations with linear models.

Guiding Reinforcement Learning Using Uncertainty-Aware Large Language Models What learning algorithm is in-context learning? Investigations with linear models

Reference 2022

Resolution
unresolved
no resolver link, observed 2026-08-12T19:34:41.661211Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T19:34:41.661211Z digest=sha256:9eb090cdcd04072fcd130d2b5e8a7d3e71743f5ac6351e3d054805b46b72362c

Pith citing papers

Observation d63f222f-670d-4c71-beff-8e272a5254c5 · inbound

LLaPipe: LLM-Guided Reinforcement Learning for Automated Data Preparation Pipeline Construction cites this paper.

LLaPipe: LLM-Guided Reinforcement Learning for Automated Data Preparation Pipeline Construction Guiding Reinforcement Learning Using Uncertainty-Aware Large Language Models

Reference 28

Resolution
verified exact
local_arxiv, observed 2026-08-06T16:23:34.168092Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-06T16:23:32.867331Z digest=sha256:e2197a0af81863f40be56181f7b80dbc24df045fa291338f1d3dc3598190e743