Pith. sign in

Paper Citation Record · LEDGER

Semi-pessimistic Reinforcement Learning

As of 9 August 2026, this Paper Citation Record lists 61 of 61 outbound references and 0 inbound Pith citation observations for arXiv:2505.19002.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2505.19002 v1

Coverage vector

measured 61 of 61 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T14:28:45.036484Z

measured 61 of 61 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

61 of 61 outbound references displayed

  • verified exact1
  • verified fuzzy14
  • unresolved46
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 99701860-d7a5-4519-bf1c-c4c62d053410 · outbound

This paper cites write newline.

Semi-pessimistic Reinforcement Learning write newline

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-07T14:28:37.936913Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:28:37.936913Z digest=sha256:8fac79157158beb03021e876f5fd6fc7daa7bf706d7d5df42237b356ca8ef62d

Observation f0e73df3-f5a1-477b-908c-d2ce1d01c67c · outbound

This paper cites an unresolved cited work.

Semi-pessimistic Reinforcement Learning Unresolved cited work

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-07T14:28:38.039135Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:28:38.039135Z digest=sha256:9e3116219039874867809ef526a40055f05704fe8a61c9279af6ac70e78a9957

Observation 9821c6b0-05cd-4c49-80dd-16b0a2aa82e8 · outbound

This paper cites an unresolved cited work.

Semi-pessimistic Reinforcement Learning Unresolved cited work

Reference 3

Resolution
unresolved
raw_fallback, observed 2026-08-07T14:28:57.250274Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T14:28:38.180791Z digest=sha256:7d6f1d6535cdc9ac1f7d37212fa692850364419603b92c3a18ae4340ccc8c8c3

Observation 71eb6981-cd84-4902-befe-9a1bdd22d63e · outbound

This paper cites an unresolved cited work.

Semi-pessimistic Reinforcement Learning Unresolved cited work

Reference 4

Resolution
unresolved
raw_fallback, observed 2026-08-07T14:28:57.019410Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T14:28:38.255599Z digest=sha256:3d19f50c8ee1fc3c5ff0bb38b045ae52ad465bee6eac790a8f20b70e5f0e78db

Observation 9c125de8-e6cf-48ff-a349-ad1348ebd310 · outbound

This paper cites an unresolved cited work.

Semi-pessimistic Reinforcement Learning Unresolved cited work

Reference 5

Resolution
unresolved
raw_fallback, observed 2026-08-07T14:28:56.799927Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T14:28:38.383579Z digest=sha256:7c598936d32ac00638b12de76f52ef535e3cc95996c10852b5585a45e1b37ec7

Observation e81d5493-7b0c-4b82-a1fc-7e758937459f · outbound

This paper cites an unresolved cited work.

Semi-pessimistic Reinforcement Learning Unresolved cited work

Reference 6

Resolution
unresolved
raw_fallback, observed 2026-08-07T14:28:56.556621Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T14:28:38.522385Z digest=sha256:9433c51519d2f4e8d3f1fa4509106491056c51e61a821d085513afcdbd0fa40b

Observation d7313cbf-daf9-446f-b91b-117085dd1357 · outbound

This paper cites an unresolved cited work.

Semi-pessimistic Reinforcement Learning Unresolved cited work

Reference 7

Resolution
unresolved
raw_fallback, observed 2026-08-07T14:28:56.302536Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T14:28:38.650485Z digest=sha256:b6b07f85e59c9b89f827825ce15e6aa23c7b010eebef9c6fd97ced4e380a8251

Observation 4ebe2b46-3254-4a7d-98a8-3f17e77346b2 · outbound

This paper cites STEEL: Singularity-aware Reinforcement Learning.

Semi-pessimistic Reinforcement Learning STEEL: Singularity-aware Reinforcement Learning

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-07T14:28:38.798613Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:28:38.798613Z digest=sha256:6fde691361ea3bbd17d73f7e9d7ab56e3d67b058ee6040b2881589bb5ef73595

Observation c170ed35-d3bc-46d4-a1ff-3871c594919f · outbound

This paper cites an unresolved cited work.

Semi-pessimistic Reinforcement Learning Unresolved cited work

Reference 9

Resolution
unresolved
raw_fallback, observed 2026-08-07T14:28:56.007740Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T14:28:38.939129Z digest=sha256:52ce6e3429be8b4c7ec20d2740b1ced613f066bf644ae8a1d619822235bdc0ec

Observation 60e32809-7430-4148-b1f3-0bcd91ecc82e · outbound

This paper cites Geurts, and L.

Semi-pessimistic Reinforcement Learning Geurts, and L

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:28:55.794274Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T14:28:39.046707Z digest=sha256:7889446023885cc152a6a799ef12ee8f6f4320e35ee5b2fdd1df8149d45f0c5d

Observation fecb2668-4647-4436-863e-7c89ae43de2f · outbound

This paper cites D4RL: Datasets for Deep Data-Driven Reinforcement Learning.

Semi-pessimistic Reinforcement Learning D4RL: Datasets for Deep Data-Driven Reinforcement Learning

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-07T14:28:39.226745Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:28:39.226745Z digest=sha256:a0f24857b55b1af907d30c0829d9922d69b4d2382b72109e8b581cf77a38154f

Observation 6429a876-1341-4dac-b32d-f10146cf9b5f · outbound

This paper cites Soft Actor-Critic Algorithms and Applications.

Semi-pessimistic Reinforcement Learning Soft Actor-Critic Algorithms and Applications

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-07T14:28:39.355009Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:28:39.355009Z digest=sha256:f60d28557eb66f61e6896077f79b8b27cab75ace367ba037cdea8b63effdefa9

Observation bad278c7-37fe-42fc-95cc-bd0370df7f71 · outbound

This paper cites an unresolved cited work.

Semi-pessimistic Reinforcement Learning Unresolved cited work

Reference 13

Resolution
unresolved
raw_fallback, observed 2026-08-07T14:28:55.507828Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T14:28:39.509393Z digest=sha256:17eeda5feb1f1b3ab93b92937c54c3b239b2cfb0218138feec2a08df295417a5

Observation 645a845d-c875-4ef4-81c5-6137f04d4656 · outbound

This paper cites an unresolved cited work.

Semi-pessimistic Reinforcement Learning Unresolved cited work

Reference 14

Resolution
unresolved
raw_fallback, observed 2026-08-07T14:28:55.270751Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T14:28:39.678987Z digest=sha256:0c5db293786233bfaddfdd01d6e0a3f052fe0ce8d7624b6d8e07bda0cd77f47e

Observation a33bc208-27f6-4cd9-b831-28489c7966f3 · outbound

This paper cites Policy learning "without" overlap: Pessimism and generalized empirical Bernstein's inequality.

Semi-pessimistic Reinforcement Learning Policy learning "without" overlap: Pessimism and generalized empirical Bernstein's inequality

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-07T14:28:39.800402Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:28:39.800402Z digest=sha256:82b7bf0484dd9162e779c24c764f4dc6af66357981886a7b1a0e1147a18e1d97

Observation 6a86d18d-3515-485e-a07f-81a6732ffad8 · outbound

This paper cites Yang, and Z.

Semi-pessimistic Reinforcement Learning Yang, and Z

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:28:55.065832Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T14:28:39.967853Z digest=sha256:a817c68ea2af4bded40e24ac81b55e21ae0456f14e7058cf5220cf1c7bfcc98b

Observation fbe3824e-8f47-479f-9246-12f0e15cf401 · outbound

This paper cites an unresolved cited work.

Semi-pessimistic Reinforcement Learning Unresolved cited work

Reference 17

Resolution
unresolved
raw_fallback, observed 2026-08-07T14:28:54.844614Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T14:28:40.102144Z digest=sha256:a5f63548425c43013880bfc51d9eb3513356cedb0fe20bf306bd05b15f1101e7

Observation 8af24646-485f-46cb-8e3c-56ba55d1887d · outbound

This paper cites an unresolved cited work.

Semi-pessimistic Reinforcement Learning Unresolved cited work

Reference 18

Resolution
unresolved
raw_fallback, observed 2026-08-07T14:28:54.590785Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T14:28:40.228709Z digest=sha256:344c84fa6f148d5c69665a7b764541a5c4a1b6791c8d57ad80970611cb4585f3

Observation c5c1d08f-7322-42bd-affd-adbc77525ee2 · outbound

This paper cites Rajeswaran, P.

Semi-pessimistic Reinforcement Learning Rajeswaran, P

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:28:54.361890Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T14:28:40.351419Z digest=sha256:f2829937ceba3b25c5ebb356a9acb2176bab3d6432553977a0715025c3283f02

Observation dcad40ad-9cdd-48ab-b6b8-00aabd78b12c · outbound

This paper cites an unresolved cited work.

Semi-pessimistic Reinforcement Learning Unresolved cited work

Reference 20

Resolution
unresolved
raw_fallback, observed 2026-08-07T14:28:54.108956Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T14:28:40.485615Z digest=sha256:7f3bb7d5d07dc42f3de2bfb5caee7b9044fa569cebf2f14396522aae252ddab5

Observation 38461ea0-756f-4039-9c4e-62738c0297ff · outbound

This paper cites Zolna, Y.

Semi-pessimistic Reinforcement Learning Zolna, Y

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:28:53.858404Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T14:28:40.615672Z digest=sha256:f86212d396962859121ee75edc433d413b49d9713467747907c921256c12a95e

Observation b1a99c68-e931-4ddd-b07e-61d80242b9ea · outbound

This paper cites an unresolved cited work.

Semi-pessimistic Reinforcement Learning Unresolved cited work

Reference 22

Resolution
unresolved
raw_fallback, observed 2026-08-07T14:28:53.672146Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T14:28:40.740155Z digest=sha256:3419070053d0088005024e011ebc3755cb3f6d7d648406350ca0af4bf29a36ef

Observation eccf748a-e09f-445b-a675-6d0793f56565 · outbound

This paper cites an unresolved cited work.

Semi-pessimistic Reinforcement Learning Unresolved cited work

Reference 23

Resolution
unresolved
raw_fallback, observed 2026-08-07T14:28:53.450113Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T14:28:40.901587Z digest=sha256:979db5c93dc1476896e3ac7442d9bce2ff0c68acff10bafa5c5f8775a234f61a

Observation c76397c6-f9e7-454c-926f-cfff88c31828 · outbound

This paper cites Offline Reinforcement Learning: Tutorial, Review, and Perspectives on Open Problems.

Semi-pessimistic Reinforcement Learning Offline Reinforcement Learning: Tutorial, Review, and Perspectives on Open Problems

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-07T14:28:41.046807Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:28:41.046807Z digest=sha256:cde8bc6176a1c7794724221fd25267ce11ae8de32f9475a7160503bd43b57a79

Observation e983a0b9-61fa-4b6b-97da-269b919d1803 · outbound

This paper cites an unresolved cited work.

Semi-pessimistic Reinforcement Learning Unresolved cited work

Reference 25

Resolution
unresolved
raw_fallback, observed 2026-08-07T14:28:53.196644Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T14:28:41.171724Z digest=sha256:9cf3ffe703a9cf3dd94f4d0a38ddc7bc9eb482f59908b77434502b7d69949e04

Observation 127d3ecf-b8fe-4854-b9aa-841769404d31 · outbound

This paper cites Zhou, and R.

Semi-pessimistic Reinforcement Learning Zhou, and R

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:28:52.925678Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T14:28:41.310645Z digest=sha256:a47d3e1599f770c3e375e8a9c0abc371d78b08d6b3dec175ff0c3c515e7f9012

Observation 33d75495-4c97-4310-9767-eea3cc7d5c2d · outbound

This paper cites an unresolved cited work.

Semi-pessimistic Reinforcement Learning Unresolved cited work

Reference 27

Resolution
unresolved
raw_fallback, observed 2026-08-07T14:28:52.722306Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T14:28:41.393759Z digest=sha256:3b268a3d8d155ecf3964fd000271aacebf693889f65cffac70b5ce51a1d879eb

Observation 6bfdb06c-e426-4153-8b9a-410f5517bef8 · outbound

This paper cites Pogosyan, S.

Semi-pessimistic Reinforcement Learning Pogosyan, S

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:28:52.468941Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T14:28:41.477973Z digest=sha256:d01c2b914fbbdb721b76793f22794134c97fbddcfd9715d69c2122322d0c62ab

Observation 25183d71-fd7f-4130-8817-1822819b9378 · outbound

This paper cites Online Estimation and Inference for Robust Policy Evaluation in Reinforcement Learning.

Semi-pessimistic Reinforcement Learning Online Estimation and Inference for Robust Policy Evaluation in Reinforcement Learning

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-07T14:28:41.567710Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:28:41.567710Z digest=sha256:c28cebbe6553278600fb8da4a634de3eea9c190e0ab85416d9fb5648aeb6e16d

Observation 37ccdb60-6a77-4422-9398-579ac402fa6e · outbound

This paper cites Swaminathan, A.

Semi-pessimistic Reinforcement Learning Swaminathan, A

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:28:52.250748Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T14:28:41.695359Z digest=sha256:51d3d68de808223f3aa20535d7af429937dbe9d11e895e57b12e816ac3edfb29

Observation 48239ab6-cb8e-4a74-9456-59210da79243 · outbound

This paper cites an unresolved cited work.

Semi-pessimistic Reinforcement Learning Unresolved cited work

Reference 31

Resolution
unresolved
raw_fallback, observed 2026-08-07T14:28:52.026860Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T14:28:41.794393Z digest=sha256:e25ea614a71075e443310bb2bc6b5c6fbd32b542b3652482d8da385870eb311d

Observation 73aa90c8-ba05-405b-b920-f852d6bc3aa3 · outbound

This paper cites On the Sample Complexity of Reinforcement Learning with Policy Space Generalization.

Semi-pessimistic Reinforcement Learning On the Sample Complexity of Reinforcement Learning with Policy Space Generalization

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-07T14:28:41.882921Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:28:41.882921Z digest=sha256:f52a6bef5b2fbbcba67a9fd70c05ffe605ba2c681047a5b39cb8447f5fcda9a9

Observation 05bb4ab3-b68c-4711-b8a2-082db11450db · outbound

This paper cites Gilron, S.

Semi-pessimistic Reinforcement Learning Gilron, S

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:28:51.782463Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T14:28:41.994961Z digest=sha256:d0cffedeff8b9ed9ffb52ec2f59edef204a37eb813a1b25a2e7de12fd5161db4

Observation dc0e4f90-7e54-4493-83f5-f1da2aec25b5 · outbound

This paper cites an unresolved cited work.

Semi-pessimistic Reinforcement Learning Unresolved cited work

Reference 34

Resolution
unresolved
raw_fallback, observed 2026-08-07T14:28:51.477087Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T14:28:42.089577Z digest=sha256:7730752f741b35279f3811b9247af1b1123764320c9f239a48a5bb5d4725aa39

Observation f21f92ed-0e4b-4bb3-abd4-4e73647403f2 · outbound

This paper cites an unresolved cited work.

Semi-pessimistic Reinforcement Learning Unresolved cited work

Reference 35

Resolution
unresolved
raw_fallback, observed 2026-08-07T14:28:51.203211Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T14:28:42.209714Z digest=sha256:599e2846caa6383654fa4288bad69a7a915ca9e17a5982ed94865c100e0fe89f

Observation 019fba24-78d8-4b4c-a10a-5d9e3bcbb7ad · outbound

This paper cites an unresolved cited work.

Semi-pessimistic Reinforcement Learning Unresolved cited work

Reference 36

Resolution
unresolved
raw_fallback, observed 2026-08-07T14:28:50.922182Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T14:28:42.290825Z digest=sha256:e0c258f853258cdf705a24ba1415c7f99c45c38f77a2f9116d2ddbf79ff8e027

Observation 0eda664c-e47f-40d6-b53b-5d9285f6e9f8 · outbound

This paper cites an unresolved cited work.

Semi-pessimistic Reinforcement Learning Unresolved cited work

Reference 37

Resolution
unresolved
raw_fallback, observed 2026-08-07T14:28:50.686027Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T14:28:42.400930Z digest=sha256:37630e39db8f3e4bb95a924034cee9a46d516d49941b3d4d35815ea81a1e7562

Observation 8a50c41a-5308-4242-a27d-dfd577f58744 · outbound

This paper cites an unresolved cited work.

Semi-pessimistic Reinforcement Learning Unresolved cited work

Reference 38

Resolution
unresolved
raw_fallback, observed 2026-08-07T14:28:50.405140Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T14:28:42.487562Z digest=sha256:6d91c9e5c8b778316ab6284d24d716f1a41a47379e9971a4e43c2ef4d8abac28

Observation adea0d2d-cbbf-413f-a857-e65d4ec5dc1b · outbound

This paper cites an unresolved cited work.

Semi-pessimistic Reinforcement Learning Unresolved cited work

Reference 39

Resolution
unresolved
raw_fallback, observed 2026-08-07T14:28:50.120874Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T14:28:42.592262Z digest=sha256:1e837d85f5a1a882c309037547ca738a6a1381cc8bf9d137968c792754402921

Observation 685d8228-8a5c-43c4-8172-798b7a0ae1b1 · outbound

This paper cites an unresolved cited work.

Semi-pessimistic Reinforcement Learning Unresolved cited work

Reference 40

Resolution
unresolved
raw_fallback, observed 2026-08-07T14:28:49.895424Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T14:28:42.712853Z digest=sha256:dfcd4e26259f4702e87ed8de6a4c2928ff10335b0322adf12ef3d16c1a2a43ef

Observation 530d3269-cb13-4011-b961-13263010636e · outbound

This paper cites Zhang, W.

Semi-pessimistic Reinforcement Learning Zhang, W

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:28:49.533109Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T14:28:42.802507Z digest=sha256:a1e3c8916258aebf17e2d34d13e3e200f13dfef2cfc475e4ced85e71f0f51986

Observation 9e3d4bc2-1f05-4d93-8ba3-924e40ecb855 · outbound

This paper cites an unresolved cited work.

Semi-pessimistic Reinforcement Learning Unresolved cited work

Reference 42

Resolution
unresolved
raw_fallback, observed 2026-08-07T14:28:49.211361Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T14:28:42.910136Z digest=sha256:9d7a2f90c807650d88006c8880c5f3ed8e36bd3273457c822888cfaea6b9be17

Observation ab69acd3-d127-4f75-b8bf-0526ca9ca715 · outbound

This paper cites an unresolved cited work.

Semi-pessimistic Reinforcement Learning Unresolved cited work

Reference 43

Resolution
unresolved
raw_fallback, observed 2026-08-07T14:28:48.943611Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T14:28:43.035009Z digest=sha256:6a9427e4a0b9b876802250611153d98a0148bac0aa66e18151143a0b2271070e

Observation 9d8bd769-3430-436e-9059-cf931d838a70 · outbound

This paper cites an unresolved cited work.

Semi-pessimistic Reinforcement Learning Unresolved cited work

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-07T14:28:43.157235Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:28:43.157235Z digest=sha256:a4756cdd3e0c43d2dee8dc6ab0f23f140fe109e97b49a53bab31cafc7f155385

Observation 1d6aa5ed-ee94-4e47-a15b-e1279b4242e3 · outbound

This paper cites Finite Sample Analysis of Minimax Offline Reinforcement Learning: Completeness, Fast Rates and First-Order Efficiency.

Semi-pessimistic Reinforcement Learning Finite Sample Analysis of Minimax Offline Reinforcement Learning: Completeness, Fast Rates and First-Order Efficiency

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-07T14:28:43.256958Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:28:43.256958Z digest=sha256:0ac93bcd6424398ebcefab2b0eb4815aac4291a4c61a220abdd5244ca930fce0

Observation 27fdadc1-b5eb-4595-b1f3-d184932d905d · outbound

This paper cites an unresolved cited work.

Semi-pessimistic Reinforcement Learning Unresolved cited work

Reference 46

Resolution
unresolved
raw_fallback, observed 2026-08-07T14:28:48.741183Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T14:28:43.344777Z digest=sha256:46927c850bc5dd8233c6d075dae530de51d0667caa9e1f135de64addb1a4d768

Observation 4e0a3fd1-f9d0-489b-b201-af777aa3ddc5 · outbound

This paper cites Pessimistic Causal Reinforcement Learning with Mediators for Confounded Offline Data.

Semi-pessimistic Reinforcement Learning Pessimistic Causal Reinforcement Learning with Mediators for Confounded Offline Data

Reference 47

Resolution
verified exact
local_arxiv, observed 2026-08-07T14:28:45.310256Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T14:28:43.440171Z digest=sha256:4c98c3c24b02f71083a2366b533c49a9feaa7f2d22955d609169ad1dedc43654

Observation 6d058ed4-02a2-4e04-b0eb-43a14f154dea · outbound

This paper cites Qi, and R.

Semi-pessimistic Reinforcement Learning Qi, and R

Reference 48

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:28:48.474622Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T14:28:43.546569Z digest=sha256:2e12a1e4934d82edd2af8c7d5c667b7a401cab20d853388311eb1ad4c8eda7b7

Observation 7cae2f61-7ad1-4077-bb0a-b904cbc3f14c · outbound

This paper cites an unresolved cited work.

Semi-pessimistic Reinforcement Learning Unresolved cited work

Reference 49

Resolution
unresolved
raw_fallback, observed 2026-08-07T14:28:48.253023Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T14:28:43.679055Z digest=sha256:adc4e3e70a07191a5be21cc37ea3361c7e34034b571e286097ef8a3e420f387e

Observation 962e6983-c9b3-4fac-aea5-a0da3b4ca402 · outbound

This paper cites an unresolved cited work.

Semi-pessimistic Reinforcement Learning Unresolved cited work

Reference 50

Resolution
unresolved
raw_fallback, observed 2026-08-07T14:28:47.968674Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T14:28:43.762845Z digest=sha256:8e5a289ee775b0812a98f334210d00eb14051afd0ce3ac6c7c8b348d1fce4028

Observation fc6ce3aa-8435-4594-931f-69aae7b9550d · outbound

This paper cites Levine, and O.

Semi-pessimistic Reinforcement Learning Levine, and O

Reference 51

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:28:47.719874Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T14:28:43.856847Z digest=sha256:74d4757f6c975b05e020ce0081e39324221fa5115366285308dc798b4b6da553

Observation f1411628-f282-483f-9c5b-f0c5438c2cea · outbound

This paper cites an unresolved cited work.

Semi-pessimistic Reinforcement Learning Unresolved cited work

Reference 52

Resolution
unresolved
raw_fallback, observed 2026-08-07T14:28:47.451523Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T14:28:43.964585Z digest=sha256:a711dc530925260a6cae99d54e721d411e1b544366b53f93c710baa276c12088

Observation 754d8c1f-ca35-4a3f-9f49-208d97dfff2b · outbound

This paper cites an unresolved cited work.

Semi-pessimistic Reinforcement Learning Unresolved cited work

Reference 53

Resolution
unresolved
raw_fallback, observed 2026-08-07T14:28:47.186402Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T14:28:44.069549Z digest=sha256:ef85595eb589763875255cf37ed09b3bccb375914958e57e42b745e17f5ad735

Observation 9e8d988d-f8f7-49c2-96f3-c00efc09b07e · outbound

This paper cites an unresolved cited work.

Semi-pessimistic Reinforcement Learning Unresolved cited work

Reference 54

Resolution
unresolved
raw_fallback, observed 2026-08-07T14:28:46.935926Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T14:28:44.213152Z digest=sha256:799db32201ff2729f318ab0065258c4c10ec6be06cf1b8fbdcb28690e9bfa510

Observation 4c0f40ec-c9d1-432f-9589-555eb41acd2d · outbound

This paper cites Kumar, Y.

Semi-pessimistic Reinforcement Learning Kumar, Y

Reference 55

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:28:46.724458Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T14:28:44.284209Z digest=sha256:0b3cc72aace8af64588d9c207cc189a060bf321243a768d199c76f9ea27cdb58

Observation f8dbd09d-c964-4f72-bdec-041b23118a72 · outbound

This paper cites Thomas, L.

Semi-pessimistic Reinforcement Learning Thomas, L

Reference 56

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:28:46.490758Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T14:28:44.392273Z digest=sha256:db80014a585a7c2a6fa5f6111bb598d300d08d8276d5ccc7b6d49ac30c95f364

Observation 8dd568f2-ae44-4cc2-a1d3-cabb155ed5e6 · outbound

This paper cites an unresolved cited work.

Semi-pessimistic Reinforcement Learning Unresolved cited work

Reference 57

Resolution
unresolved
raw_fallback, observed 2026-08-07T14:28:46.272248Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T14:28:44.518642Z digest=sha256:ec0b7051274762396e64cbd1c72695b22312d8b060e607d3e05d32f8a22d52a8

Observation 24c7a0c3-6138-41dd-9dd2-301c39ff353f · outbound

This paper cites an unresolved cited work.

Semi-pessimistic Reinforcement Learning Unresolved cited work

Reference 58

Resolution
unresolved
no resolver link, observed 2026-08-07T14:28:44.635383Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:28:44.635383Z digest=sha256:3a0878864a397c686c642946e2fe5d59943df9eaa8d82cee23a4bb6a81db1d2c

Observation 65d115b8-b914-4359-9b1d-5a742998bc2c · outbound

This paper cites an unresolved cited work.

Semi-pessimistic Reinforcement Learning Unresolved cited work

Reference 59

Resolution
unresolved
raw_fallback, observed 2026-08-07T14:28:46.040662Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T14:28:44.780597Z digest=sha256:423df66ed15fbd7c1406b9fa0cb0d3600c0ef5c3bafd5a923a19540f4e8e4536

Observation 1f167cd3-1a22-4a72-97a8-3af97810b653 · outbound

This paper cites Zhu, and A.

Semi-pessimistic Reinforcement Learning Zhu, and A

Reference 60

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:28:45.815329Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T14:28:44.917322Z digest=sha256:6a9594f2cf401e98664314ccafd4d92959a9c10b0ace20b5aa854e1bd367f62d

Observation 60538fa8-48c7-4515-967f-93c138b3bb16 · outbound

This paper cites an unresolved cited work.

Semi-pessimistic Reinforcement Learning Unresolved cited work

Reference 61

Resolution
unresolved
raw_fallback, observed 2026-08-07T14:28:45.571579Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T14:28:45.036484Z digest=sha256:1436ce8b4020e65ec4e69b23ee10293a8b8b06471841b15beaa510df5d375cd4

Pith citing papers

No inbound Pith citation observations are available.