Pith. sign in

Paper Citation Record · LEDGER

SimuDICE: Offline Policy Optimization Through World Model Updates and DICE Estimation

As of 13 August 2026, this Paper Citation Record lists 42 of 42 outbound references and 0 inbound Pith citation observations for arXiv:2412.06486.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2412.06486 v1

Coverage vector

measured 42 of 42 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-11T19:42:10.399295Z

measured 42 of 42 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-12T06:34:41.77262+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

42 of 42 outbound references displayed

  • verified exact0
  • verified fuzzy20
  • unresolved22
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation b947bd01-8fba-4ead-bbee-b2d651283573 · outbound

This paper cites Diffusion for World Modeling: Visual Details Matter in Atari.

SimuDICE: Offline Policy Optimization Through World Model Updates and DICE Estimation Diffusion for World Modeling: Visual Details Matter in Atari

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-11T19:42:10.189641Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T19:42:10.189641Z digest=sha256:f93b245a0013bd1efb405c976a9e75ea87f48c8a65318aa5c59a64ba6607dd98

Observation 973362bb-fa34-4211-98ee-95edc52a959e · outbound

This paper cites Machine Learning20(1-2), 65–81 (1995).

SimuDICE: Offline Policy Optimization Through World Model Updates and DICE Estimation Machine Learning20(1-2), 65–81 (1995)

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T19:42:11.378659Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-11T19:42:10.195132Z digest=sha256:519d377f61cfe47fc905df17fcdf18370cf329241def03c8afdfede9f459d07c

Observation 3153c975-ffaf-4e94-b376-faec6ae384ff · outbound

This paper cites Behavioral Priors and Dynamics Models: Improving Performance and Domain Transfer in Offline RL.

SimuDICE: Offline Policy Optimization Through World Model Updates and DICE Estimation Behavioral Priors and Dynamics Models: Improving Performance and Domain Transfer in Offline RL

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-11T19:42:10.199994Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T19:42:10.199994Z digest=sha256:99fab99373ad98338ebf3db4604168195f43265ea22e7fc5f8dd775f8978b05c

Observation b1cb7d9f-9616-4e94-b017-47126842144f · outbound

This paper cites Model-Based Reinforcement Learning via Meta-Policy Optimization.

SimuDICE: Offline Policy Optimization Through World Model Updates and DICE Estimation Model-Based Reinforcement Learning via Meta-Policy Optimization

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-11T19:42:10.205189Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T19:42:10.205189Z digest=sha256:88fa62751925250c53be650d7a56492d975228822edbb6d2e1d6578bd41a436b

Observation 563fc216-1dc3-4c59-8208-49ad4e05bb89 · outbound

This paper cites In: Proceedings of the 28th International Conference on Machine Learning (ICML-11).

SimuDICE: Offline Policy Optimization Through World Model Updates and DICE Estimation In: Proceedings of the 28th International Conference on Machine Learning (ICML-11)

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T19:42:11.361030Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-11T19:42:10.210366Z digest=sha256:500d41e9c98ec8a0b65b0d92658f0395dcf9254c3eaf440a08e560f970d2a3fb

Observation dfe80ac4-ac1b-4b8f-ae54-8c4ecb2dd705 · outbound

This paper cites Diffusion World Model: Future Modeling Beyond Step-by-Step Rollout for Offline Reinforcement Learning.

SimuDICE: Offline Policy Optimization Through World Model Updates and DICE Estimation Diffusion World Model: Future Modeling Beyond Step-by-Step Rollout for Offline Reinforcement Learning

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-11T19:42:10.215013Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T19:42:10.215013Z digest=sha256:4b63b73c402588aa43a82b49fd4b0f28173f48bf3d07bf171ef1ae40d7c677d4

Observation 8badf9c3-688d-42db-ad42-a6f6cbf86a5f · outbound

This paper cites Mismatched No More: Joint Model-Policy Optimization for Model-Based RL.

SimuDICE: Offline Policy Optimization Through World Model Updates and DICE Estimation Mismatched No More: Joint Model-Policy Optimization for Model-Based RL

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-11T19:42:10.220446Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T19:42:10.220446Z digest=sha256:743b411e955f99fc6dfc2cbb6088329f9cb9c965fb38e26f7723d892b77bafb4

Observation 9762876d-6ab6-4ee8-b71e-5873250cb41b · outbound

This paper cites In: Proceedings of the 36th International Conference on Ma- chine Learning.

SimuDICE: Offline Policy Optimization Through World Model Updates and DICE Estimation In: Proceedings of the 36th International Conference on Ma- chine Learning

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T19:42:11.342318Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-11T19:42:10.225768Z digest=sha256:1ee61b0dfedb4a5fcc54f7acbaaeff19db7ea3dc0298327071f8555ff5169ad9

Observation 3ba53e6a-36a8-43b3-bbd8-688010ec5272 · outbound

This paper cites In: International Conference on Learning Representations (2020).

SimuDICE: Offline Policy Optimization Through World Model Updates and DICE Estimation In: International Conference on Learning Representations (2020)

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T19:42:11.318325Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-11T19:42:10.230979Z digest=sha256:00748602cd9e31ed792f8c9cfb748716293be6aae89b259f5fea5a597014a72d

Observation a225117c-de33-4553-9202-24b7cee3fc69 · outbound

This paper cites Mastering Diverse Domains through World Models.

SimuDICE: Offline Policy Optimization Through World Model Updates and DICE Estimation Mastering Diverse Domains through World Models

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-11T19:42:10.236664Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T19:42:10.236664Z digest=sha256:485c86fd1526e831d37f401aeb80ea619dbf29ba73017bee3476375e87426666

Observation 51f4aa3f-ce4c-4de6-82f2-cd726a5ea663 · outbound

This paper cites Mastering Atari with Discrete World Models.

SimuDICE: Offline Policy Optimization Through World Model Updates and DICE Estimation Mastering Atari with Discrete World Models

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-11T19:42:10.242391Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T19:42:10.242391Z digest=sha256:8284d494b66585405eb978f9e7dd36366446ec3dd4fc2c7892a7032c197372d3

Observation cf7c2ae0-6e26-430f-9ad0-98d6376a4ef2 · outbound

This paper cites In: Proceedings of the 34th International Conference on Machine Learning (ICML 2017).

SimuDICE: Offline Policy Optimization Through World Model Updates and DICE Estimation In: Proceedings of the 34th International Conference on Machine Learning (ICML 2017)

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T19:42:11.299303Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-11T19:42:10.248617Z digest=sha256:f6ce0a494832fade8bc34f69acefdf96aaca5b732ae8201f9599d984c5280481

Observation ed5f68a8-abac-43ba-b957-1c64af657cba · outbound

This paper cites Mitigating Value Hallucination in Dyna Planning via Multistep Predecessor Models.

SimuDICE: Offline Policy Optimization Through World Model Updates and DICE Estimation Mitigating Value Hallucination in Dyna Planning via Multistep Predecessor Models

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-11T19:42:10.253605Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T19:42:10.253605Z digest=sha256:e191c1c2028d567a176cd116004f7a72ba01e267ba9df25c1b4253117205b852

Observation d4782f62-e473-4094-b20c-53db1a6892db · outbound

This paper cites When to Trust Your Model: Model-Based Policy Optimization.

SimuDICE: Offline Policy Optimization Through World Model Updates and DICE Estimation When to Trust Your Model: Model-Based Policy Optimization

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-11T19:42:10.258756Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T19:42:10.258756Z digest=sha256:e94b8c02ee2a8ef1c6d3d26dc721500c25d20bbf5ae0af1a1dd905cd1980169b

Observation 89f30253-fab4-4fdf-8a4f-47572aec66c7 · outbound

This paper cites Way Off-Policy Batch Deep Reinforcement Learning of Implicit Human Preferences in Dialog.

SimuDICE: Offline Policy Optimization Through World Model Updates and DICE Estimation Way Off-Policy Batch Deep Reinforcement Learning of Implicit Human Preferences in Dialog

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-11T19:42:10.264231Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T19:42:10.264231Z digest=sha256:51296a3d7402a7f637b359f8276ce3fb82d8f05c01df4cf420152bf02e386e81

Observation eaae9a28-e4b3-4c35-97e8-7f335715b546 · outbound

This paper cites In: Advances in Neural Information Processing Sys- tems (2020).

SimuDICE: Offline Policy Optimization Through World Model Updates and DICE Estimation In: Advances in Neural Information Processing Sys- tems (2020)

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T19:42:11.279370Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-11T19:42:10.269370Z digest=sha256:9ce976d8458865df51a20a3b0b87347c2968c34833d2081c19d24f176b7fc444

Observation 494cc83e-6a33-4642-91a6-3bd7bfcf1a35 · outbound

This paper cites Offline Reinforcement Learning with Implicit Q-Learning.

SimuDICE: Offline Policy Optimization Through World Model Updates and DICE Estimation Offline Reinforcement Learning with Implicit Q-Learning

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-11T19:42:10.274241Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T19:42:10.274241Z digest=sha256:c8a4058bed583af36933d02df397fdee8d5758b9295be80e12e13e419113c80f

Observation bec1f690-7e6b-4a96-9787-b99b1990e7c5 · outbound

This paper cites In: Wallach, H., Larochelle, H., Beygelzimer, A., d’Alché Buc, F., Fox, E., Garnett, R.

SimuDICE: Offline Policy Optimization Through World Model Updates and DICE Estimation In: Wallach, H., Larochelle, H., Beygelzimer, A., d’Alché Buc, F., Fox, E., Garnett, R

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T19:42:11.261397Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-11T19:42:10.279461Z digest=sha256:24b9ecc56eb2cf8472617410c85a29310cb5982d3b403ea163948f6a76653c4b

Observation 9fa1fdb6-41ef-424f-916b-e2d5dbe35155 · outbound

This paper cites In: Advances in Neural Information Processing Systems.

SimuDICE: Offline Policy Optimization Through World Model Updates and DICE Estimation In: Advances in Neural Information Processing Systems

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T19:42:11.242411Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-11T19:42:10.284609Z digest=sha256:f7c29641910c9ee4e1604d0e548b20036d587d1d517e6d139ec867178223b19b

Observation fe4f6e9f-1fa6-47f2-95b8-5fe7820a86a0 · outbound

This paper cites Model-Ensemble Trust-Region Policy Optimization.

SimuDICE: Offline Policy Optimization Through World Model Updates and DICE Estimation Model-Ensemble Trust-Region Policy Optimization

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-11T19:42:10.289841Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T19:42:10.289841Z digest=sha256:4044d46d4568577afff144287cbfdf0ff7bae72bc4b60966be4d194f319c4eb0

Observation ca2da1fb-b1eb-40b7-8637-eaca30cb2d68 · outbound

This paper cites Objective Mismatch in Model-based Reinforcement Learning.

SimuDICE: Offline Policy Optimization Through World Model Updates and DICE Estimation Objective Mismatch in Model-based Reinforcement Learning

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-11T19:42:10.295272Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T19:42:10.295272Z digest=sha256:c7d0836d828b8c988c26cbb7f59c50e6de81163ece896f3ff8f23c9ef0eb2529

Observation 9d147b65-0ff4-4a7a-acfb-6baa4265167e · outbound

This paper cites In: Wiering, M., van Otterlo, M.

SimuDICE: Offline Policy Optimization Through World Model Updates and DICE Estimation In: Wiering, M., van Otterlo, M

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-11T19:42:10.300617Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T19:42:10.300617Z digest=sha256:c4116d2e68541e13843084915225d11d5fe9720a2bf16a3a75971a45adb4ae45

Observation 5b82b523-a9ca-4c87-b3fe-900a634d6cb1 · outbound

This paper cites Offline Reinforcement Learning: Tutorial, Review, and Perspectives on Open Problems.

SimuDICE: Offline Policy Optimization Through World Model Updates and DICE Estimation Offline Reinforcement Learning: Tutorial, Review, and Perspectives on Open Problems

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-11T19:42:10.309676Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T19:42:10.309676Z digest=sha256:01e9bbc0c1bf65820486dfdaab0f5b519f75e6b9b4b9656833fb393cf662c5a6

Observation 0f14a637-6785-4a94-9ae0-1334d7a4db8a · outbound

This paper cites In: Advances in Neural Information Processing Systems 31 (NeurIPS 2018).

SimuDICE: Offline Policy Optimization Through World Model Updates and DICE Estimation In: Advances in Neural Information Processing Systems 31 (NeurIPS 2018)

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T19:42:11.223196Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-11T19:42:10.314105Z digest=sha256:f05a30c0cc1d5b4cee965a1f3cbb39f2899c19e55f465b8cc116649a887e016f

Observation 1bc2afc4-8fe8-44c6-a8a1-50e1494aa701 · outbound

This paper cites Nature518(7540), 529–533 (2015).

SimuDICE: Offline Policy Optimization Through World Model Updates and DICE Estimation Nature518(7540), 529–533 (2015)

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-11T19:42:10.318529Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T19:42:10.318529Z digest=sha256:798a3a3ae77920410eb288a7699b79f9efe3ee565759965f58866b3e337b3a7a

Observation 2d80d6ee-80af-4a74-a6b8-4ac62cbd84ea · outbound

This paper cites Journal of the American Statistical Associa- tion 96(456), 1410–1423 (2001).

SimuDICE: Offline Policy Optimization Through World Model Updates and DICE Estimation Journal of the American Statistical Associa- tion 96(456), 1410–1423 (2001)

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T19:42:11.191433Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-11T19:42:10.323042Z digest=sha256:2019a2798c56542191b0d4d62e5cbff9b0554d1a524ff570e26dd1a61df058e3

Observation dece943e-5bdf-4777-9766-e31ebce96df4 · outbound

This paper cites In: Advances in Neural Information Processing Systems.

SimuDICE: Offline Policy Optimization Through World Model Updates and DICE Estimation In: Advances in Neural Information Processing Systems

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T19:42:11.163423Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-11T19:42:10.327798Z digest=sha256:5f8f851be71d4ab86f3d5d62e7082dadc05b60b4d8861527dbea5ffbe540f969

Observation 8045e4a9-162c-4281-88c2-b0279028fe4e · outbound

This paper cites In: Brodley, C.E., Danyluk, A.P.

SimuDICE: Offline Policy Optimization Through World Model Updates and DICE Estimation In: Brodley, C.E., Danyluk, A.P

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T19:42:11.131171Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-11T19:42:10.332224Z digest=sha256:2404c7588556092dfae49a33c20a442b361cbc49be5e59f57f52be0c965a9602

Observation 49c848be-0d01-41b0-b351-7f9cb40f802d · outbound

This paper cites John Wiley & Sons, New York (1994).

SimuDICE: Offline Policy Optimization Through World Model Updates and DICE Estimation John Wiley & Sons, New York (1994)

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T19:42:11.108584Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-11T19:42:10.337015Z digest=sha256:41439b0c5f5b737c0cb11f0ad2d3de5a2fb85142d025021fee0b7b7bde8efd7c

Observation b20933d8-25c2-4a78-bff2-2ee9c8db376e · outbound

This paper cites Nature529(7587), 484–489 (2016).

SimuDICE: Offline Policy Optimization Through World Model Updates and DICE Estimation Nature529(7587), 484–489 (2016)

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-11T19:42:10.341398Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T19:42:10.341398Z digest=sha256:e01f5392696589d17ebadf2a7edd930b17b51007520aa90614dd6228566a850e

Observation 5d99fa4f-0f58-4a0b-8ddd-b3f4e6739738 · outbound

This paper cites Proceedings of the Seventh International Conference on Machine Learning pp.

SimuDICE: Offline Policy Optimization Through World Model Updates and DICE Estimation Proceedings of the Seventh International Conference on Machine Learning pp

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T19:42:11.091953Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-11T19:42:10.345684Z digest=sha256:c72c6ecb457d4ffca2e3c2f8fbc95cffdfaccc1f7a4c5102ca9b6c701a3ec231

Observation 4d29e233-c948-4478-b0a0-6c88b739b8b2 · outbound

This paper cites ACM SIGART Bulletin2(4), 160–163 (1991).

SimuDICE: Offline Policy Optimization Through World Model Updates and DICE Estimation ACM SIGART Bulletin2(4), 160–163 (1991)

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-11T19:42:10.349895Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T19:42:10.349895Z digest=sha256:9d99cf64c1d8b24296a86064aeaddc5d1c8d54e3099d5f37102dd9517f493a4a

Observation 97a056d3-d514-4b22-8583-21591a2ee65d · outbound

This paper cites MIT Press, Cambridge, MA, 2 edn.

SimuDICE: Offline Policy Optimization Through World Model Updates and DICE Estimation MIT Press, Cambridge, MA, 2 edn

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T19:42:11.063501Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-11T19:42:10.353994Z digest=sha256:7dd6cb2bcc6139144c64e6998ba733cca8cf4f06960eb96a0e9c0eb9a3d52874

Observation 78e12b8a-c814-4161-a79f-f307329c26cb · outbound

This paper cites https://doi.org/10.1016/j.engappai.2021.104366.

SimuDICE: Offline Policy Optimization Through World Model Updates and DICE Estimation https://doi.org/10.1016/j.engappai.2021.104366

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-11T19:42:10.358222Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T19:42:10.358222Z digest=sha256:b8d4a55eb6220ec668afadff47126fd9e6afa86bb410fea86d77a6c87cb244e6

Observation 7eaea192-8683-4ad0-9380-6c531e3cabe7 · outbound

This paper cites In: Proceedings of the 2017 IEEE/RSJ International Conference on Intelli- gent Robots and Systems (IROS).

SimuDICE: Offline Policy Optimization Through World Model Updates and DICE Estimation In: Proceedings of the 2017 IEEE/RSJ International Conference on Intelli- gent Robots and Systems (IROS)

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T19:42:11.045050Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-11T19:42:10.363370Z digest=sha256:31f8de401d0c47921eee013150ae75e30dd56e8e5a06ee25d53549b7f627b7b2

Observation 3ea4fe2d-c176-4a87-828c-ff3aeffe5add · outbound

This paper cites https://doi.org/ 10.5281/zenodo.8127026, https://zenodo.org/record/8127025.

SimuDICE: Offline Policy Optimization Through World Model Updates and DICE Estimation https://doi.org/ 10.5281/zenodo.8127026, https://zenodo.org/record/8127025

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-11T19:42:10.367616Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T19:42:10.367616Z digest=sha256:b0870e59de31281f6e4ddf5720c3e82b7af4dcc7a1d6000f9e98fe575cc7e909

Observation b00eb42e-0127-43bc-8ce0-adcf8f8b8bed · outbound

This paper cites In: Advances in Neural Information Processing Systems 9 (NIPS 1996).

SimuDICE: Offline Policy Optimization Through World Model Updates and DICE Estimation In: Advances in Neural Information Processing Systems 9 (NIPS 1996)

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T19:42:11.027373Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-11T19:42:10.371897Z digest=sha256:02d3833cceda3f91cf698127872a7302591d746463d981d89d91e3c8a3fb39cf

Observation f3a6a0fc-0ae1-46ce-b921-60853679256a · outbound

This paper cites Behavior Regularized Offline Reinforcement Learning.

SimuDICE: Offline Policy Optimization Through World Model Updates and DICE Estimation Behavior Regularized Offline Reinforcement Learning

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-11T19:42:10.376766Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T19:42:10.376766Z digest=sha256:fceb1b2a1e5f142f3e935ce0c5738e8bd5c73d015bba9490ad1f9ce5d63d3183

Observation 4a7f0cf4-38a0-4530-a354-58ecc3a5a6f7 · outbound

This paper cites In: International Conference on Learning Representa- tions.

SimuDICE: Offline Policy Optimization Through World Model Updates and DICE Estimation In: International Conference on Learning Representa- tions

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T19:42:11.008882Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-11T19:42:10.381760Z digest=sha256:710ba8b09fc084f099e5c3c4404bb4f9e6baf9b8c8968e127f61d90e5953abbc

Observation b032284e-ccd6-42e4-8147-f27bcac58350 · outbound

This paper cites In: Advances in Neural Information Processing Systems (2020).

SimuDICE: Offline Policy Optimization Through World Model Updates and DICE Estimation In: Advances in Neural Information Processing Systems (2020)

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T19:42:10.990650Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-11T19:42:10.386961Z digest=sha256:ec831487d5aa10a404983330e6fc10f0b6945c24b799a71dd1cbfea4548853c5

Observation 2bed0e1e-8349-4bf5-a6d8-8face270847d · outbound

This paper cites In: International Conference on Learning Representations (2020).

SimuDICE: Offline Policy Optimization Through World Model Updates and DICE Estimation In: International Conference on Learning Representations (2020)

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T19:42:10.971097Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-11T19:42:10.393194Z digest=sha256:4712e958cd81d23b6aeb806142f821c958a34c7b025eca19bea46e8c5f8d00e3

Observation 34fb847f-b4dc-4ae1-b526-63dd1c14575e · outbound

This paper cites GradientDICE: Rethinking Generalized Offline Estimation of Stationary Values.

SimuDICE: Offline Policy Optimization Through World Model Updates and DICE Estimation GradientDICE: Rethinking Generalized Offline Estimation of Stationary Values

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-11T19:42:10.399295Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T19:42:10.399295Z digest=sha256:cfc4bf58fe1999796a98af5e3e2ebed8dfb50bfe57b78a18af4c91e4417f9e4e

Pith citing papers

No inbound Pith citation observations are available.