Pith. sign in

Paper Citation Record · LEDGER

SimuDICE: Offline Policy Optimization Through World Model Updates and DICE Estimation

As of 13 August 2026, this Paper Citation Record lists 42 of 42 outbound references and 0 inbound Pith citation observations for arXiv:2412.06486.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2412.06486 v1

Coverage vector

measured 42 of 42 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-11T19:42:10.399295Z

measured 42 of 42 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-13T06:32:02.005865+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

42 of 42 outbound references displayed

  • verified exact0
  • verified fuzzy20
  • unresolved22
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation b947bd01-8fba-4ead-bbee-b2d651283573 · outbound

This paper cites Diffusion for World Modeling: Visual Details Matter in Atari.

SimuDICE: Offline Policy Optimization Through World Model Updates and DICE Estimation Diffusion for World Modeling: Visual Details Matter in Atari

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-11T19:42:10.189641Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T19:42:10.189641Z digest=sha256:f93b245a0013bd1efb405c976a9e75ea87f48c8a65318aa5c59a64ba6607dd98

Observation 973362bb-fa34-4211-98ee-95edc52a959e · outbound

This paper cites Machine Learning20(1-2), 65–81 (1995).

SimuDICE: Offline Policy Optimization Through World Model Updates and DICE Estimation Machine Learning20(1-2), 65–81 (1995)

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T19:42:11.378659Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-11T19:42:10.195132Z digest=sha256:6272337e127958f0f5faaf2ea2f850c748c91606f9c05483f20a4df17dab0ce0

Observation 3153c975-ffaf-4e94-b376-faec6ae384ff · outbound

This paper cites Behavioral Priors and Dynamics Models: Improving Performance and Domain Transfer in Offline RL.

SimuDICE: Offline Policy Optimization Through World Model Updates and DICE Estimation Behavioral Priors and Dynamics Models: Improving Performance and Domain Transfer in Offline RL

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-11T19:42:10.199994Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T19:42:10.199994Z digest=sha256:99fab99373ad98338ebf3db4604168195f43265ea22e7fc5f8dd775f8978b05c

Observation b1cb7d9f-9616-4e94-b017-47126842144f · outbound

This paper cites Model-Based Reinforcement Learning via Meta-Policy Optimization.

SimuDICE: Offline Policy Optimization Through World Model Updates and DICE Estimation Model-Based Reinforcement Learning via Meta-Policy Optimization

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-11T19:42:10.205189Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T19:42:10.205189Z digest=sha256:88fa62751925250c53be650d7a56492d975228822edbb6d2e1d6578bd41a436b

Observation 563fc216-1dc3-4c59-8208-49ad4e05bb89 · outbound

This paper cites In: Proceedings of the 28th International Conference on Machine Learning (ICML-11).

SimuDICE: Offline Policy Optimization Through World Model Updates and DICE Estimation In: Proceedings of the 28th International Conference on Machine Learning (ICML-11)

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T19:42:11.361030Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-11T19:42:10.210366Z digest=sha256:72131d96df60dcc968c6ea54b1ad259c0da0266c86a88f106af9e5a354e55298

Observation dfe80ac4-ac1b-4b8f-ae54-8c4ecb2dd705 · outbound

This paper cites Diffusion World Model: Future Modeling Beyond Step-by-Step Rollout for Offline Reinforcement Learning.

SimuDICE: Offline Policy Optimization Through World Model Updates and DICE Estimation Diffusion World Model: Future Modeling Beyond Step-by-Step Rollout for Offline Reinforcement Learning

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-11T19:42:10.215013Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T19:42:10.215013Z digest=sha256:d4e319488fd1fd88b2a16c81965b850038a9c96ad2ec3637079d2ed62795d088

Observation 8badf9c3-688d-42db-ad42-a6f6cbf86a5f · outbound

This paper cites Mismatched No More: Joint Model-Policy Optimization for Model-Based RL.

SimuDICE: Offline Policy Optimization Through World Model Updates and DICE Estimation Mismatched No More: Joint Model-Policy Optimization for Model-Based RL

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-11T19:42:10.220446Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T19:42:10.220446Z digest=sha256:743b411e955f99fc6dfc2cbb6088329f9cb9c965fb38e26f7723d892b77bafb4

Observation 9762876d-6ab6-4ee8-b71e-5873250cb41b · outbound

This paper cites In: Proceedings of the 36th International Conference on Ma- chine Learning.

SimuDICE: Offline Policy Optimization Through World Model Updates and DICE Estimation In: Proceedings of the 36th International Conference on Ma- chine Learning

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T19:42:11.342318Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-11T19:42:10.225768Z digest=sha256:fb9f09a1d95cbef94d44894362955595d5343c2392159604d5153882aab3a8e3

Observation 3ba53e6a-36a8-43b3-bbd8-688010ec5272 · outbound

This paper cites In: International Conference on Learning Representations (2020).

SimuDICE: Offline Policy Optimization Through World Model Updates and DICE Estimation In: International Conference on Learning Representations (2020)

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T19:42:11.318325Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-11T19:42:10.230979Z digest=sha256:1b2f797d83fbf4f24cc16ef33547034a950706f21ea1df9ecd9b0c5c4254ea07

Observation a225117c-de33-4553-9202-24b7cee3fc69 · outbound

This paper cites Mastering Diverse Domains through World Models.

SimuDICE: Offline Policy Optimization Through World Model Updates and DICE Estimation Mastering Diverse Domains through World Models

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-11T19:42:10.236664Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T19:42:10.236664Z digest=sha256:485c86fd1526e831d37f401aeb80ea619dbf29ba73017bee3476375e87426666

Observation 51f4aa3f-ce4c-4de6-82f2-cd726a5ea663 · outbound

This paper cites Mastering Atari with Discrete World Models.

SimuDICE: Offline Policy Optimization Through World Model Updates and DICE Estimation Mastering Atari with Discrete World Models

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-11T19:42:10.242391Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T19:42:10.242391Z digest=sha256:8284d494b66585405eb978f9e7dd36366446ec3dd4fc2c7892a7032c197372d3

Observation cf7c2ae0-6e26-430f-9ad0-98d6376a4ef2 · outbound

This paper cites In: Proceedings of the 34th International Conference on Machine Learning (ICML 2017).

SimuDICE: Offline Policy Optimization Through World Model Updates and DICE Estimation In: Proceedings of the 34th International Conference on Machine Learning (ICML 2017)

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T19:42:11.299303Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-11T19:42:10.248617Z digest=sha256:f7340ac9242a37e63f8212df743f5476278b786cabcffde65cd6317eadc87b13

Observation ed5f68a8-abac-43ba-b957-1c64af657cba · outbound

This paper cites Mitigating Value Hallucination in Dyna Planning via Multistep Predecessor Models.

SimuDICE: Offline Policy Optimization Through World Model Updates and DICE Estimation Mitigating Value Hallucination in Dyna Planning via Multistep Predecessor Models

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-11T19:42:10.253605Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T19:42:10.253605Z digest=sha256:e191c1c2028d567a176cd116004f7a72ba01e267ba9df25c1b4253117205b852

Observation d4782f62-e473-4094-b20c-53db1a6892db · outbound

This paper cites When to Trust Your Model: Model-Based Policy Optimization.

SimuDICE: Offline Policy Optimization Through World Model Updates and DICE Estimation When to Trust Your Model: Model-Based Policy Optimization

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-11T19:42:10.258756Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T19:42:10.258756Z digest=sha256:e94b8c02ee2a8ef1c6d3d26dc721500c25d20bbf5ae0af1a1dd905cd1980169b

Observation 89f30253-fab4-4fdf-8a4f-47572aec66c7 · outbound

This paper cites Way Off-Policy Batch Deep Reinforcement Learning of Implicit Human Preferences in Dialog.

SimuDICE: Offline Policy Optimization Through World Model Updates and DICE Estimation Way Off-Policy Batch Deep Reinforcement Learning of Implicit Human Preferences in Dialog

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-11T19:42:10.264231Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T19:42:10.264231Z digest=sha256:3d14469d6770dfb3b751d4de71db0131c49f52b48de9e1c11fc35c58b9462527

Observation eaae9a28-e4b3-4c35-97e8-7f335715b546 · outbound

This paper cites In: Advances in Neural Information Processing Sys- tems (2020).

SimuDICE: Offline Policy Optimization Through World Model Updates and DICE Estimation In: Advances in Neural Information Processing Sys- tems (2020)

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T19:42:11.279370Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-11T19:42:10.269370Z digest=sha256:e51af8aa726cb54799388249c01905fece926c4e8d53505781c0ab7903816de6

Observation 494cc83e-6a33-4642-91a6-3bd7bfcf1a35 · outbound

This paper cites Offline Reinforcement Learning with Implicit Q-Learning.

SimuDICE: Offline Policy Optimization Through World Model Updates and DICE Estimation Offline Reinforcement Learning with Implicit Q-Learning

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-11T19:42:10.274241Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T19:42:10.274241Z digest=sha256:68901b63cde1eaaae2b5ca621f767f7c042f4b68f7ad1e25e247d406e850b3ab

Observation bec1f690-7e6b-4a96-9787-b99b1990e7c5 · outbound

This paper cites In: Wallach, H., Larochelle, H., Beygelzimer, A., d’Alché Buc, F., Fox, E., Garnett, R.

SimuDICE: Offline Policy Optimization Through World Model Updates and DICE Estimation In: Wallach, H., Larochelle, H., Beygelzimer, A., d’Alché Buc, F., Fox, E., Garnett, R

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T19:42:11.261397Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-11T19:42:10.279461Z digest=sha256:aaa178853e32bae0e8a3cbdeb0ff4942b94176cdddf6f9d40ea273eca4409723

Observation 9fa1fdb6-41ef-424f-916b-e2d5dbe35155 · outbound

This paper cites In: Advances in Neural Information Processing Systems.

SimuDICE: Offline Policy Optimization Through World Model Updates and DICE Estimation In: Advances in Neural Information Processing Systems

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T19:42:11.242411Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-11T19:42:10.284609Z digest=sha256:0f0794c1f0d955823876138c7077ddfc32a9e93d78fec3b15f410533ecfc40d7

Observation fe4f6e9f-1fa6-47f2-95b8-5fe7820a86a0 · outbound

This paper cites Model-Ensemble Trust-Region Policy Optimization.

SimuDICE: Offline Policy Optimization Through World Model Updates and DICE Estimation Model-Ensemble Trust-Region Policy Optimization

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-11T19:42:10.289841Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T19:42:10.289841Z digest=sha256:4044d46d4568577afff144287cbfdf0ff7bae72bc4b60966be4d194f319c4eb0

Observation ca2da1fb-b1eb-40b7-8637-eaca30cb2d68 · outbound

This paper cites Objective Mismatch in Model-based Reinforcement Learning.

SimuDICE: Offline Policy Optimization Through World Model Updates and DICE Estimation Objective Mismatch in Model-based Reinforcement Learning

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-11T19:42:10.295272Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T19:42:10.295272Z digest=sha256:c7d0836d828b8c988c26cbb7f59c50e6de81163ece896f3ff8f23c9ef0eb2529

Observation 9d147b65-0ff4-4a7a-acfb-6baa4265167e · outbound

This paper cites In: Wiering, M., van Otterlo, M.

SimuDICE: Offline Policy Optimization Through World Model Updates and DICE Estimation In: Wiering, M., van Otterlo, M

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-11T19:42:10.300617Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T19:42:10.300617Z digest=sha256:c4116d2e68541e13843084915225d11d5fe9720a2bf16a3a75971a45adb4ae45

Observation 5b82b523-a9ca-4c87-b3fe-900a634d6cb1 · outbound

This paper cites Offline Reinforcement Learning: Tutorial, Review, and Perspectives on Open Problems.

SimuDICE: Offline Policy Optimization Through World Model Updates and DICE Estimation Offline Reinforcement Learning: Tutorial, Review, and Perspectives on Open Problems

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-11T19:42:10.309676Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T19:42:10.309676Z digest=sha256:01e9bbc0c1bf65820486dfdaab0f5b519f75e6b9b4b9656833fb393cf662c5a6

Observation 0f14a637-6785-4a94-9ae0-1334d7a4db8a · outbound

This paper cites In: Advances in Neural Information Processing Systems 31 (NeurIPS 2018).

SimuDICE: Offline Policy Optimization Through World Model Updates and DICE Estimation In: Advances in Neural Information Processing Systems 31 (NeurIPS 2018)

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T19:42:11.223196Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-11T19:42:10.314105Z digest=sha256:ca3f8bd8a11bbc6c2f4b4f4078cf6a027839f8f3f31b3dfaeb356a8913c0228d

Observation 1bc2afc4-8fe8-44c6-a8a1-50e1494aa701 · outbound

This paper cites Nature518(7540), 529–533 (2015).

SimuDICE: Offline Policy Optimization Through World Model Updates and DICE Estimation Nature518(7540), 529–533 (2015)

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-11T19:42:10.318529Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T19:42:10.318529Z digest=sha256:798a3a3ae77920410eb288a7699b79f9efe3ee565759965f58866b3e337b3a7a

Observation 2d80d6ee-80af-4a74-a6b8-4ac62cbd84ea · outbound

This paper cites Journal of the American Statistical Associa- tion 96(456), 1410–1423 (2001).

SimuDICE: Offline Policy Optimization Through World Model Updates and DICE Estimation Journal of the American Statistical Associa- tion 96(456), 1410–1423 (2001)

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T19:42:11.191433Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-11T19:42:10.323042Z digest=sha256:a6f689ccab5fa968187316be5b28c1601cc20a91b7220a4f139cdf80ad88e8f9

Observation dece943e-5bdf-4777-9766-e31ebce96df4 · outbound

This paper cites In: Advances in Neural Information Processing Systems.

SimuDICE: Offline Policy Optimization Through World Model Updates and DICE Estimation In: Advances in Neural Information Processing Systems

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T19:42:11.163423Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-11T19:42:10.327798Z digest=sha256:a8f81e9a65de110e9773dff2de1a0d78dd26406ddd71f638be46fb9c6153f6c8

Observation 8045e4a9-162c-4281-88c2-b0279028fe4e · outbound

This paper cites In: Brodley, C.E., Danyluk, A.P.

SimuDICE: Offline Policy Optimization Through World Model Updates and DICE Estimation In: Brodley, C.E., Danyluk, A.P

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T19:42:11.131171Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-11T19:42:10.332224Z digest=sha256:76c020137487c6e88ee0664d2cfbed33c8f57665fbb97649e3cd0af34671ef18

Observation 49c848be-0d01-41b0-b351-7f9cb40f802d · outbound

This paper cites John Wiley & Sons, New York (1994).

SimuDICE: Offline Policy Optimization Through World Model Updates and DICE Estimation John Wiley & Sons, New York (1994)

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T19:42:11.108584Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-11T19:42:10.337015Z digest=sha256:170981f2ae7a19613cbb089d4382973b69dfbc5f58b226a165d2836560713365

Observation b20933d8-25c2-4a78-bff2-2ee9c8db376e · outbound

This paper cites Nature529(7587), 484–489 (2016).

SimuDICE: Offline Policy Optimization Through World Model Updates and DICE Estimation Nature529(7587), 484–489 (2016)

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-11T19:42:10.341398Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T19:42:10.341398Z digest=sha256:e01f5392696589d17ebadf2a7edd930b17b51007520aa90614dd6228566a850e

Observation 5d99fa4f-0f58-4a0b-8ddd-b3f4e6739738 · outbound

This paper cites Proceedings of the Seventh International Conference on Machine Learning pp.

SimuDICE: Offline Policy Optimization Through World Model Updates and DICE Estimation Proceedings of the Seventh International Conference on Machine Learning pp

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T19:42:11.091953Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-11T19:42:10.345684Z digest=sha256:b000099daa4b554e1eb10647693a79c367054a6e2eab9b186850588a14fa5a0c

Observation 4d29e233-c948-4478-b0a0-6c88b739b8b2 · outbound

This paper cites ACM SIGART Bulletin2(4), 160–163 (1991).

SimuDICE: Offline Policy Optimization Through World Model Updates and DICE Estimation ACM SIGART Bulletin2(4), 160–163 (1991)

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-11T19:42:10.349895Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T19:42:10.349895Z digest=sha256:9d99cf64c1d8b24296a86064aeaddc5d1c8d54e3099d5f37102dd9517f493a4a

Observation 97a056d3-d514-4b22-8583-21591a2ee65d · outbound

This paper cites MIT Press, Cambridge, MA, 2 edn.

SimuDICE: Offline Policy Optimization Through World Model Updates and DICE Estimation MIT Press, Cambridge, MA, 2 edn

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T19:42:11.063501Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-11T19:42:10.353994Z digest=sha256:db68aa7502a1895369553ae4aeb375690b3732800bc2b03f8ca183525be908a1

Observation 78e12b8a-c814-4161-a79f-f307329c26cb · outbound

This paper cites https://doi.org/10.1016/j.engappai.2021.104366.

SimuDICE: Offline Policy Optimization Through World Model Updates and DICE Estimation https://doi.org/10.1016/j.engappai.2021.104366

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-11T19:42:10.358222Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T19:42:10.358222Z digest=sha256:b8d4a55eb6220ec668afadff47126fd9e6afa86bb410fea86d77a6c87cb244e6

Observation 7eaea192-8683-4ad0-9380-6c531e3cabe7 · outbound

This paper cites In: Proceedings of the 2017 IEEE/RSJ International Conference on Intelli- gent Robots and Systems (IROS).

SimuDICE: Offline Policy Optimization Through World Model Updates and DICE Estimation In: Proceedings of the 2017 IEEE/RSJ International Conference on Intelli- gent Robots and Systems (IROS)

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T19:42:11.045050Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-11T19:42:10.363370Z digest=sha256:e5f096b6f2ab02c80e0d88e1d8c43305e54600357b5467b4c9b79e89b4101870

Observation 3ea4fe2d-c176-4a87-828c-ff3aeffe5add · outbound

This paper cites https://doi.org/ 10.5281/zenodo.8127026, https://zenodo.org/record/8127025.

SimuDICE: Offline Policy Optimization Through World Model Updates and DICE Estimation https://doi.org/ 10.5281/zenodo.8127026, https://zenodo.org/record/8127025

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-11T19:42:10.367616Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T19:42:10.367616Z digest=sha256:b0870e59de31281f6e4ddf5720c3e82b7af4dcc7a1d6000f9e98fe575cc7e909

Observation b00eb42e-0127-43bc-8ce0-adcf8f8b8bed · outbound

This paper cites In: Advances in Neural Information Processing Systems 9 (NIPS 1996).

SimuDICE: Offline Policy Optimization Through World Model Updates and DICE Estimation In: Advances in Neural Information Processing Systems 9 (NIPS 1996)

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T19:42:11.027373Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-11T19:42:10.371897Z digest=sha256:85732548421a2c6155d79e32327fc5f2bb719207e866bfcb0ea4e601901ede6f

Observation f3a6a0fc-0ae1-46ce-b921-60853679256a · outbound

This paper cites Behavior Regularized Offline Reinforcement Learning.

SimuDICE: Offline Policy Optimization Through World Model Updates and DICE Estimation Behavior Regularized Offline Reinforcement Learning

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-11T19:42:10.376766Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T19:42:10.376766Z digest=sha256:fceb1b2a1e5f142f3e935ce0c5738e8bd5c73d015bba9490ad1f9ce5d63d3183

Observation 4a7f0cf4-38a0-4530-a354-58ecc3a5a6f7 · outbound

This paper cites In: International Conference on Learning Representa- tions.

SimuDICE: Offline Policy Optimization Through World Model Updates and DICE Estimation In: International Conference on Learning Representa- tions

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T19:42:11.008882Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-11T19:42:10.381760Z digest=sha256:1da67e24ca916fdfeec0c7b1f351182040f5255accc52e2d91a39341a5f8fdc7

Observation b032284e-ccd6-42e4-8147-f27bcac58350 · outbound

This paper cites In: Advances in Neural Information Processing Systems (2020).

SimuDICE: Offline Policy Optimization Through World Model Updates and DICE Estimation In: Advances in Neural Information Processing Systems (2020)

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T19:42:10.990650Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-11T19:42:10.386961Z digest=sha256:1e6aeefe289582614efa5c43f0963f80c15497cff84f5abdf03978074d14c231

Observation 2bed0e1e-8349-4bf5-a6d8-8face270847d · outbound

This paper cites In: International Conference on Learning Representations (2020).

SimuDICE: Offline Policy Optimization Through World Model Updates and DICE Estimation In: International Conference on Learning Representations (2020)

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T19:42:10.971097Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-11T19:42:10.393194Z digest=sha256:56f91655831a8b8e4a3a9769c40b03ff77856b1453a89ae3ffd608f69019781a

Observation 34fb847f-b4dc-4ae1-b526-63dd1c14575e · outbound

This paper cites GradientDICE: Rethinking Generalized Offline Estimation of Stationary Values.

SimuDICE: Offline Policy Optimization Through World Model Updates and DICE Estimation GradientDICE: Rethinking Generalized Offline Estimation of Stationary Values

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-11T19:42:10.399295Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T19:42:10.399295Z digest=sha256:cfc4bf58fe1999796a98af5e3e2ebed8dfb50bfe57b78a18af4c91e4417f9e4e

Pith citing papers

No inbound Pith citation observations are available.