Pith. sign in

Paper Citation Record · LEDGER

Fast and Robust: Task Sampling with Posterior and Diversity Synergies for Adaptive Decision-Makers in Randomized Environments

As of 18 August 2026, this Paper Citation Record lists 100 of 100 outbound references and 0 inbound Pith citation observations for arXiv:2504.19139.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2504.19139 v3

Coverage vector

measured 100 of 100 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-16T06:07:19.785138Z

measured 100 of 100 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-18T06:34:40.430872+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

100 of 100 outbound references displayed

  • verified exact2
  • verified fuzzy34
  • unresolved64
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 09e1139d-b60f-407b-8b04-691ecdb6c77c · outbound

This paper cites Sharp-maml: Sharpness-aware model-agnostic meta learning.

Fast and Robust: Task Sampling with Posterior and Diversity Synergies for Adaptive Decision-Makers in Randomized Environments Sharp-maml: Sharpness-aware model-agnostic meta learning

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-16T06:07:19.454225Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T06:07:19.454225Z digest=sha256:e6efac0ce81b882dc9c808404b03c9b4ce51bba53a54f530a5082309e6c22201

Observation a7338292-d05e-4a24-9bbc-194edf99542e · outbound

This paper cites A Bayesian Sampling Approach to Exploration in Reinforcement Learning.

Fast and Robust: Task Sampling with Posterior and Diversity Synergies for Adaptive Decision-Makers in Randomized Environments A Bayesian Sampling Approach to Exploration in Reinforcement Learning

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-16T06:07:19.458774Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T06:07:19.458774Z digest=sha256:6d5ff4e0c5c08dd4f9029f0ad5f9015b410adc6a96e780d0ab12e1cf7855dc63

Observation f3958e16-e579-48bd-a001-efba37863483 · outbound

This paper cites Finite-time analysis of the multiarmed bandit problem, 2002 a.

Fast and Robust: Task Sampling with Posterior and Diversity Synergies for Adaptive Decision-Makers in Randomized Environments Finite-time analysis of the multiarmed bandit problem, 2002 a

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-16T06:07:19.462707Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T06:07:19.462707Z digest=sha256:a0e34f4143effa1453f57d51a0c169f9a79c43cfa3a517345d94a75ddcdb311a

Observation a9393637-ad10-475d-b595-a2952282abb0 · outbound

This paper cites Using confidence bounds for exploitation-exploration trade-offs.

Fast and Robust: Task Sampling with Posterior and Diversity Synergies for Adaptive Decision-Makers in Randomized Environments Using confidence bounds for exploitation-exploration trade-offs

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-16T06:07:19.466308Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T06:07:19.466308Z digest=sha256:148e720c6ba8159b458512adb1b47e9e49004dd0a93ab3076193554c2cde3393

Observation 88de1a95-ef77-40c5-a89c-ebfc37edfea0 · outbound

This paper cites A Tutorial on Meta-Reinforcement Learning.

Fast and Robust: Task Sampling with Posterior and Diversity Synergies for Adaptive Decision-Makers in Randomized Environments A Tutorial on Meta-Reinforcement Learning

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-16T06:07:19.469725Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T06:07:19.469725Z digest=sha256:6a2e0146f8cc9e8950cb1bd2fabea6ef6ecd4176d782a3dca3b423dfe210dbd6

Observation 1b58aae1-a71b-435b-941b-1b2545c4c55f · outbound

This paper cites C., and Ye, Y.

Fast and Robust: Task Sampling with Posterior and Diversity Synergies for Adaptive Decision-Makers in Randomized Environments C., and Ye, Y

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-16T06:07:19.473581Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T06:07:19.473581Z digest=sha256:8b64b4fc77a2fb38889f8245fb23029e993a7fcf445493f13d26d91d1fd1510b

Observation db0e3d95-27dc-4f8f-9028-b0b522d7e5ba · outbound

This paper cites C., and Jordan, M.

Fast and Robust: Task Sampling with Posterior and Diversity Synergies for Adaptive Decision-Makers in Randomized Environments C., and Jordan, M

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-16T06:07:19.477281Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T06:07:19.477281Z digest=sha256:18c87bfed47c082d39d2e61b68efb56715ccfe5550b2ddc66c0fc48cb51e9dc3

Observation 09bf18a0-b6c5-4008-b68c-004d67679998 · outbound

This paper cites On Evaluating Adversarial Robustness.

Fast and Robust: Task Sampling with Posterior and Diversity Synergies for Adaptive Decision-Makers in Randomized Environments On Evaluating Adversarial Robustness

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-16T06:07:19.480792Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T06:07:19.480792Z digest=sha256:cf056856ccc339e3c6f44c5aac6fcfebea35bfdad49d2eaf1257fd60c94b443f

Observation e9d098ea-8a69-4c04-959b-5272c941ddc2 · outbound

This paper cites and Valko, M.

Fast and Robust: Task Sampling with Posterior and Diversity Synergies for Adaptive Decision-Makers in Randomized Environments and Valko, M

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-16T06:07:19.484656Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T06:07:19.484656Z digest=sha256:fb894c6831ab2725e48336749dd6c40b66835550bf7699e95ba6ab3de84b462d

Observation 084e88be-a8f7-4f9b-84ec-5a9d8acf0de6 · outbound

This paper cites Risk aversion in finite markov decision processes using total cost criteria and average value at risk.

Fast and Robust: Task Sampling with Posterior and Diversity Synergies for Adaptive Decision-Makers in Randomized Environments Risk aversion in finite markov decision processes using total cost criteria and average value at risk

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-16T06:07:19.487852Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T06:07:19.487852Z digest=sha256:80168bf4fa845a070b54b8cdff3c68a7033c7c80da4e995eed88749d8f4fea84

Observation ee3a6edc-370d-40bb-91bf-9252012b9bba · outbound

This paper cites Box2d: A 2d physics engine for games, 2007.

Fast and Robust: Task Sampling with Posterior and Diversity Synergies for Adaptive Decision-Makers in Randomized Environments Box2d: A 2d physics engine for games, 2007

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-16T06:07:19.490989Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T06:07:19.490989Z digest=sha256:596d4f326b37c4374bd47a7ff3957b2cddbe23c2a521d0a78b3456b15d1774d3

Observation 5aa9a946-f255-45e9-9c9d-39b50762b71c · outbound

This paper cites Tohan: A one-step approach towards few-shot hypothesis adaptation.

Fast and Robust: Task Sampling with Posterior and Diversity Synergies for Adaptive Decision-Makers in Randomized Environments Tohan: A one-step approach towards few-shot hypothesis adaptation

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-16T06:07:19.493963Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T06:07:19.493963Z digest=sha256:7094bd131309417b2173694bf5e30ddb431da42e600ab1032efd12043963859e

Observation 10390f7b-7b0d-4a0a-82da-9448cc321188 · outbound

This paper cites Unveiling causal reasoning in large language models: Reality or mirage? Advances in Neural Information Processing Systems, 37: 0 96640--96670, 2024 a.

Fast and Robust: Task Sampling with Posterior and Diversity Synergies for Adaptive Decision-Makers in Randomized Environments Unveiling causal reasoning in large language models: Reality or mirage? Advances in Neural Information Processing Systems, 37: 0 96640--96670, 2024 a

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-16T06:07:19.497136Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T06:07:19.497136Z digest=sha256:eb86a4804a65d7aa27ecf73add0f9c623225e841a22b34dfa2f2038868829ace

Observation 5ae36d18-d5c2-4ac7-a337-b1d2f9fa7f17 · outbound

This paper cites Does confusion really hurt novel class discovery? International Journal of Computer Vision, 132 0 (8): 0 3191--3207, 2024 b.

Fast and Robust: Task Sampling with Posterior and Diversity Synergies for Adaptive Decision-Makers in Randomized Environments Does confusion really hurt novel class discovery? International Journal of Computer Vision, 132 0 (8): 0 3191--3207, 2024 b

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-16T06:07:19.500290Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T06:07:19.500290Z digest=sha256:9cf4e4f18e39bf4b828feac51913cbee758579eaa526fd51812469aec88cd1ec

Observation 5a52b742-1efd-446d-9d0c-e764011016c3 · outbound

This paper cites Risk-sensitive and data-driven sequential decision making.

Fast and Robust: Task Sampling with Posterior and Diversity Synergies for Adaptive Decision-Makers in Randomized Environments Risk-sensitive and data-driven sequential decision making

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-16T06:07:19.503511Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T06:07:19.503511Z digest=sha256:605a426dadbfd99e5acd1f7c7e1cbee24f7479ac34f7d0e2ea1f99ce2201fa3e

Observation d48cba5d-a026-457d-bfcf-fbdb0540a6a3 · outbound

This paper cites Risk-sensitive and robust decision-making: a cvar optimization approach.

Fast and Robust: Task Sampling with Posterior and Diversity Synergies for Adaptive Decision-Makers in Randomized Environments Risk-sensitive and robust decision-making: a cvar optimization approach

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-16T06:07:19.506792Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T06:07:19.506792Z digest=sha256:08453c1bfe622cc567edba323decf3b0219a7f3cd757b366838bb32c3eeb7768

Observation 3d165292-de38-4999-ac7a-796108c24444 · outbound

This paper cites Risk-constrained reinforcement learning with percentile risk criteria.

Fast and Robust: Task Sampling with Posterior and Diversity Synergies for Adaptive Decision-Makers in Randomized Environments Risk-constrained reinforcement learning with percentile risk criteria

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-16T06:07:19.510159Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T06:07:19.510159Z digest=sha256:af4387981f97bb0eabefac466ae454b267bd947dcab8b2ad03105c5aafa1c046

Observation 08b77fd6-6cf0-4907-8477-1097cdfaff34 · outbound

This paper cites Safe policy learning for continuous control.

Fast and Robust: Task Sampling with Posterior and Diversity Synergies for Adaptive Decision-Makers in Randomized Environments Safe policy learning for continuous control

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-16T06:07:19.513287Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T06:07:19.513287Z digest=sha256:9036df24fda20ed0c458b05780435f295d9a78d4eb9df4530ebef76e95b1538e

Observation ad555537-0d1a-4d8e-85dd-5c9e0045ab09 · outbound

This paper cites A., Ghahramani, Z., and Jordan, M.

Fast and Robust: Task Sampling with Posterior and Diversity Synergies for Adaptive Decision-Makers in Randomized Environments A., Ghahramani, Z., and Jordan, M

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-16T06:07:19.516464Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T06:07:19.516464Z digest=sha256:337c99333e03251742bd9eb2c3ae40b595d34a46e4507ee03104357025b59e0f

Observation 520db437-9c7b-45ba-a628-5fb89d6311fd · outbound

This paper cites Task-robust model-agnostic meta-learning.

Fast and Robust: Task Sampling with Posterior and Diversity Synergies for Adaptive Decision-Makers in Randomized Environments Task-robust model-agnostic meta-learning

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-16T06:07:19.519552Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T06:07:19.519552Z digest=sha256:af6e5d55433078aaf24b9dbc3d0564d9e23842ff41e676b1794a89e2d4efa9ea

Observation 2f599618-15bf-4f0a-b9e9-82c98d327771 · outbound

This paper cites Bullet physics simulation.

Fast and Robust: Task Sampling with Posterior and Diversity Synergies for Adaptive Decision-Makers in Randomized Environments Bullet physics simulation

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-16T06:07:19.522707Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T06:07:19.522707Z digest=sha256:8fb4bb07bf8020d3d8ab9f932d638879022bace3e263be65295a622e7c42db42

Observation cfb1dac3-b75c-4f21-841a-4174aca1c1e3 · outbound

This paper cites Emergent complexity and zero-shot transfer via unsupervised environment design.

Fast and Robust: Task Sampling with Posterior and Diversity Synergies for Adaptive Decision-Makers in Randomized Environments Emergent complexity and zero-shot transfer via unsupervised environment design

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-16T06:07:19.525933Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T06:07:19.525933Z digest=sha256:67095e313e4bc7052a9677c7e3f5755b7c72a613bbb957510d2dea75d733b2b7

Observation c7dbbb5e-b488-41c1-91c5-920cb78e1cc4 · outbound

This paper cites RL$^2$: Fast Reinforcement Learning via Slow Reinforcement Learning.

Fast and Robust: Task Sampling with Posterior and Diversity Synergies for Adaptive Decision-Makers in Randomized Environments RL$^2$: Fast Reinforcement Learning via Slow Reinforcement Learning

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-16T06:07:19.529048Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T06:07:19.529048Z digest=sha256:217e966135e467248731c4d361bc274328f9e0df3622d15f46ce81e047680951

Observation 0bf6c1c1-77c8-416d-8117-80618e674125 · outbound

This paper cites Model-agnostic meta-learning for fast adaptation of deep networks.

Fast and Robust: Task Sampling with Posterior and Diversity Synergies for Adaptive Decision-Makers in Randomized Environments Model-agnostic meta-learning for fast adaptation of deep networks

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-16T06:07:19.532581Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T06:07:19.532581Z digest=sha256:2b8c8399b0037e7fd30df598562d25d2abdb722d71d95afbe13f464ee7581ccb

Observation 4ba7b02f-826c-411a-a9fd-26ad116d33db · outbound

This paper cites Probabilistic model-agnostic meta-learning.

Fast and Robust: Task Sampling with Posterior and Diversity Synergies for Adaptive Decision-Makers in Randomized Environments Probabilistic model-agnostic meta-learning

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T06:07:20.587274Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-16T06:07:19.535863Z digest=sha256:71437e6d3017dc0a6de01bc0afdb309b066fd23230f98ce226faf918fe4b2adf

Observation 3e10efdc-50e3-4987-acf9-069324e04e40 · outbound

This paper cites Addressing function approximation error in actor-critic methods.

Fast and Robust: Task Sampling with Posterior and Diversity Synergies for Adaptive Decision-Makers in Randomized Environments Addressing function approximation error in actor-critic methods

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-16T06:07:19.539079Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T06:07:19.539079Z digest=sha256:07ad16215888bb89ddae23e815ac6b5dfa859131be81721eae51ca6699080e0d

Observation 73b4ff04-48e1-4b27-bccc-7b3ad2b1462c · outbound

This paper cites Deep bayesian active learning with image data.

Fast and Robust: Task Sampling with Posterior and Diversity Synergies for Adaptive Decision-Makers in Randomized Environments Deep bayesian active learning with image data

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-16T06:07:19.542165Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T06:07:19.542165Z digest=sha256:9585e2f2c392f0c085613418dbb5b922ea91ad9dffba2f249099c372d2246c12

Observation c013a9d6-e318-4622-9214-ca36fda4c9c0 · outbound

This paper cites On Upper-Confidence Bound Policies for Non-Stationary Bandit Problems.

Fast and Robust: Task Sampling with Posterior and Diversity Synergies for Adaptive Decision-Makers in Randomized Environments On Upper-Confidence Bound Policies for Non-Stationary Bandit Problems

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-16T06:07:19.545422Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T06:07:19.545422Z digest=sha256:f8710e7080a7d210ee2cfa35e5f59e0b3c1548b603db61ef7f77e7871d5394ba

Observation 39d67b2e-5b1a-4a58-8bd8-dbf2cecdf521 · outbound

This paper cites Neural Processes.

Fast and Robust: Task Sampling with Posterior and Diversity Synergies for Adaptive Decision-Makers in Randomized Environments Neural Processes

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-16T06:07:19.549045Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T06:07:19.549045Z digest=sha256:7bbebd6c6c434b4736defa9a63225211dd5fe6a0a95b88d2bc12c1cece16cb30

Observation db9426d3-cb51-49c4-bda5-bc3087bb20c4 · outbound

This paper cites W., Gast, J., Ruiz, I.

Fast and Robust: Task Sampling with Posterior and Diversity Synergies for Adaptive Decision-Makers in Randomized Environments W., Gast, J., Ruiz, I

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-16T06:07:19.552344Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T06:07:19.552344Z digest=sha256:adf5df8b6fd3355904bd6e81f8f01c490bfe514992ab0ab32db7e2c83122acb9

Observation ecc75b9d-4c65-4272-868a-e61485fe8737 · outbound

This paper cites Efficient risk-averse reinforcement learning.

Fast and Robust: Task Sampling with Posterior and Diversity Synergies for Adaptive Decision-Makers in Randomized Environments Efficient risk-averse reinforcement learning

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-16T06:07:19.555550Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T06:07:19.555550Z digest=sha256:ad5f5e3bd718ae4e7e96da899eae897b8fb4ba4e6450c9a0c834220c6b6ee4ae

Observation e8b401d3-9702-4d82-bef4-917a36cf50af · outbound

This paper cites Train hard, fight easy: Robust meta reinforcement learning.

Fast and Robust: Task Sampling with Posterior and Diversity Synergies for Adaptive Decision-Makers in Randomized Environments Train hard, fight easy: Robust meta reinforcement learning

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T06:07:20.553993Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-16T06:07:19.558873Z digest=sha256:e57ec108db22a8c7be928743e5c04e44155eb9956737f2eda7cb5d7d876ebe53

Observation 72f301d9-fb70-4907-b8f6-859b32073cea · outbound

This paper cites Meta-reinforcement learning of structured exploration strategies.

Fast and Robust: Task Sampling with Posterior and Diversity Synergies for Adaptive Decision-Makers in Randomized Environments Meta-reinforcement learning of structured exploration strategies

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T06:07:20.544201Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-16T06:07:19.562580Z digest=sha256:1e0cb84f35214e3b814c62b5c484bcb8a6c50f46033a3565be36e72638ec8fec

Observation 8809c028-b3aa-4805-b1b9-8c5a26198c5d · outbound

This paper cites Beyond uniform sampling: Offline reinforcement learning with imbalanced datasets.

Fast and Robust: Task Sampling with Posterior and Diversity Synergies for Adaptive Decision-Makers in Randomized Environments Beyond uniform sampling: Offline reinforcement learning with imbalanced datasets

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T06:07:20.533631Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-16T06:07:19.565645Z digest=sha256:6150069291240b35f0705c467c696f330ba524cd9cccc02fa02d9bbd3c761967

Observation a6e17ffc-edd1-4c2a-8251-2376183c8832 · outbound

This paper cites Meta-learning in neural networks: A survey.

Fast and Robust: Task Sampling with Posterior and Diversity Synergies for Adaptive Decision-Makers in Randomized Environments Meta-learning in neural networks: A survey

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-16T06:07:19.568751Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T06:07:19.568751Z digest=sha256:043a5d368e13c6551b8e6ef58e5853319d3582fcd40e30b0949c969c1ad4995a

Observation f92a9204-8118-4b48-8ea2-1ee27e65fead · outbound

This paper cites Replay-guided adversarial environment design.

Fast and Robust: Task Sampling with Posterior and Diversity Synergies for Adaptive Decision-Makers in Randomized Environments Replay-guided adversarial environment design

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T06:07:20.517773Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-16T06:07:19.571839Z digest=sha256:222b9095d9e0e0cbaf09c871a1357999cb4a0c328f03748c74ea530eb460db89

Observation 848061a3-a3bb-463d-ba5f-2ee24d7fc1a3 · outbound

This paper cites Auto-Encoding Variational Bayes.

Fast and Robust: Task Sampling with Posterior and Diversity Synergies for Adaptive Decision-Makers in Randomized Environments Auto-Encoding Variational Bayes

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-16T06:07:19.575103Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T06:07:19.575103Z digest=sha256:18a1c3fcb53b9f9ec111fda2309d3f1d31efed7f9bb334540fa0130a50fa6f39

Observation 8b2b7e8e-e1f6-42e9-bbe1-db88d8b52e17 · outbound

This paper cites Batchbald: Efficient and diverse batch acquisition for deep bayesian active learning.

Fast and Robust: Task Sampling with Posterior and Diversity Synergies for Adaptive Decision-Makers in Randomized Environments Batchbald: Efficient and diverse batch acquisition for deep bayesian active learning

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T06:07:20.507713Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-16T06:07:19.578658Z digest=sha256:df9011ef3f86cdf5c1f9d6d34ed830c324b1edaf88d9e9e8941ef4720c2aa4d1

Observation 40efd838-94f0-4597-b41c-b036314ff133 · outbound

This paper cites W., Sagawa, S., Marklund, H., Xie, S.

Fast and Robust: Task Sampling with Posterior and Diversity Synergies for Adaptive Decision-Makers in Randomized Environments W., Sagawa, S., Marklund, H., Xie, S

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-16T06:07:19.581994Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T06:07:19.581994Z digest=sha256:3eab884afd09c5fa07c2983666a7be87cd327e0356af2c891252416e12e09e6c

Observation 628fe51d-540b-4f10-82bd-dba56e076eb1 · outbound

This paper cites D., Jansen, N., and Topcu, U.

Fast and Robust: Task Sampling with Posterior and Diversity Synergies for Adaptive Decision-Makers in Randomized Environments D., Jansen, N., and Topcu, U

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T06:07:20.491172Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-16T06:07:19.585184Z digest=sha256:f160872c0dd51487a40797fca82396a4374af5ebcdfc5c4fe78e9b3ccea81789

Observation 36d99476-cedd-4c8e-be10-4f532acf5e6f · outbound

This paper cites Offline Reinforcement Learning: Tutorial, Review, and Perspectives on Open Problems.

Fast and Robust: Task Sampling with Posterior and Diversity Synergies for Adaptive Decision-Makers in Randomized Environments Offline Reinforcement Learning: Tutorial, Review, and Perspectives on Open Problems

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-16T06:07:19.588739Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T06:07:19.588739Z digest=sha256:dda4344281b5b86fa3bf199e9cbcff1d869e44aeccfa6c899743d38df0075bc8

Observation 181ab4fc-e272-40ef-9092-d7ec4a0f3549 · outbound

This paper cites FOCAL: Efficient Fully-Offline Meta-Reinforcement Learning via Distance Metric Learning and Behavior Regularization.

Fast and Robust: Task Sampling with Posterior and Diversity Synergies for Adaptive Decision-Makers in Randomized Environments FOCAL: Efficient Fully-Offline Meta-Reinforcement Learning via Distance Metric Learning and Behavior Regularization

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-16T06:07:19.592312Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T06:07:19.592312Z digest=sha256:b4d84c81c903044ddb2bfba1e91a1783b99831e68fe49b4df2744e379bd8f474

Observation 211c40b9-32db-49a9-b287-3d629f91d480 · outbound

This paper cites an unresolved cited work.

Fast and Robust: Task Sampling with Posterior and Diversity Synergies for Adaptive Decision-Makers in Randomized Environments Unresolved cited work

Reference 43

Resolution
unresolved
raw_fallback, observed 2026-08-16T06:07:20.480834Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-16T06:07:19.595751Z digest=sha256:3beb1f906155672ce24c95fa242a7a725c8358c0ec3b4ab47c21d678ebf36e5f

Observation 59bbb7de-cfa1-4f46-b92c-7bfc49fc81bb · outbound

This paper cites Theoretical investigations and practical enhancements on tail task risk minimization in meta learning.

Fast and Robust: Task Sampling with Posterior and Diversity Synergies for Adaptive Decision-Makers in Randomized Environments Theoretical investigations and practical enhancements on tail task risk minimization in meta learning

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T06:07:20.469952Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-16T06:07:19.598914Z digest=sha256:4134eebb9b0b98eb846d98087cd189bfaa732b217dcbbfa91a236adc0c15fe0d

Observation 37f86d4f-c7ec-4872-b1b3-e52a5dfa7ce8 · outbound

This paper cites DrEureka: Language Model Guided Sim-To-Real Transfer.

Fast and Robust: Task Sampling with Posterior and Diversity Synergies for Adaptive Decision-Makers in Randomized Environments DrEureka: Language Model Guided Sim-To-Real Transfer

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-16T06:07:19.602087Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T06:07:19.602087Z digest=sha256:756581a262ceb9e920a93fb7e246a0964b118dce8727057d3e53b6ef40671ac2

Observation 932ccd72-ef9f-4421-85aa-dc657993df71 · outbound

This paper cites and Teneketzis, D.

Fast and Robust: Task Sampling with Posterior and Diversity Synergies for Adaptive Decision-Makers in Randomized Environments and Teneketzis, D

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-16T06:07:19.605532Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T06:07:19.605532Z digest=sha256:88ee69ddaefd4549a79a2eac7e9fb9233978547a1946634b02842480d887dcfc

Observation 5bd5de1a-bfa4-409b-8ba8-171371fcb1a5 · outbound

This paper cites Supported value regularization for offline reinforcement learning.

Fast and Robust: Task Sampling with Posterior and Diversity Synergies for Adaptive Decision-Makers in Randomized Environments Supported value regularization for offline reinforcement learning

Reference 47

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T06:07:20.453803Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-16T06:07:19.609039Z digest=sha256:467b045c089c3d475a777d423fc5ed313c62b810dc14e78475071257efa5866f

Observation 08e8b828-7914-4b44-9625-d7ed06d336af · outbound

This paper cites Supported trust region optimization for offline reinforcement learning.

Fast and Robust: Task Sampling with Posterior and Diversity Synergies for Adaptive Decision-Makers in Randomized Environments Supported trust region optimization for offline reinforcement learning

Reference 48

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T06:07:20.443745Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-16T06:07:19.612323Z digest=sha256:c1ed5f2ba7789c5564eb8887ccfa2312e7aec383035e95dc2a18c495340af5d9

Observation 8d134177-3bb4-4deb-9dae-a8b253696e41 · outbound

This paper cites Offline Reinforcement Learning with OOD State Correction and OOD Action Suppression.

Fast and Robust: Task Sampling with Posterior and Diversity Synergies for Adaptive Decision-Makers in Randomized Environments Offline Reinforcement Learning with OOD State Correction and OOD Action Suppression

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-16T06:07:19.615567Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T06:07:19.615567Z digest=sha256:d6379eda0a23a9f44e6cf955556b5f07762a28d7fd07adc3971635f6051b534d

Observation a3552b7b-4a38-43e3-a0c6-812e1ea38396 · outbound

This paper cites Doubly Mild Generalization for Offline Reinforcement Learning.

Fast and Robust: Task Sampling with Posterior and Diversity Synergies for Adaptive Decision-Makers in Randomized Environments Doubly Mild Generalization for Offline Reinforcement Learning

Reference 50

Resolution
verified exact
local_arxiv, observed 2026-08-16T06:07:19.968776Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-16T06:07:19.619155Z digest=sha256:963080bfe1f274ec08273f67b48f26f7d73f385496b3c2f236c2f2f5d8632c49

Observation 6315d687-cfaa-4d1e-bf3b-f147cf91b0ee · outbound

This paper cites J., and Paull, L.

Fast and Robust: Task Sampling with Posterior and Diversity Synergies for Adaptive Decision-Makers in Randomized Environments J., and Paull, L

Reference 51

Resolution
unresolved
no resolver link, observed 2026-08-16T06:07:19.622708Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T06:07:19.622708Z digest=sha256:6efc5f1ba089db9254a55c5265d5761a3a7ad4ba4571ebc92ccd484eaea2da24

Observation 7ae5c00f-4e86-45e3-ba3f-be12b02d81ce · outbound

This paper cites H., and Gal, Y.

Fast and Robust: Task Sampling with Posterior and Diversity Synergies for Adaptive Decision-Makers in Randomized Environments H., and Gal, Y

Reference 52

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T06:07:20.427747Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-16T06:07:19.625944Z digest=sha256:ca24e9d8387672d0caa59d5ad25cbfecfa763dd5c0b5aa6b8a83b611cfd41ebc

Observation 8cbba36c-a336-4627-a8be-834454a72db0 · outbound

This paper cites Domain randomization for simulation-based policy optimization with transferability assessment.

Fast and Robust: Task Sampling with Posterior and Diversity Synergies for Adaptive Decision-Makers in Randomized Environments Domain randomization for simulation-based policy optimization with transferability assessment

Reference 53

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T06:07:20.417889Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-16T06:07:19.629250Z digest=sha256:18fb48fea6ba67ddfbc75295d3064358f82639312caf1e4ea37d07b972432e2e

Observation 01d74f5d-c795-4de4-968e-58930b2a59b2 · outbound

This paper cites Data-efficient domain randomization with bayesian optimization.

Fast and Robust: Task Sampling with Posterior and Diversity Synergies for Adaptive Decision-Makers in Randomized Environments Data-efficient domain randomization with bayesian optimization

Reference 54

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T06:07:20.407725Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-16T06:07:19.632665Z digest=sha256:8dbbba517c67f81fd29841a4fa94506cc0b1b582551728283c3973d9bf49115f

Observation 3b839439-50fd-4352-86f7-964aca3c4c97 · outbound

This paper cites Variational Continual Learning.

Fast and Robust: Task Sampling with Posterior and Diversity Synergies for Adaptive Decision-Makers in Randomized Environments Variational Continual Learning

Reference 55

Resolution
unresolved
no resolver link, observed 2026-08-16T06:07:19.635795Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T06:07:19.635795Z digest=sha256:79df44c357f3e63f28e789de4acc15074d49d4384ad8085849261d53264c9f1c

Observation 551dc370-d83b-4015-ae6d-200c68d636d4 · outbound

This paper cites (more) efficient reinforcement learning via posterior sampling.

Fast and Robust: Task Sampling with Posterior and Diversity Synergies for Adaptive Decision-Makers in Randomized Environments (more) efficient reinforcement learning via posterior sampling

Reference 56

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T06:07:20.397796Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-16T06:07:19.639417Z digest=sha256:b97330f28856cc27a31ab63947f7f24e3173e86c65f5ac64367a116f790c8be1

Observation bc30fc04-e964-43ee-b0dd-38a1710cb077 · outbound

This paper cites Risk averse robust adversarial reinforcement learning.

Fast and Robust: Task Sampling with Posterior and Diversity Synergies for Adaptive Decision-Makers in Randomized Environments Risk averse robust adversarial reinforcement learning

Reference 57

Resolution
unresolved
no resolver link, observed 2026-08-16T06:07:19.642846Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T06:07:19.642846Z digest=sha256:6abbb9c8f9e5cc7ac3aa7dfcbc43764537a467ce0b25985f6a306409f016daef

Observation 81ba255b-0a9b-4753-b37e-5efd2a4783a0 · outbound

This paper cites an unresolved cited work.

Fast and Robust: Task Sampling with Posterior and Diversity Synergies for Adaptive Decision-Makers in Randomized Environments Unresolved cited work

Reference 58

Resolution
unresolved
no resolver link, observed 2026-08-16T06:07:19.646330Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T06:07:19.646330Z digest=sha256:499ce43ecf4e2b925b7d397bf26332c810d13f66b1bdd258942560da1f4b77a0

Observation c2de9c67-7809-45f9-b0d2-514e663ce01c · outbound

This paper cites Meta-learning with neural bandit scheduler.

Fast and Robust: Task Sampling with Posterior and Diversity Synergies for Adaptive Decision-Makers in Randomized Environments Meta-learning with neural bandit scheduler

Reference 59

Resolution
unresolved
no resolver link, observed 2026-08-16T06:07:19.649421Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T06:07:19.649421Z digest=sha256:fc3ff4345638703bd8e14ba90c7c19f1704d8370c67208ed9f7fc2ea254077aa

Observation d5bd3537-2d1a-44f4-92cd-f9bb40e89e8f · outbound

This paper cites Hokoff: Real game dataset from honor of kings and its offline reinforcement learning benchmarks.

Fast and Robust: Task Sampling with Posterior and Diversity Synergies for Adaptive Decision-Makers in Randomized Environments Hokoff: Real game dataset from honor of kings and its offline reinforcement learning benchmarks

Reference 60

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T06:07:20.368888Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-16T06:07:19.652730Z digest=sha256:ea0eb234d6c468915ee216c82f11fc7c7b7b2fbc3a0bf522f3c946b321a81af2

Observation c99a8ad3-082b-4c41-9227-809318140e1d · outbound

This paper cites Stop Wandering, Find the Keys: LLMs Discriminate Key States for Efficient Multi-Agent Exploration.

Fast and Robust: Task Sampling with Posterior and Diversity Synergies for Adaptive Decision-Makers in Randomized Environments Stop Wandering, Find the Keys: LLMs Discriminate Key States for Efficient Multi-Agent Exploration

Reference 61

Resolution
verified exact
local_arxiv, observed 2026-08-16T06:07:19.942774Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-16T06:07:19.656167Z digest=sha256:affefe4e036b381812e82b01c82a91cb74354bb8a042e98d7e2040ada98bb610

Observation b25e7614-0244-45db-86b5-a40bb327682d · outbound

This paper cites Latent reward: Llm-empowered credit assignment in episodic reinforcement learning.

Fast and Robust: Task Sampling with Posterior and Diversity Synergies for Adaptive Decision-Makers in Randomized Environments Latent reward: Llm-empowered credit assignment in episodic reinforcement learning

Reference 62

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T06:07:20.357963Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-16T06:07:19.659467Z digest=sha256:bdb4f26fdda51789ff0ac6b34e89b5e9bd7418a9f5b6e3d0315c21faf853d076

Observation 3d905242-76d5-4c81-ac73-0180c35e1eb2 · outbound

This paper cites M., and Levine, S.

Fast and Robust: Task Sampling with Posterior and Diversity Synergies for Adaptive Decision-Makers in Randomized Environments M., and Levine, S

Reference 63

Resolution
unresolved
no resolver link, observed 2026-08-16T06:07:19.662642Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T06:07:19.662642Z digest=sha256:7b8d007fb40cda35d8003ca39e7c02597ba00ab475e4c6fd0087d92524c77a7c

Observation 60d5ae6c-a7a6-4dca-beee-76d1eccebf43 · outbound

This paper cites Efficient off-policy meta-reinforcement learning via probabilistic context variables.

Fast and Robust: Task Sampling with Posterior and Diversity Synergies for Adaptive Decision-Makers in Randomized Environments Efficient off-policy meta-reinforcement learning via probabilistic context variables

Reference 64

Resolution
unresolved
no resolver link, observed 2026-08-16T06:07:19.665736Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T06:07:19.665736Z digest=sha256:fb1d59d9307a8d481e5b942bbe74138ebdf1931374d3714d90d9fa0bc04d293a

Observation 87b89c97-cf3f-487a-994d-a873bf6c236a · outbound

This paper cites J., Fidler, S., and Litany, O.

Fast and Robust: Task Sampling with Posterior and Diversity Synergies for Adaptive Decision-Makers in Randomized Environments J., Fidler, S., and Litany, O

Reference 65

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T06:07:20.335668Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-16T06:07:19.669179Z digest=sha256:52f143f46daa78abbce725d6a9aae6ed4e1249c5facea460b57ded380ef707fa

Observation 3b880b53-a8d7-4ca5-a884-3b728f473cec · outbound

This paper cites B., Chen, X., and Wang, X.

Fast and Robust: Task Sampling with Posterior and Diversity Synergies for Adaptive Decision-Makers in Randomized Environments B., Chen, X., and Wang, X

Reference 66

Resolution
unresolved
no resolver link, observed 2026-08-16T06:07:19.672267Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T06:07:19.672267Z digest=sha256:e8b3fe612bd101535eb432102c78e495347a81619109b0642bdb5c60c4fa9e1c

Observation f7823eb2-0ff8-439a-aa19-6b91830928f8 · outbound

This paper cites Risk-averse bayes-adaptive reinforcement learning.

Fast and Robust: Task Sampling with Posterior and Diversity Synergies for Adaptive Decision-Makers in Randomized Environments Risk-averse bayes-adaptive reinforcement learning

Reference 67

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T06:07:20.319048Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-16T06:07:19.675393Z digest=sha256:ce2927364db645eb31fddeaf18ef5eba2eb9e4979b705b756dd6fb25484d4290

Observation 5543703a-f2c1-4506-8bbd-850b2cf96ee3 · outbound

This paper cites Been there, done that: Meta-learning with episodic recall.

Fast and Robust: Task Sampling with Posterior and Diversity Synergies for Adaptive Decision-Makers in Randomized Environments Been there, done that: Meta-learning with episodic recall

Reference 68

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T06:07:20.308943Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-16T06:07:19.678600Z digest=sha256:108cd97fca60ee8bafff28413368126c74b6985be7d9ac842bbeaf3ca2e1594b

Observation 1eb0376d-39d5-47a8-b7d3-e9702776e451 · outbound

This paper cites T., Uryasev, S., et al.

Fast and Robust: Task Sampling with Posterior and Diversity Synergies for Adaptive Decision-Makers in Randomized Environments T., Uryasev, S., et al

Reference 69

Resolution
unresolved
no resolver link, observed 2026-08-16T06:07:19.681694Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T06:07:19.681694Z digest=sha256:da53699ec4652b1d142abf1557b33b996f128b8113c4ab67e332c01cb0fd1aee

Observation 92992d40-807e-4363-8f8d-82445fdaa4f8 · outbound

This paper cites and Van Roy, B.

Fast and Robust: Task Sampling with Posterior and Diversity Synergies for Adaptive Decision-Makers in Randomized Environments and Van Roy, B

Reference 70

Resolution
unresolved
no resolver link, observed 2026-08-16T06:07:19.684809Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T06:07:19.684809Z digest=sha256:13e3a5ee5b5695b3fdd9224efcd91561bef78549c1cdb8969ef70d9211e0b8e6

Observation 9ab5554e-712c-4b3a-917b-27ab353a1863 · outbound

This paper cites W., Hashimoto, T.

Fast and Robust: Task Sampling with Posterior and Diversity Synergies for Adaptive Decision-Makers in Randomized Environments W., Hashimoto, T

Reference 71

Resolution
unresolved
no resolver link, observed 2026-08-16T06:07:19.688100Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T06:07:19.688100Z digest=sha256:54467c680102f5ee21c96d12a6f61b126bdc73b364c8990a0b1b114b9ca4ce47

Observation c24aa8df-f8a0-4d30-83b2-fe21871d7f89 · outbound

This paper cites Proximal Policy Optimization Algorithms.

Fast and Robust: Task Sampling with Posterior and Diversity Synergies for Adaptive Decision-Makers in Randomized Environments Proximal Policy Optimization Algorithms

Reference 72

Resolution
unresolved
no resolver link, observed 2026-08-16T06:07:19.691253Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T06:07:19.691253Z digest=sha256:4bb0637e85b3b9b65e0ec0a6e9d5fc99e802f544f23834534d483d40bc8a9ff7

Observation b0e4148b-a0f4-43d9-9e54-251830dc707e · outbound

This paper cites Prompting is a double-edged sword: Improving worst-group robustness of foundation models.

Fast and Robust: Task Sampling with Posterior and Diversity Synergies for Adaptive Decision-Makers in Randomized Environments Prompting is a double-edged sword: Improving worst-group robustness of foundation models

Reference 73

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T06:07:20.280010Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-16T06:07:19.694623Z digest=sha256:7d2083ab08450cd1ad6d4360ea1a5187c6240c77572ab27d0ce78918f3e1c9e5

Observation c1138d65-1084-4484-8d0d-3887b3c25aed · outbound

This paper cites Counterfactual conservative q learning for offline multi-agent reinforcement learning.

Fast and Robust: Task Sampling with Posterior and Diversity Synergies for Adaptive Decision-Makers in Randomized Environments Counterfactual conservative q learning for offline multi-agent reinforcement learning

Reference 74

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T06:07:20.269237Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-16T06:07:19.697683Z digest=sha256:b6806bd6642184ab73993770ba10fc770f2bbb1d78970962b1148b688493b5c8

Observation dca42403-e6b6-408d-9b0f-51f93d2c270a · outbound

This paper cites Complementary attention for multi-agent reinforcement learning.

Fast and Robust: Task Sampling with Posterior and Diversity Synergies for Adaptive Decision-Makers in Randomized Environments Complementary attention for multi-agent reinforcement learning

Reference 75

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T06:07:20.258683Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-16T06:07:19.700660Z digest=sha256:09472950d6d640a11b9e2c6901c4dc62bc6cee9e0b3c63ad9e930959b134488d

Observation 454e847c-0d68-4e41-8062-4a923f958261 · outbound

This paper cites an unresolved cited work.

Fast and Robust: Task Sampling with Posterior and Diversity Synergies for Adaptive Decision-Makers in Randomized Environments Unresolved cited work

Reference 76

Resolution
unresolved
no resolver link, observed 2026-08-16T06:07:19.703760Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T06:07:19.703760Z digest=sha256:8aae6357a1c6fcbb6091efda28ba71993bf71bb79b252d28b553ef61385cc7bf

Observation e135e20a-24fe-43e4-ad74-5fe1fe8e0566 · outbound

This paper cites Policy gradient for coherent risk measures.

Fast and Robust: Task Sampling with Posterior and Diversity Synergies for Adaptive Decision-Makers in Randomized Environments Policy gradient for coherent risk measures

Reference 77

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T06:07:20.241938Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-16T06:07:19.706802Z digest=sha256:aa94e89a881adada412a2b0ec8f374efa92dae5e2ac39f5f32905ab5af09b853

Observation f346bfa0-9a39-4f98-b3e4-e66d4a83688d · outbound

This paper cites ManiSkill3: GPU Parallelized Robotics Simulation and Rendering for Generalizable Embodied AI.

Fast and Robust: Task Sampling with Posterior and Diversity Synergies for Adaptive Decision-Makers in Randomized Environments ManiSkill3: GPU Parallelized Robotics Simulation and Rendering for Generalizable Embodied AI

Reference 78

Resolution
unresolved
no resolver link, observed 2026-08-16T06:07:19.710213Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T06:07:19.710213Z digest=sha256:72fa89f8e4ab13dd531a0b0473d114670ddfd0ef0082e0d29e75a03dd306e4c2

Observation 7c8b863b-f3d3-47f5-97ba-99f315cfc6cf · outbound

This paper cites an unresolved cited work.

Fast and Robust: Task Sampling with Posterior and Diversity Synergies for Adaptive Decision-Makers in Randomized Environments Unresolved cited work

Reference 79

Resolution
unresolved
no resolver link, observed 2026-08-16T06:07:19.713847Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T06:07:19.713847Z digest=sha256:5096991f51f6127dde9a3941c972bf9e340541c3abc41796448f34938a573241

Observation 9b8e0d7a-63b7-4354-a78a-379b9e52ae3e · outbound

This paper cites Domain Randomization via Entropy Maximization.

Fast and Robust: Task Sampling with Posterior and Diversity Synergies for Adaptive Decision-Makers in Randomized Environments Domain Randomization via Entropy Maximization

Reference 80

Resolution
unresolved
no resolver link, observed 2026-08-16T06:07:19.717212Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T06:07:19.717212Z digest=sha256:b000306479beabd5d620b477e031c4edb08ef3eb8134f5771b694e17cb1658a7

Observation f05afa65-1ef3-4159-9de9-1b58e8cd730c · outbound

This paper cites Domain randomization for transferring deep neural networks from simulation to the real world.

Fast and Robust: Task Sampling with Posterior and Diversity Synergies for Adaptive Decision-Makers in Randomized Environments Domain randomization for transferring deep neural networks from simulation to the real world

Reference 81

Resolution
unresolved
no resolver link, observed 2026-08-16T06:07:19.720919Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T06:07:19.720919Z digest=sha256:705b0f43253b112381df6621bda33bf503967e232e84ec9c30d8911d239706ad

Observation 911e0356-f945-4776-a7d5-931688ce6c72 · outbound

This paper cites Mujoco: A physics engine for model-based control.

Fast and Robust: Task Sampling with Posterior and Diversity Synergies for Adaptive Decision-Makers in Randomized Environments Mujoco: A physics engine for model-based control

Reference 82

Resolution
unresolved
no resolver link, observed 2026-08-16T06:07:19.724132Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T06:07:19.724132Z digest=sha256:882bf2ff7811be403f07c1856b86a9237d104261b456ef308f63d9cc49904df4

Observation 4bbdb804-6c21-48aa-a9ab-60fa43844b8a · outbound

This paper cites N., Vapnik, V., et al.

Fast and Robust: Task Sampling with Posterior and Diversity Synergies for Adaptive Decision-Makers in Randomized Environments N., Vapnik, V., et al

Reference 83

Resolution
unresolved
no resolver link, observed 2026-08-16T06:07:19.727152Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T06:07:19.727152Z digest=sha256:d0b45e9dc616c31fea189793887ff1d616aeed7cfdc0f29c3f8669c5e657d211

Observation 450bf3bf-cd93-4342-baff-0af5473f97e5 · outbound

This paper cites LLM-Empowered State Representation for Reinforcement Learning.

Fast and Robust: Task Sampling with Posterior and Diversity Synergies for Adaptive Decision-Makers in Randomized Environments LLM-Empowered State Representation for Reinforcement Learning

Reference 84

Resolution
unresolved
no resolver link, observed 2026-08-16T06:07:19.730282Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T06:07:19.730282Z digest=sha256:503ba225617e77da879e0a89227553dae82541c441372f386ec61e2381c90201

Observation 932eb590-c4d9-47e8-a013-3a3f5c2caed3 · outbound

This paper cites Robust Fast Adaptation from Adversarially Explicit Task Distribution Generation.

Fast and Robust: Task Sampling with Posterior and Diversity Synergies for Adaptive Decision-Makers in Randomized Environments Robust Fast Adaptation from Adversarially Explicit Task Distribution Generation

Reference 85

Resolution
unresolved
no resolver link, observed 2026-08-16T06:07:19.733479Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T06:07:19.733479Z digest=sha256:5b698829779ef832b692574748a1577c5dc7e625c003d7631f270b320481045e

Observation de78651d-fc48-4d6f-a572-d227deb1ec3c · outbound

This paper cites and Van Hoof, H.

Fast and Robust: Task Sampling with Posterior and Diversity Synergies for Adaptive Decision-Makers in Randomized Environments and Van Hoof, H

Reference 86

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T06:07:20.207349Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-16T06:07:19.736690Z digest=sha256:669478c8265b79763276787f209ce81a692cd5c1b0ef779f15d8ebb2fb20232d

Observation 332c131a-42bd-4c21-800c-77e06e2532a3 · outbound

This paper cites and Van Hoof, H.

Fast and Robust: Task Sampling with Posterior and Diversity Synergies for Adaptive Decision-Makers in Randomized Environments and Van Hoof, H

Reference 87

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T06:07:20.196751Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-16T06:07:19.739818Z digest=sha256:b20aabbf0a4295b9ddafea5f463e92aa0bc33ecbdce2e257cba65edf81ee5f32

Observation e5755f72-9968-40e7-8150-c53782f0a346 · outbound

This paper cites and Van Hoof, H.

Fast and Robust: Task Sampling with Posterior and Diversity Synergies for Adaptive Decision-Makers in Randomized Environments and Van Hoof, H

Reference 88

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T06:07:20.186407Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-16T06:07:19.742889Z digest=sha256:5b87be1cb4003e39ec33faa9934af0c1ad9fec499ca515f4e1760acf68cdc92f

Observation 3e54db66-d416-48ce-b0a7-f000105f813a · outbound

This paper cites Bridge the inference gaps of neural processes via expectation maximization.

Fast and Robust: Task Sampling with Posterior and Diversity Synergies for Adaptive Decision-Makers in Randomized Environments Bridge the inference gaps of neural processes via expectation maximization

Reference 89

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T06:07:20.175111Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-16T06:07:19.746258Z digest=sha256:e06dc3111b19069c4f2be503d746c07720c0a04dccf4c7147bd1e71a8959947d

Observation 1819eff5-2710-4920-8ae0-32c5a4631aa6 · outbound

This paper cites Bridge the inference gaps of neural processes via expectation maximization.

Fast and Robust: Task Sampling with Posterior and Diversity Synergies for Adaptive Decision-Makers in Randomized Environments Bridge the inference gaps of neural processes via expectation maximization

Reference 90

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T06:07:20.163804Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-16T06:07:19.749224Z digest=sha256:736517c5181c3084a1c28ad934207c39323bae47845f066366667b93f0acdbfa

Observation 785659ff-e84c-45ef-a2a6-f488a576919b · outbound

This paper cites A simple yet effective strategy to robustify the meta learning paradigm.

Fast and Robust: Task Sampling with Posterior and Diversity Synergies for Adaptive Decision-Makers in Randomized Environments A simple yet effective strategy to robustify the meta learning paradigm

Reference 91

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T06:07:20.153016Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-16T06:07:19.752321Z digest=sha256:e9b874cf823b14db3431c73909b59a3e3e898ee85f6541cf2f3a55ffa1d033fe

Observation b922284c-bd28-4f68-84e1-7b097446435b · outbound

This paper cites C., Xiao, Z., Mao, Y., Qu, Y., Shen, J., Lv, Y., and Ji, X.

Fast and Robust: Task Sampling with Posterior and Diversity Synergies for Adaptive Decision-Makers in Randomized Environments C., Xiao, Z., Mao, Y., Qu, Y., Shen, J., Lv, Y., and Ji, X

Reference 93

Resolution
unresolved
no resolver link, observed 2026-08-16T06:07:19.758674Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T06:07:19.758674Z digest=sha256:033f2baca27e733060ffe52821096a81e5a0bd22fcbeae816e6c17545dfc29ee

Observation 9e805beb-f4c3-45a6-b04f-0c97540e7fb1 · outbound

This paper cites Max-min diversification with fairness constraints: Exact and approximation algorithms.

Fast and Robust: Task Sampling with Posterior and Diversity Synergies for Adaptive Decision-Makers in Randomized Environments Max-min diversification with fairness constraints: Exact and approximation algorithms

Reference 94

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T06:07:20.142842Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-16T06:07:19.762013Z digest=sha256:99738630333e63a23bc41f088976d38d45e02c89eea7e43fdaf1cf9d13c32a6c

Observation 50d7c5f2-06e2-4879-ba82-e4f9643a5733 · outbound

This paper cites Entropy-based active learning for object detection with progressive diversity constraint.

Fast and Robust: Task Sampling with Posterior and Diversity Synergies for Adaptive Decision-Makers in Randomized Environments Entropy-based active learning for object detection with progressive diversity constraint

Reference 95

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T06:07:20.132062Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-16T06:07:19.765397Z digest=sha256:e6c5a8e7e72c28587bec4b55de7673f66c075ac4f9233119db173afdf9a21471

Observation 173e4ce0-585b-4a84-ac1c-48d876fc04cd · outbound

This paper cites Enhancing context-based meta-reinforcement learning algorithms via an efficient task encoder (student abstract).

Fast and Robust: Task Sampling with Posterior and Diversity Synergies for Adaptive Decision-Makers in Randomized Environments Enhancing context-based meta-reinforcement learning algorithms via an efficient task encoder (student abstract)

Reference 96

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T06:07:20.121864Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-16T06:07:19.768519Z digest=sha256:3dbbff6996a6c2e4bb637f5067b8a8e69055fb87ad118f17e34b5b4b58ba0426

Observation 5ade9aee-9d27-4eb9-abbb-5de699f1a2ad · outbound

This paper cites Bayesian model-agnostic meta-learning.

Fast and Robust: Task Sampling with Posterior and Diversity Synergies for Adaptive Decision-Makers in Randomized Environments Bayesian model-agnostic meta-learning

Reference 97

Resolution
unresolved
no resolver link, observed 2026-08-16T06:07:19.772193Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T06:07:19.772193Z digest=sha256:0740194a74009b21971af0175671948fd52372552c2d06e795995f360ba0f84e

Observation c8f669a2-91a3-409f-93e8-6bcc611ac78f · outbound

This paper cites In-sample actor critic for offline reinforcement learning.

Fast and Robust: Task Sampling with Posterior and Diversity Synergies for Adaptive Decision-Makers in Randomized Environments In-sample actor critic for offline reinforcement learning

Reference 98

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T06:07:20.105029Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-16T06:07:19.775352Z digest=sha256:abb47db8146c7618e370a01d914edeef8c21fe2aa555804ee0743987c6e2f9ad

Observation 79df17aa-53a6-4126-bccd-7de4dd4c6752 · outbound

This paper cites Combining active learning and semi-supervised learning using gaussian fields and harmonic functions.

Fast and Robust: Task Sampling with Posterior and Diversity Synergies for Adaptive Decision-Makers in Randomized Environments Combining active learning and semi-supervised learning using gaussian fields and harmonic functions

Reference 99

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T06:07:20.094970Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-16T06:07:19.778563Z digest=sha256:391597660539141b55344c016f28db0c1adf793f947932598bfe66350e34a15d

Observation 0a639983-ad2c-4d88-b0bd-d292922863b0 · outbound

This paper cites VariBAD: A Very Good Method for Bayes-Adaptive Deep RL via Meta-Learning.

Fast and Robust: Task Sampling with Posterior and Diversity Synergies for Adaptive Decision-Makers in Randomized Environments VariBAD: A Very Good Method for Bayes-Adaptive Deep RL via Meta-Learning

Reference 100

Resolution
unresolved
no resolver link, observed 2026-08-16T06:07:19.781695Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T06:07:19.781695Z digest=sha256:e1a70de4fdfeb5b983874296cb97c1b16788a998b4654de5142a0550fce875ef

Observation d7272cf5-1381-4047-a0c9-7ee5a53ea186 · outbound

This paper cites write newline.

Fast and Robust: Task Sampling with Posterior and Diversity Synergies for Adaptive Decision-Makers in Randomized Environments write newline

Reference 101

Resolution
unresolved
no resolver link, observed 2026-08-16T06:07:19.785138Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T06:07:19.785138Z digest=sha256:9e6f2a44be9f1939074728e6f06410840bfced34345c63dafa8eed989e8b85bb

Pith citing papers

No inbound Pith citation observations are available.