Pith. sign in

Paper Citation Record · LEDGER

Understanding Behavioral Metric Learning: A Large-Scale Study on Distracting Reinforcement Learning Environments

As of 12 August 2026, this Paper Citation Record lists 79 of 79 outbound references and 1 inbound Pith citation observation for arXiv:2506.00563.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2506.00563 v2

Coverage vector

measured 79 of 79 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T12:10:08.987854Z

measured 80 of 80 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-11T06:34:44.6726+00:00

measured 1 of 1 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-05-09T20:00:55.197662Z

measured 1 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Reference resolution

79 of 79 outbound references displayed

  • verified exact3
  • verified fuzzy44
  • unresolved32
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

0
arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Outbound references

Observation e14aaef1-00af-43f1-bff2-4e5f80decd75 · outbound

This paper cites Contrastive Behavioral Similarity Embeddings for Generalization in Reinforcement Learning.

Understanding Behavioral Metric Learning: A Large-Scale Study on Distracting Reinforcement Learning Environments Contrastive Behavioral Similarity Embeddings for Generalization in Reinforcement Learning

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-07T12:09:58.708141Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:09:58.708141Z digest=sha256:f43c73f19fceeb1adb930cf0c50c1f49b2bb35f1a8e2a6ae9282db45e4fc4083

Observation ac61ba63-31a3-4367-8b37-f9181c824f60 · outbound

This paper cites Deep reinforcement learning at the edge of the statistical precipice.

Understanding Behavioral Metric Learning: A Large-Scale Study on Distracting Reinforcement Learning Environments Deep reinforcement learning at the edge of the statistical precipice

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:10:18.711466Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-08-07T12:09:58.870069Z digest=sha256:7a72bb9b2e523e75747253fb7daef5f8ec3a7c318ec62045b14d86c0296bd09a

Observation 5716f765-391f-408b-9072-d16b236b6e21 · outbound

This paper cites Layer Normalization.

Understanding Behavioral Metric Learning: A Large-Scale Study on Distracting Reinforcement Learning Environments Layer Normalization

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-07T12:09:58.996705Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:09:58.996705Z digest=sha256:d1bbe087584d35e45343a132b3d28a06ae5635e13a293da6970121c47f9618b2

Observation 86a6e4d0-061c-4591-9a1b-8bbe15e8a26a · outbound

This paper cites CrossQ: Batch Normalization in Deep Reinforcement Learning for Greater Sample Efficiency and Simplicity.

Understanding Behavioral Metric Learning: A Large-Scale Study on Distracting Reinforcement Learning Environments CrossQ: Batch Normalization in Deep Reinforcement Learning for Greater Sample Efficiency and Simplicity

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-07T12:09:59.140424Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:09:59.140424Z digest=sha256:da03d9d3084f59643992bb5fc9d1a76eac46f67108c846f7daafa8d39f07a2cf

Observation 8bb28be9-cd18-4ca9-addf-f0d8e9b72f07 · outbound

This paper cites Online Abstraction with MDP Homomorphisms for Deep Learning.

Understanding Behavioral Metric Learning: A Large-Scale Study on Distracting Reinforcement Learning Environments Online Abstraction with MDP Homomorphisms for Deep Learning

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-07T12:09:59.274698Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:09:59.274698Z digest=sha256:206cc27f76c9fd889f008bb39787b8aa763fe2d396a6fe7cccbffdaf949257d8

Observation 55d96dde-b654-44c2-9817-6697f6912d7f · outbound

This paper cites Scalable methods for computing state similarity in deterministic markov decision processes.

Understanding Behavioral Metric Learning: A Large-Scale Study on Distracting Reinforcement Learning Environments Scalable methods for computing state similarity in deterministic markov decision processes

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-07T12:09:59.392893Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:09:59.392893Z digest=sha256:7cf5611281d78a03ad3085415ec6aadffd33638218a5b3058ffac1dcf620804c

Observation bd4db822-0655-46a0-b212-3b04738494a3 · outbound

This paper cites Mico: Improved representations via sampling-based state similarity for markov decision processes.

Understanding Behavioral Metric Learning: A Large-Scale Study on Distracting Reinforcement Learning Environments Mico: Improved representations via sampling-based state similarity for markov decision processes

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:10:18.409970Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-08-07T12:09:59.461170Z digest=sha256:da621878ac943080f659266f3d57cc27c5a7a3fdf87d35c11c1441aca6849f6b

Observation bfd2714a-d32b-49e7-9c28-39585c54c841 · outbound

This paper cites A Kernel Perspective on Behavioural Metrics for Markov Decision Processes.

Understanding Behavioral Metric Learning: A Large-Scale Study on Distracting Reinforcement Learning Environments A Kernel Perspective on Behavioural Metrics for Markov Decision Processes

Reference 8

Resolution
verified exact
local_arxiv, observed 2026-08-07T12:10:09.934997Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-08-07T12:09:59.548686Z digest=sha256:1f21ee31214222936468b2b83b8d5505126c34af66bad4e8e1340bec153c74da

Observation fc4e8565-3c9c-4dc4-818c-f3b254e2ce68 · outbound

This paper cites Learning representations via a robust behavioral metric for deep reinforcement learning.

Understanding Behavioral Metric Learning: A Large-Scale Study on Distracting Reinforcement Learning Environments Learning representations via a robust behavioral metric for deep reinforcement learning

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:10:18.177019Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-08-07T12:09:59.652201Z digest=sha256:60d6c0a26b34671cfb8640200b8c6e1064db75602851551af9a797ecae251266

Observation 5f28d638-bdc7-4b0d-a90f-3d50cb07acc2 · outbound

This paper cites State chrono representation for enhancing generalization in reinforcement learning.

Understanding Behavioral Metric Learning: A Large-Scale Study on Distracting Reinforcement Learning Environments State chrono representation for enhancing generalization in reinforcement learning

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:10:17.950416Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-08-07T12:09:59.778307Z digest=sha256:e47415d689829f54e00e8dc96b1dd130899d1c3c70d0d9809ee9a23a1130d3ef

Observation 160a0d47-bb4c-42ed-b586-54d8fcf5d89c · outbound

This paper cites Offline reinforcement learning with pseudometric learning.

Understanding Behavioral Metric Learning: A Large-Scale Study on Distracting Reinforcement Learning Environments Offline reinforcement learning with pseudometric learning

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:10:17.743648Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-08-07T12:09:59.881247Z digest=sha256:65cee8c12674b849c28aff6e933d1e48ff1f37bce6841d10346323fd69919424

Observation 1ed9f1d8-353a-4f1c-ac81-b080d213cbc6 · outbound

This paper cites Bisimulation for labelled markov processes.

Understanding Behavioral Metric Learning: A Large-Scale Study on Distracting Reinforcement Learning Environments Bisimulation for labelled markov processes

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:10:17.530504Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-08-07T12:09:59.982921Z digest=sha256:ef0e4d2717d207286b0a6c89267e951e4725ac02885b5655a5b625ffb93a0c89

Observation af24be27-ebb8-4e44-96da-2f2d6a1d0137 · outbound

This paper cites Provably efficient rl with rich observations via latent state decoding.

Understanding Behavioral Metric Learning: A Large-Scale Study on Distracting Reinforcement Learning Environments Provably efficient rl with rich observations via latent state decoding

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:10:17.325024Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-08-07T12:10:00.096030Z digest=sha256:94a1d42ad4d4910b5c7bf8326c25654842cd8466b511bf03c4081102c63244b0

Observation 2963fefa-d659-4957-a574-5e8ab0b217c5 · outbound

This paper cites Differential privacy.

Understanding Behavioral Metric Learning: A Large-Scale Study on Distracting Reinforcement Learning Environments Differential privacy

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-07T12:10:00.192754Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:10:00.192754Z digest=sha256:be226da3f1ea8cc5cdc2ac3da84458a1769019b76abfc0865652eeec5ef861bc

Observation e9804ab4-9cc1-4052-93c8-2b9f2fc1691e · outbound

This paper cites Provable RL with Exogenous Distractors via Multistep Inverse Dynamics.

Understanding Behavioral Metric Learning: A Large-Scale Study on Distracting Reinforcement Learning Environments Provable RL with Exogenous Distractors via Multistep Inverse Dynamics

Reference 15

Resolution
verified exact
local_arxiv, observed 2026-08-07T12:10:09.765339Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-08-07T12:10:00.264076Z digest=sha256:96f14347442eec5dae301328a082d2eae1236af8ae6290644e3a2482cfd083ee

Observation c6e73560-ecf9-49a2-ab06-ed543ac55529 · outbound

This paper cites Metrics for finite markov decision processes.

Understanding Behavioral Metric Learning: A Large-Scale Study on Distracting Reinforcement Learning Environments Metrics for finite markov decision processes

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:10:17.118699Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-08-07T12:10:00.324820Z digest=sha256:de738cce6487316679492814ab800dd5099b8159bb6390acab5238ce07c3bbb0

Observation a353b52e-0a78-41f2-b7d3-0dc4330feb8e · outbound

This paper cites Bisimulation metrics for continuous markov decision processes.

Understanding Behavioral Metric Learning: A Large-Scale Study on Distracting Reinforcement Learning Environments Bisimulation metrics for continuous markov decision processes

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:10:16.873188Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-08-07T12:10:00.423106Z digest=sha256:a29cf1cf5937a690ee92736e58dcb939e76d9d4693877665a6b17c4b6d2c6868

Observation c0d954e0-386a-4314-8908-e142b08ddef6 · outbound

This paper cites For sale: State-action representation learning for deep reinforcement learning.

Understanding Behavioral Metric Learning: A Large-Scale Study on Distracting Reinforcement Learning Environments For sale: State-action representation learning for deep reinforcement learning

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:10:16.665792Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-08-07T12:10:00.548679Z digest=sha256:a66ae71b202c57c37973aac8bea020fd1e4567d16aa09e285041dbddb0f7f542

Observation 36c45093-8205-432d-a6a0-4303bd0e2635 · outbound

This paper cites Deepmdp: Learning continuous latent space models for representation learning.

Understanding Behavioral Metric Learning: A Large-Scale Study on Distracting Reinforcement Learning Environments Deepmdp: Learning continuous latent space models for representation learning

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:10:16.432994Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-08-07T12:10:00.635009Z digest=sha256:4621fc03c058902beaf5a26e34cbaf7fb357a01e046fc2d995379049a2891a54

Observation 6aba3156-7c65-41a0-b0ee-ad36b0a4b9c1 · outbound

This paper cites Fully homomorphic encryption using ideal lattices.

Understanding Behavioral Metric Learning: A Large-Scale Study on Distracting Reinforcement Learning Environments Fully homomorphic encryption using ideal lattices

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:10:16.249067Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-08-07T12:10:00.715862Z digest=sha256:23d2a739046f5686e12443368de63214ba875be34b3e2ff83ffb834f83c142f7

Observation 78335d36-026a-483a-9633-db4893764342 · outbound

This paper cites Equivalence notions and model minimization in markov decision processes.

Understanding Behavioral Metric Learning: A Large-Scale Study on Distracting Reinforcement Learning Environments Equivalence notions and model minimization in markov decision processes

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:10:16.023035Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-08-07T12:10:00.797333Z digest=sha256:5cd0f75c51ec4412f707eb53c3a20a2c25f0c305bb948a108d62f217f0fd7bd6

Observation 86f75616-1fab-4aab-ab42-47d76a67982c · outbound

This paper cites Measuring visual generalization in continuous control from pixels, 2020.

Understanding Behavioral Metric Learning: A Large-Scale Study on Distracting Reinforcement Learning Environments Measuring visual generalization in continuous control from pixels, 2020

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:10:15.823053Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-08-07T12:10:00.881709Z digest=sha256:b532f54c5994decef3272e2ed10f4fc833cc097c306499e083cb8794d64e0eee

Observation 133587ba-02ac-454b-92de-9184e9cae30c · outbound

This paper cites Bootstrap your own latent-a new approach to self-supervised learning.

Understanding Behavioral Metric Learning: A Large-Scale Study on Distracting Reinforcement Learning Environments Bootstrap your own latent-a new approach to self-supervised learning

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-07T12:10:00.960140Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:10:00.960140Z digest=sha256:1e8f9310d9442359864cc5f12c28d4e27d7372ad431867d0b334e50c1d3e2912

Observation 32926965-1190-44ed-b5a5-933abac577eb · outbound

This paper cites Soft actor-critic: Off-policy maximum entropy deep reinforcement learning with a stochastic actor.

Understanding Behavioral Metric Learning: A Large-Scale Study on Distracting Reinforcement Learning Environments Soft actor-critic: Off-policy maximum entropy deep reinforcement learning with a stochastic actor

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-07T12:10:01.065851Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:10:01.065851Z digest=sha256:1899d9588a620c0dd19441a11b2be5291b2c6aac4c93e21e8656294f10916bdb

Observation 97653980-44b6-452e-b514-5b9c3d02ef0d · outbound

This paper cites Generalization in reinforcement learning by soft data augmentation.

Understanding Behavioral Metric Learning: A Large-Scale Study on Distracting Reinforcement Learning Environments Generalization in reinforcement learning by soft data augmentation

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:10:15.521464Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-08-07T12:10:01.184038Z digest=sha256:21ad2ae5bd759aca479c83c7b81beb47544df71870df955f377ae8def55cba49

Observation d30c9ffa-b9a8-48f0-bf99-65f1cc9e1859 · outbound

This paper cites TD-MPC2: Scalable, Robust World Models for Continuous Control.

Understanding Behavioral Metric Learning: A Large-Scale Study on Distracting Reinforcement Learning Environments TD-MPC2: Scalable, Robust World Models for Continuous Control

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-07T12:10:01.292716Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:10:01.292716Z digest=sha256:e32cb66eb0ae9f8188143ff1a31716ae279c6a2167d323f079810011b5e8ba12

Observation 401db3f8-06ac-4e22-809c-1fb5192650db · outbound

This paper cites Bisimulation makes analogies in goal-conditioned reinforcement learning.

Understanding Behavioral Metric Learning: A Large-Scale Study on Distracting Reinforcement Learning Environments Bisimulation makes analogies in goal-conditioned reinforcement learning

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:10:15.308971Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-08-07T12:10:01.470417Z digest=sha256:bd277de12fc2d23caaf6d3f80982635fb106278c28037282e3cdc8594a28c790

Observation 9590b081-f6c2-42c2-b002-39d4d4229ea1 · outbound

This paper cites Dropout Q-Functions for Doubly Efficient Reinforcement Learning.

Understanding Behavioral Metric Learning: A Large-Scale Study on Distracting Reinforcement Learning Environments Dropout Q-Functions for Doubly Efficient Reinforcement Learning

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-07T12:10:01.582830Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:10:01.582830Z digest=sha256:113fc71d7276de10c110d40a9256805c5202457854e27f0ef4f0d1e280607c87

Observation 5e7b8d07-4517-4bae-a4f8-a34032663cfe · outbound

This paper cites Offline RL with Observation Histories: Analyzing and Improving Sample Complexity.

Understanding Behavioral Metric Learning: A Large-Scale Study on Distracting Reinforcement Learning Environments Offline RL with Observation Histories: Analyzing and Improving Sample Complexity

Reference 29

Resolution
verified exact
local_arxiv, observed 2026-08-07T12:10:09.535229Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-08-07T12:10:01.802325Z digest=sha256:bb4fbcd7e32803fe7fe8d25ea9c9d12f95c98a43e5dd4cd46cca39f1d4a503b6

Observation b9e03a9b-b85f-4c60-9a63-d11d2b49d2c0 · outbound

This paper cites Robust estimation of a location parameter.

Understanding Behavioral Metric Learning: A Large-Scale Study on Distracting Reinforcement Learning Environments Robust estimation of a location parameter

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:10:15.121679Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-08-07T12:10:01.956563Z digest=sha256:a47783d634cf8d728e546274cfac44c07e68c6d1daa314c05a96bd6fae3db67e

Observation 69f86b3d-c90e-4df1-a115-c076f20e78ff · outbound

This paper cites Dissecting Deep RL with High Update Ratios: Combatting Value Divergence.

Understanding Behavioral Metric Learning: A Large-Scale Study on Distracting Reinforcement Learning Environments Dissecting Deep RL with High Update Ratios: Combatting Value Divergence

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-07T12:10:02.108907Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:10:02.108907Z digest=sha256:5f68dd2d80ffa95867a02b2e84bc32a13c63c7f295c0de7271fdc3402aa96f52

Observation 6a92fa29-fa0a-40b3-9afe-b324da8b8596 · outbound

This paper cites Agent-Controller Representations: Principled Offline RL with Rich Exogenous Information.

Understanding Behavioral Metric Learning: A Large-Scale Study on Distracting Reinforcement Learning Environments Agent-Controller Representations: Principled Offline RL with Rich Exogenous Information

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-07T12:10:02.272630Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:10:02.272630Z digest=sha256:8cf9bc7860fd72073bea0c33a53382b7d535267ce187b0f0074ce6c726daf523

Observation 4a5fd4cf-f48b-4944-97ae-96187a520c25 · outbound

This paper cites Notes on state abstractions, 2018.

Understanding Behavioral Metric Learning: A Large-Scale Study on Distracting Reinforcement Learning Environments Notes on state abstractions, 2018

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:10:14.927607Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-08-07T12:10:02.378080Z digest=sha256:f6f4f244765b2e0dee4155ff03059a20cc2f3d74ce09b1ef81876cb93ace7d65

Observation 41f3f372-d33e-492c-884e-7fb7725a1911 · outbound

This paper cites The Kinetics Human Action Video Dataset.

Understanding Behavioral Metric Learning: A Large-Scale Study on Distracting Reinforcement Learning Environments The Kinetics Human Action Video Dataset

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-07T12:10:02.500205Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:10:02.500205Z digest=sha256:8c64027fd6a527665a58eab7c77ae2739b19089ba4a9ac57dc5e3fb20a24f223

Observation 775c2685-2da6-4591-9c80-0e17ef943d0c · outbound

This paper cites Towards robust bisimulation metric learning.

Understanding Behavioral Metric Learning: A Large-Scale Study on Distracting Reinforcement Learning Environments Towards robust bisimulation metric learning

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:10:14.772844Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-08-07T12:10:02.620055Z digest=sha256:5079d5b1105c4734233a87a3714eec8792bac29e227e91eaddc0730aad910bd8

Observation e54fcacf-ade9-48b2-86c1-e823d1dd9d15 · outbound

This paper cites Actor-critic algorithms.

Understanding Behavioral Metric Learning: A Large-Scale Study on Distracting Reinforcement Learning Environments Actor-critic algorithms

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-07T12:10:02.701330Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:10:02.701330Z digest=sha256:a536e22f59b677d7a2e19fd954218e4b99cdaa0d8bb2daed146d0bf029fa5bec

Observation 411d86a9-8572-41fd-ae03-3ea0008967ed · outbound

This paper cites On the necessity of abstraction.

Understanding Behavioral Metric Learning: A Large-Scale Study on Distracting Reinforcement Learning Environments On the necessity of abstraction

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:10:14.633089Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-08-07T12:10:02.807491Z digest=sha256:6c808091f645112ef12c2dfd8d4b8c20a143d825e722244fbdec57daf4bfcdb0

Observation b28ae5d3-817d-4871-8aa2-e3586711a174 · outbound

This paper cites Towards a unified theory of state abstraction for mdps.

Understanding Behavioral Metric Learning: A Large-Scale Study on Distracting Reinforcement Learning Environments Towards a unified theory of state abstraction for mdps

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:10:14.497410Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-08-07T12:10:02.901775Z digest=sha256:30b6611211847005138f0e31829a0a0a693ff1f4dbd4e27c926fa7fa750fef9e

Observation 2f008026-a793-43c8-8cc6-de98b8a454bd · outbound

This paper cites Normalization Enhances Generalization in Visual Reinforcement Learning.

Understanding Behavioral Metric Learning: A Large-Scale Study on Distracting Reinforcement Learning Environments Normalization Enhances Generalization in Visual Reinforcement Learning

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-07T12:10:03.069253Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:10:03.069253Z digest=sha256:65e4254f8b8e1957cdd754a034125ad414e4407e2e7a08e8fff87599512934ea

Observation 115ec923-15e0-4ac2-939b-fa8505fcac96 · outbound

This paper cites Does self-supervised learning really improve reinforcement learning from pixels? Advances in Neural Information Processing Systems, 35: 0 30865--30881, 2022.

Understanding Behavioral Metric Learning: A Large-Scale Study on Distracting Reinforcement Learning Environments Does self-supervised learning really improve reinforcement learning from pixels? Advances in Neural Information Processing Systems, 35: 0 30865--30881, 2022

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:10:14.403204Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-08-07T12:10:03.185882Z digest=sha256:9ea9304f21ecf8abb519ca04e2f28122540eb32d1af04fff108a41af7bc97bc8

Observation 72e401eb-9734-402d-ba76-e8d8684832be · outbound

This paper cites Policy-independent behavioral metric-based representation for deep reinforcement learning.

Understanding Behavioral Metric Learning: A Large-Scale Study on Distracting Reinforcement Learning Environments Policy-independent behavioral metric-based representation for deep reinforcement learning

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:10:14.212925Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-08-07T12:10:03.335536Z digest=sha256:daa5a02b35aa64143a2bb0aa2ca5668124ec7870c2a6061d0c4d3c3944f5ddf1

Observation 267e61b1-7c37-48b5-838e-01ae14620537 · outbound

This paper cites Robust representation learning by clustering with bisimulation metrics for visual reinforcement learning with distractions.

Understanding Behavioral Metric Learning: A Large-Scale Study on Distracting Reinforcement Learning Environments Robust representation learning by clustering with bisimulation metrics for visual reinforcement learning with distractions

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:10:14.066667Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-08-07T12:10:03.488758Z digest=sha256:c5f23d76814fbf858c8534a3bce7f84102f876de9c8c23e4e7b9ec92d9feaef9

Observation f5dd29b7-3fd7-4adc-9231-5630e1c5563c · outbound

This paper cites A calculus of communicating systems.

Understanding Behavioral Metric Learning: A Large-Scale Study on Distracting Reinforcement Learning Environments A calculus of communicating systems

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:10:13.862971Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-08-07T12:10:03.623907Z digest=sha256:26bfd28846548fa22870b6ba405043ea1bd50690a27add62d93befa6e3b9a100

Observation 65f2fd91-e0f9-4f6c-acc1-bbddd58b6956 · outbound

This paper cites Learning Temporal Distances: Contrastive Successor Features Can Provide a Metric Structure for Decision-Making.

Understanding Behavioral Metric Learning: A Large-Scale Study on Distracting Reinforcement Learning Environments Learning Temporal Distances: Contrastive Successor Features Can Provide a Metric Structure for Decision-Making

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-07T12:10:03.799665Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:10:03.799665Z digest=sha256:2e438bc1e5586fc98cf22faf34dcfd74fd36913b58eb3c30957abd621052b7ac

Observation 8fd96c6b-db8e-4849-a3f8-57a82ec57793 · outbound

This paper cites Overestimation, Overfitting, and Plasticity in Actor-Critic: the Bitter Lesson of Reinforcement Learning.

Understanding Behavioral Metric Learning: A Large-Scale Study on Distracting Reinforcement Learning Environments Overestimation, Overfitting, and Plasticity in Actor-Critic: the Bitter Lesson of Reinforcement Learning

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-07T12:10:03.917362Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:10:03.917362Z digest=sha256:51cca0b5920256356375a890ebce8ee47343090f9c4482a8fa9c2dd5314db10b

Observation c7d1b2e4-9bf3-443a-8c6e-ece296354c84 · outbound

This paper cites Bridging State and History Representations: Understanding Self-Predictive RL.

Understanding Behavioral Metric Learning: A Large-Scale Study on Distracting Reinforcement Learning Environments Bridging State and History Representations: Understanding Self-Predictive RL

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-07T12:10:04.252390Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:10:04.252390Z digest=sha256:d9ec7c6bebe761fe5cc9a430ea6f8598df372e16bbeffde3331664b96eabf549

Observation 931028da-d570-487c-9abd-9ba7b96635b7 · outbound

This paper cites Control-oriented model-based reinforcement learning with implicit differentiation.

Understanding Behavioral Metric Learning: A Large-Scale Study on Distracting Reinforcement Learning Environments Control-oriented model-based reinforcement learning with implicit differentiation

Reference 47

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:10:13.688357Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-08-07T12:10:04.606982Z digest=sha256:2043eca1900e147b6842887db6fccb6ccb7ff6a569bd3aead77a7f8fb23771b1

Observation 54cb92e7-247e-4cd7-92c0-07564015d10b · outbound

This paper cites Labelled Markov Processes.

Understanding Behavioral Metric Learning: A Large-Scale Study on Distracting Reinforcement Learning Environments Labelled Markov Processes

Reference 48

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:10:13.476146Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-08-07T12:10:05.193640Z digest=sha256:e49ba8bb6929011f1ae1d8c44337a347b8a8db3183c64033ecffe6e8334d7691

Observation fcedacdf-7d37-48e7-90da-33688ff09bd1 · outbound

This paper cites Policy gradient methods in the presence of symmetries and state abstractions.

Understanding Behavioral Metric Learning: A Large-Scale Study on Distracting Reinforcement Learning Environments Policy gradient methods in the presence of symmetries and state abstractions

Reference 49

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:10:13.298662Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-08-07T12:10:06.399252Z digest=sha256:c35356cf5cf1b77889bc44789a6399e89abebfb4e6e97635486601d604eef306

Observation a2831909-7709-4ddd-81ba-8b1f2cc6dca0 · outbound

This paper cites Concurrency and automata on infinite sequences.

Understanding Behavioral Metric Learning: A Large-Scale Study on Distracting Reinforcement Learning Environments Concurrency and automata on infinite sequences

Reference 50

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:10:13.122998Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-08-07T12:10:06.461937Z digest=sha256:b13d5edc9c841146ab603093c761a05fa552db67b0ad78ca66316f5f2d8947ce

Observation 25533d6b-be35-4a66-9adc-84251b81fdbb · outbound

This paper cites State-action similarity-based representations for off-policy evaluation.

Understanding Behavioral Metric Learning: A Large-Scale Study on Distracting Reinforcement Learning Environments State-action similarity-based representations for off-policy evaluation

Reference 51

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:10:12.931278Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-08-07T12:10:06.534084Z digest=sha256:c1d66c1afee3e0e47b6ce44c372b81e2ccf1bfe545c2dae747c7c93f3be5fa74

Observation e443b986-897a-4bff-8345-e822fcbbd9cc · outbound

This paper cites An algebraic approach to abstraction in reinforcement learning.

Understanding Behavioral Metric Learning: A Large-Scale Study on Distracting Reinforcement Learning Environments An algebraic approach to abstraction in reinforcement learning

Reference 52

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:10:12.706810Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-08-07T12:10:06.595261Z digest=sha256:56f3a65fd9d3d5bb17fa37351650b95913f96ad89e83fb3101b9cf5059ae24a4

Observation e20b064e-c1c4-4cf3-97e9-984df4c0f430 · outbound

This paper cites Model minimization in hierarchical reinforcement learning.

Understanding Behavioral Metric Learning: A Large-Scale Study on Distracting Reinforcement Learning Environments Model minimization in hierarchical reinforcement learning

Reference 53

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:10:12.514760Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-08-07T12:10:06.695657Z digest=sha256:79f62df5c312d97db83b9ff1a5be93a1528cff85b15c061c3805cfb6af5d298a

Observation fb3364eb-b99b-4fe9-bf9e-ef2453325940 · outbound

This paper cites Continuous mdp homomorphisms and homomorphic policy gradient.

Understanding Behavioral Metric Learning: A Large-Scale Study on Distracting Reinforcement Learning Environments Continuous mdp homomorphisms and homomorphic policy gradient

Reference 54

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:10:12.371374Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-08-07T12:10:06.791035Z digest=sha256:cb35353da6af0c5bfaa6a405367449bae83bfa3c8ada877a5b856b7cca8a9fdc

Observation 5d1e18c3-e85b-4599-8fcf-5576952ec28c · outbound

This paper cites Learning Action-based Representations Using Invariance.

Understanding Behavioral Metric Learning: A Large-Scale Study on Distracting Reinforcement Learning Environments Learning Action-based Representations Using Invariance

Reference 55

Resolution
unresolved
no resolver link, observed 2026-08-07T12:10:06.863385Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:10:06.863385Z digest=sha256:e645e0c025a2aff613e2f6b60566294fd7fe87e8f0f4e0bf789017fd110f635b

Observation 0b58ff51-e8f0-48eb-b82f-6a086d098ca2 · outbound

This paper cites Facenet: A unified embedding for face recognition and clustering.

Understanding Behavioral Metric Learning: A Large-Scale Study on Distracting Reinforcement Learning Environments Facenet: A unified embedding for face recognition and clustering

Reference 56

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:10:12.215444Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-08-07T12:10:06.946603Z digest=sha256:85d531c8761d46d72bec3f845134c8c2e39573eb99a33e7e7ceb84f6edd65a14

Observation aacff5eb-c4b3-4ba5-ac6d-3b9a6edfe064 · outbound

This paper cites Data-Efficient Reinforcement Learning with Self-Predictive Representations.

Understanding Behavioral Metric Learning: A Large-Scale Study on Distracting Reinforcement Learning Environments Data-Efficient Reinforcement Learning with Self-Predictive Representations

Reference 57

Resolution
unresolved
no resolver link, observed 2026-08-07T12:10:07.046569Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:10:07.046569Z digest=sha256:6a72dd0845c34ceee235d1f4910b474e1c3181e48881acdcdf896849d05ac0ab

Observation 5fd4873b-5137-480a-9f76-3412cba45740 · outbound

This paper cites Bisimulation metric for Model Predictive Control.

Understanding Behavioral Metric Learning: A Large-Scale Study on Distracting Reinforcement Learning Environments Bisimulation metric for Model Predictive Control

Reference 58

Resolution
unresolved
no resolver link, observed 2026-08-07T12:10:07.161039Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:10:07.161039Z digest=sha256:74416dc36d85292dd9b8223d6f05784b8f42e47a6a7e0ee9a8caaf402a600d34

Observation ffeb9102-a325-48c2-8769-e5630f47905e · outbound

This paper cites Reinforcement learning with soft state aggregation.

Understanding Behavioral Metric Learning: A Large-Scale Study on Distracting Reinforcement Learning Environments Reinforcement learning with soft state aggregation

Reference 59

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:10:12.002559Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-08-07T12:10:07.288340Z digest=sha256:e2a05e9fca27af8731877cbd07cd1e685c04673c0c1ae278bc25a1249aaf9f9a

Observation 9049565e-2b4a-44b2-8be0-bc466da837ba · outbound

This paper cites A Walk in the Park: Learning to Walk in 20 Minutes With Model-Free Reinforcement Learning.

Understanding Behavioral Metric Learning: A Large-Scale Study on Distracting Reinforcement Learning Environments A Walk in the Park: Learning to Walk in 20 Minutes With Model-Free Reinforcement Learning

Reference 60

Resolution
unresolved
no resolver link, observed 2026-08-07T12:10:07.379760Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:10:07.379760Z digest=sha256:c5d642287be19708c69b3415c5bc7189fe86a334a460ca1b3ddfc4f86ccaa5bb

Observation bbe7f5bf-f3f4-4e6c-ac12-f37294582b35 · outbound

This paper cites The Distracting Control Suite -- A Challenging Benchmark for Reinforcement Learning from Pixels.

Understanding Behavioral Metric Learning: A Large-Scale Study on Distracting Reinforcement Learning Environments The Distracting Control Suite -- A Challenging Benchmark for Reinforcement Learning from Pixels

Reference 61

Resolution
unresolved
no resolver link, observed 2026-08-07T12:10:07.468224Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:10:07.468224Z digest=sha256:889cffee0623538dda868b615ff378a4c9f21754cd6b01165471865d6f1b8431

Observation 5d38d754-8f0d-4bd0-9436-f6c5b17ea498 · outbound

This paper cites Approximate information state for approximate planning and reinforcement learning in partially observed systems.

Understanding Behavioral Metric Learning: A Large-Scale Study on Distracting Reinforcement Learning Environments Approximate information state for approximate planning and reinforcement learning in partially observed systems

Reference 62

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:10:11.853121Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-08-07T12:10:07.537316Z digest=sha256:fa74decd18f6371a457e7267aefa0232bc04e4878f98792b91964025fb3fb1c0

Observation 2bdfd756-83ac-44a5-aafa-613e25fea1e5 · outbound

This paper cites DeepMind Control Suite.

Understanding Behavioral Metric Learning: A Large-Scale Study on Distracting Reinforcement Learning Environments DeepMind Control Suite

Reference 63

Resolution
unresolved
no resolver link, observed 2026-08-07T12:10:07.623587Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:10:07.623587Z digest=sha256:8986cc61dc8b40e349ced79d93538e64e113797257569a2c905371b7631fb471

Observation 5cadb823-c646-48a9-968e-b1a8aae70913 · outbound

This paper cites Lax probabilistic bisimulation.

Understanding Behavioral Metric Learning: A Large-Scale Study on Distracting Reinforcement Learning Environments Lax probabilistic bisimulation

Reference 64

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:10:11.635694Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-08-07T12:10:07.691352Z digest=sha256:b89998e12828495fc59241b8bec2dfc21549f4d226fd044f6ba53ee4bb49f715

Observation 6d4f6ab7-e5e1-416c-b58c-f381b4871e22 · outbound

This paper cites Learning Representations for Pixel-based Control: What Matters and Why?.

Understanding Behavioral Metric Learning: A Large-Scale Study on Distracting Reinforcement Learning Environments Learning Representations for Pixel-based Control: What Matters and Why?

Reference 65

Resolution
unresolved
no resolver link, observed 2026-08-07T12:10:07.756955Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:10:07.756955Z digest=sha256:46ce12d76cb74cb2e32e40c822d3cef0f93c96560b99d304362d8811283d03ba

Observation d2ad554f-ef64-4e00-94e9-dd55c99069ee · outbound

This paper cites dm\_control: Software and tasks for continuous control.

Understanding Behavioral Metric Learning: A Large-Scale Study on Distracting Reinforcement Learning Environments dm\_control: Software and tasks for continuous control

Reference 66

Resolution
unresolved
no resolver link, observed 2026-08-07T12:10:07.869845Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:10:07.869845Z digest=sha256:8c9a50e32ee3610f70e153bfd4c3bf08bcf2d7a1511f4fc5c13779814dd5f3c7

Observation a1863e4a-5ea5-45cd-935c-f9893becef74 · outbound

This paper cites Plannable Approximations to MDP Homomorphisms: Equivariance under Actions.

Understanding Behavioral Metric Learning: A Large-Scale Study on Distracting Reinforcement Learning Environments Plannable Approximations to MDP Homomorphisms: Equivariance under Actions

Reference 67

Resolution
unresolved
no resolver link, observed 2026-08-07T12:10:07.966060Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:10:07.966060Z digest=sha256:655d568d1f4aec386e82137d25293717a48f6c9c00d2e43fb0d8ef692030bbf0

Observation 25028b49-4390-48a3-8642-48798d115c7c · outbound

This paper cites Mdp homomorphic networks: Group symmetries in reinforcement learning.

Understanding Behavioral Metric Learning: A Large-Scale Study on Distracting Reinforcement Learning Environments Mdp homomorphic networks: Group symmetries in reinforcement learning

Reference 68

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:10:11.402431Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-08-07T12:10:08.037509Z digest=sha256:8d671aefa89d33e5fc512f7d6c49af00f176503674d0b983bb31a7d29e5b5006

Observation 76f8c420-993f-4d08-8869-c3e699ab28ca · outbound

This paper cites When does Self-Prediction help? Understanding Auxiliary Tasks in Reinforcement Learning.

Understanding Behavioral Metric Learning: A Large-Scale Study on Distracting Reinforcement Learning Environments When does Self-Prediction help? Understanding Auxiliary Tasks in Reinforcement Learning

Reference 69

Resolution
unresolved
no resolver link, observed 2026-08-07T12:10:08.108971Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:10:08.108971Z digest=sha256:eea83b84b6b7e07cdd70d2a05c2e99afeeef51186812c307106c79d1eb63134a

Observation 977ac222-ea26-496e-a7a7-baedeaf9a543 · outbound

This paper cites Efficient potential-based exploration in reinforcement learning using inverse dynamic bisimulation metric.

Understanding Behavioral Metric Learning: A Large-Scale Study on Distracting Reinforcement Learning Environments Efficient potential-based exploration in reinforcement learning using inverse dynamic bisimulation metric

Reference 70

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:10:11.238699Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-08-07T12:10:08.176373Z digest=sha256:3198ee304b682fc2d5a7886fd606fe9d253fd4fb4f3328a879ed929e288e14e1

Observation e90c7c54-1f0e-42be-b1c1-c6d223bff424 · outbound

This paper cites Rethinking exploration in reinforcement learning with effective metric-based exploration bonus.

Understanding Behavioral Metric Learning: A Large-Scale Study on Distracting Reinforcement Learning Environments Rethinking exploration in reinforcement learning with effective metric-based exploration bonus

Reference 71

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:10:11.048518Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-08-07T12:10:08.271332Z digest=sha256:2eab84563c560c10b8f9f1023853d5e648dfab9536290b5b48a455b963fe6197

Observation 9c9c5dd5-7582-4680-9f1a-78961d4f25b2 · outbound

This paper cites Mastering Visual Continuous Control: Improved Data-Augmented Reinforcement Learning.

Understanding Behavioral Metric Learning: A Large-Scale Study on Distracting Reinforcement Learning Environments Mastering Visual Continuous Control: Improved Data-Augmented Reinforcement Learning

Reference 72

Resolution
unresolved
no resolver link, observed 2026-08-07T12:10:08.362608Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:10:08.362608Z digest=sha256:69f2f3b3561613d70f680d3e2f509880f22f33c4ec60760c7ddb8ee41137b1c4

Observation b681e948-fb94-4a12-8579-4c9251b77a80 · outbound

This paper cites Improving sample efficiency in model-free reinforcement learning from images.

Understanding Behavioral Metric Learning: A Large-Scale Study on Distracting Reinforcement Learning Environments Improving sample efficiency in model-free reinforcement learning from images

Reference 73

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:10:10.905731Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-08-07T12:10:08.455921Z digest=sha256:0ea0fb86e03f080788e45ed029f6fd6377bc90bab71f1a5646cb752165c1118e

Observation ca9e0fd8-5706-4d22-929f-61b9a4722915 · outbound

This paper cites Rl-vigen: A reinforcement learning benchmark for visual generalization.

Understanding Behavioral Metric Learning: A Large-Scale Study on Distracting Reinforcement Learning Environments Rl-vigen: A reinforcement learning benchmark for visual generalization

Reference 74

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:10:10.732423Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-08-07T12:10:08.524623Z digest=sha256:ed18969b0b93e435a753311c5711f49745131c10632b7b7a98a2eee0d97ac4cc

Observation 88f09758-2d04-4bc8-91c3-0a9c7bbb457c · outbound

This paper cites Simsr: Simple distance-based state representations for deep reinforcement learning.

Understanding Behavioral Metric Learning: A Large-Scale Study on Distracting Reinforcement Learning Environments Simsr: Simple distance-based state representations for deep reinforcement learning

Reference 75

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:10:10.516198Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-08-07T12:10:08.606269Z digest=sha256:754241b2a971bff5d9b246d93a590941ebded0c6e7703a06b884c4134ff0b76e

Observation 336e1ed6-9e01-4c97-991c-7509e3e44bf2 · outbound

This paper cites Understanding and addressing the pitfalls of bisimulation-based representations in offline reinforcement learning.

Understanding Behavioral Metric Learning: A Large-Scale Study on Distracting Reinforcement Learning Environments Understanding and addressing the pitfalls of bisimulation-based representations in offline reinforcement learning

Reference 76

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:10:10.296954Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-08-07T12:10:08.728222Z digest=sha256:e561f5b2cc27b5e67efe04828f4f7e6acebb6db97a198d2737e78d3dab2ca433

Observation 1b9d6aa3-31f5-480a-8477-4de65d0ab96f · outbound

This paper cites Natural Environment Benchmarks for Reinforcement Learning.

Understanding Behavioral Metric Learning: A Large-Scale Study on Distracting Reinforcement Learning Environments Natural Environment Benchmarks for Reinforcement Learning

Reference 77

Resolution
unresolved
no resolver link, observed 2026-08-07T12:10:08.787938Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:10:08.787938Z digest=sha256:c21c7ffde10c601fe7e5dec61f676bc4b4638f59f4a46e58b2050cd4ed3c503b

Observation c31911ca-51a1-4fe3-9d1f-3e4581e9afdb · outbound

This paper cites Learning Invariant Representations for Reinforcement Learning without Reconstruction.

Understanding Behavioral Metric Learning: A Large-Scale Study on Distracting Reinforcement Learning Environments Learning Invariant Representations for Reinforcement Learning without Reconstruction

Reference 78

Resolution
unresolved
no resolver link, observed 2026-08-07T12:10:08.875246Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:10:08.875246Z digest=sha256:3419dd04ae002be1a2051c31e74ce8b0d695d258e361e3f8ea0898740c373172

Observation 4ac23cb1-3a23-4855-8ef0-985cb1ad2783 · outbound

This paper cites write newline.

Understanding Behavioral Metric Learning: A Large-Scale Study on Distracting Reinforcement Learning Environments write newline

Reference 79

Resolution
unresolved
no resolver link, observed 2026-08-07T12:10:08.987854Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:10:08.987854Z digest=sha256:1205c1fe09350e7a34a39e48156b9092a45e1a616702a5893d09ad3b3ff237df

Pith citing papers

Observation e2946632-5655-4d00-b5b1-88a0e843cd76 · inbound

SAVGO: Learning State-Action Value Geometry with Cosine Similarity for Continuous Control cites this paper.

SAVGO: Learning State-Action Value Geometry with Cosine Similarity for Continuous Control Understanding Behavioral Metric Learning: A Large-Scale Study on Distracting Reinforcement Learning Environments

Reference 30

Resolution
verified exact
arxiv_id, observed 2026-05-09T20:01:36.584993Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-05-09T20:00:55.197662Z digest=sha256:c937323caf074dea2e71d4a9036308c2aefce0d54ab0b4f734070a1307834482