Pith. sign in

Paper Citation Record · LEDGER

Understanding Behavioral Metric Learning: A Large-Scale Study on Distracting Reinforcement Learning Environments

As of 7 August 2026, this Paper Citation Record lists 79 of 79 outbound references and 1 inbound Pith citation observation for arXiv:2506.00563.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2506.00563 v2

Coverage vector

measured 79 of 79 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T12:10:08.987854Z

measured 80 of 80 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00

measured 1 of 1 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-05-09T20:00:55.197662Z

measured 1 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Reference resolution

79 of 79 outbound references displayed

  • verified exact3
  • verified fuzzy44
  • unresolved32
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

0
arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Outbound references

Observation e14aaef1-00af-43f1-bff2-4e5f80decd75 · outbound

This paper cites Contrastive Behavioral Similarity Embeddings for Generalization in Reinforcement Learning.

Understanding Behavioral Metric Learning: A Large-Scale Study on Distracting Reinforcement Learning Environments Contrastive Behavioral Similarity Embeddings for Generalization in Reinforcement Learning

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-07T12:09:58.708141Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:09:58.708141Z digest=sha256:f7644fad1a52a92afc23027118646e762390fa9319287a6002ea9efad852e649

Observation ac61ba63-31a3-4367-8b37-f9181c824f60 · outbound

This paper cites Deep reinforcement learning at the edge of the statistical precipice.

Understanding Behavioral Metric Learning: A Large-Scale Study on Distracting Reinforcement Learning Environments Deep reinforcement learning at the edge of the statistical precipice

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:10:18.711466Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T12:09:58.870069Z digest=sha256:3fb3e8a57e8c5aa4d3c18f44d4965c1a2e0795672c9cd2fc0ff9d1c94e10688c

Observation 5716f765-391f-408b-9072-d16b236b6e21 · outbound

This paper cites Layer Normalization.

Understanding Behavioral Metric Learning: A Large-Scale Study on Distracting Reinforcement Learning Environments Layer Normalization

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-07T12:09:58.996705Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:09:58.996705Z digest=sha256:68394cd71fd6f32335a1d24faffe78bd2febe782995a8f5c5751e0a47bb6e9bd

Observation 86a6e4d0-061c-4591-9a1b-8bbe15e8a26a · outbound

This paper cites CrossQ: Batch Normalization in Deep Reinforcement Learning for Greater Sample Efficiency and Simplicity.

Understanding Behavioral Metric Learning: A Large-Scale Study on Distracting Reinforcement Learning Environments CrossQ: Batch Normalization in Deep Reinforcement Learning for Greater Sample Efficiency and Simplicity

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-07T12:09:59.140424Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:09:59.140424Z digest=sha256:981c7cee148820f69675e4109db0083cf9ba45271f67c0e51e6480beecf3211f

Observation 8bb28be9-cd18-4ca9-addf-f0d8e9b72f07 · outbound

This paper cites Online Abstraction with MDP Homomorphisms for Deep Learning.

Understanding Behavioral Metric Learning: A Large-Scale Study on Distracting Reinforcement Learning Environments Online Abstraction with MDP Homomorphisms for Deep Learning

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-07T12:09:59.274698Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:09:59.274698Z digest=sha256:1c1064e0d8189cf072979aec8b6db223106c8c42238d245ec1dd53332da442c1

Observation 55d96dde-b654-44c2-9817-6697f6912d7f · outbound

This paper cites Scalable methods for computing state similarity in deterministic markov decision processes.

Understanding Behavioral Metric Learning: A Large-Scale Study on Distracting Reinforcement Learning Environments Scalable methods for computing state similarity in deterministic markov decision processes

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-07T12:09:59.392893Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:09:59.392893Z digest=sha256:fa636f5dce2b0c6b478d8ef72ac3263f6fafec1f1a56ef91203cd73dd6165fc3

Observation bd4db822-0655-46a0-b212-3b04738494a3 · outbound

This paper cites Mico: Improved representations via sampling-based state similarity for markov decision processes.

Understanding Behavioral Metric Learning: A Large-Scale Study on Distracting Reinforcement Learning Environments Mico: Improved representations via sampling-based state similarity for markov decision processes

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:10:18.409970Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T12:09:59.461170Z digest=sha256:73cf64ad4d629d5cb5496c4819161b1fee5df22ce2542239b98c8990527c12fe

Observation bfd2714a-d32b-49e7-9c28-39585c54c841 · outbound

This paper cites A Kernel Perspective on Behavioural Metrics for Markov Decision Processes.

Understanding Behavioral Metric Learning: A Large-Scale Study on Distracting Reinforcement Learning Environments A Kernel Perspective on Behavioural Metrics for Markov Decision Processes

Reference 8

Resolution
verified exact
local_arxiv, observed 2026-08-07T12:10:09.934997Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T12:09:59.548686Z digest=sha256:632dd0caab1f2db7e607a403202ff2edac624e084e0c263e9adf016e9ef6b31b

Observation fc4e8565-3c9c-4dc4-818c-f3b254e2ce68 · outbound

This paper cites Learning representations via a robust behavioral metric for deep reinforcement learning.

Understanding Behavioral Metric Learning: A Large-Scale Study on Distracting Reinforcement Learning Environments Learning representations via a robust behavioral metric for deep reinforcement learning

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:10:18.177019Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T12:09:59.652201Z digest=sha256:272f91e79c5a143992b65b545080a49cfc6176eaf4f2be07547f110e8cc5ba15

Observation 5f28d638-bdc7-4b0d-a90f-3d50cb07acc2 · outbound

This paper cites State chrono representation for enhancing generalization in reinforcement learning.

Understanding Behavioral Metric Learning: A Large-Scale Study on Distracting Reinforcement Learning Environments State chrono representation for enhancing generalization in reinforcement learning

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:10:17.950416Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T12:09:59.778307Z digest=sha256:3b459fd70b3a39cf3a334297f86c1c7a33f8cddf42d5f0ef82cbbd2a21d8757e

Observation 160a0d47-bb4c-42ed-b586-54d8fcf5d89c · outbound

This paper cites Offline reinforcement learning with pseudometric learning.

Understanding Behavioral Metric Learning: A Large-Scale Study on Distracting Reinforcement Learning Environments Offline reinforcement learning with pseudometric learning

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:10:17.743648Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T12:09:59.881247Z digest=sha256:ef18c8db5c059d99afeddc58b14bf7c49d216c33563da1f372e24e3a3697f6f6

Observation 1ed9f1d8-353a-4f1c-ac81-b080d213cbc6 · outbound

This paper cites Bisimulation for labelled markov processes.

Understanding Behavioral Metric Learning: A Large-Scale Study on Distracting Reinforcement Learning Environments Bisimulation for labelled markov processes

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:10:17.530504Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T12:09:59.982921Z digest=sha256:e0620545e826b46f14c4b7f6ce804afe65883d3a93810e264dbacfccc29f768e

Observation af24be27-ebb8-4e44-96da-2f2d6a1d0137 · outbound

This paper cites Provably efficient rl with rich observations via latent state decoding.

Understanding Behavioral Metric Learning: A Large-Scale Study on Distracting Reinforcement Learning Environments Provably efficient rl with rich observations via latent state decoding

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:10:17.325024Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T12:10:00.096030Z digest=sha256:f6fda7f9ce1c97137bd57b084544c2eb77a72a8d745fed78764f066e26aae175

Observation 2963fefa-d659-4957-a574-5e8ab0b217c5 · outbound

This paper cites Differential privacy.

Understanding Behavioral Metric Learning: A Large-Scale Study on Distracting Reinforcement Learning Environments Differential privacy

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-07T12:10:00.192754Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:10:00.192754Z digest=sha256:5bb00f53bef9c9dc56a395e8fc61d9f51ca0b10be6d23a31b5876680dc1d4987

Observation e9804ab4-9cc1-4052-93c8-2b9f2fc1691e · outbound

This paper cites Provable RL with Exogenous Distractors via Multistep Inverse Dynamics.

Understanding Behavioral Metric Learning: A Large-Scale Study on Distracting Reinforcement Learning Environments Provable RL with Exogenous Distractors via Multistep Inverse Dynamics

Reference 15

Resolution
verified exact
local_arxiv, observed 2026-08-07T12:10:09.765339Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T12:10:00.264076Z digest=sha256:8fe785ee4aa63d26a19d69e1f2f883d513be5a4d30b6bd4cb63929987df86821

Observation c6e73560-ecf9-49a2-ab06-ed543ac55529 · outbound

This paper cites Metrics for finite markov decision processes.

Understanding Behavioral Metric Learning: A Large-Scale Study on Distracting Reinforcement Learning Environments Metrics for finite markov decision processes

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:10:17.118699Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T12:10:00.324820Z digest=sha256:87ca3b4a38de852689f5741912472c0221673aa2715acbe796fba831c2c9daa6

Observation a353b52e-0a78-41f2-b7d3-0dc4330feb8e · outbound

This paper cites Bisimulation metrics for continuous markov decision processes.

Understanding Behavioral Metric Learning: A Large-Scale Study on Distracting Reinforcement Learning Environments Bisimulation metrics for continuous markov decision processes

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:10:16.873188Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T12:10:00.423106Z digest=sha256:2e3fbe0a2fff81a64a4d70203a31c74a4addb902bc62b38b397adf0d07576d3b

Observation c0d954e0-386a-4314-8908-e142b08ddef6 · outbound

This paper cites For sale: State-action representation learning for deep reinforcement learning.

Understanding Behavioral Metric Learning: A Large-Scale Study on Distracting Reinforcement Learning Environments For sale: State-action representation learning for deep reinforcement learning

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:10:16.665792Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T12:10:00.548679Z digest=sha256:99f04301ec7265345f6bee455a2dcaa98aa67229706a902b4a4982018ffc9f18

Observation 36c45093-8205-432d-a6a0-4303bd0e2635 · outbound

This paper cites Deepmdp: Learning continuous latent space models for representation learning.

Understanding Behavioral Metric Learning: A Large-Scale Study on Distracting Reinforcement Learning Environments Deepmdp: Learning continuous latent space models for representation learning

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:10:16.432994Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T12:10:00.635009Z digest=sha256:5752363b962d427e099633703d29ad4669981ef2e1bb87e495b1aa711d799bbc

Observation 6aba3156-7c65-41a0-b0ee-ad36b0a4b9c1 · outbound

This paper cites Fully homomorphic encryption using ideal lattices.

Understanding Behavioral Metric Learning: A Large-Scale Study on Distracting Reinforcement Learning Environments Fully homomorphic encryption using ideal lattices

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:10:16.249067Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T12:10:00.715862Z digest=sha256:cf5ddf1d2f243609efc4346171e570994db26d52c448af22ac43823d55774ccd

Observation 78335d36-026a-483a-9633-db4893764342 · outbound

This paper cites Equivalence notions and model minimization in markov decision processes.

Understanding Behavioral Metric Learning: A Large-Scale Study on Distracting Reinforcement Learning Environments Equivalence notions and model minimization in markov decision processes

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:10:16.023035Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T12:10:00.797333Z digest=sha256:48b1e0c7de1e6f340b516ad0bf6b59a6e83f3cbea19cc9a52245b438aebdd72f

Observation 86f75616-1fab-4aab-ab42-47d76a67982c · outbound

This paper cites Measuring visual generalization in continuous control from pixels, 2020.

Understanding Behavioral Metric Learning: A Large-Scale Study on Distracting Reinforcement Learning Environments Measuring visual generalization in continuous control from pixels, 2020

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:10:15.823053Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T12:10:00.881709Z digest=sha256:850672dfd6bc33677f8d857a22a810c255b911ba9afffd9a826bc9eecd274ac7

Observation 133587ba-02ac-454b-92de-9184e9cae30c · outbound

This paper cites Bootstrap your own latent-a new approach to self-supervised learning.

Understanding Behavioral Metric Learning: A Large-Scale Study on Distracting Reinforcement Learning Environments Bootstrap your own latent-a new approach to self-supervised learning

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-07T12:10:00.960140Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:10:00.960140Z digest=sha256:c00baf0a938f080327c3e4976b27589fa03712de74ae76f4f8c2568045a7e731

Observation 32926965-1190-44ed-b5a5-933abac577eb · outbound

This paper cites Soft actor-critic: Off-policy maximum entropy deep reinforcement learning with a stochastic actor.

Understanding Behavioral Metric Learning: A Large-Scale Study on Distracting Reinforcement Learning Environments Soft actor-critic: Off-policy maximum entropy deep reinforcement learning with a stochastic actor

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-07T12:10:01.065851Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:10:01.065851Z digest=sha256:28549b9baa3d56eda875b2a2a4093f7c558ada7f4a88202a103b232841229d85

Observation 97653980-44b6-452e-b514-5b9c3d02ef0d · outbound

This paper cites Generalization in reinforcement learning by soft data augmentation.

Understanding Behavioral Metric Learning: A Large-Scale Study on Distracting Reinforcement Learning Environments Generalization in reinforcement learning by soft data augmentation

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:10:15.521464Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T12:10:01.184038Z digest=sha256:2cffb4cbf765672902bcbeb62399c9998981141a11d3a3cbe498c681995015c0

Observation d30c9ffa-b9a8-48f0-bf99-65f1cc9e1859 · outbound

This paper cites TD-MPC2: Scalable, Robust World Models for Continuous Control.

Understanding Behavioral Metric Learning: A Large-Scale Study on Distracting Reinforcement Learning Environments TD-MPC2: Scalable, Robust World Models for Continuous Control

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-07T12:10:01.292716Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:10:01.292716Z digest=sha256:dead573dd6200e1c2d33157a052af92c8b52506ed413fe232155dcdc29cd8033

Observation 401db3f8-06ac-4e22-809c-1fb5192650db · outbound

This paper cites Bisimulation makes analogies in goal-conditioned reinforcement learning.

Understanding Behavioral Metric Learning: A Large-Scale Study on Distracting Reinforcement Learning Environments Bisimulation makes analogies in goal-conditioned reinforcement learning

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:10:15.308971Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T12:10:01.470417Z digest=sha256:7fa26e81253acc2bcec114ade7f535db77d7f900ab86cfeea431a139b38d3b3a

Observation 9590b081-f6c2-42c2-b002-39d4d4229ea1 · outbound

This paper cites Dropout Q-Functions for Doubly Efficient Reinforcement Learning.

Understanding Behavioral Metric Learning: A Large-Scale Study on Distracting Reinforcement Learning Environments Dropout Q-Functions for Doubly Efficient Reinforcement Learning

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-07T12:10:01.582830Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:10:01.582830Z digest=sha256:e0dbcd95d1d9db3f396e7f1fb9254cd84f5b8a2fd4a31e9125fd4aef63fd23d0

Observation 5e7b8d07-4517-4bae-a4f8-a34032663cfe · outbound

This paper cites Offline RL with Observation Histories: Analyzing and Improving Sample Complexity.

Understanding Behavioral Metric Learning: A Large-Scale Study on Distracting Reinforcement Learning Environments Offline RL with Observation Histories: Analyzing and Improving Sample Complexity

Reference 29

Resolution
verified exact
local_arxiv, observed 2026-08-07T12:10:09.535229Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T12:10:01.802325Z digest=sha256:4326fb42c4ffe10bba955ee4b26fce2742c4c98542a46bd05939bf24d16736bf

Observation b9e03a9b-b85f-4c60-9a63-d11d2b49d2c0 · outbound

This paper cites Robust estimation of a location parameter.

Understanding Behavioral Metric Learning: A Large-Scale Study on Distracting Reinforcement Learning Environments Robust estimation of a location parameter

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:10:15.121679Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T12:10:01.956563Z digest=sha256:d4d92d409a3e6808c82fd4b36f4daa34dc8e7b24c04b1e8376019e4420db9359

Observation 69f86b3d-c90e-4df1-a115-c076f20e78ff · outbound

This paper cites Dissecting Deep RL with High Update Ratios: Combatting Value Divergence.

Understanding Behavioral Metric Learning: A Large-Scale Study on Distracting Reinforcement Learning Environments Dissecting Deep RL with High Update Ratios: Combatting Value Divergence

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-07T12:10:02.108907Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:10:02.108907Z digest=sha256:3b86d24f54d365597bc9551079565965977c8ed9b93ce74f861e5bff611e0cd3

Observation 6a92fa29-fa0a-40b3-9afe-b324da8b8596 · outbound

This paper cites Agent-Controller Representations: Principled Offline RL with Rich Exogenous Information.

Understanding Behavioral Metric Learning: A Large-Scale Study on Distracting Reinforcement Learning Environments Agent-Controller Representations: Principled Offline RL with Rich Exogenous Information

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-07T12:10:02.272630Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:10:02.272630Z digest=sha256:85b7f1811467429f17500415b322145dd07c10145c8a734ea3b94ddd5a33dee1

Observation 4a5fd4cf-f48b-4944-97ae-96187a520c25 · outbound

This paper cites Notes on state abstractions, 2018.

Understanding Behavioral Metric Learning: A Large-Scale Study on Distracting Reinforcement Learning Environments Notes on state abstractions, 2018

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:10:14.927607Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T12:10:02.378080Z digest=sha256:bc915b8b9c8f00a7d1a43b880682e910e3a910610a22e9114442e3efeff6573b

Observation 41f3f372-d33e-492c-884e-7fb7725a1911 · outbound

This paper cites The Kinetics Human Action Video Dataset.

Understanding Behavioral Metric Learning: A Large-Scale Study on Distracting Reinforcement Learning Environments The Kinetics Human Action Video Dataset

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-07T12:10:02.500205Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:10:02.500205Z digest=sha256:232a13772baa0b61522dbfbc4abf5d5313d6e60361a38bd898bb5072852fc7c6

Observation 775c2685-2da6-4591-9c80-0e17ef943d0c · outbound

This paper cites Towards robust bisimulation metric learning.

Understanding Behavioral Metric Learning: A Large-Scale Study on Distracting Reinforcement Learning Environments Towards robust bisimulation metric learning

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:10:14.772844Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T12:10:02.620055Z digest=sha256:afc6fc118afa00f4962f8c1ce06e3bbbca49dbf985336f282492e99096e16190

Observation e54fcacf-ade9-48b2-86c1-e823d1dd9d15 · outbound

This paper cites Actor-critic algorithms.

Understanding Behavioral Metric Learning: A Large-Scale Study on Distracting Reinforcement Learning Environments Actor-critic algorithms

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-07T12:10:02.701330Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:10:02.701330Z digest=sha256:fed12523d1cc499447b3e28b434d7b3b503ea607493412c46bd7874bb8471711

Observation 411d86a9-8572-41fd-ae03-3ea0008967ed · outbound

This paper cites On the necessity of abstraction.

Understanding Behavioral Metric Learning: A Large-Scale Study on Distracting Reinforcement Learning Environments On the necessity of abstraction

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:10:14.633089Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T12:10:02.807491Z digest=sha256:7362b06f2447fc52130208984fc750bca3d35b33f60e4206732aec0a9a93be0e

Observation b28ae5d3-817d-4871-8aa2-e3586711a174 · outbound

This paper cites Towards a unified theory of state abstraction for mdps.

Understanding Behavioral Metric Learning: A Large-Scale Study on Distracting Reinforcement Learning Environments Towards a unified theory of state abstraction for mdps

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:10:14.497410Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T12:10:02.901775Z digest=sha256:8f628cee696dca1c0c2b4a4bf4a3f670394337ad942ddc057e5ca5f6ac65ad12

Observation 2f008026-a793-43c8-8cc6-de98b8a454bd · outbound

This paper cites Normalization Enhances Generalization in Visual Reinforcement Learning.

Understanding Behavioral Metric Learning: A Large-Scale Study on Distracting Reinforcement Learning Environments Normalization Enhances Generalization in Visual Reinforcement Learning

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-07T12:10:03.069253Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:10:03.069253Z digest=sha256:f8de57dc7fe0f93adec258831ee4f626324d984ea45b04e7089bfc372d5efe88

Observation 115ec923-15e0-4ac2-939b-fa8505fcac96 · outbound

This paper cites Does self-supervised learning really improve reinforcement learning from pixels? Advances in Neural Information Processing Systems, 35: 0 30865--30881, 2022.

Understanding Behavioral Metric Learning: A Large-Scale Study on Distracting Reinforcement Learning Environments Does self-supervised learning really improve reinforcement learning from pixels? Advances in Neural Information Processing Systems, 35: 0 30865--30881, 2022

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:10:14.403204Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T12:10:03.185882Z digest=sha256:f8868555345cb4fe9d2be4b3651a6b92218d1106bfff99835182552e934dbe14

Observation 72e401eb-9734-402d-ba76-e8d8684832be · outbound

This paper cites Policy-independent behavioral metric-based representation for deep reinforcement learning.

Understanding Behavioral Metric Learning: A Large-Scale Study on Distracting Reinforcement Learning Environments Policy-independent behavioral metric-based representation for deep reinforcement learning

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:10:14.212925Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T12:10:03.335536Z digest=sha256:cafaca2715af626bd61f854c0885a87a031dc6a12fd028e7a3cd15d2f41e04aa

Observation 267e61b1-7c37-48b5-838e-01ae14620537 · outbound

This paper cites Robust representation learning by clustering with bisimulation metrics for visual reinforcement learning with distractions.

Understanding Behavioral Metric Learning: A Large-Scale Study on Distracting Reinforcement Learning Environments Robust representation learning by clustering with bisimulation metrics for visual reinforcement learning with distractions

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:10:14.066667Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T12:10:03.488758Z digest=sha256:84c95fd01d4c3afb9cde3d01a8cdf43929c110f9e301ec053f18223273b2d9d4

Observation f5dd29b7-3fd7-4adc-9231-5630e1c5563c · outbound

This paper cites A calculus of communicating systems.

Understanding Behavioral Metric Learning: A Large-Scale Study on Distracting Reinforcement Learning Environments A calculus of communicating systems

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:10:13.862971Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T12:10:03.623907Z digest=sha256:1f4cc80aee7269e292821b72eafe5435bc054fb7a84b69429bf00de744225c34

Observation 65f2fd91-e0f9-4f6c-acc1-bbddd58b6956 · outbound

This paper cites Learning Temporal Distances: Contrastive Successor Features Can Provide a Metric Structure for Decision-Making.

Understanding Behavioral Metric Learning: A Large-Scale Study on Distracting Reinforcement Learning Environments Learning Temporal Distances: Contrastive Successor Features Can Provide a Metric Structure for Decision-Making

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-07T12:10:03.799665Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:10:03.799665Z digest=sha256:4d5aabbb40e46818f326d26d4f884834f9d64f60445178e4e593326af8b839f8

Observation 8fd96c6b-db8e-4849-a3f8-57a82ec57793 · outbound

This paper cites Overestimation, Overfitting, and Plasticity in Actor-Critic: the Bitter Lesson of Reinforcement Learning.

Understanding Behavioral Metric Learning: A Large-Scale Study on Distracting Reinforcement Learning Environments Overestimation, Overfitting, and Plasticity in Actor-Critic: the Bitter Lesson of Reinforcement Learning

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-07T12:10:03.917362Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:10:03.917362Z digest=sha256:f30370b931dda8e8c3646585e105c69a71d71de7bf2dea99fb0596d06c23c74f

Observation c7d1b2e4-9bf3-443a-8c6e-ece296354c84 · outbound

This paper cites Bridging State and History Representations: Understanding Self-Predictive RL.

Understanding Behavioral Metric Learning: A Large-Scale Study on Distracting Reinforcement Learning Environments Bridging State and History Representations: Understanding Self-Predictive RL

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-07T12:10:04.252390Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:10:04.252390Z digest=sha256:011c0df7418a2f7fdf8635673599ca9102218a9312f10d77a871c3ee97e4349c

Observation 931028da-d570-487c-9abd-9ba7b96635b7 · outbound

This paper cites Control-oriented model-based reinforcement learning with implicit differentiation.

Understanding Behavioral Metric Learning: A Large-Scale Study on Distracting Reinforcement Learning Environments Control-oriented model-based reinforcement learning with implicit differentiation

Reference 47

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:10:13.688357Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T12:10:04.606982Z digest=sha256:1acea3f81739418de35291e72fba98cda708f680fb3fe157d3293a62dd5fdca4

Observation 54cb92e7-247e-4cd7-92c0-07564015d10b · outbound

This paper cites Labelled Markov Processes.

Understanding Behavioral Metric Learning: A Large-Scale Study on Distracting Reinforcement Learning Environments Labelled Markov Processes

Reference 48

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:10:13.476146Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T12:10:05.193640Z digest=sha256:af38c9b7a5fa62bd9e6b8d818f2a8704b1ddb1c2779a74aa29d386f6678e5bbf

Observation fcedacdf-7d37-48e7-90da-33688ff09bd1 · outbound

This paper cites Policy gradient methods in the presence of symmetries and state abstractions.

Understanding Behavioral Metric Learning: A Large-Scale Study on Distracting Reinforcement Learning Environments Policy gradient methods in the presence of symmetries and state abstractions

Reference 49

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:10:13.298662Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T12:10:06.399252Z digest=sha256:dea4069f3d8c18f55ee0b5c86cb6b4b87670707c26e351c1c8b7c3dd5e430191

Observation a2831909-7709-4ddd-81ba-8b1f2cc6dca0 · outbound

This paper cites Concurrency and automata on infinite sequences.

Understanding Behavioral Metric Learning: A Large-Scale Study on Distracting Reinforcement Learning Environments Concurrency and automata on infinite sequences

Reference 50

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:10:13.122998Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T12:10:06.461937Z digest=sha256:62c644ccd34a3f579cc40f4db72a82433fc0ab20121b5323debc011a3d6d8302

Observation 25533d6b-be35-4a66-9adc-84251b81fdbb · outbound

This paper cites State-action similarity-based representations for off-policy evaluation.

Understanding Behavioral Metric Learning: A Large-Scale Study on Distracting Reinforcement Learning Environments State-action similarity-based representations for off-policy evaluation

Reference 51

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:10:12.931278Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T12:10:06.534084Z digest=sha256:5f28b47286115e34e37bdd2eff4207827095e26ff78da0b005b074b7cf458c26

Observation e443b986-897a-4bff-8345-e822fcbbd9cc · outbound

This paper cites An algebraic approach to abstraction in reinforcement learning.

Understanding Behavioral Metric Learning: A Large-Scale Study on Distracting Reinforcement Learning Environments An algebraic approach to abstraction in reinforcement learning

Reference 52

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:10:12.706810Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T12:10:06.595261Z digest=sha256:958dcee32e6234b3d7044bea24a035f06eedcd83cc1522faa0221e71b6f4b87c

Observation e20b064e-c1c4-4cf3-97e9-984df4c0f430 · outbound

This paper cites Model minimization in hierarchical reinforcement learning.

Understanding Behavioral Metric Learning: A Large-Scale Study on Distracting Reinforcement Learning Environments Model minimization in hierarchical reinforcement learning

Reference 53

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:10:12.514760Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T12:10:06.695657Z digest=sha256:138f62603a4dc686075abe2a287ceac382573904510e49f30411fcea4031e4cb

Observation fb3364eb-b99b-4fe9-bf9e-ef2453325940 · outbound

This paper cites Continuous mdp homomorphisms and homomorphic policy gradient.

Understanding Behavioral Metric Learning: A Large-Scale Study on Distracting Reinforcement Learning Environments Continuous mdp homomorphisms and homomorphic policy gradient

Reference 54

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:10:12.371374Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T12:10:06.791035Z digest=sha256:b726a948be5b23930ce95c370bdb93ffb64675dfe389c8bc06f39f48ad9c459d

Observation 5d1e18c3-e85b-4599-8fcf-5576952ec28c · outbound

This paper cites Learning Action-based Representations Using Invariance.

Understanding Behavioral Metric Learning: A Large-Scale Study on Distracting Reinforcement Learning Environments Learning Action-based Representations Using Invariance

Reference 55

Resolution
unresolved
no resolver link, observed 2026-08-07T12:10:06.863385Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:10:06.863385Z digest=sha256:246485d753f6432ed56ca988e9c270a17548fabb07f07ea7579e1be7edc15a9c

Observation 0b58ff51-e8f0-48eb-b82f-6a086d098ca2 · outbound

This paper cites Facenet: A unified embedding for face recognition and clustering.

Understanding Behavioral Metric Learning: A Large-Scale Study on Distracting Reinforcement Learning Environments Facenet: A unified embedding for face recognition and clustering

Reference 56

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:10:12.215444Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T12:10:06.946603Z digest=sha256:d6190dd2c7f3a4d601b32d752703f5aaac461f83bed39a1d271135c7940d61a1

Observation aacff5eb-c4b3-4ba5-ac6d-3b9a6edfe064 · outbound

This paper cites Data-Efficient Reinforcement Learning with Self-Predictive Representations.

Understanding Behavioral Metric Learning: A Large-Scale Study on Distracting Reinforcement Learning Environments Data-Efficient Reinforcement Learning with Self-Predictive Representations

Reference 57

Resolution
unresolved
no resolver link, observed 2026-08-07T12:10:07.046569Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:10:07.046569Z digest=sha256:34f61d2ef7375615649985b535b529976c96ecebe5ca0f3f502df6ea2c00c5c3

Observation 5fd4873b-5137-480a-9f76-3412cba45740 · outbound

This paper cites Bisimulation metric for Model Predictive Control.

Understanding Behavioral Metric Learning: A Large-Scale Study on Distracting Reinforcement Learning Environments Bisimulation metric for Model Predictive Control

Reference 58

Resolution
unresolved
no resolver link, observed 2026-08-07T12:10:07.161039Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:10:07.161039Z digest=sha256:93aa47aa3b40283bdef29a8eb5701406c323bfe1e94f3196adbf8691b124c3da

Observation ffeb9102-a325-48c2-8769-e5630f47905e · outbound

This paper cites Reinforcement learning with soft state aggregation.

Understanding Behavioral Metric Learning: A Large-Scale Study on Distracting Reinforcement Learning Environments Reinforcement learning with soft state aggregation

Reference 59

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:10:12.002559Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T12:10:07.288340Z digest=sha256:5c6e1f2c268797a72848418829d90678e1c670c5dbab9dfdaea631daa16f45c9

Observation 9049565e-2b4a-44b2-8be0-bc466da837ba · outbound

This paper cites A Walk in the Park: Learning to Walk in 20 Minutes With Model-Free Reinforcement Learning.

Understanding Behavioral Metric Learning: A Large-Scale Study on Distracting Reinforcement Learning Environments A Walk in the Park: Learning to Walk in 20 Minutes With Model-Free Reinforcement Learning

Reference 60

Resolution
unresolved
no resolver link, observed 2026-08-07T12:10:07.379760Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:10:07.379760Z digest=sha256:5ee91ba7744fe2f20685ae6f1ef40a6929cc204b78482565c10c72073ddbb36b

Observation bbe7f5bf-f3f4-4e6c-ac12-f37294582b35 · outbound

This paper cites The Distracting Control Suite -- A Challenging Benchmark for Reinforcement Learning from Pixels.

Understanding Behavioral Metric Learning: A Large-Scale Study on Distracting Reinforcement Learning Environments The Distracting Control Suite -- A Challenging Benchmark for Reinforcement Learning from Pixels

Reference 61

Resolution
unresolved
no resolver link, observed 2026-08-07T12:10:07.468224Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:10:07.468224Z digest=sha256:0d67ea7a53b46ee3bf5919964a2b2f79e751455dad84025dc5c8e5ff1202893d

Observation 5d38d754-8f0d-4bd0-9436-f6c5b17ea498 · outbound

This paper cites Approximate information state for approximate planning and reinforcement learning in partially observed systems.

Understanding Behavioral Metric Learning: A Large-Scale Study on Distracting Reinforcement Learning Environments Approximate information state for approximate planning and reinforcement learning in partially observed systems

Reference 62

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:10:11.853121Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T12:10:07.537316Z digest=sha256:d3d9f0e840dc74186bd66bca7537b3b49e46296da2e956d37f79d32c4c3da3d4

Observation 2bdfd756-83ac-44a5-aafa-613e25fea1e5 · outbound

This paper cites DeepMind Control Suite.

Understanding Behavioral Metric Learning: A Large-Scale Study on Distracting Reinforcement Learning Environments DeepMind Control Suite

Reference 63

Resolution
unresolved
no resolver link, observed 2026-08-07T12:10:07.623587Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:10:07.623587Z digest=sha256:d716dc54e3f30ee3959ebb572a28da2cc2b57f58ad33613b3893427957dfd31b

Observation 5cadb823-c646-48a9-968e-b1a8aae70913 · outbound

This paper cites Lax probabilistic bisimulation.

Understanding Behavioral Metric Learning: A Large-Scale Study on Distracting Reinforcement Learning Environments Lax probabilistic bisimulation

Reference 64

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:10:11.635694Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T12:10:07.691352Z digest=sha256:d6c0d4db0bae19fb94e8e5f068ffa073821798d83437594a51b61519437b3ba9

Observation 6d4f6ab7-e5e1-416c-b58c-f381b4871e22 · outbound

This paper cites Learning Representations for Pixel-based Control: What Matters and Why?.

Understanding Behavioral Metric Learning: A Large-Scale Study on Distracting Reinforcement Learning Environments Learning Representations for Pixel-based Control: What Matters and Why?

Reference 65

Resolution
unresolved
no resolver link, observed 2026-08-07T12:10:07.756955Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:10:07.756955Z digest=sha256:4628f7e2a6d3ca28a4e988c193fc526849cbd1ca30729bda15240e87c95f4487

Observation d2ad554f-ef64-4e00-94e9-dd55c99069ee · outbound

This paper cites dm\_control: Software and tasks for continuous control.

Understanding Behavioral Metric Learning: A Large-Scale Study on Distracting Reinforcement Learning Environments dm\_control: Software and tasks for continuous control

Reference 66

Resolution
unresolved
no resolver link, observed 2026-08-07T12:10:07.869845Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:10:07.869845Z digest=sha256:a3e16f28640be3177f29c923c9a6a623bce2566c8aca43427391e82bf23b5e41

Observation a1863e4a-5ea5-45cd-935c-f9893becef74 · outbound

This paper cites Plannable Approximations to MDP Homomorphisms: Equivariance under Actions.

Understanding Behavioral Metric Learning: A Large-Scale Study on Distracting Reinforcement Learning Environments Plannable Approximations to MDP Homomorphisms: Equivariance under Actions

Reference 67

Resolution
unresolved
no resolver link, observed 2026-08-07T12:10:07.966060Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:10:07.966060Z digest=sha256:cf82de42a6410adbd6511c2369510887a80fd00753b54c9b9e6210608c2847bf

Observation 25028b49-4390-48a3-8642-48798d115c7c · outbound

This paper cites Mdp homomorphic networks: Group symmetries in reinforcement learning.

Understanding Behavioral Metric Learning: A Large-Scale Study on Distracting Reinforcement Learning Environments Mdp homomorphic networks: Group symmetries in reinforcement learning

Reference 68

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:10:11.402431Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T12:10:08.037509Z digest=sha256:a4d83beeedae39dfacdfb9711b308a3372f4d3ae14925ca2d496a36ef00320fc

Observation 76f8c420-993f-4d08-8869-c3e699ab28ca · outbound

This paper cites When does Self-Prediction help? Understanding Auxiliary Tasks in Reinforcement Learning.

Understanding Behavioral Metric Learning: A Large-Scale Study on Distracting Reinforcement Learning Environments When does Self-Prediction help? Understanding Auxiliary Tasks in Reinforcement Learning

Reference 69

Resolution
unresolved
no resolver link, observed 2026-08-07T12:10:08.108971Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:10:08.108971Z digest=sha256:cbc2f2a413dd136584e2887c10ee56cfa50c53f1e85efedc7f9ec53b72a245b7

Observation 977ac222-ea26-496e-a7a7-baedeaf9a543 · outbound

This paper cites Efficient potential-based exploration in reinforcement learning using inverse dynamic bisimulation metric.

Understanding Behavioral Metric Learning: A Large-Scale Study on Distracting Reinforcement Learning Environments Efficient potential-based exploration in reinforcement learning using inverse dynamic bisimulation metric

Reference 70

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:10:11.238699Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T12:10:08.176373Z digest=sha256:d952981ba7ec6acce99cc52b337944fb67031ba373a8db442658a9235fe5ba02

Observation e90c7c54-1f0e-42be-b1c1-c6d223bff424 · outbound

This paper cites Rethinking exploration in reinforcement learning with effective metric-based exploration bonus.

Understanding Behavioral Metric Learning: A Large-Scale Study on Distracting Reinforcement Learning Environments Rethinking exploration in reinforcement learning with effective metric-based exploration bonus

Reference 71

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:10:11.048518Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T12:10:08.271332Z digest=sha256:fb620340056b95d8de5ab108ecbbf58b69f95bea2cd8aa1ce4c845e2a6dacdf8

Observation 9c9c5dd5-7582-4680-9f1a-78961d4f25b2 · outbound

This paper cites Mastering Visual Continuous Control: Improved Data-Augmented Reinforcement Learning.

Understanding Behavioral Metric Learning: A Large-Scale Study on Distracting Reinforcement Learning Environments Mastering Visual Continuous Control: Improved Data-Augmented Reinforcement Learning

Reference 72

Resolution
unresolved
no resolver link, observed 2026-08-07T12:10:08.362608Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:10:08.362608Z digest=sha256:96251385b2285ae5a24234f975a58f86d2c3425c2297934bac76b23b522bccfa

Observation b681e948-fb94-4a12-8579-4c9251b77a80 · outbound

This paper cites Improving sample efficiency in model-free reinforcement learning from images.

Understanding Behavioral Metric Learning: A Large-Scale Study on Distracting Reinforcement Learning Environments Improving sample efficiency in model-free reinforcement learning from images

Reference 73

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:10:10.905731Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T12:10:08.455921Z digest=sha256:af8899de2e0193227fc63aceca899d7ea7944a600e0412cc0e649eae6740d4ce

Observation ca9e0fd8-5706-4d22-929f-61b9a4722915 · outbound

This paper cites Rl-vigen: A reinforcement learning benchmark for visual generalization.

Understanding Behavioral Metric Learning: A Large-Scale Study on Distracting Reinforcement Learning Environments Rl-vigen: A reinforcement learning benchmark for visual generalization

Reference 74

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:10:10.732423Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T12:10:08.524623Z digest=sha256:93b3743cfefb78786e582efee06fefe9057b751447059673d273ef68dfaec69c

Observation 88f09758-2d04-4bc8-91c3-0a9c7bbb457c · outbound

This paper cites Simsr: Simple distance-based state representations for deep reinforcement learning.

Understanding Behavioral Metric Learning: A Large-Scale Study on Distracting Reinforcement Learning Environments Simsr: Simple distance-based state representations for deep reinforcement learning

Reference 75

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:10:10.516198Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T12:10:08.606269Z digest=sha256:5b985089be0ddc1951b06b664f3cd19ad10d420183c36aefa72535687a17ceb6

Observation 336e1ed6-9e01-4c97-991c-7509e3e44bf2 · outbound

This paper cites Understanding and addressing the pitfalls of bisimulation-based representations in offline reinforcement learning.

Understanding Behavioral Metric Learning: A Large-Scale Study on Distracting Reinforcement Learning Environments Understanding and addressing the pitfalls of bisimulation-based representations in offline reinforcement learning

Reference 76

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:10:10.296954Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T12:10:08.728222Z digest=sha256:3d50d66a3f41512db633da81c3d8c5cdba091cb42a6bbdce075c83d38fe56702

Observation 1b9d6aa3-31f5-480a-8477-4de65d0ab96f · outbound

This paper cites Natural Environment Benchmarks for Reinforcement Learning.

Understanding Behavioral Metric Learning: A Large-Scale Study on Distracting Reinforcement Learning Environments Natural Environment Benchmarks for Reinforcement Learning

Reference 77

Resolution
unresolved
no resolver link, observed 2026-08-07T12:10:08.787938Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:10:08.787938Z digest=sha256:a2abcc5dbfdcf2fc3f2e573c953dc48669e19552a42251460709dfe2ab47e0e0

Observation c31911ca-51a1-4fe3-9d1f-3e4581e9afdb · outbound

This paper cites Learning Invariant Representations for Reinforcement Learning without Reconstruction.

Understanding Behavioral Metric Learning: A Large-Scale Study on Distracting Reinforcement Learning Environments Learning Invariant Representations for Reinforcement Learning without Reconstruction

Reference 78

Resolution
unresolved
no resolver link, observed 2026-08-07T12:10:08.875246Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:10:08.875246Z digest=sha256:7452a68694a750dc021bfee9f2039514c9152c174ef7e34ef4e82c589d327b71

Observation 4ac23cb1-3a23-4855-8ef0-985cb1ad2783 · outbound

This paper cites write newline.

Understanding Behavioral Metric Learning: A Large-Scale Study on Distracting Reinforcement Learning Environments write newline

Reference 79

Resolution
unresolved
no resolver link, observed 2026-08-07T12:10:08.987854Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:10:08.987854Z digest=sha256:194239e4ec16f67178dfdd59af7ac1aee0ab772fb2526cf9c8182c68528e4ed0

Pith citing papers

Observation e2946632-5655-4d00-b5b1-88a0e843cd76 · inbound

SAVGO: Learning State-Action Value Geometry with Cosine Similarity for Continuous Control cites this paper.

SAVGO: Learning State-Action Value Geometry with Cosine Similarity for Continuous Control Understanding Behavioral Metric Learning: A Large-Scale Study on Distracting Reinforcement Learning Environments

Reference 30

Resolution
verified exact
arxiv_id, observed 2026-05-09T20:01:36.584993Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-05-09T20:00:55.197662Z digest=sha256:514fac0b5a196880e1ddaaa905f9576c2243567868280a5f5c3e1453ed0e4af3