Pith. sign in

Paper Citation Record · LEDGER

Inverse Reinforcement Learning with Switching Rewards and History Dependency for Characterizing Animal Behaviors

As of 11 August 2026, this Paper Citation Record lists 42 of 42 outbound references and 3 inbound Pith citation observations for arXiv:2501.12633.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2501.12633 v3

Coverage vector

measured 42 of 42 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-10T17:05:29.071278Z

measured 45 of 45 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-11T06:34:44.6726+00:00

measured 3 of 3 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-04T12:42:58.307441Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-06-29T19:23:53.741950Z

Reference resolution

42 of 42 outbound references displayed

  • verified exact2
  • verified fuzzy20
  • unresolved20
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation e31c651a-b48a-4060-8030-3f71ccc54dd5 · outbound

This paper cites an unresolved cited work.

Inverse Reinforcement Learning with Switching Rewards and History Dependency for Characterizing Animal Behaviors Unresolved cited work

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-10T17:05:28.935285Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-10T17:05:28.935285Z digest=sha256:d2592baafc1becdb4eda06ad5fc0b5c8c4d659cb3d5713cb6d9f0e1b1da10a16

Observation 2babc4da-bd78-4ce4-a79a-dd5003996cba · outbound

This paper cites an unresolved cited work.

Inverse Reinforcement Learning with Switching Rewards and History Dependency for Characterizing Animal Behaviors Unresolved cited work

Reference 2

Resolution
unresolved
raw_fallback, observed 2026-08-10T17:05:29.851954Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-08-10T17:05:28.939335Z digest=sha256:334ef0b896c72150bfefd2395d3350d36198ccb5bd2a110e1334a4b2de5c820a

Observation 06928f22-99a3-4fb0-955e-5b7b5fe850c1 · outbound

This paper cites Ashwood, Nicholas A.

Inverse Reinforcement Learning with Switching Rewards and History Dependency for Characterizing Animal Behaviors Ashwood, Nicholas A

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-10T17:05:28.942657Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-10T17:05:28.942657Z digest=sha256:1ff0f53b37672d0b42ceb6627363c35e95c64539f6e899f81ebdd26d3e665373

Observation 890e67a3-8726-4d3a-afb5-bf6a446529a2 · outbound

This paper cites an unresolved cited work.

Inverse Reinforcement Learning with Switching Rewards and History Dependency for Characterizing Animal Behaviors Unresolved cited work

Reference 4

Resolution
unresolved
raw_fallback, observed 2026-08-10T17:05:29.841725Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-08-10T17:05:28.946216Z digest=sha256:dd02825c270f9ee1305440d61e8d46b093b087f46bf09b47db35c2a9fab72cfd

Observation f8c7b837-2bf0-49fe-ab1f-a1f58ea797e5 · outbound

This paper cites Reinforcement learning with lstm in non-markovian tasks with longterm dependencies.

Inverse Reinforcement Learning with Switching Rewards and History Dependency for Characterizing Animal Behaviors Reinforcement learning with lstm in non-markovian tasks with longterm dependencies

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T17:05:29.831416Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-08-10T17:05:28.949756Z digest=sha256:503bfe46db1122183db05fad6fab9ac5f14f6c06eab975d085a212cf4f5a0565

Observation 1282d7ee-2698-4720-bfb5-35e74d3ac728 · outbound

This paper cites Option-aware adversarial inverse reinforcement learning for robotic control.

Inverse Reinforcement Learning with Switching Rewards and History Dependency for Characterizing Animal Behaviors Option-aware adversarial inverse reinforcement learning for robotic control

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-10T17:05:28.953251Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-10T17:05:28.953251Z digest=sha256:06a2d76ab9b3083e19119fff9b05b9a30a6936f1a77978720abf82b051b82324

Observation 43de4f6b-770f-4ca3-9b3f-37463e1cea94 · outbound

This paper cites Learning robust rewards with adverserial inverse reinforcement learning.

Inverse Reinforcement Learning with Switching Rewards and History Dependency for Characterizing Animal Behaviors Learning robust rewards with adverserial inverse reinforcement learning

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-10T17:05:28.957197Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-10T17:05:28.957197Z digest=sha256:d79b0b97414ac79d298e8941cdaa24042bf378d6615042600baaf99d3894d92c

Observation dffccea0-607c-405b-ae09-dcaaaf4276f6 · outbound

This paper cites Iq-learn: Inverse soft-q learning for imitation.

Inverse Reinforcement Learning with Switching Rewards and History Dependency for Characterizing Animal Behaviors Iq-learn: Inverse soft-q learning for imitation

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T17:05:29.814666Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-08-10T17:05:28.960380Z digest=sha256:2a1b70fbd4ad751d342e305bfcbdbbddbfe8ce08b4b93ed45ede0f5787fd5f03

Observation 6f25f3a1-887d-4b1c-a0c9-9960b0861ff1 · outbound

This paper cites Reinforcement learning with deep energy-based policies.

Inverse Reinforcement Learning with Switching Rewards and History Dependency for Characterizing Animal Behaviors Reinforcement learning with deep energy-based policies

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-10T17:05:28.963293Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-10T17:05:28.963293Z digest=sha256:711c61b3ea8cf894eb452cbc3c576c6e778e520b7c07287aeea0daf414ae7f0b

Observation 5c6dff0c-d7d9-4947-8b4f-64ba44489e1d · outbound

This paper cites Soft actor-critic: Off-policy maximum entropy deep reinforcement learning with a stochastic actor.

Inverse Reinforcement Learning with Switching Rewards and History Dependency for Characterizing Animal Behaviors Soft actor-critic: Off-policy maximum entropy deep reinforcement learning with a stochastic actor

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T17:05:29.798644Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-08-10T17:05:28.966469Z digest=sha256:e18c34858ac105ae97bd6f0f1185944aab91c0e878b5280fa14e423fcf33a236

Observation 427144c5-0c2f-43ce-88b5-91edca1dbbf3 · outbound

This paper cites Area-specificity and plasticity of history-dependent value coding during learning.

Inverse Reinforcement Learning with Switching Rewards and History Dependency for Characterizing Animal Behaviors Area-specificity and plasticity of history-dependent value coding during learning

Reference 11

Resolution
verified exact
doi, observed 2026-08-10T17:05:29.111183Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-08-10T17:05:28.969606Z digest=sha256:8421faacd05040cad2c5b4e8016ac72c28df77d89a0741cf03b0db248d0484c4

Observation c5c50c21-94f8-4111-8cc6-0c453c3277bd · outbound

This paper cites Deep recurrent q-learning for partially observable mdps.

Inverse Reinforcement Learning with Switching Rewards and History Dependency for Characterizing Animal Behaviors Deep recurrent q-learning for partially observable mdps

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T17:05:29.788619Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-08-10T17:05:28.972933Z digest=sha256:806abb96d7ee36c299ba25ceb7fe991330f74fac17a426470f7e60550d01aa90

Observation fa6ed0f0-96c9-453a-8637-935da7e403ff · outbound

This paper cites Can ai predict animal movements? filling gaps in animal trajectories using inverse reinforcement learning.

Inverse Reinforcement Learning with Switching Rewards and History Dependency for Characterizing Animal Behaviors Can ai predict animal movements? filling gaps in animal trajectories using inverse reinforcement learning

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T17:05:29.778941Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-08-10T17:05:28.975881Z digest=sha256:597f93feeb569426b67025f6b34cc18060c3e8d8d6a3db03f201cac9674b9d29

Observation 2f2eb347-f708-452c-b819-0ec3daf43d64 · outbound

This paper cites Vime: Variational information maximizing exploration.

Inverse Reinforcement Learning with Switching Rewards and History Dependency for Characterizing Animal Behaviors Vime: Variational information maximizing exploration

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T17:05:29.768677Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-08-10T17:05:28.978951Z digest=sha256:6b2eb3f76567f5b8feb7bc0a1356748e1414633c93f756608e0fa24b2ef80edb

Observation 4103dd3f-95b1-4c01-ac9c-6248f1fafbc8 · outbound

This paper cites The what, how, and why of naturalistic behavior.

Inverse Reinforcement Learning with Switching Rewards and History Dependency for Characterizing Animal Behaviors The what, how, and why of naturalistic behavior

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-10T17:05:28.982112Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-10T17:05:28.982112Z digest=sha256:859f139d1096820c7176898201f470fabc5f7c34de1ffb63cd117b042b0499bc

Observation 9d79c6b3-84ee-4c1a-94cb-793bb61daef1 · outbound

This paper cites Recurrent switching linear dynamical systems.

Inverse Reinforcement Learning with Switching Rewards and History Dependency for Characterizing Animal Behaviors Recurrent switching linear dynamical systems

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-10T17:05:28.984970Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-10T17:05:28.984970Z digest=sha256:34be8c676b22d3d487d47d0b3a6e8024a9b134b4b218c17dd250f0b4f883b8c0

Observation 1de4615a-6cf6-4980-a309-d20ee000f3e8 · outbound

This paper cites Spontaneous behaviour is structured by reinforcement without explicit reward.

Inverse Reinforcement Learning with Switching Rewards and History Dependency for Characterizing Animal Behaviors Spontaneous behaviour is structured by reinforcement without explicit reward

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T17:05:29.758259Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-08-10T17:05:28.988119Z digest=sha256:81e587d05ffb76fff82e4d521aea0c3415dd20b979b07badfcfeec6a3457d52f

Observation 897886ba-5686-4cec-9104-bde448a1cdd1 · outbound

This paper cites Neural mechanisms underlying the temporal organization of naturalistic animal behavior.

Inverse Reinforcement Learning with Switching Rewards and History Dependency for Characterizing Animal Behaviors Neural mechanisms underlying the temporal organization of naturalistic animal behavior

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T17:05:29.748647Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-08-10T17:05:28.991419Z digest=sha256:a7588815a2995e3e3211d7cd0bc109cb57700efff47a5652b0a23a4c313c88cb

Observation 8ba32143-773c-44da-b6c7-1827de1a3ed3 · outbound

This paper cites Ng and Stuart J.

Inverse Reinforcement Learning with Switching Rewards and History Dependency for Characterizing Animal Behaviors Ng and Stuart J

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T17:05:29.737675Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-08-10T17:05:28.994406Z digest=sha256:826d33d4221f0273e141eccf977cf1578970d0975c0763ebc6154523530825d6

Observation e0baeb92-cfd2-447d-8c1a-7ee6aad6784e · outbound

This paper cites Inverse reinforcement learning with locally consistent reward functions.

Inverse Reinforcement Learning with Switching Rewards and History Dependency for Characterizing Animal Behaviors Inverse reinforcement learning with locally consistent reward functions

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T17:05:29.727171Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-08-10T17:05:28.997453Z digest=sha256:38d37fec5e078fd30d07b0c036cdd15b97af976693c9eac824c481ace5ef32bf

Observation 0ab91280-6518-42cc-8831-eaf8884f6c9b · outbound

This paper cites Neural Map: Structured Memory for Deep Reinforcement Learning.

Inverse Reinforcement Learning with Switching Rewards and History Dependency for Characterizing Animal Behaviors Neural Map: Structured Memory for Deep Reinforcement Learning

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-10T17:05:29.000160Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-10T17:05:29.000160Z digest=sha256:4747b09b0f500143bd9dec98c4a2cbb75abb206352d1dbdc692d583af73fcd6e

Observation ba0d4822-8f84-4231-aa2a-2259be80926c · outbound

This paper cites Inverse reinforcement learning of bird flocking behavior.

Inverse Reinforcement Learning with Switching Rewards and History Dependency for Characterizing Animal Behaviors Inverse reinforcement learning of bird flocking behavior

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T17:05:29.717394Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-08-10T17:05:29.003408Z digest=sha256:df06b517f8e7f77d33cc14678ed8eeb5c2c547c38b8c36a89a243d8e61370013

Observation ea2be7f1-1628-401d-88af-fd0ee16ec21e · outbound

This paper cites Mice in a labyrinth show rapid learning, sudden insight, and efficient exploration.

Inverse Reinforcement Learning with Switching Rewards and History Dependency for Characterizing Animal Behaviors Mice in a labyrinth show rapid learning, sudden insight, and efficient exploration

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-10T17:05:29.006505Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-10T17:05:29.006505Z digest=sha256:3345cb6a5669c289fe4cd160ae8a30e79503cb1318f8faadc94c53875924ce68

Observation dd59d97c-f7dc-4b11-a245-a941a9a686f6 · outbound

This paper cites Obtaining reward functions of rats using inverse reinforcement learning.

Inverse Reinforcement Learning with Switching Rewards and History Dependency for Characterizing Animal Behaviors Obtaining reward functions of rats using inverse reinforcement learning

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T17:05:29.707238Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-08-10T17:05:29.009760Z digest=sha256:03a6ff02144e611005893e76b259347a0558d26673fa6c42d2e2c2349bb64ad5

Observation a564e214-214c-4827-a054-5d47c3763555 · outbound

This paper cites Active sensing with predictive coding and uncertainty minimization.

Inverse Reinforcement Learning with Switching Rewards and History Dependency for Characterizing Animal Behaviors Active sensing with predictive coding and uncertainty minimization

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-10T17:05:29.012758Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-10T17:05:29.012758Z digest=sha256:bb72bc843b0ffc9680af3b15acfcf73b0b1d6f706df547f13936880ea286ffed

Observation 299d3dd6-e37f-4c77-9a1c-ba5ec95589c8 · outbound

This paper cites Latent Variable Models for Characterizing the Dynamic Structure Underlying Complex Behaviors.

Inverse Reinforcement Learning with Switching Rewards and History Dependency for Characterizing Animal Behaviors Latent Variable Models for Characterizing the Dynamic Structure Underlying Complex Behaviors

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T17:05:29.697431Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-08-10T17:05:29.016014Z digest=sha256:b7e2ed0178e2ec61f5ad9b2e5bb3f2a279f4a88a37daf8cbb40677d0b287bb8a

Observation 9bd69652-71cb-4681-b458-9cec253884de · outbound

This paper cites Bayesian nonparametric inverse reinforcement learning for switched markov decision processes.

Inverse Reinforcement Learning with Switching Rewards and History Dependency for Characterizing Animal Behaviors Bayesian nonparametric inverse reinforcement learning for switched markov decision processes

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T17:05:29.687576Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-08-10T17:05:29.019844Z digest=sha256:2bd29d0365316124b8dccb365c2d075d9b0a8e65538f6c5ff63ce5beb1082f05

Observation ed2aa13e-5143-443e-be81-8636624f3247 · outbound

This paper cites Dyna, an integrated architecture for learning, planning, and reacting.

Inverse Reinforcement Learning with Switching Rewards and History Dependency for Characterizing Animal Behaviors Dyna, an integrated architecture for learning, planning, and reacting

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-10T17:05:29.022764Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-10T17:05:29.022764Z digest=sha256:283682719cc963ee0119d05cdb65f9d3a6998b2bf6d3c67220aa7c95813e5bb1

Observation adfcf920-13b8-418e-8918-e4843a75b2b6 · outbound

This paper cites Keypoint-moseq: parsing behavior by linking point tracking to pose dynamics.

Inverse Reinforcement Learning with Switching Rewards and History Dependency for Characterizing Animal Behaviors Keypoint-moseq: parsing behavior by linking point tracking to pose dynamics

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T17:05:29.670162Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-08-10T17:05:29.025723Z digest=sha256:7f0de3e8364b4d0257892baa697e36d57d7715510d637e5f902775181ffc5e6e

Observation 1f5fe048-d2d1-414f-9316-5b995af711ea · outbound

This paper cites Mapping sub-second structure in mouse behavior.

Inverse Reinforcement Learning with Switching Rewards and History Dependency for Characterizing Animal Behaviors Mapping sub-second structure in mouse behavior

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T17:05:29.659566Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-08-10T17:05:29.028562Z digest=sha256:320eba707c2d2de40b2ff5f5910cc20564e1306183c791a610e3fb6210724fe2

Observation 46462f06-c652-4c83-b02d-3fb2003331b5 · outbound

This paper cites an unresolved cited work.

Inverse Reinforcement Learning with Switching Rewards and History Dependency for Characterizing Animal Behaviors Unresolved cited work

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-10T17:05:29.031667Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-10T17:05:29.031667Z digest=sha256:322be4781f8315f6ec65f5e5271ca75c28e3ae91fb68084b83aa4e11575cc22b

Observation ba47349e-543f-4d1f-96c3-9fda901104bf · outbound

This paper cites Inverse reinforcement learning with the average reward criterion.

Inverse Reinforcement Learning with Switching Rewards and History Dependency for Characterizing Animal Behaviors Inverse reinforcement learning with the average reward criterion

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T17:05:29.650153Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-08-10T17:05:29.034608Z digest=sha256:46d200b0ce8f80c094e0b321aab55cce4be1ebe63d19e2d5b3952acefc4b083c

Observation 1b85156d-480c-4efa-a8a7-02b44159334a · outbound

This paper cites Imitating Language via Scalable Inverse Reinforcement Learning.

Inverse Reinforcement Learning with Switching Rewards and History Dependency for Characterizing Animal Behaviors Imitating Language via Scalable Inverse Reinforcement Learning

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-10T17:05:29.037500Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-10T17:05:29.037500Z digest=sha256:bc12b14118baf198b7506dac4452de893e188a1b6301555cbca45f6e81099aa4

Observation 4c744ff2-8032-44f2-a197-3d9a3014733f · outbound

This paper cites Identification of animal behavioral strategies by inverse reinforcement learning.

Inverse Reinforcement Learning with Switching Rewards and History Dependency for Characterizing Animal Behaviors Identification of animal behavioral strategies by inverse reinforcement learning

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T17:05:29.640503Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-08-10T17:05:29.041100Z digest=sha256:cf7d19fe1a8f7d85cf346b989516f88355d3219e75ed5e93a48b89e875bcf8ca

Observation 464930bb-32d5-420a-957f-c511a27d0462 · outbound

This paper cites Maximum-likelihood inverse reinforcement learning with finite-time guarantees.

Inverse Reinforcement Learning with Switching Rewards and History Dependency for Characterizing Animal Behaviors Maximum-likelihood inverse reinforcement learning with finite-time guarantees

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T17:05:29.630114Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-08-10T17:05:29.043980Z digest=sha256:9fca83ede54938c7c4597f1089f37aaa4e9fe0ed42bbd5e2a47aa773f1b529e4

Observation 4c32f061-0b69-4036-be9b-1301f034a06c · outbound

This paper cites Multi-intention inverse q-learning for interpretable behavior representation.

Inverse Reinforcement Learning with Switching Rewards and History Dependency for Characterizing Animal Behaviors Multi-intention inverse q-learning for interpretable behavior representation

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T17:05:29.618872Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-08-10T17:05:29.046980Z digest=sha256:7c31c3fb03eca113143ec9237733bb9b2526f4f26e3bf964c29479caf1538043

Observation c4a94482-262b-48d2-8fef-da36566aea0a · outbound

This paper cites Ziebart, Andrew Maas, J.

Inverse Reinforcement Learning with Switching Rewards and History Dependency for Characterizing Animal Behaviors Ziebart, Andrew Maas, J

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-10T17:05:29.050233Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-10T17:05:29.050233Z digest=sha256:62987647b52db1c10b9081e5188a935449fe6dac288c43cd9582ab85b7e3bad8

Observation 4c2a85c1-0248-42ef-9331-fea43efe26bb · outbound

This paper cites Ziebart, J.

Inverse Reinforcement Learning with Switching Rewards and History Dependency for Characterizing Animal Behaviors Ziebart, J

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-10T17:05:29.053428Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-10T17:05:29.053428Z digest=sha256:e986304e5bbccab5fd042eb464b79e63d6bf79a462d111582706628d65c89d1b

Observation b2a863eb-0719-42b0-8ee7-863662020543 · outbound

This paper cites write newline.

Inverse Reinforcement Learning with Switching Rewards and History Dependency for Characterizing Animal Behaviors write newline

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-10T17:05:29.056307Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-10T17:05:29.056307Z digest=sha256:6765d12090115e66ba9a882b5d0814f2aecfce14bdd73f883f2726efa50e3ee5

Observation d3d62832-d692-48eb-9628-5bce9ccc937d · outbound

This paper cites @esa (Ref.

Inverse Reinforcement Learning with Switching Rewards and History Dependency for Characterizing Animal Behaviors @esa (Ref

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-10T17:05:29.060869Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-10T17:05:29.060869Z digest=sha256:409710ee15cf8405dc44e64b0c4876070d31f31c5aac3be4e6a2a2eb7088b5ba

Observation 1d8df172-0eec-49d6-ad23-ebccad2ed9a3 · outbound

This paper cites an unresolved cited work.

Inverse Reinforcement Learning with Switching Rewards and History Dependency for Characterizing Animal Behaviors Unresolved cited work

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-10T17:05:29.064277Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-10T17:05:29.064277Z digest=sha256:8e3ea1968afa4b8ec5f6eaf6a5aca33c7e4ec6071937eefab519f429500a512e

Observation 7db36db6-7191-4c21-a088-56f2739c9be3 · outbound

This paper cites an unresolved cited work.

Inverse Reinforcement Learning with Switching Rewards and History Dependency for Characterizing Animal Behaviors Unresolved cited work

Reference 42

Resolution
verified exact
raw_fallback, observed 2026-08-10T17:05:29.219493Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-08-10T17:05:29.071278Z digest=sha256:b1bd6f1535c2ed2a02c8d00b8af8570ec2f35452aa246640488b19aa77f50d4b

Pith citing papers

Observation c51944cb-860f-43b7-bdd3-aa00e00e696b · inbound

Distributional Inverse Reinforcement Learning cites this paper.

Distributional Inverse Reinforcement Learning Inverse Reinforcement Learning with Switching Rewards and History Dependency for Characterizing Animal Behaviors

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-04T12:42:58.307441Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T12:42:58.307441Z digest=sha256:2363266aec11f8def46fa80c6b73ebdbb8f77e8266b2ec1f2b7328215ee46dff

Observation 5dd3d431-9932-4ed4-acfd-6a61b135a375 · inbound

Improving Zero-Shot Offline RL via Behavioral Task Sampling cites this paper.

Improving Zero-Shot Offline RL via Behavioral Task Sampling Inverse Reinforcement Learning with Switching Rewards and History Dependency for Characterizing Animal Behaviors

Reference 12

Resolution
verified exact
arxiv_id, observed 2026-05-11T23:41:18.778357Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-05-07T16:27:15.347522Z digest=sha256:eecef8e9b64435a70f090fed86c55dbb4c590458ed87c2ad77679dbe97f472e0

Observation da554326-c531-43e7-bf51-d82c82b18576 · inbound

Probabilistic Recurrent Intention Switching Model cites this paper.

Probabilistic Recurrent Intention Switching Model Inverse Reinforcement Learning with Switching Rewards and History Dependency for Characterizing Animal Behaviors

Reference 2

Resolution
metadata mismatch
arxiv_id, observed 2026-06-29T19:23:53.744266Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-06-29T19:22:57.970141Z digest=sha256:271e65becbd95f2af3f521bda149f8bbd991ef760b3f5b617e172c324a1a50f3