Pith. sign in

Paper Citation Record · LEDGER

Inverse Reinforcement Learning with Switching Rewards and History Dependency for Characterizing Animal Behaviors

As of 11 August 2026, this Paper Citation Record lists 42 of 42 outbound references and 3 inbound Pith citation observations for arXiv:2501.12633.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2501.12633 v3

Coverage vector

measured 42 of 42 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-10T17:05:29.071278Z

measured 45 of 45 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-10T06:31:04.303077+00:00

measured 3 of 3 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-04T12:42:58.307441Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-06-29T19:23:53.741950Z

Reference resolution

42 of 42 outbound references displayed

  • verified exact2
  • verified fuzzy20
  • unresolved20
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation e31c651a-b48a-4060-8030-3f71ccc54dd5 · outbound

This paper cites an unresolved cited work.

Inverse Reinforcement Learning with Switching Rewards and History Dependency for Characterizing Animal Behaviors Unresolved cited work

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-10T17:05:28.935285Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-10T17:05:28.935285Z digest=sha256:d2592baafc1becdb4eda06ad5fc0b5c8c4d659cb3d5713cb6d9f0e1b1da10a16

Observation 2babc4da-bd78-4ce4-a79a-dd5003996cba · outbound

This paper cites an unresolved cited work.

Inverse Reinforcement Learning with Switching Rewards and History Dependency for Characterizing Animal Behaviors Unresolved cited work

Reference 2

Resolution
unresolved
raw_fallback, observed 2026-08-10T17:05:29.851954Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-10T17:05:28.939335Z digest=sha256:348a6bb5f031d64c2f1a7765c7bac96c3bd355260c81ecc48b795dff9020887c

Observation 06928f22-99a3-4fb0-955e-5b7b5fe850c1 · outbound

This paper cites Ashwood, Nicholas A.

Inverse Reinforcement Learning with Switching Rewards and History Dependency for Characterizing Animal Behaviors Ashwood, Nicholas A

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-10T17:05:28.942657Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-10T17:05:28.942657Z digest=sha256:1ff0f53b37672d0b42ceb6627363c35e95c64539f6e899f81ebdd26d3e665373

Observation 890e67a3-8726-4d3a-afb5-bf6a446529a2 · outbound

This paper cites an unresolved cited work.

Inverse Reinforcement Learning with Switching Rewards and History Dependency for Characterizing Animal Behaviors Unresolved cited work

Reference 4

Resolution
unresolved
raw_fallback, observed 2026-08-10T17:05:29.841725Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-10T17:05:28.946216Z digest=sha256:a86b75dc86f7d802b0b23d5d23b557307d60b6d6ab5d9769117ccc0cc437365d

Observation f8c7b837-2bf0-49fe-ab1f-a1f58ea797e5 · outbound

This paper cites Reinforcement learning with lstm in non-markovian tasks with longterm dependencies.

Inverse Reinforcement Learning with Switching Rewards and History Dependency for Characterizing Animal Behaviors Reinforcement learning with lstm in non-markovian tasks with longterm dependencies

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T17:05:29.831416Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-10T17:05:28.949756Z digest=sha256:c9ff03d2753155632cb972f4db088c92062944d55b0e75c7fc0c91970818a261

Observation 1282d7ee-2698-4720-bfb5-35e74d3ac728 · outbound

This paper cites Option-aware adversarial inverse reinforcement learning for robotic control.

Inverse Reinforcement Learning with Switching Rewards and History Dependency for Characterizing Animal Behaviors Option-aware adversarial inverse reinforcement learning for robotic control

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-10T17:05:28.953251Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-10T17:05:28.953251Z digest=sha256:06a2d76ab9b3083e19119fff9b05b9a30a6936f1a77978720abf82b051b82324

Observation 43de4f6b-770f-4ca3-9b3f-37463e1cea94 · outbound

This paper cites Learning robust rewards with adverserial inverse reinforcement learning.

Inverse Reinforcement Learning with Switching Rewards and History Dependency for Characterizing Animal Behaviors Learning robust rewards with adverserial inverse reinforcement learning

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-10T17:05:28.957197Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-10T17:05:28.957197Z digest=sha256:d79b0b97414ac79d298e8941cdaa24042bf378d6615042600baaf99d3894d92c

Observation dffccea0-607c-405b-ae09-dcaaaf4276f6 · outbound

This paper cites Iq-learn: Inverse soft-q learning for imitation.

Inverse Reinforcement Learning with Switching Rewards and History Dependency for Characterizing Animal Behaviors Iq-learn: Inverse soft-q learning for imitation

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T17:05:29.814666Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-10T17:05:28.960380Z digest=sha256:f48e8aa566d5ada556b14df19dcd089ec9102cf795e7908aab5261d02e9868ba

Observation 6f25f3a1-887d-4b1c-a0c9-9960b0861ff1 · outbound

This paper cites Reinforcement learning with deep energy-based policies.

Inverse Reinforcement Learning with Switching Rewards and History Dependency for Characterizing Animal Behaviors Reinforcement learning with deep energy-based policies

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-10T17:05:28.963293Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-10T17:05:28.963293Z digest=sha256:711c61b3ea8cf894eb452cbc3c576c6e778e520b7c07287aeea0daf414ae7f0b

Observation 5c6dff0c-d7d9-4947-8b4f-64ba44489e1d · outbound

This paper cites Soft actor-critic: Off-policy maximum entropy deep reinforcement learning with a stochastic actor.

Inverse Reinforcement Learning with Switching Rewards and History Dependency for Characterizing Animal Behaviors Soft actor-critic: Off-policy maximum entropy deep reinforcement learning with a stochastic actor

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T17:05:29.798644Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-10T17:05:28.966469Z digest=sha256:d2c5f0af37ebe5a49ad2a0ca494e9a8263839b208400cbb7dc4afc87d8976aa5

Observation 427144c5-0c2f-43ce-88b5-91edca1dbbf3 · outbound

This paper cites Area-specificity and plasticity of history-dependent value coding during learning.

Inverse Reinforcement Learning with Switching Rewards and History Dependency for Characterizing Animal Behaviors Area-specificity and plasticity of history-dependent value coding during learning

Reference 11

Resolution
verified exact
doi, observed 2026-08-10T17:05:29.111183Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-10T17:05:28.969606Z digest=sha256:e916aaf766f94ae3c1f57efbb1de88b887bfc82fc6b7179fc79867dc7fb7e48e

Observation c5c50c21-94f8-4111-8cc6-0c453c3277bd · outbound

This paper cites Deep recurrent q-learning for partially observable mdps.

Inverse Reinforcement Learning with Switching Rewards and History Dependency for Characterizing Animal Behaviors Deep recurrent q-learning for partially observable mdps

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T17:05:29.788619Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-10T17:05:28.972933Z digest=sha256:f080622cbd74bb598ef7172120cd7855588aa5b63a76602c20742767260076f5

Observation fa6ed0f0-96c9-453a-8637-935da7e403ff · outbound

This paper cites Can ai predict animal movements? filling gaps in animal trajectories using inverse reinforcement learning.

Inverse Reinforcement Learning with Switching Rewards and History Dependency for Characterizing Animal Behaviors Can ai predict animal movements? filling gaps in animal trajectories using inverse reinforcement learning

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T17:05:29.778941Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-10T17:05:28.975881Z digest=sha256:964c806823b9e1845681f8a2bf8ce5c6884f0b95df2c517483a3e062fad5ea4c

Observation 2f2eb347-f708-452c-b819-0ec3daf43d64 · outbound

This paper cites Vime: Variational information maximizing exploration.

Inverse Reinforcement Learning with Switching Rewards and History Dependency for Characterizing Animal Behaviors Vime: Variational information maximizing exploration

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T17:05:29.768677Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-10T17:05:28.978951Z digest=sha256:b279b7f9a1710cd808b6e906ce5ba326e5044da45af4b04f2aadfc95604fa59e

Observation 4103dd3f-95b1-4c01-ac9c-6248f1fafbc8 · outbound

This paper cites The what, how, and why of naturalistic behavior.

Inverse Reinforcement Learning with Switching Rewards and History Dependency for Characterizing Animal Behaviors The what, how, and why of naturalistic behavior

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-10T17:05:28.982112Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-10T17:05:28.982112Z digest=sha256:859f139d1096820c7176898201f470fabc5f7c34de1ffb63cd117b042b0499bc

Observation 9d79c6b3-84ee-4c1a-94cb-793bb61daef1 · outbound

This paper cites Recurrent switching linear dynamical systems.

Inverse Reinforcement Learning with Switching Rewards and History Dependency for Characterizing Animal Behaviors Recurrent switching linear dynamical systems

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-10T17:05:28.984970Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-10T17:05:28.984970Z digest=sha256:34be8c676b22d3d487d47d0b3a6e8024a9b134b4b218c17dd250f0b4f883b8c0

Observation 1de4615a-6cf6-4980-a309-d20ee000f3e8 · outbound

This paper cites Spontaneous behaviour is structured by reinforcement without explicit reward.

Inverse Reinforcement Learning with Switching Rewards and History Dependency for Characterizing Animal Behaviors Spontaneous behaviour is structured by reinforcement without explicit reward

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T17:05:29.758259Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-10T17:05:28.988119Z digest=sha256:407ce8bee1a7c51181647c6e345e9818dc58fa221ed6c99044dab56ddaa681ff

Observation 897886ba-5686-4cec-9104-bde448a1cdd1 · outbound

This paper cites Neural mechanisms underlying the temporal organization of naturalistic animal behavior.

Inverse Reinforcement Learning with Switching Rewards and History Dependency for Characterizing Animal Behaviors Neural mechanisms underlying the temporal organization of naturalistic animal behavior

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T17:05:29.748647Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-10T17:05:28.991419Z digest=sha256:50860bec26b938654aa6d16b69fc4f93af9b93c64024d3b82e92da023dea55e1

Observation 8ba32143-773c-44da-b6c7-1827de1a3ed3 · outbound

This paper cites Ng and Stuart J.

Inverse Reinforcement Learning with Switching Rewards and History Dependency for Characterizing Animal Behaviors Ng and Stuart J

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T17:05:29.737675Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-10T17:05:28.994406Z digest=sha256:26edc6882faa6b359d328d85c5c5869eece4c864dfac1c90b45ca232a388c788

Observation e0baeb92-cfd2-447d-8c1a-7ee6aad6784e · outbound

This paper cites Inverse reinforcement learning with locally consistent reward functions.

Inverse Reinforcement Learning with Switching Rewards and History Dependency for Characterizing Animal Behaviors Inverse reinforcement learning with locally consistent reward functions

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T17:05:29.727171Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-10T17:05:28.997453Z digest=sha256:aafebf7672bf84f8c960a2c8123b4674581dd15a346b1b4ceaee1ce661a0dc18

Observation 0ab91280-6518-42cc-8831-eaf8884f6c9b · outbound

This paper cites Neural Map: Structured Memory for Deep Reinforcement Learning.

Inverse Reinforcement Learning with Switching Rewards and History Dependency for Characterizing Animal Behaviors Neural Map: Structured Memory for Deep Reinforcement Learning

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-10T17:05:29.000160Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-10T17:05:29.000160Z digest=sha256:4747b09b0f500143bd9dec98c4a2cbb75abb206352d1dbdc692d583af73fcd6e

Observation ba0d4822-8f84-4231-aa2a-2259be80926c · outbound

This paper cites Inverse reinforcement learning of bird flocking behavior.

Inverse Reinforcement Learning with Switching Rewards and History Dependency for Characterizing Animal Behaviors Inverse reinforcement learning of bird flocking behavior

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T17:05:29.717394Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-10T17:05:29.003408Z digest=sha256:a813b9a0f7a49a8f7eb9bebb2e0a5032ac12c039f8b5ab2a2c1de5f06404e0b2

Observation ea2be7f1-1628-401d-88af-fd0ee16ec21e · outbound

This paper cites Mice in a labyrinth show rapid learning, sudden insight, and efficient exploration.

Inverse Reinforcement Learning with Switching Rewards and History Dependency for Characterizing Animal Behaviors Mice in a labyrinth show rapid learning, sudden insight, and efficient exploration

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-10T17:05:29.006505Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-10T17:05:29.006505Z digest=sha256:3345cb6a5669c289fe4cd160ae8a30e79503cb1318f8faadc94c53875924ce68

Observation dd59d97c-f7dc-4b11-a245-a941a9a686f6 · outbound

This paper cites Obtaining reward functions of rats using inverse reinforcement learning.

Inverse Reinforcement Learning with Switching Rewards and History Dependency for Characterizing Animal Behaviors Obtaining reward functions of rats using inverse reinforcement learning

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T17:05:29.707238Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-10T17:05:29.009760Z digest=sha256:870bea85da7c7722de9be32159206df10dea458bb16aaaf315734584dcde6afd

Observation a564e214-214c-4827-a054-5d47c3763555 · outbound

This paper cites Active sensing with predictive coding and uncertainty minimization.

Inverse Reinforcement Learning with Switching Rewards and History Dependency for Characterizing Animal Behaviors Active sensing with predictive coding and uncertainty minimization

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-10T17:05:29.012758Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-10T17:05:29.012758Z digest=sha256:bb72bc843b0ffc9680af3b15acfcf73b0b1d6f706df547f13936880ea286ffed

Observation 299d3dd6-e37f-4c77-9a1c-ba5ec95589c8 · outbound

This paper cites Latent Variable Models for Characterizing the Dynamic Structure Underlying Complex Behaviors.

Inverse Reinforcement Learning with Switching Rewards and History Dependency for Characterizing Animal Behaviors Latent Variable Models for Characterizing the Dynamic Structure Underlying Complex Behaviors

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T17:05:29.697431Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-10T17:05:29.016014Z digest=sha256:1a2619031443184f33eebd4ce7cb23bffa5fb5ad9e30e3eb29f3dc499fdbf16e

Observation 9bd69652-71cb-4681-b458-9cec253884de · outbound

This paper cites Bayesian nonparametric inverse reinforcement learning for switched markov decision processes.

Inverse Reinforcement Learning with Switching Rewards and History Dependency for Characterizing Animal Behaviors Bayesian nonparametric inverse reinforcement learning for switched markov decision processes

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T17:05:29.687576Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-10T17:05:29.019844Z digest=sha256:b471adebb8b2cc54872263fbfca18f0189d1a3f17945841309ab5339579516f0

Observation ed2aa13e-5143-443e-be81-8636624f3247 · outbound

This paper cites Dyna, an integrated architecture for learning, planning, and reacting.

Inverse Reinforcement Learning with Switching Rewards and History Dependency for Characterizing Animal Behaviors Dyna, an integrated architecture for learning, planning, and reacting

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-10T17:05:29.022764Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-10T17:05:29.022764Z digest=sha256:283682719cc963ee0119d05cdb65f9d3a6998b2bf6d3c67220aa7c95813e5bb1

Observation adfcf920-13b8-418e-8918-e4843a75b2b6 · outbound

This paper cites Keypoint-moseq: parsing behavior by linking point tracking to pose dynamics.

Inverse Reinforcement Learning with Switching Rewards and History Dependency for Characterizing Animal Behaviors Keypoint-moseq: parsing behavior by linking point tracking to pose dynamics

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T17:05:29.670162Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-10T17:05:29.025723Z digest=sha256:48fd850f24a21886a4754e053491b215cf948d368194024f33970c05b3f67b83

Observation 1f5fe048-d2d1-414f-9316-5b995af711ea · outbound

This paper cites Mapping sub-second structure in mouse behavior.

Inverse Reinforcement Learning with Switching Rewards and History Dependency for Characterizing Animal Behaviors Mapping sub-second structure in mouse behavior

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T17:05:29.659566Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-10T17:05:29.028562Z digest=sha256:428f7638aaf0814b16c74c5dac340881696aa3578c2312f8f7f42b9adaac39cb

Observation 46462f06-c652-4c83-b02d-3fb2003331b5 · outbound

This paper cites an unresolved cited work.

Inverse Reinforcement Learning with Switching Rewards and History Dependency for Characterizing Animal Behaviors Unresolved cited work

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-10T17:05:29.031667Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-10T17:05:29.031667Z digest=sha256:322be4781f8315f6ec65f5e5271ca75c28e3ae91fb68084b83aa4e11575cc22b

Observation ba47349e-543f-4d1f-96c3-9fda901104bf · outbound

This paper cites Inverse reinforcement learning with the average reward criterion.

Inverse Reinforcement Learning with Switching Rewards and History Dependency for Characterizing Animal Behaviors Inverse reinforcement learning with the average reward criterion

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T17:05:29.650153Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-10T17:05:29.034608Z digest=sha256:5b823d31ed334f0de91bbbf8de9ba5eb4e97205a447acdc40a188a0018b57343

Observation 1b85156d-480c-4efa-a8a7-02b44159334a · outbound

This paper cites Imitating Language via Scalable Inverse Reinforcement Learning.

Inverse Reinforcement Learning with Switching Rewards and History Dependency for Characterizing Animal Behaviors Imitating Language via Scalable Inverse Reinforcement Learning

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-10T17:05:29.037500Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-10T17:05:29.037500Z digest=sha256:bc12b14118baf198b7506dac4452de893e188a1b6301555cbca45f6e81099aa4

Observation 4c744ff2-8032-44f2-a197-3d9a3014733f · outbound

This paper cites Identification of animal behavioral strategies by inverse reinforcement learning.

Inverse Reinforcement Learning with Switching Rewards and History Dependency for Characterizing Animal Behaviors Identification of animal behavioral strategies by inverse reinforcement learning

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T17:05:29.640503Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-10T17:05:29.041100Z digest=sha256:c1d2e6385913740e6cd0d348dad8f50fadbb0384f7a7adb7df396d65f1d467c0

Observation 464930bb-32d5-420a-957f-c511a27d0462 · outbound

This paper cites Maximum-likelihood inverse reinforcement learning with finite-time guarantees.

Inverse Reinforcement Learning with Switching Rewards and History Dependency for Characterizing Animal Behaviors Maximum-likelihood inverse reinforcement learning with finite-time guarantees

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T17:05:29.630114Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-10T17:05:29.043980Z digest=sha256:7a974538b8c118bfdbcedf29e280737fee065f737f34786fdbdbe43f020b59d0

Observation 4c32f061-0b69-4036-be9b-1301f034a06c · outbound

This paper cites Multi-intention inverse q-learning for interpretable behavior representation.

Inverse Reinforcement Learning with Switching Rewards and History Dependency for Characterizing Animal Behaviors Multi-intention inverse q-learning for interpretable behavior representation

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T17:05:29.618872Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-10T17:05:29.046980Z digest=sha256:15b98f95b92eb15547142d783a0df727ba9bd9eb3d28f8a30426c7a48f8ed04a

Observation c4a94482-262b-48d2-8fef-da36566aea0a · outbound

This paper cites Ziebart, Andrew Maas, J.

Inverse Reinforcement Learning with Switching Rewards and History Dependency for Characterizing Animal Behaviors Ziebart, Andrew Maas, J

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-10T17:05:29.050233Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-10T17:05:29.050233Z digest=sha256:62987647b52db1c10b9081e5188a935449fe6dac288c43cd9582ab85b7e3bad8

Observation 4c2a85c1-0248-42ef-9331-fea43efe26bb · outbound

This paper cites Ziebart, J.

Inverse Reinforcement Learning with Switching Rewards and History Dependency for Characterizing Animal Behaviors Ziebart, J

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-10T17:05:29.053428Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-10T17:05:29.053428Z digest=sha256:e986304e5bbccab5fd042eb464b79e63d6bf79a462d111582706628d65c89d1b

Observation b2a863eb-0719-42b0-8ee7-863662020543 · outbound

This paper cites write newline.

Inverse Reinforcement Learning with Switching Rewards and History Dependency for Characterizing Animal Behaviors write newline

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-10T17:05:29.056307Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-10T17:05:29.056307Z digest=sha256:6765d12090115e66ba9a882b5d0814f2aecfce14bdd73f883f2726efa50e3ee5

Observation d3d62832-d692-48eb-9628-5bce9ccc937d · outbound

This paper cites @esa (Ref.

Inverse Reinforcement Learning with Switching Rewards and History Dependency for Characterizing Animal Behaviors @esa (Ref

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-10T17:05:29.060869Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-10T17:05:29.060869Z digest=sha256:409710ee15cf8405dc44e64b0c4876070d31f31c5aac3be4e6a2a2eb7088b5ba

Observation 1d8df172-0eec-49d6-ad23-ebccad2ed9a3 · outbound

This paper cites an unresolved cited work.

Inverse Reinforcement Learning with Switching Rewards and History Dependency for Characterizing Animal Behaviors Unresolved cited work

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-10T17:05:29.064277Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-10T17:05:29.064277Z digest=sha256:8e3ea1968afa4b8ec5f6eaf6a5aca33c7e4ec6071937eefab519f429500a512e

Observation 7db36db6-7191-4c21-a088-56f2739c9be3 · outbound

This paper cites an unresolved cited work.

Inverse Reinforcement Learning with Switching Rewards and History Dependency for Characterizing Animal Behaviors Unresolved cited work

Reference 42

Resolution
verified exact
raw_fallback, observed 2026-08-10T17:05:29.219493Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-10T17:05:29.071278Z digest=sha256:9ffa1f264794beb23dfdea774cd4dcfab9e94697431bdab48a784ad693e0eb43

Pith citing papers

Observation c51944cb-860f-43b7-bdd3-aa00e00e696b · inbound

Distributional Inverse Reinforcement Learning cites this paper.

Distributional Inverse Reinforcement Learning Inverse Reinforcement Learning with Switching Rewards and History Dependency for Characterizing Animal Behaviors

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-04T12:42:58.307441Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T12:42:58.307441Z digest=sha256:2363266aec11f8def46fa80c6b73ebdbb8f77e8266b2ec1f2b7328215ee46dff

Observation 5dd3d431-9932-4ed4-acfd-6a61b135a375 · inbound

Improving Zero-Shot Offline RL via Behavioral Task Sampling cites this paper.

Improving Zero-Shot Offline RL via Behavioral Task Sampling Inverse Reinforcement Learning with Switching Rewards and History Dependency for Characterizing Animal Behaviors

Reference 12

Resolution
verified exact
arxiv_id, observed 2026-05-11T23:41:18.778357Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-07T16:27:15.347522Z digest=sha256:2cab566fafbb7c4ca61fd9a41355c43306d4a228d21b226855f94af1591f50d5

Observation da554326-c531-43e7-bf51-d82c82b18576 · inbound

Probabilistic Recurrent Intention Switching Model cites this paper.

Probabilistic Recurrent Intention Switching Model Inverse Reinforcement Learning with Switching Rewards and History Dependency for Characterizing Animal Behaviors

Reference 2

Resolution
metadata mismatch
arxiv_id, observed 2026-06-29T19:23:53.744266Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-06-29T19:22:57.970141Z digest=sha256:c83e8c9206691ecb113c4a8360cf0b588686fa1905827192e11e268910040dac