Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-10T17:05:29.071278Z
Paper Citation Record · LEDGER
As of 11 August 2026, this Paper Citation Record lists 42 of 42 outbound references and 3 inbound Pith citation observations for arXiv:2501.12633.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-10T17:05:29.071278Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-10T06:31:04.303077+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-04T12:42:58.307441Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-06-29T19:23:53.741950Z
42 of 42 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation e31c651a-b48a-4060-8030-3f71ccc54dd5 · outbound
Inverse Reinforcement Learning with Switching Rewards and History Dependency for Characterizing Animal Behaviors Unresolved cited work
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2babc4da-bd78-4ce4-a79a-dd5003996cba · outbound
Inverse Reinforcement Learning with Switching Rewards and History Dependency for Characterizing Animal Behaviors Unresolved cited work
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 06928f22-99a3-4fb0-955e-5b7b5fe850c1 · outbound
Inverse Reinforcement Learning with Switching Rewards and History Dependency for Characterizing Animal Behaviors Ashwood, Nicholas A
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 890e67a3-8726-4d3a-afb5-bf6a446529a2 · outbound
Inverse Reinforcement Learning with Switching Rewards and History Dependency for Characterizing Animal Behaviors Unresolved cited work
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation f8c7b837-2bf0-49fe-ab1f-a1f58ea797e5 · outbound
Inverse Reinforcement Learning with Switching Rewards and History Dependency for Characterizing Animal Behaviors Reinforcement learning with lstm in non-markovian tasks with longterm dependencies
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 1282d7ee-2698-4720-bfb5-35e74d3ac728 · outbound
Inverse Reinforcement Learning with Switching Rewards and History Dependency for Characterizing Animal Behaviors Option-aware adversarial inverse reinforcement learning for robotic control
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 43de4f6b-770f-4ca3-9b3f-37463e1cea94 · outbound
Inverse Reinforcement Learning with Switching Rewards and History Dependency for Characterizing Animal Behaviors Learning robust rewards with adverserial inverse reinforcement learning
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation dffccea0-607c-405b-ae09-dcaaaf4276f6 · outbound
Inverse Reinforcement Learning with Switching Rewards and History Dependency for Characterizing Animal Behaviors Iq-learn: Inverse soft-q learning for imitation
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 6f25f3a1-887d-4b1c-a0c9-9960b0861ff1 · outbound
Inverse Reinforcement Learning with Switching Rewards and History Dependency for Characterizing Animal Behaviors Reinforcement learning with deep energy-based policies
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5c6dff0c-d7d9-4947-8b4f-64ba44489e1d · outbound
Inverse Reinforcement Learning with Switching Rewards and History Dependency for Characterizing Animal Behaviors Soft actor-critic: Off-policy maximum entropy deep reinforcement learning with a stochastic actor
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 427144c5-0c2f-43ce-88b5-91edca1dbbf3 · outbound
Inverse Reinforcement Learning with Switching Rewards and History Dependency for Characterizing Animal Behaviors Area-specificity and plasticity of history-dependent value coding during learning
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation c5c50c21-94f8-4111-8cc6-0c453c3277bd · outbound
Inverse Reinforcement Learning with Switching Rewards and History Dependency for Characterizing Animal Behaviors Deep recurrent q-learning for partially observable mdps
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation fa6ed0f0-96c9-453a-8637-935da7e403ff · outbound
Inverse Reinforcement Learning with Switching Rewards and History Dependency for Characterizing Animal Behaviors Can ai predict animal movements? filling gaps in animal trajectories using inverse reinforcement learning
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 2f2eb347-f708-452c-b819-0ec3daf43d64 · outbound
Inverse Reinforcement Learning with Switching Rewards and History Dependency for Characterizing Animal Behaviors Vime: Variational information maximizing exploration
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 4103dd3f-95b1-4c01-ac9c-6248f1fafbc8 · outbound
Inverse Reinforcement Learning with Switching Rewards and History Dependency for Characterizing Animal Behaviors The what, how, and why of naturalistic behavior
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9d79c6b3-84ee-4c1a-94cb-793bb61daef1 · outbound
Inverse Reinforcement Learning with Switching Rewards and History Dependency for Characterizing Animal Behaviors Recurrent switching linear dynamical systems
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1de4615a-6cf6-4980-a309-d20ee000f3e8 · outbound
Inverse Reinforcement Learning with Switching Rewards and History Dependency for Characterizing Animal Behaviors Spontaneous behaviour is structured by reinforcement without explicit reward
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 897886ba-5686-4cec-9104-bde448a1cdd1 · outbound
Inverse Reinforcement Learning with Switching Rewards and History Dependency for Characterizing Animal Behaviors Neural mechanisms underlying the temporal organization of naturalistic animal behavior
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 8ba32143-773c-44da-b6c7-1827de1a3ed3 · outbound
Inverse Reinforcement Learning with Switching Rewards and History Dependency for Characterizing Animal Behaviors Ng and Stuart J
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation e0baeb92-cfd2-447d-8c1a-7ee6aad6784e · outbound
Inverse Reinforcement Learning with Switching Rewards and History Dependency for Characterizing Animal Behaviors Inverse reinforcement learning with locally consistent reward functions
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 0ab91280-6518-42cc-8831-eaf8884f6c9b · outbound
Inverse Reinforcement Learning with Switching Rewards and History Dependency for Characterizing Animal Behaviors Neural Map: Structured Memory for Deep Reinforcement Learning
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ba0d4822-8f84-4231-aa2a-2259be80926c · outbound
Inverse Reinforcement Learning with Switching Rewards and History Dependency for Characterizing Animal Behaviors Inverse reinforcement learning of bird flocking behavior
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation ea2be7f1-1628-401d-88af-fd0ee16ec21e · outbound
Inverse Reinforcement Learning with Switching Rewards and History Dependency for Characterizing Animal Behaviors Mice in a labyrinth show rapid learning, sudden insight, and efficient exploration
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation dd59d97c-f7dc-4b11-a245-a941a9a686f6 · outbound
Inverse Reinforcement Learning with Switching Rewards and History Dependency for Characterizing Animal Behaviors Obtaining reward functions of rats using inverse reinforcement learning
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation a564e214-214c-4827-a054-5d47c3763555 · outbound
Inverse Reinforcement Learning with Switching Rewards and History Dependency for Characterizing Animal Behaviors Active sensing with predictive coding and uncertainty minimization
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 299d3dd6-e37f-4c77-9a1c-ba5ec95589c8 · outbound
Inverse Reinforcement Learning with Switching Rewards and History Dependency for Characterizing Animal Behaviors Latent Variable Models for Characterizing the Dynamic Structure Underlying Complex Behaviors
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 9bd69652-71cb-4681-b458-9cec253884de · outbound
Inverse Reinforcement Learning with Switching Rewards and History Dependency for Characterizing Animal Behaviors Bayesian nonparametric inverse reinforcement learning for switched markov decision processes
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation ed2aa13e-5143-443e-be81-8636624f3247 · outbound
Inverse Reinforcement Learning with Switching Rewards and History Dependency for Characterizing Animal Behaviors Dyna, an integrated architecture for learning, planning, and reacting
Reference 28
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation adfcf920-13b8-418e-8918-e4843a75b2b6 · outbound
Inverse Reinforcement Learning with Switching Rewards and History Dependency for Characterizing Animal Behaviors Keypoint-moseq: parsing behavior by linking point tracking to pose dynamics
Reference 29
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 1f5fe048-d2d1-414f-9316-5b995af711ea · outbound
Inverse Reinforcement Learning with Switching Rewards and History Dependency for Characterizing Animal Behaviors Mapping sub-second structure in mouse behavior
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 46462f06-c652-4c83-b02d-3fb2003331b5 · outbound
Inverse Reinforcement Learning with Switching Rewards and History Dependency for Characterizing Animal Behaviors Unresolved cited work
Reference 31
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ba47349e-543f-4d1f-96c3-9fda901104bf · outbound
Inverse Reinforcement Learning with Switching Rewards and History Dependency for Characterizing Animal Behaviors Inverse reinforcement learning with the average reward criterion
Reference 32
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 1b85156d-480c-4efa-a8a7-02b44159334a · outbound
Inverse Reinforcement Learning with Switching Rewards and History Dependency for Characterizing Animal Behaviors Imitating Language via Scalable Inverse Reinforcement Learning
Reference 33
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4c744ff2-8032-44f2-a197-3d9a3014733f · outbound
Inverse Reinforcement Learning with Switching Rewards and History Dependency for Characterizing Animal Behaviors Identification of animal behavioral strategies by inverse reinforcement learning
Reference 34
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 464930bb-32d5-420a-957f-c511a27d0462 · outbound
Inverse Reinforcement Learning with Switching Rewards and History Dependency for Characterizing Animal Behaviors Maximum-likelihood inverse reinforcement learning with finite-time guarantees
Reference 35
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 4c32f061-0b69-4036-be9b-1301f034a06c · outbound
Inverse Reinforcement Learning with Switching Rewards and History Dependency for Characterizing Animal Behaviors Multi-intention inverse q-learning for interpretable behavior representation
Reference 36
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation c4a94482-262b-48d2-8fef-da36566aea0a · outbound
Inverse Reinforcement Learning with Switching Rewards and History Dependency for Characterizing Animal Behaviors Ziebart, Andrew Maas, J
Reference 37
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4c2a85c1-0248-42ef-9331-fea43efe26bb · outbound
Reference 38
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b2a863eb-0719-42b0-8ee7-863662020543 · outbound
Inverse Reinforcement Learning with Switching Rewards and History Dependency for Characterizing Animal Behaviors write newline
Reference 39
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d3d62832-d692-48eb-9628-5bce9ccc937d · outbound
Reference 40
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1d8df172-0eec-49d6-ad23-ebccad2ed9a3 · outbound
Inverse Reinforcement Learning with Switching Rewards and History Dependency for Characterizing Animal Behaviors Unresolved cited work
Reference 41
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7db36db6-7191-4c21-a088-56f2739c9be3 · outbound
Inverse Reinforcement Learning with Switching Rewards and History Dependency for Characterizing Animal Behaviors Unresolved cited work
Reference 42
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation c51944cb-860f-43b7-bdd3-aa00e00e696b · inbound
Distributional Inverse Reinforcement Learning Inverse Reinforcement Learning with Switching Rewards and History Dependency for Characterizing Animal Behaviors
Reference 27
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5dd3d431-9932-4ed4-acfd-6a61b135a375 · inbound
Improving Zero-Shot Offline RL via Behavioral Task Sampling Inverse Reinforcement Learning with Switching Rewards and History Dependency for Characterizing Animal Behaviors
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation da554326-c531-43e7-bf51-d82c82b18576 · inbound
Probabilistic Recurrent Intention Switching Model Inverse Reinforcement Learning with Switching Rewards and History Dependency for Characterizing Animal Behaviors
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.