Pith. sign in

Paper Citation Record · LEDGER

Inverse Reinforcement Learning using Revealed Preferences and Passive Stochastic Optimization

As of 17 August 2026, this Paper Citation Record lists 23 of 23 outbound references and 0 inbound Pith citation observations for arXiv:2507.04396.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2507.04396 v1

Coverage vector

measured 23 of 23 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-06T19:58:57.173888Z

measured 23 of 23 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-16T06:30:59.297886+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

23 of 23 outbound references displayed

  • verified exact1
  • verified fuzzy14
  • unresolved5
  • parse uncertain0
  • malformed identifier1
  • metadata mismatch2

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 7b56e6eb-7c03-48ca-a495-2d65b8ae96e7 · outbound

This paper cites The construction of utility functions from expenditure data.

Inverse Reinforcement Learning using Revealed Preferences and Passive Stochastic Optimization The construction of utility functions from expenditure data

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T19:58:59.723242Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-06T19:58:55.056514Z digest=sha256:6d2091b0e38c491f92a1b6f30e1956600b32eba49656e89e9cb9e7a4bf1f9903

Observation 20a30a65-8223-46c7-b36f-0cb97649882d · outbound

This paper cites Finite-sample bounds for adaptive inverse reinforcement learn- ing using passive langevin dynamics.

Inverse Reinforcement Learning using Revealed Preferences and Passive Stochastic Optimization Finite-sample bounds for adaptive inverse reinforcement learn- ing using passive langevin dynamics

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T19:58:58.304324Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-06T19:58:56.931328Z digest=sha256:f5ef446dc30461fc0382d7e8107cb6da79304a4aefb6323fc5723ade541ce910

Observation e4fa3006-cc10-42b7-9cf6-b5b4e3e08c9f · outbound

This paper cites The strong ergodic theorem for densities: generalized Shannon-McMillan- Breiman theorem.

Inverse Reinforcement Learning using Revealed Preferences and Passive Stochastic Optimization The strong ergodic theorem for densities: generalized Shannon-McMillan- Breiman theorem

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T19:58:59.340275Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-06T19:58:55.284327Z digest=sha256:e6deb68cf7df383b17db25a3c62c1b77bc048eb94ee4515c44978ee0b4009526

Observation 4bb9b298-080e-455e-9c0c-ed38d1f1b1b8 · outbound

This paper cites Fine-Tuning Language Models from Human Preferences.

Inverse Reinforcement Learning using Revealed Preferences and Passive Stochastic Optimization Fine-Tuning Language Models from Human Preferences

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-06T19:58:57.173888Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:58:57.173888Z digest=sha256:bccfad04bb7eed1e153f75552b3ebb716ba1bdbbe61e3c49c855307a8ba5ff8c

Observation 647bfaa8-b775-4bec-963c-d2dac5be68ca · outbound

This paper cites Unifying Revealed Preference and Revealed Rational Inattention.

Inverse Reinforcement Learning using Revealed Preferences and Passive Stochastic Optimization Unifying Revealed Preference and Revealed Rational Inattention

Reference 15

Resolution
metadata mismatch
local_arxiv, observed 2026-08-06T19:58:57.425926Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-06T19:58:56.400096Z digest=sha256:c8322a4ad3acb0870c53be05fed06008e2fd59d53c39bad6f9eac9351560b26b

Observation 95b75689-967f-4dd6-9368-cff862584bfe · outbound

This paper cites Continuous Inverse Optimal Control with Locally Optimal Examples.

Inverse Reinforcement Learning using Revealed Preferences and Passive Stochastic Optimization Continuous Inverse Optimal Control with Locally Optimal Examples

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-06T19:58:56.195785Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:58:56.195785Z digest=sha256:6b1be68a32cda62c2051b70b686e42b4fdae2b44a83af0a101c8b265cc87d5b5

Observation b5386775-11de-49d5-ad02-a64e1a954616 · outbound

This paper cites Langevin-type models I: diffusions with given stationary distri- butions and their discretizations.

Inverse Reinforcement Learning using Revealed Preferences and Passive Stochastic Optimization Langevin-type models I: diffusions with given stationary distri- butions and their discretizations

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T19:58:58.104986Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-06T19:58:57.027640Z digest=sha256:23871f52dfdb6b2eb4974634c70789b7044bbb3135097bd092e247ba8f17d335

Observation 32b137b8-f760-406d-919e-53a95a647e76 · outbound

This paper cites Regularized Inverse Reinforcement Learning.

Inverse Reinforcement Learning using Revealed Preferences and Passive Stochastic Optimization Regularized Inverse Reinforcement Learning

Reference 500

Resolution
metadata mismatch
local_arxiv, observed 2026-08-06T19:58:57.919764Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-06T19:58:55.819721Z digest=sha256:e13378447cfcdf3ce013dd736ecccd57ce82456ebf352ebc0b2bfd5093452562

Observation 2569f676-3f3e-4db9-b200-5f4730b25031 · outbound

This paper cites Apprenticeship learning via inverse reinforcement learning.

Inverse Reinforcement Learning using Revealed Preferences and Passive Stochastic Optimization Apprenticeship learning via inverse reinforcement learning

Reference 1979

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T19:58:59.513129Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-06T19:58:55.198530Z digest=sha256:7add9c1c558c7d751fe8aea33fb6c56823cbc8e8065ba34442c06225f08224b6

Observation bf50c81a-a70c-4880-8b55-b709b10d04b9 · outbound

This paper cites Non-convex learning via Stochastic Gradient Langevin Dynamics: a nonasymptotic analysis.

Inverse Reinforcement Learning using Revealed Preferences and Passive Stochastic Optimization Non-convex learning via Stochastic Gradient Langevin Dynamics: a nonasymptotic analysis

Reference 1983

Resolution
unresolved
no resolver link, observed 2026-08-06T19:58:56.687880Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:58:56.687880Z digest=sha256:d66546629302ec86ed2dd7219b5c34c4d9b5c3cb48af78519a7fc3f1eb4795fb

Observation 0ee527b1-8e4e-4cef-b3b4-bdec461142b5 · outbound

This paper cites Real-Time Reinforcement Learning of Constrained Markov Decision Processes with Weak Derivatives.

Inverse Reinforcement Learning using Revealed Preferences and Passive Stochastic Optimization Real-Time Reinforcement Learning of Constrained Markov Decision Processes with Weak Derivatives

Reference 1984

Resolution
verified exact
local_arxiv, observed 2026-08-06T19:58:57.731086Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-06T19:58:55.988934Z digest=sha256:cdfe4dff2a6be22640360dbb31817f1f7d843fc6a185464057aef89a4ed167a1

Observation c55b22dd-4e31-4762-95a5-666c22ad807c · outbound

This paper cites Learning Robust Rewards with Adversarial Inverse Reinforcement Learning.

Inverse Reinforcement Learning using Revealed Preferences and Passive Stochastic Optimization Learning Robust Rewards with Adversarial Inverse Reinforcement Learning

Reference 1986

Resolution
unresolved
no resolver link, observed 2026-08-06T19:58:55.741770Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:58:55.741770Z digest=sha256:8f60f76c76c0eedf4bfdc3e87895c1b698cfc1f5c9e15ead8914cef308681028

Observation 0c777eb7-1aa4-4dae-b8cb-49b465d09ccf · outbound

This paper cites Thompson sampling for contextual bandits with linear payoffs.

Inverse Reinforcement Learning using Revealed Preferences and Passive Stochastic Optimization Thompson sampling for contextual bandits with linear payoffs

Reference 1987

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T19:58:59.616636Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-06T19:58:55.121608Z digest=sha256:6d5c1a70858847af5cfc8675d11655669f93fd73f8f6100793bb651969de2869

Observation e6cb46f9-24b5-46c4-9102-7ddb9239e337 · outbound

This paper cites Maximum margin planning.

Inverse Reinforcement Learning using Revealed Preferences and Passive Stochastic Optimization Maximum margin planning

Reference 1994

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T19:58:58.660502Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-06T19:58:56.519105Z digest=sha256:81f3b5ed51ee7c12a89cc5efe0ced454dd491cec587a99009fc4b285d5a7004a

Observation 37034726-4831-418d-942c-ad481c0e2582 · outbound

This paper cites Identifiability in inverse reinforcement learning.

Inverse Reinforcement Learning using Revealed Preferences and Passive Stochastic Optimization Identifiability in inverse reinforcement learning

Reference 1996

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T19:58:59.115812Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-06T19:58:55.530001Z digest=sha256:44def3151d00d15074911ca1be51fcf5930ba7f4a82fa202f47005fb4d16db3a

Observation 8d32fc73-a7b3-4f8e-9a60-4926d3e02098 · outbound

This paper cites Risk-constrained markov decision processes.

Inverse Reinforcement Learning using Revealed Preferences and Passive Stochastic Optimization Risk-constrained markov decision processes

Reference 1999

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T19:58:59.238305Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-06T19:58:55.371670Z digest=sha256:085aad0ed6c24ed45a0c317dbe4acf2c2c5e89e5cccd65f64091eb8e9bac2944

Observation 33cd53aa-83c8-4483-82c3-9379582fd007 · outbound

This paper cites Maximum Entropy Deep Inverse Reinforcement Learning.

Inverse Reinforcement Learning using Revealed Preferences and Passive Stochastic Optimization Maximum Entropy Deep Inverse Reinforcement Learning

Reference 2000

Resolution
unresolved
no resolver link, observed 2026-08-06T19:58:57.101600Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:58:57.101600Z digest=sha256:2fcdf64f1841a4158eb4a90c680ce265b91e12d0f08ff6b5a90bfa42756264bc

Observation 0e353394-f97b-4048-908a-7f3847e7917b · outbound

This paper cites Langevin dynamics for adaptive inverse reinforcement learning of stochastic gradient algorithms.

Inverse Reinforcement Learning using Revealed Preferences and Passive Stochastic Optimization Langevin dynamics for adaptive inverse reinforcement learning of stochastic gradient algorithms

Reference 2003

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T19:58:58.787186Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-06T19:58:56.065314Z digest=sha256:e4499ce1ef5c6926132a869d3c56b697533a8ba893a2789d4b2082ecd7c78380

Observation 8d787def-222d-47e9-8df7-e8a05bcab7dd · outbound

This paper cites A testable model of consumption with externalities.

Inverse Reinforcement Learning using Revealed Preferences and Passive Stochastic Optimization A testable model of consumption with externalities

Reference 2008

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T19:58:58.988265Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-06T19:58:55.622496Z digest=sha256:3621f434efcc1d21c275765b6e9ef515ef80a9698e5c656773efd0822acc0605

Observation 1dc32662-005d-4093-b93a-1ef086834415 · outbound

This paper cites A characterization of rationalizable consumer behavior.

Inverse Reinforcement Learning using Revealed Preferences and Passive Stochastic Optimization A characterization of rationalizable consumer behavior

Reference 2013

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T19:58:58.549024Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-06T19:58:56.604226Z digest=sha256:e773029a79625f7a55d72682395734648acf90ddcaf0bdb6676884e6fbd2f136

Observation 34e537b0-045d-46d9-bc69-c994cef7dd9c · outbound

This paper cites Apprenticeship Learning for Model Parameters of Partially Observable Environments.

Inverse Reinforcement Learning using Revealed Preferences and Passive Stochastic Optimization Apprenticeship Learning for Model Parameters of Partially Observable Environments

Reference 2016

Resolution
malformed identifier
local_arxiv, observed 2026-08-06T19:58:57.566558Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-06T19:58:56.303015Z digest=sha256:1410c5f3a85d11b97cecfc1c6295301e1fb8a7cf9af6547c9306f3da3e4b3ace

Observation c0eb391e-3ab3-4faf-83cf-738ee6a8f88b · outbound

This paper cites Implications of rational inattention.

Inverse Reinforcement Learning using Revealed Preferences and Passive Stochastic Optimization Implications of rational inattention

Reference 2018

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T19:58:58.416044Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-06T19:58:56.812338Z digest=sha256:1b2556e912b84713310cf80fd49b0391c553e049c9ecbf9f1332d903758b65a5

Observation aa3ce0cc-7668-4d19-8340-b4d61fb6912b · outbound

This paper cites Inverse game theory: learning utilities in succinct games.

Inverse Reinforcement Learning using Revealed Preferences and Passive Stochastic Optimization Inverse game theory: learning utilities in succinct games

Reference 2025

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T19:58:58.883483Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-06T19:58:55.905717Z digest=sha256:a94fc76911b90c4b842e15c908dcb57c13c55ad0f8f05884ea42f4a2ae6a2244

Pith citing papers

No inbound Pith citation observations are available.