Pith. sign in

Paper Citation Record · LEDGER

A Unified Approach to Reinforcement Learning, Quantal Response Equilibria, and Two-Player Zero-Sum Games

As of 13 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 10 inbound Pith citation observations for arXiv:2206.05825.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2206.05825 v4

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 10 of 10 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-13T06:32:02.005865+00:00

measured 10 of 10 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-08T19:53:25.787601Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-07-08T03:24:28.820322Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 114ccd43-414f-413a-be1a-983e04ccef09 · inbound

Two-Player Zero-Sum Differential Games with One-Sided Information cites this paper.

Two-Player Zero-Sum Differential Games with One-Sided Information A Unified Approach to Reinforcement Learning, Quantal Response Equilibria, and Two-Player Zero-Sum Games

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-08T19:53:25.787601Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T19:53:25.787601Z digest=sha256:57b9bd956f690669e24459998429b06dcc950804882614ed93aa1a3faac4685b

Observation 190c2580-1e85-4e8c-9c74-2aeb7f8369c0 · inbound

Incentivize without Bonus: Provably Efficient Model-based Online Multi-agent RL for Markov Games cites this paper.

Incentivize without Bonus: Provably Efficient Model-based Online Multi-agent RL for Markov Games A Unified Approach to Reinforcement Learning, Quantal Response Equilibria, and Two-Player Zero-Sum Games

Reference 52

Resolution
unresolved
no resolver link, observed 2026-08-07T20:38:32.878385Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T20:38:32.878385Z digest=sha256:bffba2bcd40ad0c3c4d4a70dd9c0ad146d976b205e945a0730c4fcddb2ab969c

Observation 2b233b10-28b9-47d9-b0cd-ab49eb3aba8d · inbound

Multiplayer Nash Preference Optimization cites this paper.

Multiplayer Nash Preference Optimization A Unified Approach to Reinforcement Learning, Quantal Response Equilibria, and Two-Player Zero-Sum Games

Reference 24

Resolution
verified exact
arxiv_id, observed 2026-05-18T13:11:23.978594Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-05-18T13:09:54.433720Z digest=sha256:0c0d0aab253353df2c545bdfc70e0d5ab6dc7ea74ad671a593c4a442cde9e602

Observation a502d113-a250-44f7-af66-af92488bbcc4 · inbound

Outbidding and Outbluffing Elite Humans: Mastering Liar's Poker via Self-Play and Reinforcement Learning cites this paper.

Outbidding and Outbluffing Elite Humans: Mastering Liar's Poker via Self-Play and Reinforcement Learning A Unified Approach to Reinforcement Learning, Quantal Response Equilibria, and Two-Player Zero-Sum Games

Reference 9

Resolution
verified exact
arxiv_id, observed 2026-05-18T01:00:34.366221Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-05-18T00:58:51.759880Z digest=sha256:ec535417acb5a2ea33303fb96a5f5f36ae07266eae1492f264d1735ff1e01211

Observation abf03fa3-2d44-4fdf-bc52-d311d7d9fb5b · inbound

GAE Falls Short in Imperfect-Information Self-Play Reinforcement Learning cites this paper.

GAE Falls Short in Imperfect-Information Self-Play Reinforcement Learning A Unified Approach to Reinforcement Learning, Quantal Response Equilibria, and Two-Player Zero-Sum Games

Reference 21

Resolution
verified exact
arxiv_id, observed 2026-05-20T07:53:10.001880Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-05-20T07:50:09.840762Z digest=sha256:2ae57ed16ec7e307cb24a74ae5571db5cae2de1b01d0360ac0c58c5c086667e6

Observation f6e8c50c-1c80-40b1-996c-54c4e83aa27b · inbound

Which Nash Equilibrium? Solver-Dependent Selection on Zero-Sum Nash Polytopes cites this paper.

Which Nash Equilibrium? Solver-Dependent Selection on Zero-Sum Nash Polytopes A Unified Approach to Reinforcement Learning, Quantal Response Equilibria, and Two-Player Zero-Sum Games

Reference 17

Resolution
verified exact
arxiv_id, observed 2026-07-01T18:45:58.589040Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=arxiv_source observed=2026-06-29T01:45:47.417624Z digest=sha256:1fd4eaae84865ac7a689bc7dfd07f3b0e8637179bc01ccdc0586775e8543eac8

Observation 3d3fc67c-431d-4c74-b948-c3c202956b11 · inbound

How Much Due Diligence Before You Bid? Learning in Intractable Takeover Auctions cites this paper.

How Much Due Diligence Before You Bid? Learning in Intractable Takeover Auctions A Unified Approach to Reinforcement Learning, Quantal Response Equilibria, and Two-Player Zero-Sum Games

Reference 13

Resolution
metadata mismatch
arxiv_id, observed 2026-06-30T07:04:20.866971Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-06-30T07:03:22.061105Z digest=sha256:85266bd635065edba76cc80669061e5d81fd03b2599e0820e5bc0a37233d4ef3

Observation 4f035eef-7dd1-4219-9ce5-f1e1a664c9e5 · inbound

FootsiesGym: A Fighting Game Benchmark for Two-Player Zero-Sum Imperfect-Information Games cites this paper.

FootsiesGym: A Fighting Game Benchmark for Two-Player Zero-Sum Imperfect-Information Games A Unified Approach to Reinforcement Learning, Quantal Response Equilibria, and Two-Player Zero-Sum Games

Reference 5

Resolution
metadata mismatch
local_arxiv, observed 2026-07-08T03:24:28.821966Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=arxiv_source observed=2026-07-08T03:17:09.608451Z digest=sha256:0a7aaad025f76ff52d1e8c1779742a0a36a469a3dde15335aa8f9e594ed98659

Observation 2a9e4b30-0f75-48d6-aa88-50ce89117c6c · inbound

The Curvature Shadow: An Apparent Failure of Maximum-Entropy Equilibrium Selection is a Removable Artifact cites this paper.

The Curvature Shadow: An Apparent Failure of Maximum-Entropy Equilibrium Selection is a Removable Artifact A Unified Approach to Reinforcement Learning, Quantal Response Equilibria, and Two-Player Zero-Sum Games

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-01T17:45:29.706313Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T17:45:29.706313Z digest=sha256:87469a3e28c430b07ebb893761ad5f661d8c8b0891a0944bc4467001f679367c

Observation 32ff7f07-4263-46a4-b9a3-51803c1ebd9e · inbound

IFlowNets: Extending Generative Samplers to Learn Strategies in Incomplete Information Games cites this paper.

IFlowNets: Extending Generative Samplers to Learn Strategies in Incomplete Information Games A Unified Approach to Reinforcement Learning, Quantal Response Equilibria, and Two-Player Zero-Sum Games

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-08T13:37:13.048094Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T13:37:13.048094Z digest=sha256:06af589571baa488b456f4cdcfffe733842b90358bf5a4d1ddc503693d0c30de