Pith. sign in

Paper Citation Record · LEDGER

CHAI: A CHatbot AI for Task-Oriented Dialogue with Offline Reinforcement Learning

As of 10 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 3 inbound Pith citation observations for arXiv:2204.08426.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2204.08426 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 3 of 3 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-10T06:31:04.303077+00:00

measured 3 of 3 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T20:57:48.525751Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-17T12:04:10.762877Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation ec576183-952e-4942-aac5-d6807e07ee37 · inbound

Training Language Models to Self-Correct via Reinforcement Learning cites this paper.

Training Language Models to Self-Correct via Reinforcement Learning CHAI: A CHatbot AI for Task-Oriented Dialogue with Offline Reinforcement Learning

Reference 175

Resolution
verified exact
arxiv_id, observed 2026-05-17T12:04:10.765332Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-05-17T12:04:10.210508Z digest=sha256:8a3fcc31d3a53c1ba3a39a25745886d8772316df96b5df53542dd4d49fcbe666

Observation 7948ea1c-4f62-44cc-9eea-1c02a8eb451c · inbound

Digi-Q: Learning Q-Value Functions for Training Device-Control Agents cites this paper.

Digi-Q: Learning Q-Value Functions for Training Device-Control Agents CHAI: A CHatbot AI for Task-Oriented Dialogue with Offline Reinforcement Learning

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-07T20:57:48.525751Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T20:57:48.525751Z digest=sha256:0ed16845da35a3614e382c740943d2a526ddfd3c75db499e6ad01a3601c9d6c0

Observation c4b31ec5-ca97-4e35-a30f-4380ee3cba57 · inbound

Reinforcement Learning for Machine Learning Engineering Agents cites this paper.

Reinforcement Learning for Machine Learning Engineering Agents CHAI: A CHatbot AI for Task-Oriented Dialogue with Offline Reinforcement Learning

Reference 52

Resolution
unresolved
no resolver link, observed 2026-08-05T12:24:02.358017Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T12:24:02.358017Z digest=sha256:aa243817244f7a9e87774b653af38c023bff1e9052e289b854a36a030090de90