Pith. sign in

Paper Citation Record · LEDGER

DialogXpert: Driving Intelligent and Emotion-Aware Conversations through Online Value-Based Reinforcement Learning with LLM Priors

As of 14 August 2026, this Paper Citation Record lists 16 of 16 outbound references and 0 inbound Pith citation observations for arXiv:2505.17795.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2505.17795 v1

Coverage vector

measured 16 of 16 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T14:45:05.543512Z

measured 16 of 16 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-14T06:32:32.682623+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

16 of 16 outbound references displayed

  • verified exact0
  • verified fuzzy3
  • unresolved11
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch2

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation e4d50b19-52a5-435b-8b5f-f5d45e776197 · outbound

This paper cites an unresolved cited work.

DialogXpert: Driving Intelligent and Emotion-Aware Conversations through Online Value-Based Reinforcement Learning with LLM Priors Unresolved cited work

Reference 1

Resolution
unresolved
raw_fallback, observed 2026-08-07T14:45:08.800899Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-07T14:45:04.360778Z digest=sha256:02d56356476a2b4cdfbaee86865f893cb874a6b817acadf283b2f9fd6bf0bce6

Observation 0e985aac-c9b1-418d-a185-fbe73d840e2e · outbound

This paper cites an unresolved cited work.

DialogXpert: Driving Intelligent and Emotion-Aware Conversations through Online Value-Based Reinforcement Learning with LLM Priors Unresolved cited work

Reference 2

Resolution
unresolved
raw_fallback, observed 2026-08-07T14:45:08.563763Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-07T14:45:04.487068Z digest=sha256:f17b3ab7da3d7ce651b9202469fe7055d45c69ea134c384181c6bf88f513d978

Observation 6274500f-ab8b-429d-81c3-47f04abfe6ea · outbound

This paper cites an unresolved cited work.

DialogXpert: Driving Intelligent and Emotion-Aware Conversations through Online Value-Based Reinforcement Learning with LLM Priors Unresolved cited work

Reference 3

Resolution
unresolved
raw_fallback, observed 2026-08-07T14:45:08.307620Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-07T14:45:04.636802Z digest=sha256:c64da6e4ef14bb8a39c33f5174889c24260d63484761538db6f15f72c67708bf

Observation c05cc2ba-7616-497e-b903-44e2ed90b435 · outbound

This paper cites Reasoning with Language Model is Planning with World Model.

DialogXpert: Driving Intelligent and Emotion-Aware Conversations through Online Value-Based Reinforcement Learning with LLM Priors Reasoning with Language Model is Planning with World Model

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-07T14:45:03.683850Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:45:03.683850Z digest=sha256:eca803211aa9daf4f9af1327c65807ae7c6eeb7e998710cfc0c821403940cc1a

Observation c64af7e9-f5b9-4365-945e-59f190d73473 · outbound

This paper cites Direct Preference Optimization: Your Language Model is Secretly a Reward Model.

DialogXpert: Driving Intelligent and Emotion-Aware Conversations through Online Value-Based Reinforcement Learning with LLM Priors Direct Preference Optimization: Your Language Model is Secretly a Reward Model

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-07T14:45:03.812358Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:45:03.812358Z digest=sha256:db0d3b38d5f90a8d6cb2c62d9fe36f2604a0f8681f1fd651665b1ee436547e22

Observation d3bc0cf5-81c3-482d-b345-59baec77a584 · outbound

This paper cites Conversational Tree Search: A New Hybrid Dialog Task.

DialogXpert: Driving Intelligent and Emotion-Aware Conversations through Online Value-Based Reinforcement Learning with LLM Priors Conversational Tree Search: A New Hybrid Dialog Task

Reference 7

Resolution
metadata mismatch
local_arxiv, observed 2026-08-07T14:45:05.841278Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-07T14:45:04.027094Z digest=sha256:d6edfc94395cd982950a6042d970b4df22be6c10d120fab51b30b78cf91bdaa5

Observation fdf8747c-3921-4203-8758-66c9266eb4a4 · outbound

This paper cites ProAgent: Building Proactive Cooperative Agents with Large Language Models.

DialogXpert: Driving Intelligent and Emotion-Aware Conversations through Online Value-Based Reinforcement Learning with LLM Priors ProAgent: Building Proactive Cooperative Agents with Large Language Models

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-07T14:45:04.170030Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:45:04.170030Z digest=sha256:f78bb0d6d60e464267fc535d7e1deb075bd2b61ee329a70b430f3e7ec030c0d9

Observation a9534df6-4938-44d9-a324-a86b09c210e0 · outbound

This paper cites Your responses are auto-saved after each item.

DialogXpert: Driving Intelligent and Emotion-Aware Conversations through Online Value-Based Reinforcement Learning with LLM Priors Your responses are auto-saved after each item

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:45:08.023757Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-07T14:45:04.823665Z digest=sha256:59f59de256e6ca5765f1b6b862289f468ed17f2ca6280d106233b66d08ca9f31

Observation 899b74b1-5d88-4a7c-aae6-6d65745cff9b · outbound

This paper cites an unresolved cited work.

DialogXpert: Driving Intelligent and Emotion-Aware Conversations through Online Value-Based Reinforcement Learning with LLM Priors Unresolved cited work

Reference 13

Resolution
unresolved
raw_fallback, observed 2026-08-07T14:45:07.642462Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-07T14:45:04.965029Z digest=sha256:a197e7da896c110fa7a88cb2e4cb9a12361dad593b99668d740f733ec3bb7fb1

Observation faca043b-f8b3-4c1c-a2aa-c8aaf0e8f077 · outbound

This paper cites an unresolved cited work.

DialogXpert: Driving Intelligent and Emotion-Aware Conversations through Online Value-Based Reinforcement Learning with LLM Priors Unresolved cited work

Reference 14

Resolution
unresolved
raw_fallback, observed 2026-08-07T14:45:07.331720Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-07T14:45:05.127621Z digest=sha256:31b265637adeae5a7439f41e1772380d03db627b60dd9d726b315b11e50035ff

Observation d136a87b-5703-47cf-b4b5-a6e16fbb9c3c · outbound

This paper cites Responses are auto-saved; log back in with the sameUser IDto resume.

DialogXpert: Driving Intelligent and Emotion-Aware Conversations through Online Value-Based Reinforcement Learning with LLM Priors Responses are auto-saved; log back in with the sameUser IDto resume

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:45:06.937990Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-07T14:45:05.325415Z digest=sha256:4379d23a7f24637faabe5d20ef099b25044545ba1d6940bc4f1a72b1891b92a7

Observation ce2da252-b646-4b40-a78f-396f6d175fa7 · outbound

This paper cites Policy LLM for {dataset}.

DialogXpert: Driving Intelligent and Emotion-Aware Conversations through Online Value-Based Reinforcement Learning with LLM Priors Policy LLM for {dataset}

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:45:06.557148Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-07T14:45:05.543512Z digest=sha256:bc037d5698424e84e30cc34f06681102982bc2ca8ac8899505ba18d1171e06b6

Observation 889ae298-7e7d-4bc6-9be1-c3038f25c6af · outbound

This paper cites Generating Emotionally Aligned Responses in Dialogues using Affect Control Theory.

DialogXpert: Driving Intelligent and Emotion-Aware Conversations through Online Value-Based Reinforcement Learning with LLM Priors Generating Emotionally Aligned Responses in Dialogues using Affect Control Theory

Reference 2017

Resolution
metadata mismatch
local_arxiv, observed 2026-08-07T14:45:06.166992Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-07T14:45:03.394545Z digest=sha256:9cd2145d84e439f387d9d6a7b220f819eda278f5c4a023b4d860af926897ef58

Observation 8521ef60-4d2c-471f-876a-f7578092d8b4 · outbound

This paper cites LLaMA: Open and Efficient Foundation Language Models.

DialogXpert: Driving Intelligent and Emotion-Aware Conversations through Online Value-Based Reinforcement Learning with LLM Priors LLaMA: Open and Efficient Foundation Language Models

Reference 2020

Resolution
unresolved
no resolver link, observed 2026-08-07T14:45:03.927445Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:45:03.927445Z digest=sha256:c75e6048433cc6e5792741786d2400cc137d4fa13f14861a4111da7fa98be93f

Observation 3f8f2f33-fe71-46f3-9bfa-30c40a2c1e0d · outbound

This paper cites Improving Multi-turn Emotional Support Dialogue Generation with Lookahead Strategy Planning.

DialogXpert: Driving Intelligent and Emotion-Aware Conversations through Online Value-Based Reinforcement Learning with LLM Priors Improving Multi-turn Emotional Support Dialogue Generation with Lookahead Strategy Planning

Reference 2022

Resolution
unresolved
no resolver link, observed 2026-08-07T14:45:03.479133Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:45:03.479133Z digest=sha256:1eaf11b09799f460fae31928bb5d700fe3e16bb26ee43ad88d1b1e5bc0142202

Observation 1cb7bb75-3b39-4219-b15f-a9e1e2969933 · outbound

This paper cites Improving Language Model Negotiation with Self-Play and In-Context Learning from AI Feedback.

DialogXpert: Driving Intelligent and Emotion-Aware Conversations through Online Value-Based Reinforcement Learning with LLM Priors Improving Language Model Negotiation with Self-Play and In-Context Learning from AI Feedback

Reference 2023

Resolution
unresolved
no resolver link, observed 2026-08-07T14:45:03.591088Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:45:03.591088Z digest=sha256:eadf31ae3bd71794f11bdcd5db24c0db6b967efdca3fa380ead69ab8cb3a7901

Pith citing papers

No inbound Pith citation observations are available.