Pith. sign in

REVIEW 1 cited by

Conversational Tree Search: A New Hybrid Dialog Task

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 2303.10227 v1 pith:ELEIARUN submitted 2023-03-17 cs.CL cs.AIcs.LG

classification cs.CLcs.AIcs.LG
keywords dialogconversationaltaskusersanswerarchitecturebaselinegoal
verification ladder T0 review T1 audit T2 compute T3 formal
0 comments
read the original abstract

Conversational interfaces provide a flexible and easy way for users to seek information that may otherwise be difficult or inconvenient to obtain. However, existing interfaces generally fall into one of two categories: FAQs, where users must have a concrete question in order to retrieve a general answer, or dialogs, where users must follow a predefined path but may receive a personalized answer. In this paper, we introduce Conversational Tree Search (CTS) as a new task that bridges the gap between FAQ-style information retrieval and task-oriented dialog, allowing domain-experts to define dialog trees which can then be converted to an efficient dialog policy that learns only to ask the questions necessary to navigate a user to their goal. We collect a dataset for the travel reimbursement domain and demonstrate a baseline as well as a novel deep Reinforcement Learning architecture for this task. Our results show that the new architecture combines the positive aspects of both the FAQ and dialog system used in the baseline and achieves higher goal completion while skipping unnecessary questions.

Discussion (0). Continue with ORCID to comment.

Forward citations

Cited by 1 Pith paper

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score. Full citation record

  1. DialogXpert: Driving Intelligent and Emotion-Aware Conversations through Online Value-Based Reinforcement Learning with LLM Priors

    cs.CL 2025-05 conditional novelty 5.0 of 10

    DialogXpert combines a frozen LLM action proposer with a lightweight online Q-network and emotion tracking, achieving sub-3-turn dialogue success rates above 94 percent in LLM-simulated benchmarks.

Pith tools