Pith. sign in

REVIEW 7 cited by

Knowledge-Infused Prompting: Assessing and Advancing Clinical Text Data Generation with Large Language Models

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 2311.00287 v2 pith:KUFXSU7Q submitted 2023-11-01 cs.CL cs.AIcs.LGq-bio.QM

classification cs.CLcs.AIcs.LGq-bio.QM
keywords clinicalclingengenerationknowledgelanguagellmstasksacross
verification ladder T0 review T1 audit T2 compute T3 formal

Signed reviews

No signed human review yet.

0 comments
read the original abstract

Clinical natural language processing requires methods that can address domain-specific challenges, such as complex medical terminology and clinical contexts. Recently, large language models (LLMs) have shown promise in this domain. Yet, their direct deployment can lead to privacy issues and are constrained by resources. To address this challenge, we delve into synthetic clinical text generation using LLMs for clinical NLP tasks. We propose an innovative, resource-efficient approach, ClinGen, which infuses knowledge into the process. Our model involves clinical knowledge extraction and context-informed LLM prompting. Both clinical topics and writing styles are drawn from external domain-specific knowledge graphs and LLMs to guide data generation. Our extensive empirical study across 7 clinical NLP tasks and 16 datasets reveals that ClinGen consistently enhances performance across various tasks, effectively aligning the distribution of real datasets and significantly enriching the diversity of generated training instances. Our code is available at \url{https://github.com/ritaranx/ClinGen}.

Discussion (0). Continue with ORCID to comment.

Forward citations

Cited by 7 Pith papers

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score. Full citation record

  1. Clinical Communication Processing with Models Trained on LLM-Generated Synthetic Data: A Structured Survey and Novel Application Case Studies

    cs.CL 2026-08 conditional novelty 6.0 of 10

    Synthetic clinical communication generated by LLMs can train clinical NLP models in thirteen case studies, but only one is tested on real patient text, leaving transfer to authentic communication unproven.

  2. From Reddit to Generative AI: Evaluating Large Language Models for Anxiety Support Fine-tuned on Social Media Data

    cs.HC 2025-05 conditional novelty 6.0 of 10

    Fine-tuning GPT-3.5 and Llama 2 on r/Anxiety posts improves readability but raises toxicity and bias while reducing empathy and reflection.

  3. FinNLI: Novel Dataset for Multi-Genre Financial Natural Language Inference Benchmarking

    cs.CL 2025-04 conditional novelty 6.0 of 10

    FinNLI is a 21,304-pair financial NLI benchmark with an expert-annotated test set where the best model reaches 78.62% macro F1, exposing a gap in current models' financial reasoning.

  4. IntellAgent: A Multi-Agent Framework for Evaluating Conversational AI Systems

    cs.CL 2025-01 conditional novelty 6.0 of 10

    IntellAgent uses a policy graph, synthetic event generation, and user simulation to automatically create benchmarks that rank conversational AI agents similarly to tau-bench.

  5. Large Language Models for Cancer Communication: Evaluating Linguistic Quality, Safety, and Accessibility in Generative AI

    cs.CL 2025-05 conditional novelty 5.0 of 10

    General-purpose LLMs outperformed specialized medical LLMs on linguistic quality and emotional engagement for breast and cervical cancer questions, while medical models were simpler to read but scored worse on safety.

  6. QBD-RankedDataGen: Generating Custom Ranked Datasets for Improving Query-By-Document Search Using LLM-Reranking with Reduced Human Effort

    cs.IR 2025-05 conditional novelty 5.0 of 10

    LLM-generated rankings for query-by-document search do not improve BM25 tuning over default parameters unless validated by human ground truth.

  7. A Comprehensive Survey in LLM(-Agent) Full Stack Safety: Data, Training and Deployment

    cs.CR 2025-04 conditional novelty 5.0 of 10

    A large collaborative survey organizes LLM and LLM-agent safety issues into a full-stack lifecycle framework from data preparation to deployment.

Pith tools