REVIEW 7 cited by
Investigating Affective Use and Emotional Well-being on ChatGPT
Not yet reviewed by Pith; the record is open.
This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.
SPECIMEN: schema-true, not a live event
T0 review · schema-true
One-sentence machine reading of the paper's core claim.
pith:XXXXXXXX · record.json · timestamp
read the original abstract
As AI chatbots see increased adoption and integration into everyday life, questions have been raised about the potential impact of human-like or anthropomorphic AI on users. In this work, we investigate the extent to which interactions with ChatGPT (with a focus on Advanced Voice Mode) may impact users' emotional well-being, behaviors and experiences through two parallel studies. To study the affective use of AI chatbots, we perform large-scale automated analysis of ChatGPT platform usage in a privacy-preserving manner, analyzing over 3 million conversations for affective cues and surveying over 4,000 users on their perceptions of ChatGPT. To investigate whether there is a relationship between model usage and emotional well-being, we conduct an Institutional Review Board (IRB)-approved randomized controlled trial (RCT) on close to 1,000 participants over 28 days, examining changes in their emotional well-being as they interact with ChatGPT under different experimental settings. In both on-platform data analysis and the RCT, we observe that very high usage correlates with increased self-reported indicators of dependence. From our RCT, we find that the impact of voice-based interactions on emotional well-being to be highly nuanced, and influenced by factors such as the user's initial emotional state and total usage duration. Overall, our analysis reveals that a small number of users are responsible for a disproportionate share of the most affective cues.
Forward citations
Cited by 7 Pith papers
-
Practicing with Language Models Cultivates Human Empathic Communication
Personalized LLM feedback after practice conversations with AI partners significantly improves human empathic communication on six preregistered dimensions without homogenizing responses, while trait empathy fails to ...
-
TRACE: Temporal Relationship-Aware Conversational Entrainment Detection in Dyadic Speech
TRACE detects synthetic emotional entrainment disruption in dyadic speech at up to 93.47% accuracy when conditioned on relationship, using windowed emotion-Whisper sequences on the new DyadEE dataset.
-
Rater State Bias in RLHF Preference Data: An Audit Framework
The paper hypothesizes that stressed RLHF raters systematically prefer emotionally validating responses, and proposes an audit framework with five falsifiable predictions to detect this bias in public models.
-
Grok in the Wild: Characterizing the Roles and Uses of Large Language Models on Social Media
People on X mostly use Grok as a question-answering oracle, but a large share of interactions cast it as a truth arbiter, advocate, or adversary in public disputes — a pattern absent from private chatbot use.
-
A clinically validated framework for auditing AI chatbot behavior in mental health interactions
Using simulated psychiatric user profiles, the authors show that AI chatbots frequently produce 'concerning behavior' that accumulates over turns, and that superficially supportive responses can amplify vulnerability—...
-
Mapping the Parasocial AI Market: User Trends, Engagement and Risks
An analysis of 110 AI companion platforms finds a large, growing market dominated by general-purpose AI and mixed-use products, with mating-oriented platforms overrepresented in the UK and weak age protection.
-
Alignment Plausibility: A New Standard for Assuring AI in Healthcare
Alignment plausibility—evidence that an AI system's values, training, and oversight cohere with safe positive health outcomes—should be the regulatory analogue of biological plausibility for LLMs in healthcare.
Discussion (0). Continue with ORCID to comment.