Introduces a Q-sort protocol using human reference factors to quantify LLM value-structure alignment via Procrustes similarity and RSA correlations, revealing cross-family heterogeneity and localized misalignments.
Beyond human norms: Unveiling unique values of large language models through interdisciplinary approaches,
4 Pith papers cite this work. Polarity classification is still indexing.
representative citing papers
EvalMORAAL evaluates moral alignment of 20 LLMs on World Values Survey and PEW data, reporting high overall correlation with human responses but a 0.21 gap between Western and non-Western regions.
A survey proposing a three-pillar framework to evaluate LLMs as tools for measuring latent psychological constructs and reviewing applications in personality and mental health.
citing papers explorer
-
Beyond Value Benchmarks: Measuring Value-Structure Alignment in Large Language Models via Symmetric Q-Sorts
Introduces a Q-sort protocol using human reference factors to quantify LLM value-structure alignment via Procrustes similarity and RSA correlations, revealing cross-family heterogeneity and localized misalignments.
-
EvalMORAAL: Interpretable Chain-of-Thought and LLM-as-Judge Evaluation for Moral Alignment in Large Language Models
EvalMORAAL evaluates moral alignment of 20 LLMs on World Values Survey and PEW data, reporting high overall correlation with human responses but a 0.21 gap between Western and non-Western regions.
-
A Survey of Large Language Models for Perception and Measurement of Human Psychology
A survey proposing a three-pillar framework to evaluate LLMs as tools for measuring latent psychological constructs and reviewing applications in personality and mental health.
- Ads in AI Chatbots? An Analysis of How Large Language Models Navigate Conflicts of Interest