Pith. sign in

REVIEW 1 cited by

Kantian Deontology Meets AI Alignment: Towards Morally Grounded Fairness Metrics

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 2311.05227 v2 pith:UN7KKIXW submitted 2023-11-09 cs.AI

classification cs.AI
keywords fairnesskantianmetricsalignmentdeontologicalframeworkapproachdeontology
verification ladder T0 review T1 audit T2 compute T3 formal
0 comments
read the original abstract

Deontological ethics, specifically understood through Immanuel Kant, provides a moral framework that emphasizes the importance of duties and principles, rather than the consequences of action. Understanding that despite the prominence of deontology, it is currently an overlooked approach in fairness metrics, this paper explores the compatibility of a Kantian deontological framework in fairness metrics, part of the AI alignment field. We revisit Kant's critique of utilitarianism, which is the primary approach in AI fairness metrics and argue that fairness principles should align with the Kantian deontological framework. By integrating Kantian ethics into AI alignment, we not only bring in a widely-accepted prominent moral theory but also strive for a more morally grounded AI landscape that better balances outcomes and procedures in pursuit of fairness and justice.

Discussion (0). Continue with ORCID to comment.

Forward citations

Cited by 1 Pith paper

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score. Full citation record

  1. Towards Friendly AI: A Comprehensive Review and New Perspectives on Human-AI Alignment

    cs.AI 2024-12 conditional novelty 2.0 of 10

    A literature review that synthesizes definitions of Friendly AI and catalogs ethical arguments and technical subfields relevant to human-AI alignment.

Pith tools