Pith. sign in

REVIEW 2 cited by

Generalized Bradley-Terry Models for Score Estimation from Paired Comparisons

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 2308.08644 v2 pith:ZH7HUZTF submitted 2023-08-16 stat.ME math.STstat.TH

Generalized Bradley-Terry Models for Score Estimation from Paired Comparisons

classification stat.ME math.STstat.TH
keywords comparisonsmodelsalternativesbradley-terrymodelscoreclassicalcomparison
verification ladder T0 review T1 audit T2 compute T3 formal T4 reserved
0 comments
read the original abstract

Many applications, e.g. in content recommendation, sports, or recruitment, leverage the comparisons of alternatives to score those alternatives. The classical Bradley-Terry model and its variants have been widely used to do so. The historical model considers binary comparisons (victory or defeat) between alternatives, while more recent developments allow finer comparisons to be taken into account. In this article, we introduce a probabilistic model encompassing a broad variety of paired comparisons that can take discrete or continuous values. We do so by considering a well-behaved subset of the exponential family, which we call the family of generalized Bradley-Terry (GBT) models, as it includes the classical Bradley-Terry model and many of its variants. Remarkably, we prove that all GBT models are guaranteed to yield a strictly convex negative log-likelihood. Moreover, assuming a Gaussian prior on alternatives' scores, we prove that the maximum a posteriori (MAP) of GBT models, whose existence, uniqueness and fast computation are thus guaranteed, varies monotonically with respect to comparisons (the more A beats B, the better the score of A) and is Lipschitz-resilient with respect to each new comparison (a single new comparison can only have a bounded effect on all the estimated scores). These desirable properties make GBT models appealing for practical use. We illustrate some features of GBT models on simulations.

discussion (0)

Sign in with ORCID, Apple, or X to comment. Anyone can read and Pith papers without signing in.

Forward citations

Cited by 2 Pith papers

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score.

  1. Pairwise Reference Alignment as a Model-Level Ordinal Observable

    cs.CL 2026-05 unverdicted novelty 5.0

    Pairwise reference alignment is formulated as an ordinal observable equal to the probability that a model score agrees with reference preferences on triples (x, y+, y-), with centered statistics, margin extensions, es...

  2. What Catches the Eye? A Conjoint Study of Infographic Design Preferences

    cs.HC 2026-05 unverdicted novelty 5.0

    Conjoint analysis of infographic preferences on unemployment data shows comparison type (scales/benchmarks) accounts for 58.5% of preference variation, graphic type 29.2%, and color 12.3%.