Pith. sign in

REVIEW 2 cited by

Bandits Warm-up Cold Recommender Systems

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 1407.2806 v1 pith:TOXXMYR3 submitted 2014-07-10 cs.LG cs.IRstat.ML

classification cs.LGcs.IRstat.ML
keywords colditemsratingsrecommendersettingsystemsuserscontextual
verification ladder T0 review T1 audit T2 compute T3 formal
0 comments
read the original abstract

We address the cold start problem in recommendation systems assuming no contextual information is available neither about users, nor items. We consider the case in which we only have access to a set of ratings of items by users. Most of the existing works consider a batch setting, and use cross-validation to tune parameters. The classical method consists in minimizing the root mean square error over a training subset of the ratings which provides a factorization of the matrix of ratings, interpreted as a latent representation of items and users. Our contribution in this paper is 5-fold. First, we explicit the issues raised by this kind of batch setting for users or items with very few ratings. Then, we propose an online setting closer to the actual use of recommender systems; this setting is inspired by the bandit framework. The proposed methodology can be used to turn any recommender system dataset (such as Netflix, MovieLens,...) into a sequential dataset. Then, we explicit a strong and insightful link between contextual bandit algorithms and matrix factorization; this leads us to a new algorithm that tackles the exploration/exploitation dilemma associated to the cold start problem in a strikingly new perspective. Finally, experimental evidence confirm that our algorithm is effective in dealing with the cold start problem on publicly available datasets. Overall, the goal of this paper is to bridge the gap between recommender systems based on matrix factorizations and those based on contextual bandits.

Discussion (0). Continue with ORCID to comment.

Forward citations

Cited by 2 Pith papers

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score. OpenAlex reports about 5 citations worldwide. Full citation record

  1. Optimization of Epsilon-Greedy Exploration

    cs.LG 2025-06 conditional novelty 6.0 of 10

    A gradient-based framework tunes epsilon-greedy exploration schedules by minimizing Bayesian regret, matching or beating heuristics in batched recommendation benchmarks.

  2. Preference-based learning for news headline recommendation

    cs.IR 2025-05 conditional novelty 4.0 of 10

    On real French news data, a preference-based greedy recommender matched neural Thompson sampling, and English translations performed nearly as well as original French embeddings.

Pith tools