Pith. sign in

REVIEW 4 cited by

On the Factory Floor: ML Engineering for Industrial-Scale Ads Recommendation Models

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 2209.05310 v1 pith:WXFGBMHW submitted 2022-09-12 cs.IR cs.LG

classification cs.IRcs.LG
keywords advertisingcaseclickengineeringindustrial-scalelearningmodelrate
verification ladder T0 review T1 audit T2 compute T3 formal

Signed reviews

No signed human review yet.

0 comments
read the original abstract

For industrial-scale advertising systems, prediction of ad click-through rate (CTR) is a central problem. Ad clicks constitute a significant class of user engagements and are often used as the primary signal for the usefulness of ads to users. Additionally, in cost-per-click advertising systems where advertisers are charged per click, click rate expectations feed directly into value estimation. Accordingly, CTR model development is a significant investment for most Internet advertising companies. Engineering for such problems requires many machine learning (ML) techniques suited to online learning that go well beyond traditional accuracy improvements, especially concerning efficiency, reproducibility, calibration, credit attribution. We present a case study of practical techniques deployed in Google's search ads CTR model. This paper provides an industry case study highlighting important areas of current ML research and illustrating how impactful new ML methods are evaluated and made useful in a large-scale industrial setting.

Discussion (0). Continue with ORCID to comment.

Forward citations

Cited by 4 Pith papers

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score. Full citation record

  1. Taming the One-Epoch Phenomenon in Online Recommendation System by Two-stage Contrastive ID Pre-training

    cs.IR 2025-08 conditional novelty 5.0 of 10

    Pre-training ID embeddings with contrastive loss in a simple model avoids one-epoch overfitting and improves Pinterest's recommendation engagement by 2.2%.

  2. Powering Job Search at Scale: LLM-Enhanced Query Understanding in Job Matching Systems

    cs.IR 2025-08 conditional novelty 5.0 of 10

    A single fine-tuned 1.5B LLM can replace multiple NER-based query understanding models in a job search engine, with measured gains in relevance and lower operational overhead.

  3. Scalable Machine Learning Training Infrastructure for Online Ads Recommendation and Auction Scoring Modeling at Google

    cs.DC 2025-01 conditional novelty 5.0 of 10

    A production-scale TPU training stack for Google Ads models combines shared input memoization, hybrid embedding partitioning, pipelining, RPC coalescing, and preemption holds to improve training throughput by 116% and...

  4. Beyond Self-Consistency: Loss-Balanced Perturbation-Based Regularization Improves Industrial-Scale Ads Ranking

    cs.IR 2025-02 conditional novelty 3.0 of 10

    Adding low-weight noisy copies of training examples improves Meta's ads ranking by about 0.1% to 0.3% relative Normalized Entropy, and slightly outperforms self-consistency regularization.

Pith tools