Pith. sign in

REVIEW 16 cited by

DHEN: A Deep and Hierarchical Ensemble Network for Large-Scale Click-Through Rate Prediction

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 2203.11014 v1 pith:J2ADB5WH submitted 2022-03-11 cs.IR cs.AIcs.LG

classification cs.IRcs.AIcs.LG
keywords dheninteractionstrainingdatasetdifferentpredictioncaptureddeep
verification ladder T0 review T1 audit T2 compute T3 formal
0 comments
read the original abstract

Learning feature interactions is important to the model performance of online advertising services. As a result, extensive efforts have been devoted to designing effective architectures to learn feature interactions. However, we observe that the practical performance of those designs can vary from dataset to dataset, even when the order of interactions claimed to be captured is the same. That indicates different designs may have different advantages and the interactions captured by them have non-overlapping information. Motivated by this observation, we propose DHEN - a deep and hierarchical ensemble architecture that can leverage strengths of heterogeneous interaction modules and learn a hierarchy of the interactions under different orders. To overcome the challenge brought by DHEN's deeper and multi-layer structure in training, we propose a novel co-designed training system that can further improve the training efficiency of DHEN. Experiments of DHEN on large-scale dataset from CTR prediction tasks attained 0.27\% improvement on the Normalized Entropy (NE) of prediction and 1.2x better training throughput than state-of-the-art baseline, demonstrating their effectiveness in practice.

Discussion (0). Continue with ORCID to comment.

Forward citations

Cited by 16 Pith papers

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score. Full citation record

  1. Triton for MTIA: Bridging the Programming Model Gaps for Custom AI Accelerators

    cs.PL 2026-07 conditional novelty 7.0 of 10

    A Triton compiler backend, TorchInductor adaptations, and small language extensions let Meta's MTIA-2i run Triton kernels competitively with expert-tuned C++ in production.

  2. ROCS: Request-Oriented Compute Sharing for Efficient Large-Scale Recommendation

    cs.LG 2026-07 conditional novelty 7.0 of 10

    ROCS restructures recommendation models so user-side computation is shared across all candidate items, yielding up to 3x serving throughput at equal or better prediction quality.

  3. HCCL: Collective Communication for Meta Training and Inference Accelerators

    cs.NI 2026-08 conditional novelty 6.0 of 10

    HCCL offloads collective communication to MTIA 300's message engines, achieving up to 940 GB/s intra-rack bandwidth and sub-6µs latency for inference.

  4. Probabilistic Residual Learning for Online Recommendations

    cs.IR 2026-07 conditional novelty 6.0 of 10

    PRL adds a cluster-aware, causality-adjusted residual correction layer to any base recommender, improving cold-start cross-domain recommendation accuracy in experiments.

  5. UniRank: Benchmarking Ranking Models for Unified Sequential Modeling and Feature Interaction

    cs.IR 2026-07 conditional novelty 6.0 of 10

    UniRank is an open benchmark that standardizes chronological autoregressive supervision, multi-task evaluation, and capacity controls for 15 unified ranking models on five large datasets.

  6. Bumblebee: Interleaved Mixed-Layer Building Blocks for Large-Scale Recommendation Systems

    cs.IR 2026-07 conditional novelty 6.0 of 10

    Interleaving sequence modeling with feature interaction in repeated blocks improves recommendation accuracy by about 0.2-1.4% NE over sequential baselines at matched parameter counts.

  7. MixFormer: Co-Scaling Up Dense and Sequence in Industrial Recommenders

    cs.IR 2026-02 conditional novelty 6.0 of 10

    MixFormer unifies dense feature interaction and user-sequence modeling in a single Transformer-style backbone with a user-item decoupling speedup, reporting accuracy and efficiency gains over stacked and parallel reco...

  8. KernelEvolve: Scaling Agentic Kernel Coding for Heterogeneous AI Accelerators at Meta

    cs.LG 2025-12 conditional novelty 6.0 of 10

    An agentic kernel-coding system combining tree search with hardware-knowledge retrieval generated optimized Triton kernels for NVIDIA, AMD, and Meta's MTIA accelerators: 100% correctness on 480 operator-platform confi...

  9. Request-Only Optimization for Recommendation Systems

    cs.IR 2025-07 conditional novelty 6.0 of 10

    A request-level training data format eliminates duplicate user features, increasing storage efficiency and training throughput while enabling larger recommendation architectures.

  10. RankMixer: Scaling Up Ranking Models in Industrial Recommenders

    cs.IR 2025-07 conditional novelty 6.0 of 10

    RankMixer scales an industrial ranking model to 1B dense parameters with 10x MFU improvement and unchanged latency, gaining 1.08% in app duration in Douyin A/B tests.

  11. Action is All You Need: Dual-Flow Generative Ranking Network for Recommendation

    cs.IR 2025-05 conditional novelty 6.0 of 10

    DFGR, a dual-flow generative ranking network, halves training cost and quarters inference cost versus HSTU-based generative ranking while reporting better offline AUC/G-AUC on public and industrial datasets.

  12. Optimus: A Generic Operator-Level PyTorch Model Transformation Framework

    cs.PF 2026-07 conditional novelty 5.5 of 10

    Optimus rewrites atomic operator patterns in PT2 graphs via greedy search, delivering large QPS, memory, and compile-time gains on production recommendation models.

  13. Large Foundation Model for Ads Recommendation

    cs.LG 2025-08 conditional novelty 5.0 of 10

    Tencent's LFM4Ads transfers user, item, and user-item cross representations from a pre-trained foundation model into downstream ad models via feature, module, and model-level mechanisms, reporting a 2.45% platform-wid...

  14. Synergizing Implicit and Explicit User Interests: A Multi-Embedding Retrieval Framework at Pinterest

    cs.IR 2025-06 conditional novelty 5.0 of 10

    A production recommender framework combines a differentiable clustering module for implicit interests and conditional retrieval for explicit followed topics, deployed at Pinterest home feed.

  15. Privacy Preserving Conversion Modeling in Data Clean Room

    cs.LG 2025-05 conditional novelty 5.0 of 10

    Batch-level aggregated gradients, LoRA adapters, and de-biased label differential privacy let advertisers and platforms train conversion models in a clean room with modest AUC loss and much lower communication cost.

  16. Decoupled Entity Representation Learning for Pinterest Ads Ranking

    cs.IR 2025-09 conditional novelty 4.0 of 10

    Pre-computed user and Pin embeddings from multi-tower models improve Pinterest ad ranking by small but statistically significant margins.

Pith tools