Pith. sign in

REVIEW 1 cited by

Collaborative Deep Learning Across Multiple Data Centers

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 1810.06877 v1 pith:4FP6Z27B submitted 2018-10-16 cs.LG cs.CVstat.ML

classification cs.LGcs.CVstat.ML
keywords datadeepmodelaveragingcenterslearningmultipleoften
verification ladder T0 review T1 audit T2 compute T3 formal
0 comments
read the original abstract

Valuable training data is often owned by independent organizations and located in multiple data centers. Most deep learning approaches require to centralize the multi-datacenter data for performance purpose. In practice, however, it is often infeasible to transfer all data to a centralized data center due to not only bandwidth limitation but also the constraints of privacy regulations. Model averaging is a conventional choice for data parallelized training, but its ineffectiveness is claimed by previous studies as deep neural networks are often non-convex. In this paper, we argue that model averaging can be effective in the decentralized environment by using two strategies, namely, the cyclical learning rate and the increased number of epochs for local model training. With the two strategies, we show that model averaging can provide competitive performance in the decentralized mode compared to the data-centralized one. In a practical environment with multiple data centers, we conduct extensive experiments using state-of-the-art deep network architectures on different types of data. Results demonstrate the effectiveness and robustness of the proposed method.

Discussion (0). Sign in to comment.

Forward citations

Cited by 1 Pith paper

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score.

  1. ScaleAcross: Designing Multi-Data-Center Infrastructure for Geo-Distributed AI Training

    cs.NI 2026-06 unverdicted novelty 4.0 of 10

    Presents an EVPN-VXLAN emulation framework with ECMP, BFD, and queue-pair traffic distribution for studying AllReduce and Parameter Server patterns in geo-distributed AI training.

Pith tools