Pith. sign in

REVIEW 3 cited by

Heterogeneous Matrix Factorization: When Features Differ by Datasets

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 2305.17744 v2 pith:WJITFGLP submitted 2023-05-28 stat.ME

classification stat.ME
keywords factorsheterogeneoussharedalgorithmfactorizationfeaturematrixsources
verification ladder T0 review T1 audit T2 compute T3 formal
0 comments
read the original abstract

In myriad statistical applications, data are collected from related but heterogeneous sources. These sources share some commonalities while containing idiosyncratic characteristics. One of the most fundamental challenges in such scenarios is to recover the shared and source-specific factors. Despite the existence of a few heuristic approaches, a generic algorithm with theoretical guarantees has yet to be established. In this paper, we tackle the problem by proposing a method called Heterogeneous Matrix Factorization to separate the shared and unique factors for a class of problems. HMF maintains the orthogonality between the shared and unique factors by leveraging an invariance property in the objective. The algorithm is easy to implement and intrinsically distributed. On the theoretic side, we show that for the square error loss, HMF will converge into the optimal solutions, which are close to the ground truth. HMF can be integrated auto-encoders to learn nonlinear feature mappings. Through a variety of case studies, we showcase HMF's benefits and applicability in video segmentation, time-series feature extraction, and recommender systems.

Discussion (0). Continue with ORCID to comment.

Forward citations

Cited by 3 Pith papers

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score. Full citation record

  1. Estimating shared subspace with AJIVE: the power and limitation of multiple data matrices

    stat.ML 2025-01 conditional novelty 7.0 of 10

    In high-SNR settings AJIVE's shared-subspace error is minimax-optimal and decays like 1/√K in the number of matrices, while in low-SNR settings a non-diminishing error floor appears even for an oracle-aided spectral e...

  2. Personalized Coupled Tensor Decomposition for Multimodal Data Fusion: Uniqueness and Algorithms

    cs.LG 2024-12 conditional novelty 7.0 of 10

    Introduces a personalized coupled tensor decomposition with uniqueness guarantees based on uni-mode uniqueness and a multilinear measurement model, with semi-algebraic and ALS algorithms.

  3. A Collaborative Process Parameter Recommender System for Fleets of Networked Manufacturing Machines -- with Application to 3D Printing

    cs.LG 2025-06 conditional novelty 5.0 of 10

    Collaborative matrix completion lets a fleet of 3D printers borrow data from each other to find machine-specific optimal print settings in roughly 40 percent fewer trials than independent tuning.

Pith tools