Which Factorization Machine Modeling Is Better: A Theoretical Answer with Optimal Guarantee

Ming Lin; Shuang Qiu; Jieping Ye; Xiaomin Song; Qi Qian; Liang Sun; Shenghuo Zhu; Rong Jin

doi:10.1609/aaai.v33i01.33014312

Authors

Ming Lin Alibaba Group
Shuang Qiu University of Michigan
Jieping Ye University of Michigan
Xiaomin Song Alibaba Group
Qi Qian Alibaba Group
Liang Sun Alibaba Group
Shenghuo Zhu Alibaba Group
Rong Jin Alibaba Group

DOI:

https://doi.org/10.1609/aaai.v33i01.33014312

Abstract

Factorization machine (FM) is a popular machine learning model to capture the second order feature interactions. The optimal learning guarantee of FM and its generalized version is not yet developed. For a rank k generalized FM of d dimensional input, the previous best known sampling complexity is O[k³d · polylog(kd)] under Gaussian distribution. This bound is sub-optimal comparing to the information theoretical lower bound O(kd). In this work, we aim to tighten this bound towards optimal and generalize the analysis to sub-gaussian distribution. We prove that when the input data satisfies the so-called τ-Moment Invertible Property, the sampling complexity of generalized FM can be improved to O[k²d · polylog(kd)/τ²]. When the second order self-interaction terms are excluded in the generalized FM, the bound can be improved to the optimal O[kd · polylog(kd)] up to the logarithmic factors. Our analysis also suggests that the positive semi-definite constraint in the conventional FM is redundant as it does not improve the sampling complexity while making the model difficult to optimize. We evaluate our improved FM model in real-time high precision GPS signal calibration task to validate its superiority.

Which Factorization Machine Modeling Is Better: A Theoretical Answer with Optimal Guarantee

Authors

DOI:

Abstract

Downloads

Published

How to Cite

Issue

Section

Information

Subscription