DocumentCode
33642
Title
Nonparametric Tikhonov Regularized NMF and Its Application in Cancer Clustering
Author
Mirzal, Andri
Author_Institution
Comput. Sci. Dept., Dhofar Univ., Salalah, Oman
Volume
11
Issue
6
fYear
2014
fDate
Nov.-Dec. 1 2014
Firstpage
1208
Lastpage
1217
Abstract
The Tikhonov regularized nonnegative matrix factorization (TNMF) is an NMF objective function that enforces smoothness on the computed solutions, and has been successfully applied to many problem domains including text mining, spectral data analysis, and cancer clustering. There is, however, an issue that is still insufficiently addressed in the development of TNMF algorithms, i.e., how to develop mechanisms that can learn the regularization parameters directly from the data sets. The common approach is to use fixed values based on a priori knowledge about the problem domains. However, from the linear inverse problems study it is known that the quality of the solutions of the Tikhonov regularized least square problems depends heavily on the choosing of appropriate regularization parameters. Since least squares are the building blocks of the NMF, it can be expected that similar situation also applies to the NMF. In this paper, we propose two formulas to automatically learn the regularization parameters from the data set based on the L-curve approach. We also develop a convergent algorithm for the TNMF based on the additive update rules. Finally, we demonstrate the use of the proposed algorithm in cancer clustering tasks.
Keywords
bioinformatics; cancer; convergence of numerical methods; data analysis; data mining; inverse problems; least squares approximations; matrix decomposition; pattern clustering; L-curve approach; NMF objective function; Tikhonov regularized least square problems; Tikhonov regularized nonnegative matrix factorization; additive update rules; building blocks; cancer clustering application; cancer clustering tasks; computed solutions; convergent algorithm; data sets; linear inverse problems; nonparametric Tikhonov regularized NMF; priori knowledge; problem domains; regularization parameters; spectral data analysis; text mining; Algorithm design and analysis; Approximation error; Bioinformatics; Cluster approximation; Computational biology; Linear programming; Cancer clustering; L-curve; Tikhonov regularization; nonnegative matrix factorization; nonparametric learning;
fLanguage
English
Journal_Title
Computational Biology and Bioinformatics, IEEE/ACM Transactions on
Publisher
ieee
ISSN
1545-5963
Type
jour
DOI
10.1109/TCBB.2014.2328342
Filename
6824755
Link To Document