• DocumentCode
    33642
  • Title

    Nonparametric Tikhonov Regularized NMF and Its Application in Cancer Clustering

  • Author

    Mirzal, Andri

  • Author_Institution
    Comput. Sci. Dept., Dhofar Univ., Salalah, Oman
  • Volume
    11
  • Issue
    6
  • fYear
    2014
  • fDate
    Nov.-Dec. 1 2014
  • Firstpage
    1208
  • Lastpage
    1217
  • Abstract
    The Tikhonov regularized nonnegative matrix factorization (TNMF) is an NMF objective function that enforces smoothness on the computed solutions, and has been successfully applied to many problem domains including text mining, spectral data analysis, and cancer clustering. There is, however, an issue that is still insufficiently addressed in the development of TNMF algorithms, i.e., how to develop mechanisms that can learn the regularization parameters directly from the data sets. The common approach is to use fixed values based on a priori knowledge about the problem domains. However, from the linear inverse problems study it is known that the quality of the solutions of the Tikhonov regularized least square problems depends heavily on the choosing of appropriate regularization parameters. Since least squares are the building blocks of the NMF, it can be expected that similar situation also applies to the NMF. In this paper, we propose two formulas to automatically learn the regularization parameters from the data set based on the L-curve approach. We also develop a convergent algorithm for the TNMF based on the additive update rules. Finally, we demonstrate the use of the proposed algorithm in cancer clustering tasks.
  • Keywords
    bioinformatics; cancer; convergence of numerical methods; data analysis; data mining; inverse problems; least squares approximations; matrix decomposition; pattern clustering; L-curve approach; NMF objective function; Tikhonov regularized least square problems; Tikhonov regularized nonnegative matrix factorization; additive update rules; building blocks; cancer clustering application; cancer clustering tasks; computed solutions; convergent algorithm; data sets; linear inverse problems; nonparametric Tikhonov regularized NMF; priori knowledge; problem domains; regularization parameters; spectral data analysis; text mining; Algorithm design and analysis; Approximation error; Bioinformatics; Cluster approximation; Computational biology; Linear programming; Cancer clustering; L-curve; Tikhonov regularization; nonnegative matrix factorization; nonparametric learning;
  • fLanguage
    English
  • Journal_Title
    Computational Biology and Bioinformatics, IEEE/ACM Transactions on
  • Publisher
    ieee
  • ISSN
    1545-5963
  • Type

    jour

  • DOI
    10.1109/TCBB.2014.2328342
  • Filename
    6824755