• Title of article

    Applying regression models to query-focused multi-document summarization

  • Author/Authors

    You Ouyang، نويسنده , , Wenjie Li، نويسنده , , Sujian Li، نويسنده , , Qin Lu، نويسنده ,

  • Issue Information
    دوماهنامه با شماره پیاپی سال 2011
  • Pages
    11
  • From page
    227
  • To page
    237
  • Abstract
    Most existing research on applying machine learning techniques to document summarization explores either classification models or learning-to-rank models. This paper presents our recent study on how to apply a different kind of learning models, namely regression models, to query-focused multi-document summarization. We choose to use Support Vector Regression (SVR) to estimate the importance of a sentence in a document set to be summarized through a set of pre-defined features. In order to learn the regression models, we propose several methods to construct the “pseudo” training data by assigning each sentence with a “nearly true” importance score calculated with the human summaries that have been provided for the corresponding document set. A series of evaluations on the DUC data sets are conducted to examine the efficiency and the robustness of the proposed approaches. When compared with classification models and ranking models, regression models are consistently preferable.
  • Keywords
    Query-focused summarization , Training data construction , Support vector regression
  • Journal title
    Information Processing and Management
  • Serial Year
    2011
  • Journal title
    Information Processing and Management
  • Record number

    1229107