Title :
Simultaneous feature selection and parameter optimization for training of dialog policy by reinforcement learning
Author :
Misu, Teruhisa ; Kashioka, Hideki
Author_Institution :
Nat. Inst. of Inf. & Commun. Technol. (NICT), Kyoto, Japan
Abstract :
This paper addresses the problem of feature selection in the reinforcement learning (RL) of the dialog policies of spoken dialog systems. A statistical dialog manager selects the system actions the system should take based on the features derived from the current dialog state and/or the system´s belief state. When defining the features used by the system for training the dialog policy, however, finding a set of actually effective features from potentially useful ones is not obvious. In addition, the selection should be done simultaneously with the optimization of the dialog policy. In this paper, we propose an incremental feature selection method for the optimization of a dialog policy by RL, in which improvement of the dialog policy and the feature selection are conducted simultaneously. Experiments in dialog policy optimization by RL with a user simulator demonstrated the following: 1) that the proposed method can find a better dialog policy with fewer policy iterations and 2) the learning speed is comparable with the case where feature selection is conducted in advance.
Keywords :
interactive systems; learning (artificial intelligence); natural language processing; RL; dialog policy training; incremental feature selection method; reinforcement learning; simultaneous parameter optimization; spoken dialog systems; Entropy; Feature extraction; Learning; Optimization; Training; Training data; Vectors; Dialog management; Feature selection; Reinforcement learning; Spoken dialog systems;
Conference_Titel :
Spoken Language Technology Workshop (SLT), 2012 IEEE
Conference_Location :
Miami, FL
Print_ISBN :
978-1-4673-5125-6
Electronic_ISBN :
978-1-4673-5124-9
DOI :
10.1109/SLT.2012.6424160