• DocumentCode
    2461674
  • Title

    People-LDA: Anchoring Topics to People using Face Recognition

  • Author

    Jain, Vidit ; Learned-Miller, Erik ; McCallum, Andrew

  • Author_Institution
    Univ. of Massachusetts Amherst, Amherst
  • fYear
    2007
  • fDate
    14-21 Oct. 2007
  • Firstpage
    1
  • Lastpage
    8
  • Abstract
    Topic models have recently emerged as powerful tools for modeling topical trends in documents. Often the resulting topics are broad and generic, associating large groups of people and issues that are loosely related. In many cases, it may be desirable to influence the direction in which topic models develop. In this paper, we explore the idea of centering topics around people. In particular, given a large corpus of images featuring collections of people and associated captions, it seems natural to extract topics specifically focussed on each person. What words are most associated with George Bush? Which with Condoleezza Rice? Since people play such an important role in life, it is natural to anchor one topic to each person. In this paper, we present People-LDA, which uses the coherence efface images in news captions to guide the development of topics. In particular, we show how topics can be refined to be more closely related to a single person (like George Bush) rather than describing groups of people in a related area (like politics). To do this we introduce a new graphical model that tightly couples images and captions through a modern face recognizer. In addition to producing topics that are people specific (using images as a guiding force), the model also performs excellent soft clustering efface images, using the language model to boost performance. We present a variety of experiments comparing our method to recent developments in topic modeling and joint image-language modeling, showing that our model has lower perplexity for face identification than competing models and produces more refined topics.
  • Keywords
    face recognition; face identification; face recognition; joint image-language modeling; people-LDA; Coherence; Face recognition; Focusing; Graphical models; Image recognition; Refining;
  • fLanguage
    English
  • Publisher
    ieee
  • Conference_Titel
    Computer Vision, 2007. ICCV 2007. IEEE 11th International Conference on
  • Conference_Location
    Rio de Janeiro
  • ISSN
    1550-5499
  • Print_ISBN
    978-1-4244-1630-1
  • Electronic_ISBN
    1550-5499
  • Type

    conf

  • DOI
    10.1109/ICCV.2007.4409055
  • Filename
    4409055