COOKIES: By using this website you agree that we can place Google Analytics Cookies on your device for performance monitoring. |
University of Cambridge > Talks.cam > NLIP Seminar Series > Interpreting Document Collections Using Topic Models
Interpreting Document Collections Using Topic ModelsAdd to your list(s) Download to your calendar using vCal
If you have a question about this talk, please contact Tamara Polajnar. Topic models are a set of statistical methods for interpreting the contents of document collections. These models automatically learn sets of topics from words frequently co-occurring in documents. Topics learned often represent abstract thematic subjects, i.e Sports or Politics. Topics are also associated with relevant documents. These characteristics make topic models a useful tool for organising large digital libraries. Hence, these methods have been used to develop browsing systems allowing users to navigate through and identify relevant information in document collections by providing users with sets of topics that contain relevant documents. The aim of this talk is to present methods for post-processing the output of topic models, making them more comprehensible and useful to humans. First, we look at the problem of identifying incoherent topics. We show that our methods work better than previously proposed approaches. Next, we propose novel methods for efficiently identifying semantically related topics which can be used for topic recommendation. Finally, we look at the problem of alternative topic representations to topic keywords. We propose approaches that provide textual or image labels which assist in topic interpretability. We also compare different topic representations within a document browsing system. This talk is part of the NLIP Seminar Series series. This talk is included in these lists:
Note that ex-directory lists are not shown. |
Other listsGraduate Women's Network Office of Scholary Communication New Results in X-ray Astronomy 2009 Anglia Ruskin University - Community Engagement CCFMarine Seminars Sandars Lectures in BibliographyOther talksUncertainty Quantification with Multi-Level and Multi-Index methods CANCELLED in solidarity with strike action: Permanent Sovereignty over Natural Resources and the Unsettling of Mainstream Narratives of International Legal History My Life in Science Seminar “Publishing in Science: an Inside Look" Developing joint research between a UK university and and INGO on disability and education: opportunities and challenges Bayesian optimal design for Gaussian process model Bayesian optimal design for Gaussian process model |