COOKIES: By using this website you agree that we can place Google Analytics Cookies on your device for performance monitoring. |
University of Cambridge > Talks.cam > MRC Biostatistics Unit Seminars > Virtual Seminar: 'Bayesian pyramids: Identifying interpretable deep structure underlying high-dimensional data’
Virtual Seminar: 'Bayesian pyramids: Identifying interpretable deep structure underlying high-dimensional data’Add to your list(s) Download to your calendar using vCal
If you have a question about this talk, please contact Alison Quenault. If you would like to join this virtual seminar, please email alison.quenault@mrc-bsu.cam.ac.uk for more information. High dimensional categorical data are routinely collected in biomedical and social sciences. It is of great importance to build interpretable models that perform dimension reduction and uncover meaningful latent structures from such discrete data. Identifiability is a fundamental requirement for valid modeling and inference in such scenarios, yet is challenging to address when there are complex latent structures. We propose a class of interpretable discrete latent structure models for discrete data and develop a general identifiability theory. Our theory is applicable to various types of latent structures, ranging from a single latent variable to deep layers of latent variables organized in a sparse graph (termed a Bayesian pyramid). The proposed identifiability conditions can ensure Bayesian posterior consistency under suitable priors. As an illustration, we consider the two-latent-layer model and propose a Bayesian shrinkage estimation approach. Simulation results for this model corroborate identifiability and estimability of the model parameters. Applications of the methodology to DNA nucleotide sequence data uncover discrete latent features that are both interpretable and highly predictive of sequence types. The proposed framework provides a recipe for interpretable unsupervised learning of discrete data, and can be a useful alternative to popular machine learning methods. Joint work with Yuqi Gu This talk is part of the MRC Biostatistics Unit Seminars series. This talk is included in these lists:
Note that ex-directory lists are not shown. |
Other listsCambridge eScience Centre The Cambridge Trust for New Thinking in Economics Institute of Astronomy Extra TalksOther talksWKB, Eigenvalue Problems and Quantisation in QM Welcome back meeting - plant display and social evening Soliton resolution on wormholes Spatial and temporal variability of the Antarctic Slope Current in an eddying ocean-sea ice model Finite element modelling of hot compression testing of titanium alloys Sodium channel complexes and cardiac arrhythmia |