Global ETD Search

Return to search

Effect fusion using model-based clustering

In social and economic studies many of the collected variables are measured on a nominal scale, often with a large number of categories. The definition of categories can be ambiguous and different classification schemes using either a finer or a coarser grid are possible. Categorization has an impact when such a variable is included as covariate in a regression model: a too fine grid will result in imprecise estimates of the corresponding effects, whereas with a too coarse grid important effects will be missed, resulting in biased effect estimates and poor predictive performance.

To achieve an automatic grouping of the levels of a categorical covariate with essentially the same effect, we adopt a Bayesian approach and specify the prior on the level effects as a location mixture of spiky Normal components. Model-based clustering of the effects during MCMC sampling allows to simultaneously detect categories which have essentially the same effect size and identify variables with no effect at all. Fusion of level effects is induced by a prior on the mixture weights which encourages empty components. The properties of this approach are investigated in simulation studies. Finally, the method is applied to analyse effects of high-dimensional categorical predictors on income in Austria.

Identifer	oai:union.ndltd.org:VIENNA/oai:epub.wu-wien.ac.at:6064
Date	01 April 2018
Creators	Malsiner-Walli, Gertraud, Pauger, Daniela, Wagner, Helga
Publisher	Sage
Source Sets	Wirtschaftsuniversität Wien
Language	English
Detected Language	English
Type	Article, PeerReviewed
Format	application/pdf
Rights	Creative Commons: Attribution 4.0 International (CC BY 4.0)
Relation	https://doi.org/10.1177/1471082X17739058, http://journals.sagepub.com/, http://epub.wu.ac.at/6064/

Page generated in 0.0023 seconds

Effect fusion using model-based clustering

Description

Links & Downloads

Tags

Additional Fields