Pitkin, J;
Ross, G;
Manolopoulou, I;
(2019)
Dirichlet process mixtures of order statistics with applications to retail analytics.
Journal of the Royal Statistical Society: Series C (Applied Statistics)
, 68
(1)
pp. 3-28.
10.1111/rssc.12296.
Preview |
Text
Manolopoulou VoR Pitkin_et_al-2019-Journal_of_the_Royal_Statistical_Society__Series_C_(Applied_Statistics).pdf - Published Version Download (9MB) | Preview |
Abstract
The rise of ‘big data’ has led to the frequent need to process and store data sets containing large numbers of high dimensional observations. Because of storage restrictions, these observations might be recorded in a lossy‐but‐sparse manner, with information collapsed onto a few entries which are considered important. This results in informative missingness in the observed data. Our motivating application comes from retail analytics, where the behaviour of product sales is summarized by the price elasticity of each product with respect to a small number of its top competitors. The resulting data are vectors of order statistics, because only the top few entries are observed. Interest lies in characterizing the behaviour of a product's competitors, and clustering products based on how their competition is spread across the market. We develop non‐parametric Bayesian methodology for modelling vectors of order statistics that utilizes a Dirichlet process mixture model with an exponentiated Weibull kernel. Our approach allows us added flexibility for the distribution of each vector, while providing parameters that characterize the decay of the leading entries. We implement our methods on a retail analytics data set of the cross‐elasticity coefficients, and our analysis reveals distinct types of behaviour across the different products of interest.
Type: | Article |
---|---|
Title: | Dirichlet process mixtures of order statistics with applications to retail analytics |
Open access status: | An open access version is available from UCL Discovery |
DOI: | 10.1111/rssc.12296 |
Publisher version: | https://doi.org/10.1111/rssc.12296 |
Language: | English |
Additional information: | © 2018 The Authors Journal of the Royal Statistical Society: Series C (Applied Statistics) Published by John Wiley & Sons Ltd on behalf of the Royal Statistical Society. This is an open access article under the terms of the Creative Commons Attribution License (https://creativecommons.org/licenses/by/4.0/). |
Keywords: | Bayesian non‐parametrics, Censoring, Cross‐elasticity |
UCL classification: | UCL UCL > Provost and Vice Provost Offices UCL > Provost and Vice Provost Offices > UCL BEAMS UCL > Provost and Vice Provost Offices > UCL BEAMS > Faculty of Maths and Physical Sciences UCL > Provost and Vice Provost Offices > UCL BEAMS > Faculty of Maths and Physical Sciences > Dept of Statistical Science |
URI: | https://discovery-pp.ucl.ac.uk/id/eprint/10050994 |
Archive Staff Only
View Item |