Detecting clusters in multivariate response regression. (3rd February 2021)
- Record Type:
- Journal Article
- Title:
- Detecting clusters in multivariate response regression. (3rd February 2021)
- Main Title:
- Detecting clusters in multivariate response regression
- Authors:
- Price, Bradley S.
Allenbrand, Corban
Sherwood, Ben - Abstract:
- Abstract: Multivariate regression, which can also be posed as a multitask machine learning problem, is used to better understand multiple outputs based on a given set of inputs. Many methods have been proposed on how to utilize shared information about responses with applications in fields such as economics, genomics, advanced manufacturing, and precision medicine. Interest in these areas coupled with the rise of large data sets ("big data") has generated interest in how to make the computations more efficient, but also to develop methods that account for the heterogeneity that may exist between responses. One way to exploit this heterogeneity between responses is to use methods that detect groups, also called clusters, of related responses. These methods provide a framework that can increase computational speed and account for complexity of relationships of a large number of responses. With this flexibility, comes additional challenges such as how to identify these clusters of responses, model selection, and the development of more complex algorithms that combine concepts from both the supervised and unsupervised learning literature. We explore current state of the art methods, present a framework to better understand methods that utilize or detect clusters of responses, and provide insights on the computational challenges associated with this framework. Specifically we present a simulation study that discusses the challenges with model selection when detecting clusters ofAbstract: Multivariate regression, which can also be posed as a multitask machine learning problem, is used to better understand multiple outputs based on a given set of inputs. Many methods have been proposed on how to utilize shared information about responses with applications in fields such as economics, genomics, advanced manufacturing, and precision medicine. Interest in these areas coupled with the rise of large data sets ("big data") has generated interest in how to make the computations more efficient, but also to develop methods that account for the heterogeneity that may exist between responses. One way to exploit this heterogeneity between responses is to use methods that detect groups, also called clusters, of related responses. These methods provide a framework that can increase computational speed and account for complexity of relationships of a large number of responses. With this flexibility, comes additional challenges such as how to identify these clusters of responses, model selection, and the development of more complex algorithms that combine concepts from both the supervised and unsupervised learning literature. We explore current state of the art methods, present a framework to better understand methods that utilize or detect clusters of responses, and provide insights on the computational challenges associated with this framework. Specifically we present a simulation study that discusses the challenges with model selection when detecting clusters of responses of interest. We also comment on extensions and open problems that are of interest to both the research and practitioner communities. This article is categorized under: Statistical Learning and Exploratory Methods of the Data Sciences > Clustering and Classification Statistical Learning and Exploratory Methods of the Data Sciences > Exploratory Data Analysis Statistical Learning and Exploratory Methods of the Data Sciences > Modeling Methods Abstract : A visual representation of the different options for detecting clusters in responses when using multivariate regression. … (more)
- Is Part Of:
- Wiley interdisciplinary reviews. Volume 14:Number 3(2022)
- Journal:
- Wiley interdisciplinary reviews
- Issue:
- Volume 14:Number 3(2022)
- Issue Display:
- Volume 14, Issue 3 (2022)
- Year:
- 2022
- Volume:
- 14
- Issue:
- 3
- Issue Sort Value:
- 2022-0014-0003-0000
- Page Start:
- n/a
- Page End:
- n/a
- Publication Date:
- 2021-02-03
- Subjects:
- Clustering -- Multivariate Regression -- Machine Learning -- Multi‐task Learning -- Optimization
Mathematical statistics -- Data processing -- Periodicals
Science -- Data processing -- Periodicals
Social sciences -- Data processing -- Periodicals
Mathematical statistics -- Periodicals
519.50285 - Journal URLs:
- http://onlinelibrary.wiley.com/journal/10.1002/(ISSN)1939-0068 ↗
http://www3.interscience.wiley.com/journal/122458798/home ↗
http://onlinelibrary.wiley.com/ ↗ - DOI:
- 10.1002/wics.1551 ↗
- Languages:
- English
- ISSNs:
- 1939-5108
- Deposit Type:
- Legaldeposit
- View Content:
- Available online (eLD content is only available in our Reading Rooms) ↗
- Physical Locations:
- British Library DSC - BLDSS-3PM
British Library HMNTS - ELD Digital store - Ingest File:
- 21485.xml