Finite-sample inference with monotone incomplete multivariate normal data, III: Hotelling’s T2-statistic

The DNA microarray technologies permit scientists to depict the expression of genes for related samples. This relationship between genes is analysed using Hotelling’s T2 as a multivariate test statistic but the disadvantage of this test, when used in microarray studies is the number of samples is larger than the number of variables. This study discovers the potential of the shrinkage approach to estimate the covariance matrix specifically when the high dimensionality problem happened. Consequently, the sample covariance matrix in Hotelling’s T2 statistic is not positive definite and become singular thus cannot be inverted. In this research, the Hotelling’s T2 statistic is combined with a shrinkage approach as an alternative estimation to estimate the covariance matrix to detect significant gene sets. The multivariate test statistic of classical Hotelling's T2 is used to integrate the correlation when assessing changes in activity level across biological conditions. The performances of the proposed methods were assessed using real data study. Shrinkage covariance matrix approach indicates a better result for detection of differentially expressed gene sets as compared to other methods.

Download Full-text

Finite-sample inference with monotone incomplete multivariate normal data, I

Journal of Multivariate Analysis ◽

10.1016/j.jmva.2009.05.003 ◽

2009 ◽

Vol 100 (9) ◽

pp. 1883-1899 ◽

Cited By ~ 21

Author(s):

Wan-Ying Chang ◽

Donald St.P. Richards

Keyword(s):

Multivariate Normal ◽

Finite Sample ◽

Finite Sample Inference

Download Full-text

On the use of the Hotelling's T2 statistic for the hierarchical clustering of hyperspectral data

2013 5th Workshop on Hyperspectral Image and Signal Processing: Evolution in Remote Sensing (WHISPERS) ◽

10.1109/whispers.2013.8080647 ◽

2013 ◽

Cited By ~ 1

Author(s):

M.A. Veganzones ◽

J. Frontera-Pons ◽

J. Chanussot ◽

J.P. Ovarlez

Keyword(s):

Hierarchical Clustering ◽

Hyperspectral Data ◽

Hotelling's T2 ◽

Hotelling’S T2 ◽

Hotelling’S T2 Statistic

Download Full-text

Robust Hotelling’s T2 statistic based on M-estimator

Journal of Physics Conference Series ◽

10.1088/1742-6596/1988/1/012116 ◽

2021 ◽

Vol 1988 (1) ◽

pp. 012116

Author(s):

Mohd Aizat Ahlam Mohamad Mokhtar ◽

Nur Syahidah Yusoff ◽

Chuan Zun Liang

Keyword(s):

Hotelling's T2 ◽

M Estimator ◽

Hotelling’S T2 ◽

Hotelling’S T2 Statistic

Download Full-text

Likelihood Based Finite Sample Inference for Singly Imputed Synthetic Data Under the Multivariate Normal and Multiple Linear Regression Models

Journal of Privacy and Confidentiality ◽

10.29012/jpc.v7i1.645 ◽

2015 ◽

Vol 7 (1) ◽

Cited By ~ 2

Author(s):

Martin Klein ◽

Bimal Sinha

Keyword(s):

Linear Regression ◽

Multiple Linear Regression ◽

Synthetic Data ◽

Multiple Linear Regression Model ◽

Original Data ◽

Multivariate Normal ◽

Finite Sample ◽

Unknown Parameters ◽

Finite Sample Inference ◽

Multiply Imputed

In this paper we develop likelihood-based finite sample inference based on singly imputed partially synthetic data, when the original data follow either a multivariate normal or a multiple linear regression model. We assume that the synthetic data are generated by using the plug-in sampling method, where unknown parameters in the data model are set equal to observed values of their point estimators based on the original data, and synthetic data are drawn from this estimated version of the model. Empirical studies are presented to show that the proposed methods do indeed perform as the theory predicts, and to compare the proposed methods for singly imputed synthetic data with the combining rules that are used to analyze multiply imputed partially synthetic data. Some theoretical comparisons between singly and multiply imputed partially synthetic data inference are also provided. A data analysis example and disclosure risk evaluation of singly and multiply imputed partially synthetic data is presented based on public use data from the Current Population Survey. We discuss the specific conditions under which the proposed methodology will yield valid inference, and evaluate the performance of the methodology when certain conditions do not hold. We outline some ways to extend the proposed methodology for certain scenarios where the required set of conditions do not hold.

Download Full-text