Paper #1857
- Título:
- The chi-square standardization, combined with Box-Cox transformation, is a valid alternative to transforming to logratios in compositional data analysis
- Autor:
- Michael Greenacre
- Fecha:
- Enero 2023
- Resumen:
- The approach to analysing compositional data with a fixed sum constraint has been dominated by the use of logratio transformations, to ensure exact subcompositional coherence and, in some situations, exact isometry as well. A problem with this approach is that data zeros, found in most applications, have to be replaced to permit the logarithmic transformation. A simpler approach is to use the chi-square standardization that is inherent in correspondence analysis. Combined with the Box-Cox power transformation, this standardization defines chi-square distances that tend to logratio distances for strictly positive data as the power parameter tends to zero, and can thus be considered equivalent to transforming to logratios. For data with zeros, a value of the power can be identified that brings the chi-square standardization as close as possible to transforming by logratios, without having to substitute the zeros. Especially in the field of high-dimensional "omics" data, this alternative presents such a high level of coherence and isometry as to be a valid, and much simpler, approach to the analysis of compositional data.
- Palabras clave:
- Box-Cox transformation, chi-square distance, correspondence analysis, isometry, logratios, Procrustes analysis, subcompositional coherence
- Códigos JEL:
- C19, C88
- Área de investigación:
- Estadística, Econometría y Métodos Cuantitativos
Descargar el paper en formato PDF