12.07.2015 Views

Jolliffe I. Principal Component Analysis (2ed., Springer, 2002)(518s)

Jolliffe I. Principal Component Analysis (2ed., Springer, 2002)(518s)

Jolliffe I. Principal Component Analysis (2ed., Springer, 2002)(518s)

SHOW MORE
SHOW LESS
  • No tags were found...

You also want an ePaper? Increase the reach of your titles

YUMPU automatically turns print PDFs into web optimized ePapers that Google loves.

11.4. Physical Interpretation of <strong>Principal</strong> <strong>Component</strong>s 297In practice, there may be a circularity of argument when PCA is usedto search for physically meaningful modes in atmospheric data. The formof these modes is often assumed known and PCA is used in an attempt toconfirm them. When it fails to do so, it is ‘accused’ of being inadequate.However, it has very clear objectives, namely finding uncorrelated derivedvariables that in succession maximize variance. If the physical modes arenot expected to maximize variance and/or to be uncorrelated, PCA shouldnot be used to look for them in the first place.One of the reasons why PCA ‘fails’ to find the expected physical modesin some cases in atmospheric science is because of its dependence on thesize and shape of the spatial domain over which observations are taken.Buell (1975) considers various spatial correlation functions, both circular(isotropic) and directional (anisotropic), together with square, triangularand rectangular spatial domains. The resulting EOFs depend on the size ofthe domain in the sense that in a small domain with positive correlationsbetween all points within it, the first EOF is certain to have all its elementsof the same sign, with the largest absolute values near the centre of thedomain, a sort of ‘overall size’ component (see Section 13.2). For largerdomains, there may be negative as well as positive correlations, so that thefirst EOF represents a more complex pattern. This gives the impressionof instability, because patterns that are present in the analysis of a largedomain are not necessarily reproduced when the PCA is restricted to asubregion of this domain.The shape of the domain also influences the form of the PCs. For example,if the first EOF has all its elements of the same sign, subsequentones must represent ‘contrasts,’ with a mixture of positive and negative values,in order to satisfy orthogonality constraints. If the spatial correlationis isotropic, the contrast represented by the second PC will be betweenregions with the greatest geographical separation, and hence will be determinedby the shape of the domain. Third and subsequent EOFs canalso be predicted for isotropic correlations, given the shape of the domain(see Buell (1975) for diagrams illustrating this). However, if the correlationis anisotropic and/or non-stationary within the domain, things are lesssimple. In any case, the form of the correlation function is important indetermining the PCs, and it is only when it takes particularly simple formsthat the nature of the PCs can be easily predicted from the size and shapeof the domain. PCA will often give useful information about the sourcesof maximum variance in a spatial data set, over and above that availablefrom knowledge of the size and shape of the spatial domain of the data.However, in interpreting PCs derived from spatial data, it should not beforgotten that the size and shape of the domain can have a strong influenceon the results. The degree of dependence of EOFs on domain shape hasbeen, like the use of rotation and the possibility of physical interpretation,a source of controversy in atmospheric science. For an entertaining andenlightening exchange of strong views on the importance of domain shape,

Hooray! Your file is uploaded and ready to be published.

Saved successfully!

Ooh no, something went wrong!