The findings capitalize on a well-known statistical phenomena that within-group correlations and between-group correlations are independent and that the total correlation is the combination of both types of correlations. Put differently, correlations between variables (within groups) are statistically independent from the mean differences between groups. The correlation between all variables across all groups can be predicted if the variances and covariances are known. To their credit they discuss this phenomenon and present some simulations in this direction (without acknowledging that these issues have been discussed for at least 30 years in methodological circles, but also fail to realize the importance of separating the within-group from the between-group effects).
What really concerns me is their interpretation of this pattern. They argue that cultural differences are valid and can not be reduced to individual level differences. Where is the problem here?
In their starting example, they refer to both IQ and personality. For both variables, reliable and stable instruments had been constructed first and subsequently, differences were also found between ethnic and cultural groups. Therefore, we have a clear idea of what is happening at the individual level and then need to figure how between-group differences can be understood (which is an on-going debate). Following on from this, they now proceed to two vaguely defined and heterogeneously measured domains, namely social orientation (defined as independence and interdependence) and cognitive style (broadly defined as analytic versus holistic). The cross-cultural literature is full of studies that show small to moderate differences between different societies (most typically US student samples versus East Asian student samples). Yet, there is no consensus on definitions, measurement or a clear understanding of what these variables really are. Indeed, variables are lumped together that are studied as different phenomena in different areas of psychology. For example, happiness is certainly not the same as self-construals or intensity of emotions. Attributions are not the same construct as thematic versus taxonomic categorization tasks. The tests also involve various different methods (which can be seen as strength or weakness depending on the viewpoint). A stroop task is often capturing different psychological processes compared to self-reports, as the heated debate on implicit versus explicit attitudes demonstrates. Looking at this array of tests, I would simply not expect a strong correlation. We are not dealing with a coherent psychological phenomenon.
The mean differences between the two groups could be due to a large number of variables. Some of the causal variables may be strongly related: education, opportunities in life, and various other variables related to wealth immediately spring to mind. In a famous quote, one of the pioneers on cognitive styles Herman Witkin explicitly argued that field independence (as measured by the FLT) is associated with formal education. Wealth is probably also the most important variable influencing social orientation, a fact that is widely known since the famous study by Hofstede published in 1980. Hence, nothing new here. But the problem is how wealth is influencing these variables and this most likely happening through different psychological processes. Hence, there is no reason to assume that these group differences are actually related rather than being probably influenced by a host of similar but distal variables.
In the latter parts of the paper, the authors then sink in a mess of tautological arguments. Because previous studies have shown these differences, then these differences need to be real and valid and hence, the individual and group level are distinct. Remember, this was what they wanted to test!
Yet, the authors also fail to notice the major shortcomings in many of the previous studies. For example, many of them did not test whether the instruments were equivalent or whether there were instrument, method, sampling or administration biases that could have influenced these results. Furthermore, these previous studies did not link potential explanatory variables to the observed differences, which is a standard practice in comparative psychological research (see for example http://pps.sagepub.com/content/1/3/234.abstract).
The current paper also fails to provide evidence on the equivalence of the measures. More importantly, it is not discussed how the two different groups are actually defined. What is social class and how were people grouped and on what criteria? The current study essentially is non-replicable.
This is a scientifically more productive endeavor than setting up straw arguments based on statistical artefacts. One only wonders how something like this could get into a prestigious journal like PNAS....