Tuesday, March 24, 2009
Distribution outside Belgium and the Netherlands
Until these problems are solved, readers outside Belgium and the Netherlands can order directly from the publisher, Johannes van Kessel Publishing at: Publishing@jvank.nl. They charge only local mailing costs (euro 6.20). Please contact them by email. To order via bol.com, van Stockum or Selexyz see: www.jvank.nl/ARMHome
The proceedings of the 2007 KNAW symposium `Advising on research methods' can also be ordered by contacting the publisher.
Herman
Proceedings of the KNAW colloquium on Advising now accessible via Google Books
Herman
Saturday, November 22, 2008
Two announcements
- November 11, 2008, the proceedings of the 2007 KNAW Colloquium `Advising on research methods' were published. The Colloquium was organized by Adèr and Mellenbergh and was funded by the Royal Netherlands Academy of Arts and Sciences (KNAW). The full reference is: Adèr & Mellenbergh (2008). Advising on research methods. Proceedings of the 2007 KNAW Colloquium. Huizen, the Netherlands: Johannes van Kessel (ISBN: 978-90-79418-03-9). The table of contents can be found at: http://www.knaw.nl/colloquia/advising/index.cfm . The commercial edition is sold at 40 euros. It can be ordered via http://www.bol.com/ , http://www.vanstockum.nl/ and http://www.selexyz.nl/ .
- We (Don Mellenbergh and myself) are organizing a public course on Advising, using the ARM-book (http://www.jvank.nl/ARMHome) as material. It will be held May 13-20, 2009 nearby the Amsterdam central station. Pre-registration is already possible. For more information, see: www.jvank.nl/ARMCourse
Herman Adèr
Friday, June 6, 2008
Several Announcements
First, a few announcements:
- Last February (2008), we (Don Mellenbergh, David Hand and myself) have published a (hand-/text-) book on methodological advising. The full reference is: Adèr, H. J., G. J. Mellenbergh and D. J. Hand (2008). Advising on research methods: A consultant's companion. ISBN 978907941801-5. Huizen: Johannes van Kessel. The book has its own webpage: www.jvank.nl/ARMHome . It is also accessible via Google books (http://www.books.google.com/ ; type: `Advising on research methods' or something of the kind).
- During a colloquium, also called `Advising on research methods', held in March 2007 in Amsterdam and sponsored by the Royal Netherlands Academy of Arts and Sciences (KNAW), a masterclass was held during which participants of the colloquium functioned both as advisors and as clients. This masterclass was recorded and is now available on DVD (free of charge). It represents unique material of excellent quality. To access the colloquium website, click: http://www.knaw.nl/colloquia/advising/index.cfm . To order the DVD click `DVD of the consultation interviews'.
- We are working on the proceedings of the above colloquium. If it becomes available, it will be announced here.
Herman Adèr.
Saturday, June 23, 2007
An alternative modelling procedure based on a strong Theory
The usual procedure would be to test whether the data D are consistent with models M1, M2 and/or M3.
But we could also go about as follows:
Generate data according to the models M1, M2 and M3, resulting in three data sets D1, D2 and D3 and test whether these data sets could have resulted from the same population as D.
Remarks:
- The above is only possible if we have a strong theory T on which we can base our models beforehand.
- An methodological advantage is that the researcher is forced to formulate his/her theoretical concepts and translate them into models before (s)he starts his or her experiment.
- A second advantage is that deviations between D and Di (i= 1, 2, 3) give information both on the relationships between variables and on the influence of the underlying (possibly multivariate) distributions (this is assuming that our models are based on known theoretical distributions like the normal distribution, which is common practice).
- A third advantage seems to be that we can directly test the alternative hypothesis.
- This procedure can not be combined with crossvalidation (randomly splitting the data in two parts, one part to find models consistent with the data, another part to test those models), because in the first part, models are formulated that are consistent with (possibly multivariate) distribution violations in the data: the same violations are present in the second part of the data, too.
Questions:
- Does a weak theory simply translates into a larger set of models?
- Simulating data based on models M1, M2 and M3 may not be trivial. Can we use similar procedures as are used in MCMC (Markov chain Monte Carlo) ?
- Can we use a Bayesian perspective, for instance by assuming that D1, D2 and D3 are based on prior distributions for the data D?
- Is the above approach known and described in the `simulation community'?
Monday, June 11, 2007
Paul de Boeck: Always do a PCA
Comments:
- PCA is an abbreviation of `Principal component analysis'. It is essentially a data reduction technique requiring no assumptions about the distribution of the variables. In a nutshell, the technique results in a representation of the data relative an orthogonal coordinate system. Data reduction is obtained by considering only a few axes of the coordinate system.
- Methodologically, the drawback of the technique lays in the orthogonality, which in most cases is not realistic in view of the substantive meaning of the data. To mend this, a promax rotation can be used which allows to obtain non-orthogonal axes.
- As an alternative to PCA, confirmatory factor analysis (CFA) can be used in an exploratory way, in particular, if some assumtions can be made about the relation between factors and items (note that CFA does have several assumptions on the distribution of the observed variables, notably multinormality.)
- Although Paul had to endure heavy critique on his fifthst rule, in my opinion he had a point. In fact, it is common practice among data analysts to use PCA as a quick and dirty technique to explore the data, even if they know how to apply CFA. If promax rotation is used instead of varimax rotation, some of the objections against the orthogonality assumption are mitigated, although not completely met: a CFA on the other half of the data using the factor structure found with PCA may result in completely different estimated angles between the factors.
- For those who heard Paul's talk, the recommendation to use PCA was not supprising: he did put heavy emphasis on data exploration as an antidote to the often theory-centered approach that prevails in social science and behavioral research, and PCA can very well used in an exploratory way. As holds for all exploratoration, the truth is never ascertained. The analysis has to be confirmed either on a another part of the data or by doing a new, carefully designed experiment, that allows for unequivocal confirmation.
- As a last remark, I want to stress the fact that the use of PCA is not so straightforward as it seems (in particular, if one wants to have some confidence in the results). For a recent article on PCA see Costello and Osborne (2005).
References
Costella, A. B. and J. W. Osborne (2005). Best Practices in Exploratory Factor Analysis: Four Recommendations for Getting the Most From YourAnalysis. Practical Assessment Research & Evaluation, 10 (7).
Monday, June 4, 2007
Paul de Boeck: rules during consultation
http://www.knaw.nl/colloquia/advising/index.cfm#proceedings
I give the rules below and will comment on some of them in this and the next few blogs:
- Not everything is worth being measured or can be measured, often the data are more interesting than the concept.
- Always reflect on which type of covariation is meant when the relationship between concepts is considered. All too often, automatically the covariation over persons is used as the basis, without good reasons.
- Measurement, reliability and validity testing, and hypothesis testing don’t need to be sequential steps, they can all be done simultaneously.
- So-called psychometric criteria are not theory-independent, and sometimes the theoretical implications of the psychometric criteria are wrong.
- Always do a PCA, it tells you about sources of differences in the data and about the interaction between the two modes of the data set.
- One does not necessarily have to care about the scale level of the data.
- Don’t construct indices of concepts, unless for descriptive summaries.
Ad rule 1 (Not everything is worth being measured or can be measured, often the data are more interesting than the concept). It should be stressed that this rule is thought to be most relevant during a consultation session. Let's take it apart: the first part (`Not everything is worth being measured or can be measured') is difficult to `sell' during consultation, because it means that during the study data were collected that were not worth collecting: this is particularly painful when it has to be said about the primary variables of an investigation. It is difficult to see what the second part (`often the data are more interesting than the concept') has to do with the first part: one can hardly say: `your study design started from a wrong idea and thus the data collection is worthless, but let's look at the data'. However (and this was clearly demonstrated during the presentation), a case can be made for a much looser connection between the data and the concepts to which they refer, because much can go wrong during implementation.
In particular (my addition): if enough data are available, a crossvalidation strategy can be useful, in which the data are randomly split in two parts. The first part is used for exploration and the emphasis is on the data (and their relation to study design and implementation), the second part is to investigate all worthwhile findings/hypotheses that came out of the exploration phase. Of course, in many cases not enough data were collected to allow this strategy. In this case two other strategies are available: (1) One may use the expected crossvalidation index (ECVI) given in Kaplan (page 117 e.v) which gives an impression of the crossvalidation adequacy of a model. (2) One may split the sample in unequal parts, using the first part for exploration as before and the (smaller) second part to test the findings of the exploration as before but now using small sample techniques like bootstrapping, if needed.
Kaplan, D. (2000). Structural Equation Modeling. Foundations and Extensions. Thousand Oaks London New Delhi: Sage Publications.