Meta-Analysis of Problem-Based Learning Effects
Meta-Analysis of Problem-Based Learning Effects
[Link]/locate/learninstruc
Abstract
This meta-analysis has two aims: (a) to address the main effects of problem based learning
on two categories of outcomes: knowledge and skills; and (b) to address potential moderators
of the effect of problem based learning. We selected 43 articles that met the criteria for
inclusion: empirical studies on problem based learning in tertiary education conducted in real-
life classrooms. The review reveals that there is a robust positive effect from PBL on the
skills of students. This is shown by the vote count, as well as by the combined effect size.
Also no single study reported negative effects. A tendency to negative results is discerned
when considering the effect of PBL on the knowledge of students. The combined effect size
is significantly negative. However, this result is strongly influenced by two studies and the
vote count does not reach a significant level. It is concluded that the combined effect size for
the effect on knowledge is non-robust. As possible moderators of PBL effects, methodological
factors, expertise-level of students, retention period and type of assessment method were inves-
tigated. This moderator analysis shows that both for knowledge- and skills-related outcomes
the expertise-level of the student is associated with the variation in effect sizes. Nevertheless,
the results for skills give a consistent positive picture. For knowledge-related outcomes the
results suggest that the differences encountered in the first and the second year disappear later
on. A last remarkable finding related to the retention period is that students in PBL gained
slightly less knowledge, but remember more of the acquired knowledge.
2003 Elsevier Science Ltd. All rights reserved.
∗
Corresponding author. Tel.: 32 16 325914.
E-mail address: [Link]@[Link] (F. Dochy).
0959-4752/03/$ - see front matter 2003 Elsevier Science Ltd. All rights reserved.
doi:10.1016/S0959-4752(02)00025-7
534 F. Dochy et al. / Learning and Instruction 13 (2003) 533–568
1. Introduction
3. Research questions
Two sets of research questions guided this meta-analysis. First, we addressed the
main effects of PBL on two broad categories of outcomes: knowledge and skills
(i.e., application of knowledge). Secondly, potential moderators of the effect of PBL
536 F. Dochy et al. / Learning and Instruction 13 (2003) 533–568
are addressed. A first category of moderators are design aspects of the reviewed
research. In the second category of moderators, we examined whether the effect of
PBL differs according to various levels of student expertise. Third, we looked more
closely at different types of assessment methods. Fourth, we investigated the influ-
ence of the insertion of a retention period.
4. Method
Before searching the literature for work pertaining to the effects of PBL, we
determined the criteria for inclusion in our analysis.
1. The work had to be empirical. Although non empirical literature and literature
reviews were selected as sources of relevant research, this literature was not
included in the analysis.
2. The characteristics of the learning environment had to fit the previously described
core model of PBL.
3. The dependent variables used in the study had to be an operationalization of the
knowledge and/or skills (i.e., knowledge application) of the students.
4. The subjects of study had to be students in tertiary education.
5. To maximize ecological validity, the study had to be conducted in a real-life
classroom or programmatic setting rather than under more controlled laboratory
conditions.
The review and integration of research literature begins with the identification of
the literature. Locating studies is the stage at which the most serious form of bias
enters a meta-analysis (Glass, McGaw, & Smith, 1981): “How one searches deter-
mines what one finds; and what one finds is the basis of the conclusions of one’s
integration” (Glass, 1976, p. 6).
The best protection against this source of bias is a thorough description of the
procedure used to locate the studies.
A first literature search was started in 1997. A wide variety of computerized data-
bases were screened: the Educational Resources Information Center (ERIC) cata-
logue, PsycLIT, ADION, LIBIS. Also, the Current Contents (for Social Sciences)
was searched. The following keywords were used: problem-solving, learning, prob-
lem-based learning, higher education, college(s), high school, research, and review.
The literature was selected based on reading the abstracts. This reading resulted in
the selection of 14 publications that met the above criteria. Next, we employed the
“snowball method” and reviewed the references in the selected articles for additional
works. Review articles and theoretical overviews were also gathered to check their
references. This method yielded 17 new studies. A second literature search, started
F. Dochy et al. / Learning and Instruction 13 (2003) 533–568 537
Using other literature reviews as a guide (Albanese & Mitchell, 1993; Dochy,
Segers & Buehl, 1999; Vernon & Blake, 1993), we defined the characteristics central
to our review and analyzed the articles we selected on the basis of these character-
istics. Specifically, the following information was recorded in tables:
As a result of a first analysis of the studies, it became clear that other variables
could also be of importance. The coding sheet was completed with the following
information:
1. the year of the study in which the assessment of the dependent variable was done;
2. if there was a retention period;
3. the name of the PBL institute.
With respect to the dependent variable, we must note that only the outcomes
related to knowledge and skills (i.e., knowledge application) were coded. Some stud-
ies have examined other effects of PBL, but those were not included in the analysis.
The dependent variable was used to distinguish tests that assess knowledge from
tests that assess knowledge application. The following operational definitions were
used in making this distinction. A knowledge test primarily measures the knowledge
of facts and the meaning of concepts and principles (Segers, 1997). This type of
knowledge is often defined as declarative knowledge (Dochy & Alexander, 1995).
A test that assesses skills (i.e. knowledge application) measures to what extent stu-
dents can apply their knowledge (Glaser, 1990). It is important to remark that there
is a continuum between knowledge and skills rather than a dichotomy. Some studies
treat both aspects. In coding those studies, both aspects were separated and categor-
ized under different headings.
Two condensed tables were created (one for knowledge and one for skills) that
contain potential critical characteristics. These tables are included in Appendices A
and B (legend in Appendix C). In the tables, the statistical values were, as much as
possible, summarized and reported as effect size (ES) and p-values.
538 F. Dochy et al. / Learning and Instruction 13 (2003) 533–568
[Link]. Metric for expressing effect sizes The metric that we used to estimate and
describe the effects of PBL on knowledge and skills was the standardized mean
difference (d-index) effect size. This metric is appropriate when the means of two
groups are being compared (Cooper, 1989; Glass, McGaw, & Smith, 1981). The d-
index expresses the distance between the two group means in terms of their common
standard deviation. This common standard deviation is calculated by using the stan-
dard deviation of the control group since it is not affected by the treatment.
F. Dochy et al. / Learning and Instruction 13 (2003) 533–568 539
[Link]. Combining effect sizes across studies Once an effect size had been calcu-
lated for each study or comparison, the effects testing the same hypothesis were
averaged. Unweighted and weighted procedures were used. In the unweighted pro-
cedure, each effect size was weighted equally in calculating the average effect. In
the weighted procedure, more weight is given to effect sizes with larger samples
(factor w=inverse of the variance), based on the assumption that the larger samples
more closely approximate actual effects (Cooper, 1989; Hedges & Olkin, 1985).
These weighted combined effect sizes were tested for statistical significance by cal-
culating the 95% confidence interval (Cooper, 1989).
[Link]. Analyzing variance in effect sizes across studies The last step was to
examine the variability of the effect sizes via a homogeneity analysis (Cooper, 1989;
Hedges & Olkin, 1985; Hunter & Schmidt, 1990). This can lead to a search for
potential moderators. So, we can gain insight into the factors that affect relationship
strengths even though these factors may have never been studied in a single experi-
ment (Cooper, 1989).
Homogeneity analysis compares the variance exhibited by a set of effect sizes
with the variance expected by sampling error. If the result of homogeneity analysis
suggests that the variance in a set of effect sizes can be attributed to sampling error
alone, one can assume the data represent a population of students (Hunter,
Schmidt, & Jackson, 1982).
To test whether a set of effect sizes is homogeneous, a Qt statistic (Chi-square
distribution, N-1 degrees of freedom) is computed. A statistically significant Qt sug-
gests the need for further grouping of the data. The between-groups statistic (Qb) is
used to test whether the average effect of the grouping is homogeneous. A statisti-
cally significant Qb indicates that the grouping factor contributes to the variance in
effect sizes, in other words, the grouping factor has a significant effect on the out-
come measure analyzed (Springer, Stanne, & Donovan, 1999).
5. Results
Forty-three studies met the inclusion criteria for the meta-analysis. Of the 43 stud-
ies, 33 (76.7%) presented data on knowledge effects and 25 (58.1%) reported data
540 F. Dochy et al. / Learning and Instruction 13 (2003) 533–568
The main effect of PBL on knowledge and skills is differentiated. The results of
the analysis is summarized in Table 1.
In general, the results of both the vote count and the combined effect size were
statistically significant. These results suggest that students in PBL are better in apply-
ing their knowledge (skills). None of the studies reported significant negative find-
ings.
However, Table 1 would indicate that PBL has a negative effect on the knowledge
base of the students, compared with the knowledge of students in a conventional
learning environment. The vote count shows a negative tendency with 14 studies
yielding a significant negative effect and only seven studies yielding a significant
positive effect. This negative effect becomes significant for the weighted combined
effect size. However, this significant negative result is mainly due to two outliers
(Eisenstaedt, Bary, & Glanz, 1990; Baca, Mennin, Kaufman, & Moore-West, 1990).
When these two studies are left aside, the combined effect sizes approaches zero
(unweighted ES=-0.051; weighted ES=⫺0.107, CI:+/⫺ 0.058).
Table 1
Main effects of PBL
a
Two-sided sign-test is significant at the 5% level.
b
All weighted effect sizes are statistically significant.
c
+/⫺ number of studies with a significance (at the 5% level) positive/negative finding.
d
the number of total nonindependent outcomes measured.
F. Dochy et al. / Learning and Instruction 13 (2003) 533–568 541
[Link]. Research design The studies included in the meta-analysis can all be cate-
gorized as quasi-experimental (cf., criteria for inclusion). Studies with a randomized
design deliver the most trustworthy data. Studies based on a comparison between
different institutes or between different tracks are less reliable because randomization
is not guaranteed. Some studies attempt to compensate for this shortcoming by con-
trolling (e.g., Antepohl & Herzig, 1997; Lewis & Tamblyn, 1987) or matching the
subjects (Anthepohl & Herzig, 1997; Baca, Mennin, Kaufman & Moore-West, 1990)
for substantial variables. Most problematic are those studies having a historical
design (Martenson, Eriksson, & Ingelman-Sundberg, 1985). Some studies comparing
the PBL-outcomes with national means were also included.
The results of the homogeneity analysis reported in Table 2 suggest no significant
variation in effect sizes for knowledge-related outcomes can be attributed to method-
related influences (Qb=7.261, p=0.063). However, the most reliable comparisons
(random) suggest that there is almost no negative effect on knowledge acquisition.
Contrary to the data concerning knowledge, the variation in effect sizes for skills
outcomes was associated with the methodological factor research design (Qb=7.177,
p=0.027). The weighted combined effect sizes of the designs “between institutes”
or “elective tracks” are higher than the combined effect size emanating from a histori-
cal-controlled research design.
Table 2
Research design as moderating variable
Knowledge 7.261
(p=0.063)
Between 0 3 2 ⫺0.242 ⫺0.049
(+/⫺0.152)ns
Random 3 3 4 ⫺1.277 ⫺0.085
(+/⫺0.187)ns
Historical 1 2 2 ⫺0.680 ⫺0.202
(+/⫺0.082)
Elective 3 6 10 ⫺0.722 ⫺0.283
(+/⫺0.112)
National 0 1
Skills 7.177
(p=0.027)
Between 4 0 4 +0.864 +0.360
(+/⫺0.137)
Elective 8 0a 10 +0.567 +0.317
(+/⫺0.103)
Historical 2 0 3 +0.685 +0.173
(+/⫺0.083)
a
Two-sided sign-test is significant at the 5% level.
b
Unless noted ns , all weighted effect sizes are statistically significant.
Table 3
Scope of Implementation as moderating variableb
a
Two-sided sign-test is significant at the 5% level.
b
All weighted effect sizes are statistically significant.
vs zero negative) and the combined effect size (ES=0.390) suggest a positive effect.
Students in the fourth year show a negative effect of PBL on knowledge: a negative
tendency in the vote-counting method and a negative combined effect size
(ES=⫺0.496). On the contrary, this negative effect is not found for students who
graduated.
These results suggest that the differences arising in the first and the second year
disappear if the reproduction of knowledge is assessed when the broader context
asks all the students to apply their knowledge (both in the conventional and the PBL
environment). The only exception is the results in the last year of the curriculum.
The effects of PBL on skills (i.e., application of knowledge), differentiated for
expertise-level of students give a rather consistent picture. On all levels, there is a
strong positive effect of PBL on the skills of the students.
Table 4
Expertise-level of students as moderating variable
Knowledge 125.845
(p=0.000)
1e year 1 1 3 ⫺0.205 ⫺0.153
(+/⫺0.186)ns
2e year 0 6a 12 ⫺1.489 ⫺0.315
(+/⫺0.067)
3e year 2 0 5 +0.338 +0.390
(+/⫺0.129)
4e year 0 1 2 ⫺1.009 ⫺0.138
(+/⫺0.199)ns
5e year 0 0 1 ⫺0.037 ⫺0.037
(+/⫺0.233)ns
Last year 0 4a 3 ⫺0.523 ⫺0.496
(+/⫺0.166)
All 0 1 1 ⫺0.919 ⫺0.919
(+/⫺0.467)
Graduated 2 0 4 +0.193 +0.174
(+/⫺0.204)ns
Skills 20.630
(p=0.009)
1e year 1 0 2 +0.414 +0.433
(+/⫺0.340)
2e year 1 0 4 +0.473 +0.318
(+/⫺0.325)ns
3e year 4 0a 11 +0.280 +0.183
(+/⫺0.093)
4e year 1 0 1 +0.238 +0.235
(+/⫺0.512)ns
5e year 1 0 1 +0.732 +0.722
(+/⫺0.536)
Last year 4 0a 3 +0.679 +0.444
(+/⫺0.174)
All 0 0 1 +0.310 +0.310
(+/⫺0.161)
Graduated 1 0 1 +1.193 +1.271
(+/⫺0.630)
a
Two-sided sign-test is significant at the 5% level.
b
Unless noted ns , all weighted effect sizes are statistically significant.
F. Dochy et al. / Learning and Instruction 13 (2003) 533–568 545
Table 5
Retention period as moderating variable
Knowledge 28.683
(p=0.000)
Retention 4 2 9 +0.003 +0.139
(+/⫺0.116)
No Retention 3 13a 24 ⫺0.826 ⫺0.209
(+/⫺0.053)
Skills 1.474
(p=0.223)
Retention 3 0 5 +0.511 +0.320
(+/⫺0.198)
No Retention 11 0a 22 +0.500 +0.224
(+/⫺0.057)
a
Two-sided sign-test is significant at the 5% level.
b
All weighted effect sizes are statistically significant.
not know as many facts), their knowledge has been elaborated more and consequently
they have better recall of that knowledge.
For tests assessing skills, the results suggest that no significant variation in effect
sizes can be attributed to the presence or absence of a retention period. The positive
effect of PBL on the skills (knowledge application) of students seems to be immedi-
ately and lasting.
If the “Key feature approach” is used, questions are asked on only the core aspects
of the case (Bordage, 1987).
The results of the statistical meta-analysis are presented in Table 6. In this analysis,
we did not use the “shifting units” method to identify the independent hypothesis
tests, but the “samples as units” method (Cooper, 1989). This approach permits a
single study to contribute more than one hypothesis test, if the hypothesis test is
carried out on separate samples of people. In this way it was possible to gain more
information on certain operationalizations of the dependent variable.
The results of the homogeneity analysis (Table 6) suggest that significant variation
in effect sizes as well as effects on knowledge (Qb=254.501, p=0.000) and skills
(Qb=25.039, p=0.001) can be attributed to the specific operationalization of the
dependent variable.
The results in the domain of the effects on skills are more coherent than the results
for knowledge. The effects found with the different operationalizations of skills are
all positive. A ranking of the operationalization based on the size of the weighted
combined effect sizes, gives the following:
NBME Step II (0.080); Essay (0.165); NBME III (0.263); Oral (0.366); Simulation
(0.413); Case(s) (0.416); Rating (0.431); MEQ (0.476).
If this classification is compared with a continuum showing to what degree the
tests assess the application of knowledge, rather than the reproduction of knowledge,
the following picture emerges: the better an instrument is capable of evaluating stu-
F. Dochy et al. / Learning and Instruction 13 (2003) 533–568 547
Table 6
Type of assessment method as moderating variable
Knowledge 254.501
(p=0.000)
NBME part I 0 6a 5 ⫺1.740 ⫺0.961
(+/⫺0.152)
Short-answer 2 1 3 +0.050 ⫺0.123
(+/⫺0.080)
MCQ 3 7 12 ⫺1.138 ⫺0.309
(+/⫺0.109)
Rating 1 0 4 +0.209 ⫺0.301
(+/⫺0.162)
Oral 0 0 2 ⫺0.334 ⫺0.350
(+/⫺0.552)ns
Progress 0 1 6 +0.011 ⫺0.005
(+/⫺0.097)ns
Free recall 1 0 1 +2.171 +2.171
(+/⫺0.457)
Skills 25.039
(p=0.001)
NBME part II 1 0 4 +0.094 +0.080
(+/⫺0.125)ns
NBME part III 1 0 2 +0.265 +0.263
(+/⫺0.153)
Case(s) 5 0a 11 +0.708 +0.416
(+/⫺0.119)
MEQ 1 0 1 +0.476 +0.476
(+/⫺0.321)
Simulation 1 0 2 +0.854 +0.413
(+/⫺0.311)
Oral 1 0 2 +0.349 +0.366
(+/⫺0.554)ns
Essay 2 0 3 +0.415 +0.165
(+/⫺0.083)
Rating 2 0 3 +0.387 +0.431
(+/⫺0.182)
a
Two-sided sign-test is significant at the 5% level.
b
Unless noted ns, all weighted effect sizes are statistically significant.
dents’ skills (i.e., application of knowledge), the larger the ascertained effects of
PBL (compared with a conventional learning environment).
Effects found with the NBME Step II are negligible. However, the NMBE Step
II is also the least suitable instrument to examine the skills of the students. It assesses
548 F. Dochy et al. / Learning and Instruction 13 (2003) 533–568
clinical knowledge rather than clinical performance (Vernon & Blake, 1993). The
essay questions give some opportunities to evaluate the integration of knowledge
(Swanson, Case, & van der Vleuten, 1997), but they are not often used to make an
application of knowledge. This is also the case for oral examination. In this case,
however, there was a clear distinction between questions that examined knowledge
and questions that assessed problem-solving skills (Goodman et al., 1991). Only the
latter are categorized as skills-related outcomes.
Excepting the NBME Step II, essay questions, and the oral examination of NBME
Step III, all the other instruments can be classified as measuring the skills of the
students to apply their knowledge in an authentic situation. On these tests, the stu-
dents in PBL score consistently higher (ES between 0.416 and 0.476). The only
exception is the results on the NBME step III (ES=0.265). It should be noted that
the exam consists only partially of authentic cases.
A rating was made for the knowledge-related outcomes:
NBME I (⫺0.961); Oral (⫺0.350); MCQ (⫺0.309); Rating (⫺0.301); Short-
answer (⫺0.123); Progress test (⫺0.005); Free recall (+2.171)
The results suggest a similar conclusion as the result presented in the Retention
period section. If the test makes a strong appeal to retrieval strategies, students in
PBL do at least as well as the students in a conventional learning environment. A
rating context, short-answer questions, or free recall tests make a stronger appeal to
retrieval strategies than a recognition task (NBME step I and MCQ) (Tans, Schmidt,
Schade-Hoogeveen, & Gijselaers, 1986). Also the progress test examines “rooted”
knowledge. The fact that students in a conventional learning environment score better
on the NBME step I and on the MCQ (see vote count), suggests that they have more
knowledge. The fact that the difference between students in conventional learning
environments and students in PBL diminishes or even disappears on a test appealing
to retrieval strategies, suggests a better organization of the students’ knowledge in
PBL. However, this conclusion is rather tentative.
6. Conclusion
The first research question in this meta-analysis dealt with the influence of PBL
on the acquisition of knowledge and the skills to apply that knowledge. The vote
count as well as the combined effect size (ES=0.460) suggest a robust positive effect
from PBL on the skills of students. Also no single study reported negative effects.
A tendency to negative results is discerned when considering the effect of PBL
on the knowledge of students. The combined effect size is significantly negative
(ES=⫺0.223). However, this result is strongly influenced by two studies. Also the
vote count does not reach a significant level.
Evaluating the practical significance of the effects requires additional interpret-
ation. Researchers in education and other fields continue to discuss how to evaluate
the practical significance of an effect size (Springer, Stanne & Donovan, 1999).
F. Dochy et al. / Learning and Instruction 13 (2003) 533–568 549
Cohen (1988) and Kirk (1996) recommend that d=0.20 (small effect), d=0.50
(moderate effect) and d=0.80 (large effect) serve as general guidelines across disci-
plines. Within education, conventional measures of the practical significance of an
effect size range from 0.25 (Tallmadge, 1977 in Springer, Stanne & Donovan, 1999)
to 0.50 (Rossi & Wright, 1977). Many education researchers (Gall, Borg, & Gall,
1996) consider an effect size of 0.33 as the minimum to establish practical signifi-
cance.
If we compare the main effects with these guidelines, it can be concluded that
the combined effect size for skills is moderate, but of practical significance. The
effect on knowledge, already described as non-robust, is also small and not practi-
cally significant.
that the more an instrument is capable of evaluating the skills of the student, the
larger the ascertained effect of PBL.
Although it is not so clear, an analogue tendency is acknowledged for the knowl-
edge-related outcomes. Students do better on a test if the test makes a stronger appeal
on retrieval strategies. This could be due to a better structured knowledge base, a
consequence of the attention for knowledge elaboration in PBL. This is in line with
the conclusion presented previously in the Retention period.
The interest in the effects of PBL has already produced two good and often cited
reviews (Albenese & Mitchell, 1993; Vernon & Blake, 1993). These reviews were
published in a short period and mostly rely on the same literature. The two reviews
used a different methodology. Albanese and Mitchell relied on a narrative integration
of the literature, while Vernon and Blake used statistical methods. Methodologically,
this analysis is more similar to Vernon and Blake. Both reviews, however, concluded
that at that moment there was not enough research to draw reliable conclusions.
The main results of this meta-analysis are similar to the conclusions of the two
reviews. They had found a robust positive effect of PBL on skills. Vernon and Blake
(1993, p. 560) express it as follows:
“Our analysis suggests that the clinical performance and skills of students exposed
to PBL are superior to those of students educated in a traditional curriculum.”
The reviews also drew similar conclusions about the effects of PBL on the knowl-
edge base of students. Albanese and Mitchell (1993, p.57) concluded very carefully:
“While the expectation that pbl students not do as well as conventional students
on basic science tests appears to be generally true, it is not always true.”
Vernon and Blake (1993) specified this doubt with their statistical meta-analysis:
And
This meta-analysis also made similar conclusions about the effect of PBL on
knowledge and provides a further validation of the findings from the two mentioned
F. Dochy et al. / Learning and Instruction 13 (2003) 533–568 551
Acknowledgements
The authors are grateful to Eduardo Cascallar at the University of New York and
Neville Bennett at the University of Exeter, UK for their comments on earlier drafts.
See Table 7
See Table 8
Study
First author and the year of publication.
Pbl-Insitute
Institute in which the experimental condition has taken place.
University of New Mexico; Universiteit van Maastricht; McMaster University;
University of NewCastle; Temple University; Michigan State University; Karolinska
Institutet; University of Kentucky; Rush Medical College; Mercer University; McGill
University; University of Rochester; Universiteit van Keulen; University of New
Brunswick; Michener Institute; University of Alberta; Harvard Medical School;
Wake Forest University; Southern Illinois University.
When no institute was mentioned, the institute was described.
Bv. ‘Institute for higher professional education’ (Tans, Schmidt, Schade-Hooge-
veen, & Gijselaers, 1986).
Level
Participants’ level.
1=first year
Table 7 552
Studies measuring knowledge
ES p-value
biochemistry + ns
pathology - ns
microbiology - p⬍0.05
Pharmacology - ns
Behavorial + ns
(continued on next page)
Table 7 (continued)
ES p-value
ES p-value
ES p-value
Son and Van Sickle, 2000 2 colleges ? S Non- 72/68 N Instrument: 0.381 p=0.05
economics equivalent
com-
parison
group
design
–16 MCQ
–8 correct/incorr
–1 short answer
Y idem 0.384 p=0.05
72/80
Aaron et al., 1998 Alberta 2 S H 113/121 N MCQ ⫺0.440 p⬍0.05
17/12 N MCQ ⫺0.769 p⬎0.05
Block and Moore, 1994 Harvard Medical 2 C R 62/63 N NBME part I = n.s.
School
–behavorial science + sign
subtest
Richards et al., 1996 Wake Forest 3 C K 88/364 Y Clinical rating scale 1) 0.5 p=0.0001
amount of factual
knowledge
F. Dochy et al. / Learning and Instruction 13 (2003) 533–568
Doucet et al., 1998 Dalhousie CME S K, C 34/29 N MCQ (40 items) 0.434 p=0.05
University,
Halifax
Distlehorst and Robbs, 1998 Southern Illinois 2 C K, C 47/154 N NBME part I 0.18 p=0.6528
Imbos et al., 1984 Maastricht A C I N Anatomy (progress test) Van jaar
1 tot 4:=
Jaar 5
en 6:+
555
Table 7 (continued) 556
ES p-value
ES p-value
Study Pbl-institute Level Scope Design Subj. pbl/conv Ret. Operat. AV Result
ES p-value
Study Pbl-institute Level Scope Design Subj. pbl/conv Ret. Operat. AV Result
ES p-value
Kaufman et al., 1989 New Mexico 3 C K Total: 120/318 N NBME II +0.224 p⬍0.01
in 1983 - ns
1984 - ns
1985 + p⬍0.1
1986 + p⬍0.1
1987 + ns
1988 + ns
clinical rotations / ns
in 1983 ⫺ ns
1984 ⫺ ns
1985 ⫺ ns
1986 + ns
1987 + ns
1988 + ns
1989 + ns
clinical subscores of clin. + sign.
rot.
in 1983 ⫺ ns
1984 + ns
1985 / ns
1986 + ns
F. Dochy et al. / Learning and Instruction 13 (2003) 533–568
1987 + p⬍0.1
1988 + ns
1989 + ns
Schwartz et al., 1997 Kentucky 3 C K N Standardized Patient + /
simulation
MEQ + /
NBME II / ns
surgery subsection + sign.
559
Table 8 (continued) 560
Study Pbl-institute Level Scope Design Subj. pbl/conv Ret. Operat. AV Result
ES p-value
Study Pbl-institute Level Scope Design Subj. pbl/conv Ret. Operat. AV Result
ES p-value
Richards et al., 1996 Wake Forest 3 C K, C 88/364 Y - Clinical rating scale +0.426
2) take history and +0.425 p=0.002
perform physical
3) derive differential +0.462 p=0.0005
diagnosis
4) organize and express +0.390 p=0.004
information
–NBME medicine shelf +0.073 p=0.80
test (~part II)
Doucet et al., 1998 Continuing CME S K 21/26 Y Key Feature Problem +1.293 p=0.001
Medical examination (28 cases)
Education
Distlehorst and Robbs, Southern 3 C (first K, C 47/154 Y –USMLE step II +0.390 p=0.0518
1998 Illinois 2 years)
–Rating +0.5 p=0.0028
–Standardized patient
simulations:
–Overall +0.3 p=0.0596
–post station encounter +0.33 p=0.0742
–patient checklist ratings +0.14 p=0.1669
Jones et al., 1984 Michigan 4 C (first K 60/142 Y FLEX weighted average + n.s.
F. Dochy et al. / Learning and Instruction 13 (2003) 533–568
State 2 years)
561
Table 8 (continued) 562
Study Pbl-institute Level Scope Design Subj. pbl/conv Ret. Operat. AV Result
ES p-value
2=second year
3=third year
4=fourth year
5=fifth year
L=last year
A=in every year
Jg=Just graduated
Scope
Scope of PBL-implementation
C=curriculum-wide
S=single course
Design
Subjects (subj.)
Number of subjects in the experimental condition (PBL) / number of subjects in
the control condition (conv)
Retention (Ret.)
Is there a retention period between treatment and test?
Y=Yes
N=No
Operationalization dependent variable (Operat. AV)
MCQ=Multiple choice question
Ratings, questionnaires
NBME (USMLE) I, II of III=National Board of Medical Examiners Part I, Part
II of Part III exam
Case
Standardized Patient (simulation)
Clinical Rotations
Oral
Essay questions
Key feature problem examination
First try pass rate
Progress testing
Result
Effect size (ES): The sign of the ES shows if the Pbl-result is greater (+) or
smaller (⫺).
564 F. Dochy et al. / Learning and Instruction 13 (2003) 533–568
If it was not possible to compute an ES, than only the sign of the results is given.
If there was no effect found than ‘/’ is indicated.
p-value: ns=not significant /=no p-value given.
References
Aaron, S., Crocket, J., Morrish, D., Basualdo, C., Kovithavongs, T., Mielke, B., & Cook, D. (1998).
Assessment of exam performance after change to problem-based learning: Differential effects by ques-
tion type. Teaching-and-Learning-in-Medicine, 10(2), 86–91.
Albanese, M. A., & Mitchell, S. (1993). Problem-based learning: A review of literature on its outcomes
and implementation issues. Academic Medicine, 68, 52–81.
Albano, M. G., Cavallo, F., Hoogenboom, R., Magni, F., Majoor, G., Mananti, F., Schuwirth, L., Stiegler,
I., & Van Der Vleuten, C. (1996). An international comparison of knowledge levels of medical stu-
dents: The Maastricht Progress Test. Medical Education, 30, 239–245.
Antepohl, W., & Herzig, S. (1997). Problem-based learning supplementing in the course of basic pharma-
cology-results and perspectives from two medical schools. Naunyn-Schmiedeberg’s Archives of Phar-
macology, 355, R18.
Antepohl, W., & Herzig, S. (1999). Problem-based learning versus lecture-based learning in a course of
basic pharmacology: A controlled, randomized study. Medical Education, 33(2), 106–113.
Ausubel, D., Novak, J., & Hanesian, H. (1978). Educational psychology: A cognitive view (2nd ed.). New
York: Holt, Rinehart & Winston.
Baca, E., Mennin, S. P., Kaufman, A., & Moore-West, M. (1990). Comparison between a problem-based,
community-oriented track and a traditional track within one medical school. In Z. M. Nooman, H. G.
Schmidt, & E. S. Ezzat (Eds.), Innovation in medical education: An evaluation of its present status
(pp. 9–26). New York: Springer.
Barrows, H. S. (1984). A Specific, problem-based, self-directed learning method designed to teach medical
problem-solving skills, self-learning skills and enhance knowledge retention and recall. In H. G.
Schmidt, & M. L. de Volder (Eds.), Tutorials in problem-based learning. A new direction in teaching
the health profession. Assen: Van Gorcum.
Barrows, H. S. (1996). Problem-based learning in medicine and beyond: a brief overview. In L. Wilker-
son, & W. H. Gijselaers (Eds.), New directions for teaching and learning, Nr.68 (pp. 3–11). San
Francisco: Jossey-Bass Publishers.
Barrows, H. S., & Tamblyn, R. M. (1976). An evaluation of problem-based learning in small groups
utilizing a simulated patient. Journal of Medical Education, 51, 52–54.
Baxter, G. P., & Shavelson, R. J. (1994). Science performance assessments: Benchmarks and surrogates.
International Journal of Educational Research, 21(3), 279–299.
Bickley, H., Donner, R. S., Walker, A. N., & Tift, J. P. (1990). Pathology education in a problem-based
medical curriculum. Teaching and Learning in Medicine, 2(1), 38–41.
Birenbaum, M. (1996). Assessment 2000: Towards a pluralistic approach to assessment. In M.
Birenbaum, & F. J. R. C. Dochy (Eds.), Alternatives in assessment of achievements, learning processes
and prior knowledge. Boston/Dordrecht/London: Kluwer Academic Publishers.
Birenbaum, M., & Dochy, F. (Eds.) (1996). Alternatives in assessments, learning processes and prior
knowledge. Boston: Kluwer Academic.
Block, S. D., & Moore, G. T. (1994). Project evaluation. In D. C. Tosteson, S. J. Adelstein, & S. T.
Carver (Eds.), New pathways to medical education: Learning to learn at Harvard Medical School.
Cambridge, MA: Harvard University Press.
Boekaerts, M. (1999a). Self-regulated learning: Where are we today? International Journal of Educational
Research, 31, 445–457.
Boekaerts, M. (1999b). Motivated learning: The study of student x situation transactional units. European
Journal of Psychology of Education, 14(4), 41–55.
Boradge, G. (1987). An alternative approach to PMP’s: The ‘key-features’ concept. In I. R. Hart, & R.
F. Dochy et al. / Learning and Instruction 13 (2003) 533–568 565
Harden (Eds.), Proceedings of the second ottawa conference. Further developments in assessing clini-
cal competences (pp. 59–75). Montreal: Can-Heal Publications Inc.
Boshuizen, H. P. A., Schmidt, H. G., & Wassamer, A. (1993). Curriculum style and the integration of
biomedical and clinical knowledge. In P. A. J. Bouhuys, H. G. Schmidt, & H. J. M. van Berkel (Eds.),
Problem-based learning as an educational strategy (pp. 33–41). Maastricht: Network Publications.
Bruner, J. S. (1959). Learning and thinking. Harvard Educational Review, 29, 184–192.
Bruner, J. S. (1961). The act of discovery. Harvard Educational Review, 31, 21–32.
Chi, M. T., Glaser, R., & Rees, E. (1982). Expertise in problem solving. In R. Sternberg (Ed.), Advances
in the psychology of human intelligence (pp. 7–76). Hillsdale, NJ: Erlbaum.
Cohen, J. (1988). Statistical power analysis for the behavioral sciences (2nd ed.). Hillsdale, NJ: Erlbaum.
Cooper, H. M. (1989). Integrating research. A guide for literature reviews (Applied Social Research
Methods Series, Vol. 2). London: Sage Publications.
De Corte, E. (1990a). A State-of-the-art of research on learning and teaching. Keynote lecture presented
at the first European Conference on the First Year Experience in Higher Education, Aalborg University,
Denmark, April 23-25.
De Corte, E. (1990b). Toward powerful learning environments for the acquisition of problem-solving
skills. European Journal of Psychology of Education, 5(1), 5–19.
De Corte, E. (1995). Fostering cognitive growth: A perspective from research on mathematics learning
and instruction. Educational Psychologist, 30(1), 37–46.
De Corte, E. (2000). Marrying theory building and the improvement of school practice: A permanent
challenge for instructional psychology. Learning & Instruction, 10(3), 249–266.
Dewey, J. (1910). How we think. Boston: Health & Co.
Dewey, J. (1944). Democracy and education. New York: Macmillan Publishing Co.
Distlehorst, L. H., & Robbs, R. S. (1998). A comparison of problem-based learning and standard curricu-
lum students: Three years of retrospective data. Teaching and Learning in Medicine, 10(3), 131–137.
Dochy, F., Segers, M., & Buehl, M. M. (1999). The relation between assessment practices and outcomes
of studies: The case of research on prior knowledge. Review of Educational Research, 69(2), 145–186.
Dochy, F. J. R. C., & Alexander, P. A. (1995). Mapping prior knowledge: A framework for discussion
among researchers. European Journal for Psychology of Education, X(3), 225–242.
Donner, R. S., & Bickley, H. (1990). Problem-based learning: An assessment of its feasibility and cost.
Human Pathology, 21, 881–885.
Doucet, M. D., Purdy, R. A., Kaufman, D. M., & Langille, D. B. (1998). Comparison of problem-based
learning and lecture format in continuing medical education on headache diagnosis and management.
Medical Education, 32(6), 590–596.
Eisenstaedt, R. S., Barry, W. E., & Glanz, K. (1990). Problem-based learning: Cognitive retention and
cohort traits of randomly selected participants and decliners. Academic Medicine, 65(9, September
suppl), 11–12.
Engel, C. E. (1997). Not just a method but a way of learning. In D. Bound, & G. Feletti (Eds.), The
challenge of problem based learning (2nd ed.) (pp. 17–27). London: Kogan Page.
Farquhar, L. J., Haf, J., & Kotabe, K. (1986). Effect of two preclinical curricula on NMBE part I examin-
ation performance. Journal of Medical Education, 61, 368–373.
Finch, P. M. (1999). The effect of problem-based learning on the academic performance of students
studying pediatric medicine in Ontario. Medical Education, 33(6), 411–417.
Gagné, E. D. (1978). Long-term retention of information following learning from prose. Review of Edu-
cational Research, 48, 629–665.
Gall, M. D., Borg, W. R., & Gall, J. P. (1996). Educational research (6th ed.). White Plains, NY: Long-
man.
Glaser, R. (1990). Toward new models for assessment. International Journal of Educational Research,
14, 475–483.
Glass, G. V. (1976). Primary, secondary and meta-analysis. Educational Researcher, 5, 3–8.
Glass, G. V., McGaw, B., & Smith, M. L. (1981). Meta-analysis in social research. London: Sage Publi-
cations.
Goodman, L. J., Brueschke, E. E., Bone, R. C., Rose, W. H., Williams, E. J., & Paul, H. A. (1991). An
566 F. Dochy et al. / Learning and Instruction 13 (2003) 533–568
experiment in medical education: A critical analysis using traditional criteria. Journal of the American
Medical Education, 265, 2373–2376.
Hedges, L. V., & Olkin, I. (1980). Vote counting methods in research synthesis. Psychological Bulletin,
88, 359–369.
Hedges, L. V., & Olkin, I. (1985). Statistical methods for meta-analysis. Orlando, FL: Academic Press.
Hmelo, C. E. (1998). Problem-based learning: Effects on the early acquisition of cognitive skill in medi-
cine. The Journal of the Learning Sciences, 7, 173–236.
Hmelo, C. E., Gotterer, G. S., & Bransford, J. D. (1997). A theory-driven approach to assessing the
cognitive effects of PBL. Instructional Science, 25, 387–408.
Honebein, P. C., Duffy, T. M., & Fishman, B. J. (1993). Constructivism and the design of learning
environments: Context and authentic activities for learning. In T. M. Duffy, J. Lowyck, & D. H.
Jonassen (Eds.), Designing environments for constructive learning. Berlin: Springer Verlag.
Hunter, J. E., & Schmidt, F. L. (1990). Methods of meta-analysis. Correcting error and bias in research
findings. California: Sage Publications.
Hunter, J. E., Schmidt, F. L., & Jackson, G. B. (1982). Meta-analysis: Cumulating research findings
across studies. Beverley Hills, CA: Sage.
Imbos, T., & Verwijnen, G. M. (1982). Voortgangstoetsing aan de medische faculteit Maastricht [Progress
testing on the Faculty of Medicine of Maastricht]. in Dutch In H. G. Schmidt (Ed.), Probleemgestuurd
Onderwijs: Bijdragen tot onderwijsresearchdagen 1981 (pp. 45–56). Stichting voor Onderzoek van
het Onderwijs.
Imbos, T., Drukker, J., van Mameren, H., & Verwijnen, M. (1984). The growth in knowledge of anatomy
in a problem-based curriculum. In H. G. Schmidt, & M. L. de Volder (Eds.), Tutorials in problem-
based learning. New direction in training for the health professions (pp. 106–115). Assen: Van Gor-
cum.
Jones, J. W., Bieber, L. L., Echt, R., Scheifley, V., & Ways, P. O. (1984). A Problem-based curriculum,
ten years of experience. In H. G. Scmidt, & M. L. de Volder (Eds.), Tutorials in problem-based
learning. New direction in training for the health professions (pp. 181–198). Assen: Van Gorcum.
Kaufman, A., Mennin, S., Waterman, R., Duban, S., Hansbarger, C., Silverblatt, H., Obenshain, S. S.,
Kantrowitz, M., Becker, T., Samet, J., & Wiese, W. (1989). The New Mexico experiment: Educational
innovation and institutional change. Academic Medicine, 64, 285–294.
Knox, J. D. E. (1989). What is… a modified essay question? Medical Teacher, 11(1), 51–55.
Kirk, R. E. (1996). Practical significance: A concept whose time has come. Educational and Psychological
Measurement, 56, 746–759.
Kulik, J. A., & Kulik, C. L. (1989). The concept of meta-analysis. International Journal of Educational
Research, 13(3), 227–234.
Lewis, K. E., & Tamblyn, R. M. (1987). The problem-based learning approach in Baccalaureate nursing
education: How effective is it? Nursing Papers, 19(2), 17–26.
Mandl, H., Gruber, H., & Renkl, A. (1996). Communities of practice toward expertise: Social foundation
of university instruction. In P. B. Bates, & U. M. Staudinger (Eds.), Interactive minds. Life-span
perspectives on the social foundation of cognition (pp. 394–412). Cambridge: Cambridge Univer-
sity Press.
Martenson, D., Eriksson, H., & Ingelman-Sundberg, M. (1985). Medical chemistry: Evaluation of active
and problem-oriented teaching methods. Medical Education, 19, 34–42.
Mennin, S. P., Friedman, M., Skipper, B., Kalishman, S., & Snyder, J. (1993). Performances on the
NMBE I, II, III by medical students in the problem-based learning and conventional tracks at the
University of New Mexico. Academic Medicine, 68, 616–624.
Mehrens, W. A., & Lehmann, I. J. (1991). Measurement and evaluation in education and psychology.
New York: Holt, Rinehard and Winston.
Moore, G. T., Block, S. D., Briggs-Style, C., & Mitchell, R. (1994). The influence of the New Pathway
curriculum on Harvard medical students. Academic Medicine, 69, 983–989.
Morgan, H. R. (1977). A problem-oriented independent studies programme in basic medical sciences.
Medical Education, 11, 394–398.
Neufeld, V., & Sibley, J. (1989). Evaluation of health sciences education programs: Program and student
assessment at McMaster University. In H. G. Scmidt, M. Lipkinjr, M. W. Vries, & J. M. Greep (Eds.),
F. Dochy et al. / Learning and Instruction 13 (2003) 533–568 567
New directions for medical education: Problem-based learning and community-oriented medical edu-
cation (pp. 165–179). New-York: Springer Verlag.
Neufeld, V. R., & Barrows, H. S. (1974). The ‘McMaster philosophy’: An approach to medical education.
Journal of Medical Education, 49, 1040–1050.
Neufeld, V. R., Woodward, C. A., & MacLeod, S. M. (1989). The McMaster M.D. Program: A case
study of renewal in medical education. Academic Medicine, 64, 423–432.
Nonaka, I., & Takeuchi, H. (1995). The knowledge-creating company. New York: Oxford University
Press.
Owen, E., Stephens, M., Moskowitz, J., & Guillermo, G. (2000). From ‘horse race’ to educational
improvement: The future of international educational assessments. In: INES (OECD), The INES com-
pendium (pp. 7-18), Tokyo, Japan. ([Link]
Patel, V. L., Groen, G. J., & Norman, G. R. (1991). Effects of conventional and problem-based medical
curricula on problem-solving. Academic Medicine, 66, 380–389.
Piaget, J. (1954). The construction of reality in the child. New York: Basic Books.
Poikela, E., & Poikela, S. (1997). Conceptions of learning and knowledge-impacts on the implementation
of problem-based learning. Zeitschrift fur Hochschuldidactik, 1, 8–21.
Quinn, J. B. (1992). Intelligent enterprise, a knowledge and service based paradigm for industry. New
York: The Free Press.
Richards, B. F., Ober, P., Cariaga-Lo, L., Camp, M. G., Philp, J., McFarlane, M., Rupp, R., & Zaccaro,
D. J. (1996). Rating of students’ performances in a third-year internal medicine clerkship: A compari-
son between problem-based and lecture-based curricula. Academic Medicine, 71(2), 187–189.
Rogers, C. R. (1969). Freedom to learn. Colombus, Ohio: Charles E. Merill Publishing Company.
Rossi, P., & Wright, S. (1977). Evaluation resarch: An assessment of theory, practice and politics. Evalu-
ation Quarterly, 1, 5–52.
Salganik, L. H., Rychen, D. S., Moser, U., & Konstant, J. W. (1999). Projects on competencies in the
OECD context: Analysis of theoretical and conceptual foundations. Neuchâtel: SFSO, OECD, ESSI.
Santos-Gomez, L., Kalishman, S., Rezler, A., Skipper, B., & Mennin, S. P. (1990). Residency performance
of graduates from a problem-based and a conventional curriculum. Medical Education, 24, 366–377.
Saunders, N. A., Mcintosh, J., Mcpherson, J., & Engel, C. E. (1990). A comparison between University
of Newcastle and University of Sydney final-year students: Knowledge and competence. In Z. H.
Nooman, H. G. Schmidt, & E. S. Ezzat (Eds.), Innovation in medical education: An evaluation of its
present status (pp. 50–54). New York: Springer.
Schmidt, H. G. (1990). Innovative and conventional curricula compared: What can be said about their
effects? In Z. H. Nooman, H. G. Schmidt, & E. S. Ezzat (Eds.), Innovation in medical education: An
evaluation of its present status (pp. 1–7). New York: Springer.
Schmidt, H. G., Machiels-Bongaerts, M., Hermans, H., ten Cate, T. J., Venekamp, R., & Boshuizen, H.
P. A. (1996). The development of diagnostic competence: Comparison of a problem-based, an inte-
grated, and a conventional medical curriculum. Academic Medicine, 71, 658–664.
Schwartz, R. W., Burgett, J. E., Blue, A. V., Donnelly, M. B., & Sloan, D. A. (1997). Problem-based
learning and performance-based testing: Effective alternatives for undergraduate surgical education
and assessment of student performance. Medical Teacher, 19, 19–23.
Segers, M. S. R. (1996). Assessment in a problem-based economics curriculum. In M. Birenbaum, & F.
Dochy (Eds.), Alternatives in assessment of achievements, learning processes and prior learning (pp.
201–226). Boston: Kluwer Academic Press.
Segers, M. S. R. (1997). An alternative for assessing problem-solving skills: The OverAll Test. Studies
in Educational Evaluation, 23(4), 373–398.
Segers, M., Dochy, F., & De Corte, E. (1999). Assessment practices and students’ knowledge profiles in
a problem-based curriculum. Learning Environments Research, 12(2), 191–213.
Shavelson, R. J., Gao, X., & Baxter, G. P. (1996). On the content validity of performance assessments:
Centrality of domain specification. In M. Birenbaum, & F. Dochy (Eds.), Alternatives in assessment of
achievements, learning processes and prior learning (pp. 131–143). Boston: Kluwer Academic Press.
Son, B., & Van Sickle, R. L. (2000). Problem-solving instruction and students’ acquisition, retention
and structuring of economics knowledge. Journal of Research and Development in Education, 33(2),
95–105.
568 F. Dochy et al. / Learning and Instruction 13 (2003) 533–568
Springer, L., Stanne, M. E., & Donovan, S. S. (1999). Effects of small-group learning on undergraduates
in science, mathematics, engineering, and technology: A meta-analysis. Review of Educational
Research, 69(1), 21–51.
Swanson, D. B., Case, S. M., & van der Vleuten, C. P. M. (1997). Strategies for student assessment. In
D. Boud, & G. Feletti (Eds.), The challenge of problem-based learning [2nd ed.] (pp. 269–282).
London: Kogan Page.
Tans, R. W., Schmidt, H. G., Schade-Hoogeveen, B. E. J., & Gijselaers, W. H. (1986). Sturing van het
onderwijsleerproces door middel van problemen: een veldexperiment [Guiding the learning process
by means of problems: A field experiment]. Tijdschrift voor Onderwijsresearch, 11, 35–46.
Tynjälä, P. (1999). Towards expert knowledge? A comparison between a constructivist and a traditional
learning environment in the University. International Journal of Educational Research, 33, 355–442.
Van Hessen, P. A. W., & Verwijnen, G. M. (1990). Does problem-based learning provide other knowl-
edge. In W. Bender, R. J. Hiemstra, A. J. J. A. Scherpbier, & R. P. Zwierstra (Eds.), Teaching and
assessing clinical competence (pp. 446–451). Groningen: Boekwerk Publications.
Van Ijzendoorn, M. H. (1997). Meta-analysis in early childhood education: Progress and problems. In B.
Spodek, A. D. Pellegrini, & O. N. Saracho (Eds.), Issues in early childhood education yearbook in
early childhood education. New York: Teachers College Press.
Verhoeven, B. H., Verwijnen, G. M., Scherpbier, A. J. J. A., Holdrinet, R. S. G., Oeseburg, B., Bulte,
J. A., & Van Der Vleuten, C. P. M. (1998). An analysis of progress test results of PBL and non-PBL
students. Medical Teacher, 20(4), 310–316.
Vernon, D. T. A., & Blake, R. L. (1993). Does problem-based learning work? A meta-analysis of evalu-
ative research. Academic Medicine, 68, 550–563.
Verwijnen, G. M., Pollemans, M. C., & Wijnen, W. H. F. W. (1995). Voortgangstoetsing [Progress
testing]. In J. C. M. Melz, A. J. J. A. Scherpbier, & C. P. M. Van der Vleuten (Eds.), Medisch
Onderwijs in de Praktijk (pp. 225–232). Assen: Van Gorcum.
Verwijnen, M., Imbos, T., Snellen, H., Stalenhoef, B., Sprooten, Y., & Van der Vleuten, C. (1982). The
evaluation system at the medical school of Maastricht. In H. G. Schmidt, M. Vries, & J. M. Greep
(Eds.), New directions for medical education: Problem-based learning and community-oriented medi-
cal education (pp. 165–179). New York: Springer.
Verwijnen, M., Van Der Vleuten, C., & Imbos, T. (1990). A comparison of an innovative medical school
with traditional schools: An analysis in the cognitive domain. In Z. H. Nooman, H. G. Schmidt, &
E. S. Ezzat (Eds.), Innovation in medical education: An evaluation of its present status (pp. 41–49).
New York: Springer.
Wittrock, M. C. (1989). Generative processes of comprehension. Educational Psychologist, 24, 345–376.