Understanding Building and Using Ontologies
Understanding Building and Using Ontologies
293–310, 1997
Nicola Guarino
LADSEB-CNR, National Research Council
Corso Stati Uniti 4, I-35127 Padova, Italy
guarino@[Link]
1. Introduction
In their paper on “Using Explicit Ontologies in KBS Development”, van
Heijst and colleagues1 seem to take for granted Bylander and Chan-
drasekaran’s hypothesis on the strong dependence of knowledge represesen-
tation on the nature and the inference strategy of the problem at hand, the so-
called interaction problem:
Representing knowledge for the purpose of solving some problem is strongly affected
by the nature of the problem and the inference strategy to be applied to the problem.
[Bylander and Chandrasekaran 1988]
The fact that the van Heijst and colleagues don’t attempt to explore in de-
tail the arguments sustaining this hypothesis is particularly puzzling, since
they admit that it contradicts one of the main assumptions of their well-known
KADS approach [Schreiber et al. 1993], namely the separation between do-
main knowledge and problem-solving knowledge. They report two reasons
brought by Bylander and Chandrasekaran to support their hypothesis:
“Firstly, the application task determines to a large extent which kinds of
knowledge should be encoded. (...) Secondly, the knowledge must be en-
coded in such a way that the inference strategy used can reason efficiently”.
In fact, at a closer inspection, the statement from Bylander and Chan-
drasekaran reported above mentions the problem of representing knowledge,
and it is related therefore to the symbol level. Now, it is certainly true that the
interaction problem exists at this level, but it seems plausible to assume that
its importance decreases at the knowledge level, to which the whole issue of
ontology belongs. Let us try for instance to re-state Bylander and Chan-
drasekaran’s statement at the knowledge level: “the knowledge required to
solve some problem is strongly affected by the nature of the problem...”. Put
in this way, it sounds even trivial. Notice however that this formulation
doesn’t refer to the way this knowledge is encoded, but simply to the rele-
vance relationship between the knowledge and the problem. In other words,
at the knowledge level the interaction problem reduces to the first of the two
“reasons” reported above. Of course, a specific piece of knowledge may be
1 In the following, I shall refer to this paper as “the paper”, and to its authors as “the
authors”.
more or less relevant for a particular task, but nothing tells us that this knowl-
edge is peculiar, specific of such task.
I will defend here the thesis of the independence of domain knowledge.
This thesis should not be intended in a rigid sense, since it is clear that – more
or less – ontological commitments always reflect particular points of view;
rather, what I would like to stress is the fact that reusability across multiple
tasks or methods should be systematically pursued even when modeling
knowledge related to a single task or method: the more this reusability is pur-
sued, the closer we get to the intrinsic, task-independent aspects of a given
piece of reality (at least, in the commonsense perception of a human agent).
In this systematic quest for reusability, the potential role of a discipline
like formal ontology appears evident. I have explored elsewhere [Guarino
1995] how a stronger connection between formal ontology, conceptual analy-
sis and knowledge engineering can contribute to establish the foundations of
the emerging field of “ontological engineering”. Following the lines of the
paper by van Heijst and colleagues, I shall discuss here how the principles of
ontological engineering can be used in the practice of knowledge-based sys-
tems building, focusing in particular on the interplay between ontologies and
problem-solving knowledge and on the ways to build and update ontologies.
I will first analyze in section 2 the various definitions of the term “ontology”
proposed by the authors, trying to make clear the problems bound to the for-
mal relationships between ontologies and conceptualizations. Then, in section
3, I will address the role of ontologies in the knowledge engineering process.
A crucial issue in this respect is the relationship between the ontology library
and the application ontology, and the role played by the latter in the update of
the former. The vision I will defend is that of application ontologies as spe-
cializations of a more general library, which includes task and method ontolo-
gies [Falasconi and Stefanelli 1994, Gennari et al. 1994] as well as domain
ontologies. The ontology-based knowledge modelling methodology proposed
by the authors will be discussed in detail in section 4, following their example
taken from the medical field. Finally, in the conclusions I will stress the role
of domain analysis, often absent in current methodological proposals where
the task analysis is strongly privileged.
2. Understanding Ontologies
2.1 Ontologies and Conceptualizations
Before discussing the principles for ontology libraries construction, the
authors report various definitions of the term "ontology" appeared in the lit-
erature, trying to establish a comprehensive definition. Together with Pier-
daniele Giaretta, I have analyzed this terminological problem in detail in
[Guarino and Giaretta 1995], focusing in particular on the possibility of giv-
ing a formal interpretation to the most cited definition of an ontology in the
knowledge sharing community, i.e. Gruber's definition:
a c
b d a d
c e b e
(a) (b)
2 In the original example also a function is considered, but we omit it here for the sake of
simplicity.
on the table (Fig. 1b). The corresponding structure would be different from
the previous one, generating therefore a different conceptualization. Of course
there is nothing wrong in such a view, if one is only interested in isolated
snapshots of the block world. But the meanings of the terms used to denote
the relevant relations are still the same, since they are invariant with respect to
the possible configurations of blocks. In fact, in the metalanguage adopted in
their book, Genesereth and Nilsson would adopt the same symbols (on,
above, clear, table) to denote the new conceptualization. We prefer to say in
this case that the states of affairs are different, but the conceptualization is the
same. The structure proposed by Genesereth and Nilsson seems to be more
apt to represent a state of affairs rather than a conceptualization.
In order to capture such intuitions, the linguistic terms we have used to
denote the relevant relations cannot be thought of as mere comments, informal
extra-information. Rather, the formal structure used for a conceptualization
should somehow account for their meaning. As the logico-philosophical lit-
erature teaches us, such a meaning cannot coincide with an extensional rela-
tion. In [Guarino and Giaretta 1995] we have presented a way to represent
this meaning in terms of an intensional structure inspired to Montague’s se-
mantics. According to this intensional interpretation, a conceptualization ac-
counts for the intended meanings of the terms used to denote the relevant re-
lations. These meanings are supposed to remain the same if the actual exten-
sions of the relations change due to different states of affairs. This means
that, for instance, the actual extensions of the relation on in the two examples
of Fig. 1a and 1b belong to the same conceptualization. Intuitively, we can
see a conceptualization as a set of informal rules that constrain the structure of
a piece of reality, which an agent uses in order to isolate and organize relevant
objects and relevant relations: the rules which tell us whether a certain block is
on another one remain the same, independently of the particular arrangement
of the blocks.
(2) A (AI-) ontology is a theory of what entities can exist in the mind of a knowledgeable
agent.
[Wielinga and Schreiber 1993]
(3) An ontology for a body of knowledge concerning a particular task or domain describes
a taxonomy of concepts for that task or domain that define the semantic interpretation
of the knowledge.
[Alberts 1993]
3 At least, in the sense that the ontological commitment of a logical theory is intended as a
set of individuals. Expressions like “can exist in the mind” are however extraneous to
Quine.
ory is ontologically committed to the entities it quantifies over. As discussed
in [Guarino et al. 1994], such a notion however is too weak for our pur-
poses, since we want not only an account of what exists, but also an account
of the structure of what exists. This structure is implied in the language we
use: this is the reason why, as noticed by the authors, the term "ontology" is
often used as a synonym of "terminology" in the AI community.
The definition (3) is more problematic. Although Van Heijst and col-
leagues correctly observe that it is the semantic interpretation of the terms of a
domain that constitutes an ontology, the formulation reported is misleading,
since "the semantic interpretation of the knowledge ... concerning a particular
task or domain" doesn’t regard the taxonomy only, but it also involves the
factual situations holding in that domain. The distinction between domain
knowledge and domain ontology made by the authors (p. 12) is therefore not
caputred by this definition. Moreover, according to definitions (1) and (2), an
ontology can be much more than a taxonomy of concepts, involving in par-
ticular constraints and interrelations among concepts. Hopefully, it should
also concern more than one particular task or domain. Alberts’ definition
seems therefore both partial and inaccurate, and I cannot see how the authors
consider it as “not contradictory” with (1) and (2), coming up with the fol-
lowing “unifying” definition:
(7) An ontology is a logical theory that constrains the indended models of a logical lan-
guage.
4 By the way, van Heijst already use the technique of mapping rules for the task of knowl-
edge-based integration (section 6), and I do not see why they don’t use the the same tech-
nique to link the application ontology to the method ontology.
A different conception of the application ontology is reported in Fig. 2.
The application ontology is built by specializing both the domain and the
method ontology. The concept CASNET-cost is the cost of an observation
leading to a pathophysiological state, which plays the role of the cost of an
hypothesis within the method “ranking by weight-to-cost-ratio”. What moti-
vates its presence in the application ontology is the fact that it plays a specific
role in CASNET’s ranking procedure. The ontological requirements of this
procedure are represented explicitly in the method ontology, and the
CASNET-specific cost satisfies the range restriction of the attribute has-
cost of the concept hypothesis belonging to the method ontology.
has-cost
pathophysiological state hypothesis cost
evidence-for
cost
Ontology library
CASNET-
Hypothesis
costs
Application ontology
CASNET-cost
Fig. 2. Application ontology as a specialization of the ontology library. Thick arrows rep-
resents subsumption links.
5 More in general, any concept of the application ontology would be either a specialization
or an instance of a (meta)concept in the ontology library. Of course, this choice would im-
ply a richer structure both in the application ontology and the ontology library.
Notice that, while the left part of Fig. 2 (the domain ontology) can be
considered as relatively static, the bottom and right parts change when the
problem solving strategy changes. In this way, the ISA arcs linking the appli-
cation ontology to the task ontology “can be seen as attributing context-
specific semantics to domain knowledge elements” ([Schreiber et al. 1995],
section 3.1). However, once the application ontology is fixed, it has a rigor-
ous model-theoretic semantics, in contrast with the approach based on syn-
tactical mapping relations discussed in [Schreiber et al. 1995].
In sum, I believe that the specific role played by single items of domain
knowledge into the decision-making process should be made explicit in the
application ontology. The indexing mechanism proposed by the authors ap-
pears to be to coarse for this purpose. As admitted by the authors (p. 25), it
makes the task of populating a largely incomplete ontology library by scoring
the newly defined concepts in the application ontology particulary difficult, or
almost impossible in presence of strong interaction problems.
Diagnostic hypotheses
Here me must find, in the domain ontology, the specializations of the concept
“hypothesis” (belonging to the method ontology) which are relevant for the
domain at hand. The method ontology imposes general constraints on such
concepts (say, an hypothesis must be a disorder). We look here at the domain
vocabulary. If we find a term suitable to be considered as an hypothesis, we
try to find it in the domain ontology, either by direct syntactic matching or by
means of suitable linguistic tools like thesauri (Wordnet, UMLS). If we don’t
find it, we may consider to introduce a definition for it, and to classify the
new concept in the domain ontology according to such definition. Then, we
must check whether this concept can be also classified as an “hypothesis”. If
this test fails, we must either change the definition of the newly created con-
cept or consider some other term as a candidate.
In the example given, the misleading term “syndrome” used in section
7.3.1 does not appear in the domain description of section 7.1. Sticking to the
originally elicited vocabulary helps therefore to avoid the introduction of un-
controlled terms. In fact, a simple linguistic analysis of the last two sentences
of section 7.1. allows us to conclude that the “therapeutic action” (which is
the goal of the task at hand) is driven by “the gradings of the lesions to each
of the systems”. The term “grading of a lesion” is therefore a good candidate
for an hypothesis6. Since “grading” is obviously an attribute of “lesion”, we
look for the main concept “lesion” in the domain ontology, which has, say,
the attribute “severity”, restricted to “grading”. We consider therefore
“severity of a lesion” as a candidate. But the severity of a lesion is nothing
else than a grading, i.e. a qualitative value. In the method ontology, the hy-
potheses of a diagnostic task are restricted to be “disorders”, and these are
disjoint from gradings. We exclude therefore “severity of a lesion”, and con-
sider the the possible specializations of “disorder”. Here we find “lesion”,
which admits any lesion of a particular severity as a subconcept. Since
“severity” is an attribute of “lesion”, and not a slot, all the lesions whith dif-
ferent severity are mutually exclusive due to the semantics of the representa-
tion primitives used (defined in the frame ontology). In conclusion, the hy-
potheses for the ARS applications are specializations of “lesion” (in fact,
“organ-system lesions”). Notice that no “lesion-subtype” relation is used; no-
tice also that “finding” is immediately excluded as a candidate since it does not
satisfy the restrictions on “hypothesis”.
Patient findings
Since the term “patient findings” appears explicitly in the task model, the cor-
responding concept belongs in my opinion to the method ontology and not to
6 We may use “lesion-grading”, but hyphenation may be misleading in this stage. The term
“grading of a lesion” makes clear what is the main concept and what is the attribute. See
[Guarino 1992].
the domain ontology as assumed by the authors (see however their note at
page 82). None of the terms used for the informal domain description (section
7.1) appears to be suitable as a candidate. The method ontology is thererore
used to elicit the concepts needed for the application ontology. As discussed
in the paper, particular expressions named “lesion-indications” are assumed to
be specializations of the concept “finding”.
Diagnostic data
The authors observe that “one aspect that distinguishes raw data from find-
ings in general is that data are directly observable”. This fact should be mod-
eled in the method ontology, which can be used therefore to elicit the concepts
related to diagnostic data in the application ontology. It is also important, in
my opinion, that these concepts reflect the information contained in the raw
data available for the application, without any implicit abstraction process. In
the example given, it seems to me a much better strategy to represent a single
datum as (ars-datum (erythema (location head-and-neck)
(degree 2))) rather than (ars-datum head-and-neck-
erythema = 2).Among other things, this choice blocks the possibility to
exploit general knowledge related to the location of a finding.
Diagnostic abduction
Due to the choice made when determining the diagnostic hypotheses, the
authors are not able to specialize the relation manifestation-of sup-
posed to exist in the domain ontology,since it holds for disorders while
ars-lesion-gradings have been defined as expressions. We don’t
have this problem, since - as we have seen - hypotheses are restricted to be
disorders by the method ontology. However, what deserves attention is the
way out adopted by the authors: they introduce an application-specific concept
called ars-manifestation-of, which is modification of manifesta-
[Link] this approach of modifying a concept as a result of a type
mismatch, introducing a new concept with a very similar name, seems to me
extremely dangerous. First, I believe that a rigorous naming discipline should
be part of the methodology. A very natural criterion in this respect is the fol-
lowing: any concept whose name is X-C should be a specialization of C8. It is
easy to see how violations to this criterion generate confusion and compro-
mise readability. Second, the introduction of an ad-hoc concept in the appli-
cation ontology which is not a specialization of an existing concept in the do-
main ontology violates the (refined) definition of application ontology that I
have proposed.
5. Conclusions
I will try here to summarize the observations made in this commentary
paper, giving at the same time an overall assessment of the main issues related
to ontology-based knowledge modelling.
Acknowledgements
This work has been done in the framework of a Special CNR Project on
ontological and linguistic tools for conceptual modelling (Progetto Coordinato
"Strumenti Ontologico-Linguistici per la Modellazione Concettuale”). I am
grateful to Pierdaniele Giaretta and Massimiliano Carrara for their precious
comments.
Bibliography
Alberts, L. K. 1993. YMIR: an Ontology for Engineering Design. University
of Twente
Bateman, J. A., Kasper, R. T., Moore, J. D., and Whitney, R. A. 1990. A
General Organization of Knowledge for Natural Language Processing: the
PENMAN upper model. USC/Information Sciences Institute, Marina del
Rey, CA.
Bylander, T. and Chandrasekaran, B. 1988. Generic tasks in knowledge-
based reasoning: The right level of abstraction for knowledge acquisition.
In B. R. Gaines and J. H. Boose (eds.), Knowledge Acquisition for
Knowledge Based Systems. Academic Press, London.
Falasconi, S. and Stefanelli, M. 1994. A Library of Medical Ontologies. In
Proceedings of ECAI94 Workshop on Comparison of Implemented On-
tologies. Amsterdam, The Nederlands, European Coordinating Committee
for Artificial Intelligence (ECCAI): 81-92.
Genesereth, M. R. and Nilsson, N. J. 1987. Logical Foundation of Artificial
Intelligence. Morgan Kaufmann, Los Altos, California.
Gennari, J. H., Tu, S. W., Rothenfluh, T. E., and Musen, M. A. 1994.
Mapping Domains to Methods in Support of Reuse. IJHCS, 41: 399-424.
Gruber, T. 1994. Toward Principles for the Design of Ontologies Used for
Knowledge Sharing. IJHCS, 43(5/6): 907-928.
Gruber, T. R. 1993. A translation approach to portable ontology specifica-
tions. Knowledge Acquisition, 5: 199-220.
Guarino, N. 1992. Concepts, Attributes and Arbitrary Relations: Some Lin-
guistic and Ontological Criteria for Structuring Knowledge Bases. Data &
Knowledge Engineering, 8: 249-261.
Guarino, N. 1994. The Ontological Level. In R. Casati, B. Smith and G .
White (eds.), Philosophy and the Cognitive Science. Hölder-Pichler-
Tempsky, Vienna: 443-456.
Guarino, N. 1995. Formal Ontology, Conceptual Analysis and Knowledge
Representation. International Journal of Human and Computer Studies,
43(5/6): 625-640.
Guarino, N. and Boldrin, L. 1993. Ontological Requirements for Knowledge
Sharing. In Proceedings of IJCAI workshop on Knowledge Sharing and
Information Interchange. Chambery, France: 1-5.
Guarino, N., Carrara, M., and Giaretta, P. 1994. Formalizing Ontological
Commitment. In Proceedings of National Conference on Artificial Intelli-
gence (AAAI-94). Seattle, Morgan Kaufmann.
Guarino, N., Carrara, M., and Giaretta, P. 1994. An Ontology of Meta-Level
Categories. In D. J., E. Sandewall and P. Torasso (eds.), Principles of
Knowledge Representation and Reasoning: Proceedings of the Fourth In-
ternational Conference (KR94). Morgan Kaufmann, San Mateo, CA: 270-
280.
Guarino, N. and Giaretta, P. 1995. Ontologies and Knowledge Bases: To-
wards a Terminological Clarification. In N. Mars (ed.) Towards Very
Large Knowledge Bases: Knowledge Building and Knowledge Sharing
1995. IOS Press, Amsterdam: 25-32.
Knight, K. and Luk, S. 1994. Building a Large Knowledge Base for Ma-
chine Translation. In Proceedings of American Association of Artificial
Intelligence Conference (AAAI-94). Seattle, WA.
Mahesh, K. 1996. Ontology Development for Machine Translation: Ideology
and Methodology. New Mexico State University, Computing Research
Laboratory MCCS-96-292.
McDermott, J. 1988. Preliminary steps toward a taxonomy of problem-
solving methods. In S. Marcus (ed.) Automating Knowledge Acquisition
for Expert Systems. Kluwer Academic Publishers.
Quine, W. O. 1961. From a Logical Point of View, Nine Logico-
Philosophical Essays. Harvard University Press, Cambridge, Mass.
Rector, A. L., Nowlan, W. A., Kay, S., Goble, C. A., and Howkins, T. J .
1993. A Framework for Modelling the Electronic Medical Record. Meth-
ods of Information in Medicine, 32: 109-119.
Schreiber, A. T. and Terpstra, P. 1996. Sysyphus-VT: a CommonKADS
Solution. IJHCS, 44: 373-402.
Schreiber, G., Wielinga, B., and Breuker, J. 1993. KADS: A Principled Ap-
proach to Knowledge-Based System Development. Academic Press, Lon-
don.
Schreiber, G., Wielinga, B., and Jansweijer, W. 1995. The KAKTUS View
on the 'O' Word. In Proceedings of IJCAI95 Workshop on Basic Onto-
logical Issues in Knowledge Sharing. Montreal, Canada.
Uschold, M. and Gruninger, M. 1996. Ontologies: Principles, Methods and
Applications. The Knowledge Engineering Review, (in press).
van Heijst, G., Schreiber, A. T., and Wielinga, B. J. 1996. Using Explicit
Ontologies in KBS Development. International Journal of Human and
Computer Studies(this issue).
Wielinga, B., Schreiber, A. T., Jansweijer, W., Anjewierden, A., and van
Harmelen, F. 1994. Framework and formalism for expressing ontologies.
ESPRIT Project 8145 KACTUS, Free University of Amsterdam deliver-
able DO1b.1.
Wielinga, B. J. and Schreiber, A. T. 1993. Reusable and sharable knowledge
bases: a European perspective. In Proceedings of Proceedings of First In-
ternational Conference on Building and Sharing of Very Large-Scaled
Knowledge Bases. Tokyo, Japan Information Processing Development
Center.