Beginning Syntax
Beginning Syntax
Beginning Syntax
An Introduction to Syntactic Analysis
IanRoberts_9781316519493_C000.indd 2 10-09-2022 16:54:50
In this series:
Beginning Syntax
An Introduction to Syntactic Analysis
IAN ROBERTS
Downing College, University of Cambridge
IanRoberts_9781316519493_C000.indd 5 10-09-2022 16:54:50
[Link]
Information on this title: [Link]/9781316519493
DOI: 10.1017/9781009023849
© Ian Roberts 2023
This publication is in copyright. Subject to statutory exception and to the provisions
of relevant collective licensing agreements, no reproduction of any part may take
place without the written permission of Cambridge University Press & Assessment.
First published 2023
Printed in <country> by <printer>
A catalogue record for this publication is available from the British Library.
Library of Congress Cataloging-in-Publication Data
ISBN 978-1-316-51949-3 Hardback
ISBN 978-1-009-01058-0 Paperback
Cambridge University Press & Assessment has no responsibility for the persistence
or accuracy of URLs for external or third-party internet websites referred to in this
publication and does not guarantee that any content on such websites is, or will
remain, accurate or appropriate.
Contents
vii
Exercises 48
Further Reading 49
4 X′-theory 78
4.1 Introduction 78
4.2 Possible and Impossible PS-Rules 78
4.3 Introducing X′-theory 80
4.4 X′-theory and Functional Categories: The Structure of the Clause 87
4.5 Adjuncts in X′-theory 95
4.6 Conclusion 96
Exercises 97
For Further Discussion 98
Further Reading 98
5 Movement 99
5.1 Introduction and Recap 99
5.2 Types of Movement Rules 100
5.3 Wh-Movement: The Basics 104
5.4 Subject Questions, Indirect Questions and Long-Distance
Wh-Movement 112
5.5 The Nature of Movement Rules 118
5.6 Limiting Wh-Movement 122
5.7 Conclusion 126
Exercises 127
For Further Discussion 128
Further Reading 128
6 Binding 130
6.1 Introduction and Recap 130
6.2 Pronouns and Anaphors 130
6.3 Anaphors 132
6.4 Pronouns 142
6.5 R-expressions 143
6.6 The Binding Principles 144
6.7 Variables, Principle C and Movement 146
6.8 Conclusion 153
Exercises 154
For Discussion 154
Further Reading 155
Conclusion 215
Glossary 219
Index 234
IanRoberts_9781316519493_C000.indd 9 10-09-2022 16:54:51
IanRoberts_9781316519493_C000.indd 10 10-09-2022 16:54:51
Preface
The purpose of this book is to present the essential elements of the theory of
syntax, presupposing no prior knowledge of either syntax or linguistics more
generally. The first chapter introduces the thinking behind modern formal
linguistics, defining the core background concepts. The second chapter is an
extended demonstration of linguistic competence, revealing to an English
speaker their tacit, untutored knowledge of the intricacies of the syntax of the
language. The next four chapters introduce the core technical concepts of syn
tax: phrase structure, constituency, movement rules and construal rules. Chapter
7 is a very brief introduction to comparative syntax, illustrating how the theory
developed in the earlier chapters can apply to languages other than English.
Finally, Chapter 8 discusses the overall architecture of grammar, introducing
several levels of representation, including the interface levels of Phonological
Form and Logical Form.
The book is based on a first-year lecture course I have taught at Cambridge
for most of the past ten years, called ‘Structures’. That course has successfully
laid the groundwork for more advanced study; several professional
syntacticians had their first exposure to the subject in that context.
I do not attempt to review either the history or the state of the art in contem
porary generative syntax. Instead, the book is intended to provide a solid basis
from which the student can move on to more advanced topics. In the decades
since its inception, generative theory has yielded a range of core insights and
concepts which can be studied independently of the details of a given
framework and which form the core of any formal approach to syntax (and core
background to the formal study of other areas of linguistics, in particular
semantics). These are the focus of this introduction, where the theoretical
orientation of the book is most clearly set out. This volume is intended as the
first in a series; future vol
umes will build on the material presented here, as well as both adding an explic
itly comparative and typological dimension and engaging more directly with the
minimalist programme for linguistic theory. The overall aim of the series is to
provide a complete course in syntax, taking the student from beginner level to
being able to engage with contemporary research literature.
This volume is intended for students starting out in syntax and assumes no
prior knowledge. As mentioned above, it corresponds to a first-year course
taught at Cambridge. It can be taught over a single term or semester with
tutorial
xi
support (as I have done at Cambridge) or over an entire year as a lecture course.
The book tries to introduce the main concepts of syntactic theory in an engaging
way with a minimum of technicalities. Each chapter features exercises, some
straightforwardly testing the material covered in the chapter, others raising fur
ther questions for reflection and discussion (e.g. in tutorials), as well as sugges
tions for further reading. Model answers are provided separately for the first type
of exercise. The second type presupposes student participation in discussions.
Chapter 1 is somewhat different from the others in that it introduces some fairly
complex data; the point at this stage of the exposition is to demonstrate tacit
knowledge, i.e. native-speaker competence; the technical issues are, as appropri
ate, revisited and introduced more gradually later. Instructors may wish to skip
over some of this material and/or revisit it later (for example, the Italian data can
supplement the material covered in Chapter 7).
I would like to thank the other people who have taught this course in whole or
in part over the years: Dora Alexopoulou, Theresa Biberauer, Craig Sailor and
Jenneke van der Wal, as well as the many doctoral students, postdocs and others
who have given tutorials for the course. Most of all, I’d like to thank the students
themselves: you ultimately make everything happen. Thanks also to Bob Freidin
and Dalina Kallulli for comments on early drafts. And of course, thanks to the
two Helens at Cambridge University Press: Helen Barton and Helen Shannon,
whose tolerance of missed deadlines fully matches my capacity to bring them
about.
IanRoberts_9781316519493_C000.indd 12 10-09-2022 16:54:51
Abbreviations
xiii
intelligent Martian, you’d soon recognise that there is some connection between
the bipeds’ noises, their ability to build machines and their control over all the
other species. You’d also notice that only the featherless bipeds make these
noises and that the young of the species spontaneously start making the noises
when they’re very small and still dependent on their parents for survival. After
abducting some bipeds and looking inside their brains, you’d realise that there’s
just a couple of areas of the biped brain which seem to control the noise-making
ability. The noises are language (Martians would probably first notice phonol
ogy, but being super-intelligent they’d quickly realise that there was more to
language than just that). By looking at ourselves through Martian ‘eyes’ (which
might of course be infrared heat sensors attached to their feet, but never mind),
we can appreciate language as a natural object.
The second thing about science is that it’s not just about observing things;
it’s also about putting the observations together to make a theory of things. One
of the most famous of all scientific theories is Einstein’s Theory of Relativity.
This is hardly the place to go into that, but the point is that Einstein and other
physicists wanted to construct a general understanding, based on observation, of
certain aspects of nature: the Theory of Relativity is mostly about space, time
and gravity. A good theory, like relativity, makes sense of the observations we
have already made and also predicts new ones that we should be able to make.
So saying linguistics is scientific doesn’t necessarily mean that we have to put
on white coats and start a laboratory. But it does mean that we should try to make
observations about language and put those observations together to make a theory
of language. Since we’re going to focus on syntax, we’re going to make obser
vations about syntax and make a theory of syntax. So the goal of this book is to
present a particular theory of syntax, a scientific theory of what it is that links sound
and meaning (as you can imagine, there are separate but connected theories of pho
nology and semantics). The theory that I will introduce here is called generative
grammar. The reasons for that name will emerge over the next couple of chapters.
Approaching language, or syntax, in this way means that there are aspects
of the more traditional, humanities-style approaches to language that we leave
behind, since they do not contribute to the goal of making observations and
building theories. First, it implies that the goal of linguistics is not to set stand
ards of ‘good’ speaking or writing. This is not to imply that such standards
should not be set, or that linguists, or others with particular expertise in lan
guage, should not be those responsible for setting them. Instead, it means that
setting these standards falls outside the scientific study of language, since doing
this doesn’t involve making observations about what people say, write or under
stand but involves recommending that people should speak or write in a cer
tain way. Zoologists don’t tell slugs how to be slimy; linguists don’t tell people
how to speak or write. Ideas of ‘good’ syntax have no place here. Prescriptive
grammatical statements of a kind once common in the teaching of English in
schools in most of the English-speaking world (see below for some examples)
are similarly irrelevant to our concerns here.
species, but a defining property of it. The official biological name of our species
(in the Linnean taxonomy) is Homo sapiens, Latin for ‘wise man’. Since a
glance at a history book or a newspaper leads one to question the wisdom of our
species, one could think that Homo loquens (‘talking man’) might have been a
better term to define us.
As just mentioned, the general concept of language, roughly defined above,
should be distinguished from individual languages. An individual language
can be taken to be a specific variant of the uniquely human capacity defined
above, usually part of a given culture: English, Navajo, Warlpiri, Basque, etc.
Concepts such as ‘English’ are thus, at least in part, cultural concepts. The dis
tinction between language (in general) and languages (individual languages as
partly cultural entities) distinguishes two variants of a single word in English.
This is rather like the way we can distinguish the general concept of cheese
from individual cheeses such as Brie, Cheddar, Parmigiano Reggiano and so on.
Language and cheese are mass nouns; they denote general concepts. Individual
cultures in different parts of the world produce their local variants of the gen
eral thing, associated with countable versions of the same nouns, i.e. individual
languages and cheeses. In several European languages, different words are used
to make the distinction between language and languages. French distinguishes
langage (language in general) from langue (individual languages); in Italian lin
guaggio is distinguished from lingua in a similar way, and Spanish distinguishes
lenguaje from lengua.
writing). One of the striking features of natural languages, both spoken and
signed, is that their salient structural properties are independent of their medium
of transmission: sign-language syntax, for example, has the same structural
properties as English, French and other spoken languages.
Constructed languages, or conlangs, are also relevant here. Conlangs are of
two main kinds. First, there are languages which were deliberately constructed
as ‘ideal’ languages intended to serve for more efficient communication than
could be afforded by seemingly imperfect already existing languages. Esperanto
is probably the best known, although far from the only, language of this kind.
Esperanto was invented by L. L. Zamenhof in 1887, at a time when there was
no clear international lingua franca in the Western world (since 1945, English
has played this role; at earlier times Latin and French did). The vocabulary of
Esperanto is a mixture of Romance, Germanic and Slavic elements, and the
grammatical system is a highly simplified and artificially regularised. The sys
tem is, however, closely based on various European, principally Romance,
languages. As such it is arguably a natural language, which happens to have a
particular origin and purpose. It is debatable, however, whether Esperanto has
any true native speakers, although it is certainly used as a second language by
up to 2 million people all around the world (compare this with the estimated 2
billion second-language speakers of English). Many other invented languages
(Interlingua, Volapük, etc.) have the same status as Esperanto, although these
days they are hardly spoken at all.
The other principal kind of constructed languages are those invented for
fictional purposes, in order to give linguistic realism to invented worlds and
their denizens. Among the best-known examples are the languages invented by
J. R. R. Tolkien in his extensive mythological writings (The Lord of the Rings,
The Hobbit, The Silmarillion, etc.). Tolkien was a professional philologist, and
so the languages he constructed have an air of authenticity. Like Esperanto,
however, they are largely based on existing, mainly European, languages. They
are thus natural languages which are unusual in their origin and purpose; they
are also barely spoken at all and certainly have no native speakers. Above all for
this last reason, we leave such languages aside here; our interest is in languages
which are acquired naturally.
Two other kinds of ‘language’ should also be mentioned. First, there is ‘body
language’: frowning, shaking or nodding one’s head, crossing one’s arms or
legs, etc. Body language can certainly communicate various attitudes or emo
tional states (friendliness, aggression, etc.), but it is not a language in the sense
that it lacks the structural properties of natural, spoken and signed, languages;
it has no discernible syntax, for example. Second, there is music. Music is often
said to be a language, and indeed it has been shown to have structural, syntac
tic properties which are akin to, perhaps even identical to, natural language.
However, although music may have a profound emotional impact, it lacks a clear
propositional semantics, in that it cannot communicate true or false statements.
It may be that music is a cognitive capacity which shares some, but not all, of
1
See R. Robins (1967), A Short History of Linguistics from Plato to Chomsky, London: Long
man, and V. Law (2003), The History of Linguistics in Europe, Cambridge: Cambridge Univer
sity Press, for discussion.
2
See R. Freidin (2020), Adventures in English Syntax, Cambridge: Cambridge University Press,
for a very illuminating discussion.
Natural Language
Leaving aside the sociocultural dimensions, then, we concentrate on
looking at natural language from a scientific perspective. More precisely, we
3
M. Atkinson (1992), Children’s Syntax, Oxford: Blackwell.
will look at language from a cognitive perspective. This approach treats lan
guage as a form of knowledge; the central idea is that our uniquely human lan
guage capacity is part of our cognitive make-up. Language is really a kind of
instinct humans, and no other species, are pre-programmed with (in fact, The
Language Instinct is the title of an influential and important book by Stephen
Pinker in which this approach is explained in detail; see Further Reading at the
end of this Introduction). This is the view that forms the basis of the theory of
generative grammar and so this is the view I will adopt here. This approach, and
most of the concepts introduced in this section, are due to Noam Chomsky (see
Further Reading at the end of the Introduction for more details).
The cognitive approach treats our capacity for language as an aspect of human
psychology which is ultimately rooted in biology. In this way, it clearly treats
language, and our capacity for language, as natural objects. On this view, there
are three factors in language design. The first is contributed by genetics: an
innate aptitude for language that is unique to our species, our ‘language instinct’.
The second factor is experience, particularly in early life. The language we are
exposed to as children represents the crucial linguistic experience; when I was
growing up in England, everyone around me spoke English and so I acquired
English. My great-grandfather was surrounded by Welsh speakers when he
was growing up in nineteenth-century North Wales, so he acquired Welsh.
Whichever language, or languages, we are exposed to as small children, our
innate language-learning capacity is brought to bear on this experience in such a
way as to result in our competence as native speakers of our first language (I will
say more about language acquisition below). Third, general cognitive capac
ities, not specific to language and perhaps not specific to humans, clearly play
a role in shaping our knowledge and use of language, although the exact role
these capacities play and how they interact with (and can be distinguished from)
the language-specific aspects of our linguistic capacities are difficult questions.
Together, these three factors constitute the human language faculty. It can be
extremely difficult to distinguish the specifically linguistic aspects that contrib
ute to forming the language faculty from more general cognitive abilities, but
the distinction can certainly be made in principle and is of course very important
for our general cognitive theory of language.
In these terms, Universal Grammar (UG) is the theory of the first factor
which makes up the language faculty: our innate genetic endowment for lan
guage. As already mentioned, UG is assumed to play a central role in language
acquisition. It is also vital in helping us to understand how we can make sense
of the idea that specific languages are actually variants of a single type of entity,
language. Our goal here, then, is to look at one part of UG: how words combine
to form sentences. In so doing, we will construct the theory of syntax.
We can now make an important distinction. Language can be seen from an
internal, individual perspective, I-language, or it can be seen as external to the
individual, E-language. Here we are going to focus on I-language, which arises
from the interaction of the three factors just introduced. This is a natural approach,
IanRoberts_9781316519493_Introduction.indd 10 05-09-2022 11:17:15
Natural Language 11
given our interest in how syntactic structures arise in the mind. E-language is
in fact a more complicated notion, involving society, culture, history and so on.
Concepts like ‘English’ and ‘French’ in their everyday senses are E-language
concepts; it is for this reason that they do not, strictly speaking, form part of our
object of study. I-language is a natural object; E-language is not.
Taking I-language as the central notion treats language as part of individ
ual psychology. This is a cognitive theory of language, because it intrinsically
involves the human mind. In fact, this theory of language could be part of an
overall theory of the mind. However, we would really like our theory of language
to be part of an overall theory of the brain; being an obviously physical object,
the brain is a natural object, and it is a clearer notion than ‘mind’ (in fact, philos
ophers have worried for centuries about whether ‘mind’ is something physical,
but we don’t need to). Unfortunately, though, at present our understanding of
how brain tissue supports cognitive processes such as language, thought and
memory is extremely limited, and so we are unable to say very much about the
relation between the physical brain and cognitive processes. That’s why I will
continue to use the older, strictly speaking vaguer, term ‘mind’.
What does it mean to say our theory is a formal theory? A formal approach to
any kind of problem or phenomenon assumes discrete, systematic ways of form
ing complex things out of simple things. For example, the alphabet is formal: its
twenty-six letters can be combined in various different, but more or less system
atic, ways to form a very large number of words. Arithmetic is a formal system
combining numbers of various kinds and functions such as addition, subtraction,
multiplication, etc.
In formal syntax and formal semantics, we combine simple elements (roughly
words and their meanings) to form more complex elements (sentences and their
meanings). The modes of combination must be systematic and precise; as we
will see, this is a major part of the challenge of constructing such a theory.
A further natural question to ask is why we want a formal and cognitive theory.
The answer to this lies in certain general trends of thought both in psychology
and in philosophy of mind (at least in the English-speaking world). It is widely
believed that the best way to understand the mind is to think of it as a kind of
computer. Computers manipulate symbols according to formal instructions, i.e.
algorithms and programs. We can think of I-language in these terms as a piece
of cognitive software, a program run on the hardware of the brain (this raises
the intriguing question, which I will not go into here, of who or what wrote the
program). Therefore, our theory of I-language must be a formal theory. What we
are interested in is how our knowledge of I-language is represented in our mind.
If we can get an idea of this, which I believe we can from studying syntax, then
we gain a very important insight into the human mind; what it is to be human,
what it is to be you.
To conclude this general discussion, we will concentrate here on developing
one aspect of a formal, cognitive theory of I-language: the theory of syntax.
As already mentioned, syntax is concerned with how relatively simple units,
words, are combined to form more complex units, sentences. This is taken to
be a cognitive capacity all humans have, as a reflex of their genetic endowment
which includes UG. UG interacts with linguistic experience in early life, and
with domain-general ‘third factors’ as described above, to give rise to mature
adult competence in one’s native language. This competence manifests itself
in the ability to produce and understand an unlimited number of sentences, and
to make judgements regarding both the syntax and the semantics of those sen
tences. In the next chapter, we will look in detail at a concrete example of this
competence in action, for native speakers of English.
Just a final note: given what I’ve said, strictly speaking I shouldn’t talk about
‘English’, as it is not a scientific term since it does not designate a natural object.
Instead of talking about ‘English speakers’, I should really say something like
‘individuals who identify themselves and are identified in their cultural milieu
as possessing an I-language which corresponds to the E-concept “English”’. For
brevity I will gloss over this more accurate formulation and continue to talk about
‘English speakers’; the same goes for other E-language names (Italian, etc.).
More generally, the term ‘language’ will refer to that aggregate of I-languages
whose speakers recognise themselves and each other as belonging to the same
E-language community.
Now we can move from the rather general matters we have been considering
here to actually starting out on the study of syntax. As we delve more and more
into the detailed and intricate nuts and bolts of the theory of syntax in the chapters
to follow, the background issues we have discussed here should be kept in mind
as they form the overall conceptual underpinning to the theory we’ll develop.
But now for the nuts and bolts. Or, actually, the fish.
Exercises
1. Write a short paragraph (not more than half a page as an absolute
maximum; ten–twelve lines per concept would be ideal) to explain
in your own words what the following notions mean to linguists:
• The language faculty
• Formal approaches to the study of language
• ‘Language’ vs ‘languages’
• English
Further Reading
Adger, D. 2019. Language Unlimited. Oxford: Oxford University Press.
Pinker, S. 1994. The Language Instinct. New York: Harper Perennial Modern Classics.
Roberts, I. 2017. The Wonders of Language, or How to Make Noises and Influence
People. Cambridge: Cambridge University Press.
These three books all provide general introductions to language and linguistics assum
ing no prior knowledge on the reader’s part. Each book has its own perspective: Adger
concentrates on how syntax is fundamental to human language and cognition, and so is
perhaps most in line with our concerns here. Pinker is a classic; it is a witty and engaging
introduction to linguistics with the emphasis on the relation between language and mind.
Roberts offers a comprehensive introduction to several different subfields of linguistics,
ranging from phonetics to historical linguistics; syntax is covered in Chapter 4.
Larson, R. 2010. Grammar as Science. Cambridge, MA: MIT Press, Unit 1. This book
is an introduction to syntax whose central idea is to present the field as an exercise in the
construction of a scientific theory. In this first unit, the central ideas regarding knowledge
of language and Universal Grammar are introduced in an attractive and original way.
Carnie, A. 2013. Syntax: A Generative Introduction. Oxford: Blackwell, Chapter 1.
Another very sound and well-written introduction to generative grammar. This first
chapter presents the central ideas behind syntactic theory, along with a section on dif
ferent approaches to syntax (something I do not attempt here). Isac, D. & C. Reiss.
2008. I-language. Oxford: Oxford University Press, Chapters 1 and [Link] the title sug
gests, the focus of this book is on the I-language approach to linguistics and syntax.
The first chapter starts with a presentation of linguistic data in order to illustrate the
approach. The third chapter considers different notions of language, similar to what has
been presented here but with a different overall slant; reading that chapter will comple
ment this one nicely.
Freidin, R. 2012. Syntax: Basic Concepts and Applications. Cambridge: Cambridge
University Press, Chapter [Link] chapter contains an excellent introduction to I-language,
the nature of grammaticality, native speakers’ ability to recognise deviant sentences as
an indication of knowledge of language, language acquisition and the argument from
the poverty of the stimulus, and, in particular, language production and comprehension.
1.1 Introduction
In this chapter, we begin our study of the theory of syntax. As we
will see later, there is no end to syntax. It is also rather difficult to discern
where it begins. Here we will make a start by doing two things. First, you’ll
find out some rather surprising facts about your knowledge of English, things
you didn’t know you knew. This will give you a concrete illustration of your
competence in English. Second, by carrying out a little translation exercise,
we’ll get a first glimpse of what makes languages similar to one another and
what makes them differ.
We begin with very simple sentences (see (1)) and build up to rather strange
and complex ones. It’s not necessary to follow every detail of the discussion
here; all the technical ideas are presented again in later chapters. Our goal here
is to demonstrate the nature of your tacit knowledge of English syntax. We can
do it with just one word: fish. This chapter can also be skipped and returned to
later (e.g. after Chapter 7).
Like quite a few basic words in English, fish is actually ambiguous. It can be
understood either as a verb or as a noun. As a noun, it refers to a class of
aquatic animals; as a verb, it refers to the activity of hunting those aquatic
animals. We indicate this ambiguity in the standard way, by surrounding the
word with square brackets, and writing the ‘category label’ (Noun/Verb) as a
subscript to the left bracket. So, the sentence in (1) actually has two distinct
representations:
1
Thanks to my good friend and colleague Professor Robert Freidin of Princeton University for
these examples. The implications of the fish sentences are discussed and explored in more
detail in Chapter 1 of Freidin (2012), . Freidin’s discussion is more detailed than here,
although it focuses exclusively on English. The exercises given there are also well worth
trying.
15
The brackets are just a way of saying ‘what is inside here is a Noun/Verb’. A
representation like that in (2) is called a labelled bracketing.
The fact that fish, like many other words including cook, book, police, report,
promise and many others, is ambiguous between a verb and a noun is a conse
quence of the fact that English has very few inflectional endings to mark gram
matical information of various kinds. If we compare English with a more richly
inflected language such as Italian, we see that the two versions of fish correspond
to two differently inflected words:
(3) a. [Noun fish ] = Italian pesce
b. [Verb fish ] = Italian pescare
It is quite easy to see that the Italian words share the root pesc- and distinct
inflections: -e, indicating a singular noun; and -are, indicating the infinitive of
a first-conjugation verb (the infinitive is the basic form of a verb with no tense
marking; ‘first conjugation’ refers to an arbitrary morphological class of verbs
in Italian: there are four conjugations altogether, as we will see below). To cut
a very long historical story short, English has largely lost its equivalents of -e
and -are, and so we are left with the equivalent of the ambiguous root pesc- (you
may also note a family resemblance between pesc- and fish; this is because both
are ultimately derived from an ancient root in the common ancestor language of
English and Italian, Indo European).
The ambiguous single word in (1) and (2) constitutes an entire sentence on its
own. More precisely, each interpretation of fish constitutes a sentence, so really
there are two distinct sentences here. We can represent them as follows (here,
again following standard conventions, ‘Noun’ is abbreviated as N and ‘Verb’ as
V, and S stands for ‘Sentence’):
(4) a. [S [N Fish ]]
b. [S [V Fish ]]
Interpreted as in (4a), the sentence draws attention to the presence of a single fish or
group of fish (note that fish is one of a relatively small group of English nouns that
is identical in the singular and the plural: the plural form fishes exists, but it denotes
several species of fish, not several individual fish of the same species). Interpreted
as in (4b), it is an imperative, indicating an order given to go fishing. Here there
is an implicit second-person pronoun, since an order is naturally understood as
being addressed to an interlocutor, so we could elaborate (4b) as in (5):
(5) [S [N You ] [V fish ]]
Here the pronoun (a kind of noun, hence N) is ‘understood’; we take that to
mean that it is present in the syntactic and semantic representations of the sen
tence, but not in the phonological representation, what you actually hear or say.
This is our first indication that there is more to syntax than meets the ear.
The natural interpretation of this sentence is to treat the first fish as a noun and
the second one as a verb, the combination again forming a sentence. We repre
sent this as in (7):
(7) [S [N fish ] [V fish ]]
In this sentence, the noun fish is the subject, understood as carrying out the
action performed by the verb. The verb fish indicates the action the subject car
ries out. So the sentence means “Fish fish stuff”. As this rough gloss indicates,
there is an implicit direct object here, indicating, somewhat vaguely, what is
being fished for, what undergoes the action of fishing.
Alternatively, we can understand the first fish as a verb, and the second fish
as a noun. This is easier to see if we add an exclamation mark to (6) (Fish fish!),
corresponding to a different intonation pattern in speech. Then the sentence has
the structure in (8):
(8) [S [V fish ] [N fish ]]
Here, as in the verbal interpretation of the single-fish example in (1), the verb
is understood as an imperative and so there is a deleted second-person pronoun
(you) as subject:
(9) [S YouN [V fish ] [N fish ]]
The second fish, the noun, is an explicit direct object, indicating that fish are
what is fished.
In Italian, where the ambiguities of the English roots are clarified by inflec
tional endings, the two interpretations of (6) can be rendered as in (10):
(10) a. I pesci pescano. (= (7))
b. Pesca pesci! (= (8/9))
Second, the verb pescano consists of the root pesc- and the ending -ano (again
compare the infinitive ending -are in (3b)). This is not the place to go into the
full details of Italian verbal morphology, but this ending contains the infor
mation that the verb is first conjugation, present tense, indicative mood (it
makes a statement of fact) and third-person plural (‘they’). That’s quite a bit
of information in just three overt phonemes. The -a- part of the ending recurs
in the infinitive ending, and this can be seen as the marker of first conjuga
tion. The -no part of the ending shows up in many other tenses and is clearly
marking third-person plural. Present tense and indicative mood are arguably
Furthermore, we can, and for the sake of consistency must, put labelled brackets
around each part of the structure:
(14) [V [V pesc- ] [Conj -a- ] [M Indicative ] [T Present ] [Agr -no ]]
The labels are fairly straightforward: ‘V’ indicates that both the root and the
whole thing form a verb; ‘Conj’ indicates the conjugation-class marker ‘M’ indi
cates mood (the traditional grammatical category indicating whether a sentence
describes a state of affairs believed to be true or not), ‘T’ indicates tense and
‘Agr’ indicates agreement: the verb is third-person plural because the subject, i
pesci, is third-person plural.
Now, if the Italian verb has a structure like (14), and the Italian sentence in
(10a) is an accurate translation of the English sentence in (7) and our assumption
that English and Italian verbs should be minimally different means that we want
structures to ‘line up’ across languages, our logic leads us to posit something
like (15) as the structure of the verb fish in (7):
(15) [V [V fish- ] [Conj ?? ] [M Indicative ] [T Present ] [Agr 3Pl ]]
Here, Conj, M, T and Agr are all silent. But our logic leads us to conclude that
they are all structurally present. Now we come up against the limits to this lin
ing-up-in-the-name-of-UG approach. It seems reasonable to take the verb fish in
(7) to be indicative in mood, present in tense and third-plural in agreement, but
the conjugation-marking seems to impose an idiosyncracy of Italian morphology
onto English, moreover an idiosyncracy that appears to have no semantic corre
late. So perhaps we should eliminate ‘Conj’ from (15). The other silent endings
are, however, justified by the semantics and our universalising methodology.
Both tense and agreement endings do audibly appear on English verbs (mood
may too, but this is a little trickier and so I’ll leave it aside). If we make the sub
ject of our example in (12) singular, but keep the present tense, we have:
(16) The boy fishes.
Here the verb has the ending -es. If we make the verb past tense, we have:
(17) The boy fished.
Here the verb has the ending -ed. So we could give the verbs in (16) and (17) the
structures in (18a) and (18b) respectively:
The fact that where boys is plural as in (12) no audible determiner is required,
combined with the unique, known, existing ‘natural kind’ interpretation assigned
to boys here, supports the idea that there is a silent determiner. This in turn sup
ports the lining-up of English and Italian seen in (11).
We can now replace (11) with a fuller lined-up, (nearly) uniform English and
Italian structure, in which the determiners, silent and overt, are represented by
the category D:
(21) [S [NP [D the ] [N [N fish-[Num Pl ]]] [V [V fish- ] [M Indicative ] [T Present ]
[Agr 3Pl ]] ]
[S [NP [D I ] [N [N pesc-[Num i ]]] [V [V pesc- ] [Conj -a- ] [M Indicative ] [T Present ]
[Agr -no ]] ]
Here I’ve added the specification of number on the noun (the ending Num);
plural number is silent in the case of fish in English as we have observed. We
could add further specifications: Person on the noun (3rd here), number marking
on the Italian article i (implying, by our now-familiar reasoning, that English
the has a silent plural ending). Furthermore, the Italian articles show gender
marking: article i is masculine. So we should add that specification to the Italian
sentence. Should we add it to the English one? The Italian gender is grammati
cal, not semantic. Pesce is a masculine noun, but not all fish are male (or there
would be no more fish, since they reproduce sexually). In Italian, all nouns have
masculine or feminine gender, and this has no semantic basis at all in many
cases: ‘sun’ is masculine (il sole); ‘moon’ is feminine (la luna), ‘sincerity’ is
The full representations of both examples, with silent endings indicated, are
given in (22):
(22) a. [S [NP you ] [VP [V [V fish- ] [M Imperative ] [T Present ] [Agr 2 ]] [[D some ]
[N [N fish-[Num Pl ] ]
b. [S [NP tu ] [V [V pesc- ] [Conj -a- ] [M Imperative ] [T Present ] [Agr 2Sg ]] ]
[NP [D i ] [N [N pesc-[Num i ]]]
Both of these are full sentences where the first ‘fish’ is the subject, the second the
verb and the third the direct object. Clearly, what is added to the two-fish exam
ples discussed above is an explicit direct object, compared to (7), and an explicit
subject, compared to (9). Traditional grammars tell us that a complete sentence
consists of a subject, what the sentence is about, and a predicate, which says
something about the subject. The predicate contains the main verb and the direct
object. Thus we could represent (23) as follows (leaving aside for the moment
the representation of the silent inflections):
(25) [S [N Fish ] [Predicate [V fish ] [N fish ]]]
This sentence is both grammatical and interpretable, but, unlike the two-fish and
three-fish examples it requires a little thought. The intonation with which it is to
be read is important: it should be ‘ish FISH fish [short pause] FISH’, with stress
on the second fish and the strongest stress on the final fish. The structure here
is significantly more complex than our earlier examples. For this reason, from
now on I will mainly give simplified representations of these more complex
examples, leaving aside the silent inflectional material (although we should not
forget that it is there).
What does (28) mean and what is its structure? Again, the Italian translation
can give us an important clue:
(29) I pesci che i pesci pescano pescano.
The markers that and che introduce a relative clause, a clause which modifies a
noun; for our purposes, the only real difference between that and che is that che
appears obligatorily in this context, while that can be ‘dropped’ (i.e. silent). So
2
We will look in detail at the relation between grammatical categories like NP and VP and gram
matical functions like subject and predicate in Chapter 2 of Volume II.
the second two fish, the sequence noun-verb, constitute a relative clause mod
ifying the first noun fish (known as the head of the relative). The fourth fish is
the verb of the main clause, with an implicit direct object just as in the two-fish
example with the structure in (7). We can thus make a first pass at a structure for
(28) as follows:
(31) [S [N fish ] [Relative clause fish fish ] [VP [V fish ]]]
The relative clause is a kind of subordinate clause. In fact, we can tell that the first
fish inside the relative clause is the subject and the second fish is the verb of the
predicate. So we have an NP-VP structure here (remember that the presence of the
overt definite article in Italian shows us that the subject is an NP, not just a noun):
(32) [S [N fish ] [Relative clause [NP [N fish ]] [VP [V fish ]]] [VP [V fish ]]]
The main clause, i.e. the independent clause containing the relative clause, indi
cated as S here, must also have a subject NP, not just N. This NP includes the
relative clause, since the relative clause modifies the noun, so we have (33):
(33) [S [NP [N fish ] [Relative clause [NP [N fish ]] [VP [V fish ]]]] [VP [V fish ]]]
Both the main clause and the relative have a subject and a predicate. Identifying
predicates with VP again, (33) tells us that each VP contains just the verb fish. In
the case of the main clause, this may be correct, since the object here is implicit
and has a rather vague interpretation (roughly ‘stuff’). But the object of the rel
ative has a completely different, and very precise, interpretation: ‘fish’. Here
we see an important characteristic of relative clauses: there is always something
missing, often referred to as a ‘gap’, which corresponds semantically to the head
of the relative (which we could then call the ‘filler’). Since we know that syntac
tic representations contain silent elements (actually, we are beginning to see that
they mostly contain silent elements) and we want to keep the syntax-semantics
mapping as straightforward as possible, the best way to account for the interpreta
tion of the gap is that it is a silent copy of the filler. So we replace (33) with (34):
(34) [S [NP [N fish ] [Relative clause [NP [N fish ]] [VP [V fish ] [NP [N fish ]]]]] [VP [V fish ]]]
We said above that English that, unlike its Italian counterpart che, can be
‘dropped’. This implies that sentences with and without that are synonymous: that
and che are just meaningless syntactic markers (saying ‘what comes next is a rel
ative clause’ in our examples). However, (30) has another interpretation that (28)
doesn’t have. In (28) the gap inside the relative clause is the direct object, as (34)
shows; it means ‘fish that get fished fish stuff’. But (30) allows a further interpre
tation, where the gap can be interpreted as the subject of the relative clause: ‘fish
that fish other fish fish stuff’. This example has the following structure:
(35) [S [NP [N fish ] that [Relative clause [NP [N fish ]] [VP [V fish ] [NP [N fish ]]]]] [VP [V fish ]]]
Here it is the subject of the relative that is the silent copy of the head, and
the marker that is obligatory. The Italian translation of this English sentence is
unambiguous and should be compared to the one in (29):
IanRoberts_9781316519493_C001.indd 24 06-09-2022 11:48:49
1.4 Really Challenging Fish 25
As we saw, the four-fish example in (28) had an implicit direct object in the main
clause (‘Fish fish fish fish stuff’). The five-fish example in (37) has an explicit
direct object in the main clause. We can therefore give the structure as in (38),
and the Italian translation in (39):
(38) [S [NP [N fish ] [Relative clause [NP [N fish ]] [VP [V fish ] [NP [N fish ]]]]] [VP [V fish ]
[NP [N fish]]]]
By now we may begin to wonder just how many repetitions of the word fish (in
the phonology, the subpart of the syntactic structure that is actually pronounced)
can give rise to a grammatical, interpretable sentence. It is clear that the step
from three fish to four fish, involving as it did the introduction of a relative
clause into the structure, places a burden on comprehension. But this burden can
be readily overcome and the sentence is fully comprehensible; in fact we can
recognise that it describes a rather odd state of affairs, the very silliness of the
sentence proves that we can understand it. We see, then, that the comprehension
burden does not increase linearly as we add extra words; the structural complex
ity of the four-fish sentence is significantly greater than the three-fish one, and
this is reflected in our initial hesitation in understanding it.
In the next section, we’ll see some really challenging examples as we increase
the number of fish still further.
Most people are completely stumped by this sentence at first sight; it appears to
be simply a random repetition of the word fish with no meaning (and so, perhaps,
no structure) at all. But in fact this isn’t true, as we’ll see below. What is striking
though is that seven fish is easier to interpret than six:
(41) Fish fish fish fish fish fish fish.
We know from (28) that a sequence of three fish can be a relative clause with
an object gap, ‘fish that are fished by other fish’. In (28), the relative clause is in
the subject position. But relative clauses can be in object position too (in fact,
(28) can be interpreted that way too, a point that was left aside above). So (41)
has the approximate structure in (42a), given in fuller detail in (42b) (where RC
stands for relative clause, and the structure is presented as a tree diagram, as this
makes it easier to see the relations among the fish; we’ll look at tree diagrams
more systematically in the next chapter):
(42) a. [ [ Fish fish fish ] [ fish [ fish fish fish ]] ]
b.
Pronounced with the right intonation (‘Fish FISH fish [pause] fish fish FISH
fish’), this example is fairly easy to understand, especially in the light of the
four- and five-fish examples. But none of this makes (40), with just six fish, any
easier to understand.
In order to see what the structure, and hence the interpretation of (40) is, we
must go back to our earlier, simpler examples. Here, once again, is the two-fish
example with noun-verb interpretation, repeated from (7):
(7) [S [N fish ] [V fish ]]
Of course, we now know that N and V here are contained in their respective
NPs and VPs. Relative clauses are really a kind of sentence, as we pointed out
above. So let us label them as S from now on. We go from two fish to four fish
by putting the structure in (7) inside the subject NP:
(43) [S [NP [N fish ] [S [N fish ] [V fish ]] ] [VP [V fish ]]]
Actually it’s not quite accurate to say we put (7) inside the subject NP; really, we
put the three-fish sentence, with the structure seen in (26), inside the subject NP
and make the object a silent version of the head of the relative:
(44) [S [NP [N fish ] [S [NP [N fish ]] [VP [V fish ] [NP [N fish ]]]]] [VP [V fish ]]]
On the basis of (44), we can add two more pronounced fish by inserting a relative
clause inside the subject NP of the relative clause, i.e. after the second [N fish ].
This gives the string of six fish in (41), and shows us what the structure is:
(45) [S [NP [N fish ] [S [NP [N fish [S [NP [N fish ] [VP [V fish ] [NP [N fish ]]]]]]] [VP [V fish
] [NP [N fish ]]]]] [VP [V fish ]]]
If we add some relative markers, preferably which (another relative marker in
English alongside that), we can just about start to make sense of (45):
(46) Fish which fish which fish fish fish fish.
By now, we know a relative clause when we see one. The essential structure of
(47) has a relative clause inside the subject containing an object gap:
(48) [S [NP the mouse [S [NP the cat [VP chased [ the mouse ]]]]] [VP died ]]
(48) is quite easy to understand. The analogue to (41), but with different words, is (49):
(49) [S [NP the mouse [S [NP the cat [S [NP the dog [VP bit [NP the cat ]]]]]]] [VP chased
[NP the mouse ]]]]] [VP died ]]
Again, the jump from (48) to (49), caused by simply adding another relative
clause modifying the cat, poses severe difficulties of comprehension. Adding
some relative markers (that again) makes (49) considerably easier to understand:
(50) The mouse that the cat that the dog bit chased died.
green ideas sleep furiously; most people find this sentence unacceptable for the
perfectly good reason that it doesn’t make any sense. It is, however, syntac
tically well-formed, as its exact formal counterpart Revolutionary new ideas
spread quickly shows. Both sentences are grammatical in that they conform to
the rules of English syntax; the first one is semantically anomalous (in fact, it
is self-contradictory) and so unacceptable. This contrast also shows that syntax
and semantics are distinct in that a sentence can be syntactically well-formed but
semantically ill-formed; we will see more examples of this type in Section 8.3.
Of course, examples with centre-embedding like (50) become still more
difficult to understand when the same phonological word occurs six times over,
as in (40). There is, in principle, no difference in grammaticality between (40)
and (50) but a real difference in acceptability (they just go from bad to worse).
But now we can line (40) up with the variant of (50) without the relative mark
ers, and we can see how (40) works, and see that it is in fact grammatical:
(40) Fish fish fish fish fish fish.
(51) The mouse the cat the dog bit chased died.
Finally, here are the easier versions complete with the relative pronouns which
and that inserted in the appropriate places:
(50) The mouse that the cat that the dog bit chased died.
These examples have the structures shown in (49) and (45) respectively.
Relative clauses can be put one inside another without limit. They can just go
on and on. Consider (52):
(52) This is [NP1 the guy [ S1 who loved [NP2 the girl [S2 who befriended [NP3 the boy
[S3 who lived ]]]]]].
For those familiar with the Harry Potter stories, each NP here describes one
of the central protagonists by means of modification by a relative clause, each
relative clause is inside the NP it modifies, and each NP – except NP1 – is
embedded in the next relative clause. NP3 describes Harry Potter himself, the
boy who lived. NP2 describes Hermione Granger, the girl who befriended the
boy who lived, and NP1 describes Ron Weasley, the guy who loved the girl who
befriended the boy who lived. Relative clauses give language a great deal of its
expressive power, by making it possible to modify nouns using modifiers that
can be as complicated as we want: in principle, there is no upper limit to the
number of times one relative clause can be embedded inside another. This is
reflected in nursery rhymes like ‘The House that Jack Built’: the dog who bit
the cat who ate the rat who lived in the house that Jack built. Centre-embedded
relatives can also, in principle, be constructed without limit. The problem is
that, in practice, as we have seen, they quickly become very hard to understand.
Compare the following centre-embedded sequence with the excerpt from ‘The
House that Jack Built’ just given:
(53) [ The rat [ the cat [ the dog [ bit ]] [ ate ]] [ lived in the house that Jack built ]
To answer (54a), we could keep on trying longer and longer fish sentences.
There is certainly no reason to take seven, the most we have seen here (see (41)
and the structures in (42)), to be an upper limit. If we look at the examples with
even numbers of fish, (28) and (40), we see that they have the general form in
(55):
(55) a. N1 + N2 + V2 + V1 (four fish, (28))
b. N1 + N2 + N3 + V3 + V2 + V1 (six fish, (40))
(This is not their exact structure, as we have seen, but these simplified representa
tions make our point here.) Each pair Nn + Vn, for n = n forms a relative clause
modifying Nn-1. Hence, as Freidin (2012:13) points out ‘any sentence containing
an even number of fish from four onwards will have at least one grammatical
representation’. Furthermore, although each relative clause (i.e. each pair Nn +
Vn in (55) for n > 1) has an object gap, as we saw above (see (44) and (45)), V1
does not have an object. But of course it could have, and this object could be fish.
This is what got us from four fish to five fish and from six fish to seven fish. To
quote Freidin (2012:13) again: ‘As this procedure [adding an object to V1, IR]
can be applied to any sentence with an even number (n) of fish, there will be a
corresponding sentence with an odd number (n+1) of fish … Hence any number
of fish will correspond to a sentence of English.’ In principle, then, a sentence
containing any number of repetitions of the word fish is grammatical. Our lin
guistic competence allows us to understand such sentences, although more than
five iterations, with the exception of seven, give rise to sentences that are difficult
to understand in practice, i.e. sentences which may be unacceptable to varying
degrees. We are able to assign syntactic representations, more technically known
as structural descriptions, to the sentences, and from there recover the meaning.
The full representations of these sentences, as we saw in our discussion of the
Italian counterparts of the two- and three-fish examples, are quite complex. The
full representation of the seven-fish example in English would be (56):
(56) [S [NP [D the ] [N [N fish-[Num Pl ]]] [S [NP [D the ] [N [N fish-[Num Pl ]]] [VP [V [V
fish- ] [M Indicative ] [T Present ] [Agr 3Pl ]] [NP [D the ] [N [N fish-[Num Pl ]]]]]]
[VP [V [V fish- ] [M Indicative ] [T Present ] [Agr 3Pl ]] [NP [D some] [N [N fish-
[Num Pl ]]] [S [NP [D the ] [N [N fish-[Num Pl ]]] [VP [V [V fish- ] [M Indicative ]
[T Present ] [Agr 3Pl ]] [NP [D the ] [N [N fish-[Num Pl ]]]]]]
You don’t need to pick your way through all the details of this representation in
order to see two things. First, most of the structure is silent. Second, sentences
of this kind are highly complex objects. But remember, and this is the really
important thing, we quite unconsciously compute such complex sentence struc
tures all the time.
Concerning question (54b), it is highly unlikely that anyone would come
across sentences of this kind (outside of a linguistics class). But, as we have
seen, they can be fairly readily understood. It is extremely difficult to see how
any account of linguistic knowledge based on simply imitating, learning and
generalising routines (perhaps through processes of stimulus, response and rein
forcement, as in behaviourist psychology) could account for our general capac
ity to produce and understand, mostly instantaneously, sentences we have never
heard before and which may never have been uttered before.
Question (54c) is the question that goes to the heart of the matter. As native
speakers of English, we have very rich tacit knowledge, of even the silliest and
most awkward sentences, which we were almost certainly never ‘taught’ in any
meaningful sense of the word. Our experience in early childhood, combined with
our cognitive capacities (both language-specific and not), has made it possible
for us to perform the kinds of mental computations required to extract structure
and meaning from what appear to be merely strings of repetitions of the same
word. Of course, this ability is not restricted to native speakers of English: native
speakers of Italian can just as easily make sense of sentences like (24), (29), (36)
and (39), and the same exercise can be repeated, in principle, with any speaker
of any language.
The fish sentences clearly illustrate the distinction between competence
and performance. As we have mentioned, competence refers to a state of
knowledge, knowledge of an I-language, that a native speaker of a given lan
guage (in roughly the ‘E’-sense as discussed in the Introduction) possesses.
Performance refers to the application of I-language in a given situation: com
bined with other cognitive and social capacities, performance makes normal
speaking and comprehension possible. Freidin’s conclusion that ‘any number
of fish will correspond to a sentence of English’ clearly concerns competence;
the rules of English syntax are such that sentences of this kind can exist, but
of course, owing to performance limitations, no actual speaker could produce
or understand a sentence consisting of an infinite number of occurrences of
the word fish and nothing else. Since placing an upper bound on the number
of times one relative clause could be embedded in another would be arbitrary,
we regard the rules of English (and other languages) as allowing for this in
principle. Of course, the same applies to numbers: no individual can ever write
out an infinite number, although our mathematical competence tells us, and
this can of course be proved, that the series of natural numbers is infinite. In
Chapters 3 and 4, we will see the precise structure-building mechanisms that
achieve this result.
In this chapter we have tried to outline and illustrate, with the con
crete help of the fish sentences, what linguistic theory investigates. The central
goal is to elucidate what it means to be a competent speaker of a language. In
other words, linguistic theory is primarily about knowledge: what is it that a
person described as a native speaker of English (for example) knows? What
are the properties of individual I-languages and what are the properties of UG
that underlie them? Clearly this must involve the operations that are capable
of producing the kinds of structures we saw in relation to the fish examples. A
further central question is where this knowledge comes from: how do children
acquire their first language? Related to this is the question of how similar and
different languages are. We saw that we can ‘line up’ English and Italian quite
well, especially if we assume that English has a number of silent inflectional
‘satellite’ categories many of whose Italian counterparts are overtly pronounced.
In language-acquisition terms, this suggests that children are looking for evi
dence – semantic, syntactic, morphological or phonological – for the presence
of these categories. Many complex issues arise as we try to pursue these ideas,
but we saw in connection with the Italian fish examples that the combination
of semantic evidence and UG predispositions of various kinds can lead us to
significantly limit the differences between English and Italian. As we said there:
semantics is universal, syntax is near-universal and languages differ primarily in
their morphophonologies. We will return to these ideas repeatedly in the chap
ters to follow.
This is a formal, cognitive theory of linguistic knowledge. Each individual
over the age of five or so has a specific I-language which makes possible their
competence in their native language. This I-language corresponds more or less to
a sociocultural E-language construct such as English, Italian and so on. We also
saw that English and Italian speakers are I-alike but E-different: the I-languages
may turn out to be quite similar to each other, while many aspects of the cultures
borne by the corresponding E-languages are strikingly different (in ways that
people generally find fascinating). Given how little we understand about the
brain, as well as ethical restrictions on what kinds of experiments we can do on
healthy children or adults, we can only find out about what is inside the mind,
how an individual’s grammatical system works, from the observed outputs of
the I-language. This includes, quite simply, the things we say and hear (language
production and comprehension) along with judgements about grammaticality,
interpretation and so on.
We saw in the Introduction that there are three factors that make an I-language.
These are, first, the genetic endowment. Here we take this to be UG, containing
the various kinds of rules and systems of rules that build and interpret syntactic,
semantic and phonological representations. The details of what constitutes UG
will occupy us for much of what follows. The second factor is experience, in
Each of these questions relates I-language to one of the three factors in language
design that underlies it. In this connection, our fish sentences are very important.
They show us that syntactic structures are potentially infinite and yet built out
of very simple elements. In particular, the structures are formed by repeating
the same operations again and again. The next few chapters, as well as much of
Volume II, are devoted to demonstrating this in full detail. The fish have given
us a first inkling of the nature of the syntactic component of I-language, but now
it is time to look more systematically into the details.
Exercises
1. Consider again the four-fish example in (28), repeated here as (i):
(i) Fish fish fish fish.
We assigned the structure in (34), here (ii), to this sentence:
(ii) [S [NP [N fish ] [Relative clause [NP [N fish ]] [VP [V fish ] [NP [N fish
]]]]] [VP [V fish ]]].
We also observed that inserting the relative marker that between the
first two fish made possible a different interpretation, which gave in
(35) (= (iii)):
(iii) [S [NP [N fish ] that [Relative clause [NP [N fish ]] [VP [V fish ] [NP [N
fish ]]]]] [VP [V fish ]]].
In all of (i–iii) there is an implicit main-clause object, meaning
something vague like ‘stuff’. It is, however, possible to interpret (i)
as having an overt object. Try to see that interpretation, indicate its
structure (in a rough way, along the lines of (ii) and (iii)), its interpre
tation and, if you can, its approximate intonation. (HINT: look again
at the ambiguity of the two-fish example discussed in (7–9)).
2. Now look again at the five-fish example from (37) (= (i)):
(i) Fish fish fish fish fish.
We said that (i) is the four-fish example plus an overt direct object
for the main-clause verb, giving the structure in (38) (= (ii)):
(ii) [S [NP [N fish ] [Relative clause [NP [N fish ]] [VP [V fish ] [NP [N fish
]]]]] [VP [V fish ] [NP [N fish]]]]
But in fact (i) has a further interpretation, where the main-clause
object is a relative clause. Give the structure for this interpretation
of (i).
Furthermore, the relative clauses in both interpretations can be
marked with that, and in each case this gives rise to a further ambi
guity inside the relative clause. Explain the ambiguity and give the
relevant structures.
3. Give the structure for the six-fish example in (40) (= (i)):
(i) Fish fish fish fish fish fish.
Now try to add a further level of embedding to (40), using any lexical
items you like (except perhaps fish). This will correspond to an eight
fish sentence; give the structure for this one.
Further Reading
R. Freidin. 2012. Syntax: Basic Concepts and Applications. Cambridge: Cambridge
University Press.
IanRoberts_9781316519493_C001.indd 34 06-09-2022 11:48:50
There are ten! = 3,628,800 possible orders for this ten-word sentence, the vast
majority of which are ungrammatical, such as the following:
(2) a. *Hopes Goona that Vurdy will be ready for his dinner.
b. *Hopes that Goona Vurdy will be ready for his dinner.
35
Our theory must make the distinction between grammatical and ungrammat
ical strings and structures; it must tell us how and why (1) is grammatical
while the examples in (2) are ungrammatical. Furthermore it should tell us
when two sentences consisting of different words (and therefore meaning
different things) have the same structure. This is true for (1) and the exam
ples in (3):
(3) a. Ron said that Hermione would be happy with her homework.
b. Boris thinks that Emile must be angry with his behaviour.
c. Alex believes that Eric should be worthy of a medal.
The basic notions we use to do this are those of categories and constituents,
which we now look at in turn.
2.2 Categories
2.2.1 Introduction: Lexical and Functional Categories
Like many approaches to morphology and syntax, including tradi
tional grammar, we make a first distinction between two main types of catego
ries, which we will call lexical categories and functional categories.
The principal lexical categories correspond to some of the ‘parts of speech’
of traditional grammar: they are noun, verb, adjective, adverb and preposition,
abbreviated standardly as N, V, Adj, Adv and P (we saw nouns and verbs in
our fish examples in Chapter 1). These lexical categories are open: people can
and do invent nouns and verbs all the time, and even the most up-to-date online
dictionaries struggle to keep up. Words belonging to these categories have clear
non-linguistic semantic content: as we saw, fish, as a noun, denotes a certain
class of aquatic animals, while fish, as a verb, denotes the activity of hunting
these animals. New adjectives and adverbs are less readily coined, but they do
arise, e.g. downloadable, googlable, etc. Prepositions are the one lexical class it
is difficult to add to; in this and certain other respects, prepositions are more sim
ilar to functional categories. We will return below to the possibility of defining
syntactic categories semantically.
Furthermore, lexical categories do not vary greatly across languages. The dis
tinction between nouns and verbs, in particular, seems to recur in all known
languages (although there is some debate about this in certain cases). Adjectives
are not found in every language; some indigenous languages of North America
may lack them, for example. In other languages, e.g. the West African languages
Hausa and Igbo, they form a small, closed class. In languages where the class
of adjectives is small or non-existent, the semantic work of adjectives, roughly
describing qualities, is usually done by relative clauses (a bus which reds). It
is probable that prepositions are not universal: the World Atlas of Language
Structures (WALS) lists thirty languages lacking adpositions (a cover term
for pre- and postpositions), including several indigenous languages of North
theme over the coming chapters, and return to it armed with more empirical and
theoretical knowledge in Volume III.
How can we distinguish the various categories? This is quite a tricky matter,
and there are no absolute hard-and-fast diagnostic tests. For lexical categories in
particular, we can make use of four kinds of properties, relating to the main areas
of linguistic structure. Hence we can distinguish categories on the basis of their
morphology, their syntax, their phonology and aspects of their meaning. Let us
look at each of these in turn.
All of this makes it very easy to identify verbs in terms of their morphology. What
we see in Italian is typical of the Romance languages, and, with variations, com
mon across the Indo-European languages as a whole. Some languages show case
marking on nouns, which indicates the function of the noun (or really of the NP) in
the clause: nominative case marking typically marks subjects, accusative direct
objects, dative indirect objects, etc. So in Latin dominus (‘master’) has the -us
ending for nominative singular (in this class of nouns, known as the second declen
sion), -um for accusative singular, -o for dative singular and in fact three more cases.
In languages of this type, nouns can be identified by case marking (adjectives agree
with the nouns they modify, so they can too).
In Modern English, common adjectives can be identified by their ability to
form the comparative with the ending -er (tall – taller) and the superlative in -est
(tallest). This applies only to adjectives of two syllables or less; hence there is no
comparative beautifuller or superlative beautifullest. This constraint applies to
the root form of adjectives, so unhappier exists, since the root is the two-syllable
happy. As usual, there is a handful of irregular adjectives, e.g. good – better –
best, bad – worse – worst.
Regular adverbs are formed from the corresponding adjective by adding -ly,
e.g. beautifully, happily, badly. There are some irregular adverbs, e.g. well, and
some which simply do not add -ly, e.g. fast. Conversely, there are some adjec
tives which end in -ly, e.g. friendly. Prepositions are invariant: they do not inflect
at all in English.
So morphological criteria can identify lexical categories in English. In most
of the Romance and the other Germanic languages, where there is typically more
inflection than in English, these criteria are correspondingly more useful and
reliable. As we saw with our fish sentences in Chapter 1, English allows many
apparently inflectionless, and hence morphologically ambiguous, forms (we
suggested that this may be due to a large amount of silent inflection in English).
In (4a), the declarative has the order Subject – Auxiliary – Verb … . In (4b),
the interrogative shows the ‘inverted’ order of subject and auxiliary. Here the
auxiliary is will, indicating (roughly) future tense. If there is no auxiliary in the
In the declarative (5a), there is no auxiliary and past tense is marked by the -ed
ending on the verb. In (5b), the auxiliary do appears, marked for past tense as
did, and precedes the subject, while the verb has no ending. Placing the inflected
verb in front of the subject is ungrammatical (indicated by the asterisk * preced
ing the sentence):
(6) *Flooded Cambridge after all the heavy rain last spring?
For this reason, readers familiar with Shakespeare and other writers from the
seventeenth century or earlier may find that examples like (6) have a certain
Shakespearean ring to them, but they are not really part of Modern English.
Following traditional grammar, we have said that clauses are divided into a
subject and a predicate. The predicate is usually a Verb Phrase, VP. Other cate
gories are able to function as predicates, though, providing we include the aux
iliary be. We can see this in examples like the following, where the predicative
category is labelled in the way we saw for the fish sentences in Chapter 1 (AP
stands for Adjective Phrase, PP for Prepositional Phrase):
(8) a. John is [AP nice ].
b. John is [AP interesting ].
c. John is [PP in a bad mood ].
d. John is [NP a nice person ].
e. John is [VP sleeping ].
In (8e), the verb sleep appears in its ‘progressive’ form marked with -ing, indi
cating an ongoing situation.
In addition to following be, predicative categories can also follow certain
verbs, such as seem.
An important distributional property of VPs is that, unlike predicative APs,
PPs or NPs, they are unable to directly follow verbs like seem, as the ungram
maticality of (9e) shows:
(9) a. John seems [AP nice ].
b. John seems [AP interesting ].
c. John seems [PP in a bad mood ].
d. John seems [NP a nice person ].
e. *John seems [VP sleeping ].
So this test picks out VPs, since only VPs are ungrammatical following seem.
One important distributional test which picks out NPs is that only NPs can be
subjects. Thus only an NP (which may consist of just a single noun or pronoun)
can appear in the blank in (10):
(10) ___ can be a pain in the neck.
From (10), we can form the sentences in (11a) by inserting a single noun or pro
noun (representing a whole NP), or those in (11b) where we see more complex
NPs, but not those in (11c), where the words inserted are not nouns:
(11) a. You/kids/injections/syntax/Dave can be a pain in the neck.
b. Professors of Linguistics/other people’s kids/injections which go wrong/
fish fish fish can be a pain in the neck.
c. *Walk/tall/in can be a pain in the neck.
Similarly, only VPs (which can be just a single verb) can appear between an
auxiliary and a manner adverb, i.e. in the slot in (12):
(12) Students can ___ quickly.
In (13a), we have single-verb VPs in that position and in (13b) more com
plex VPs. In (13c) the words inserted are not VPs and so the sentences are
ungrammatical:
(13) a. Students can talk/write/learn/understand quickly.
b. Students can dissolve in sulphuric acid/get married/conclude that you’re not
worth listening to/fish fish quickly.
c. Students can *Olly/*kids/*injections/*syntax/*tall/*in quickly.
The distributional tests we have seen so far allow us to identify NPs, VPs and
auxiliaries.
Here is a distributional test for APs. Only (gradable) APs can appear in the
blank slot in (14):
(14) The students are very ____ .
In (15a) we see simple, one-adjective APs in that slot, in (15b) we have complex
APs of various kinds, while (15c) shows that other categories cannot appear there:
(15) a. The students are very intelligent/diligent/nice/eager.
b. The students are very much more intelligent than I expected/diligent in
handing in their essays/nice to talk to/eager to please.
c. The students are very *from/*walk/*Olly.
Finally, a test for PPs is that they can be identified by their characteristic intensifiers
such as straight and right. Thus only PPs can appear in the blank slot in (16):
(16) John walked straight ____ .
In (17), we have the same pattern as in (11), (13) and (15). Example (17a) gives
simple PPs (intransitive PPs, containing just a P); (17b) illustrates PPs of various
kinds, almost always with an NP object, and (17c) shows that APs, NPs and VPs
cannot appear in this slot:
(17) a. John walked straight out/on/up/in.
b. John walked straight out of the room/on to his destiny/up the hill/into the
pub.
c. John walked straight *red/*talk/*Olly.
We see that syntactic, distributional tests can isolate the main lexical categories,
NPs, PPs, VPs and APs. In an inflectionally poor language like English, these
are probably the most reliable category diagnostics available.
In (18a), increase is a verb and has stress on the second syllable (shown by the
bold letters). In (18b), increase is a noun and is stressed on the initial syllable.
Another case where phonology indicates an aspect of syntactic structure concerns
the difference between word stress and phrasal stress. Hence blackbird, with
stress on the first syllable, is a word, in fact a compound noun. This noun denotes
a particular species of bird. On the other hand, black bird, with stress on bird, is an
NP (or part of an NP; strictly speaking there should be a determiner or plural mark
ing). This NP denotes any bird which happens to be black, following the usual rules
for attributive modification in English (so, red bus denotes any bus which happens
to be red, and so on). Thus one can say That black bird is not a blackbird without
self-contradiction, which would be impossible if the different stress patterns did not
indicate different structures (NP vs N) and therefore different meanings.
There is something to these definitions, although they are rather vague and lim
ited. On their own, they scarcely suffice to identify categories. Take, for example,
‘what someone does’ as part of the first definition just given, or ‘action’ in the
second and third definitions (after all, actions are what people do). The word
action clearly denotes ‘action’, but it’s a noun. Similarly for occurrence, also a
noun, or ‘state of being’ (existence is a noun). Or consider the following example:
(19) The economy worries John.
Here worries is a verb, but it’s neither an action nor a ‘state of being’ (either
of the economy, or of John). Is it an occurrence? That seems a rather difficult
question to answer, but it’s not clear that answering either way would shed much
light on things. Similarly, nothing is a noun, but it doesn’t denote a thing, and it
certainly doesn’t denote a place or a person.
Despite the vagueness of these definitions, we can discern some semantic
component in category distinctions. This can be seen relatively clearly in English
where, as we have repeatedly noted by now, words can belong to different cate
gories without changing form. For example, consider the different meanings of
round in the following examples:
(20) a. the round church
b. Round the rugged rock the ragged rascal ran.
c. These cars round corners very nicely.
d. Time for another round.
2.2.6 Conclusion
We have now seen several different ways of identifying syntactic
categories. None of them is absolutely foolproof, but taken together, they are
fairly reliable. Arguably, the distributional syntactic criteria are the most reliable
for English. In languages with richer inflectional morphology, such as Italian,
morphology can be more informative than in English.
There is little doubt that the diagnostics for categories vary from language
to language. But do the categories themselves vary? Does every language have
the same category inventory? In discussing this question at the beginning of
this section, we pointed out that the noun–verb distinction may be universal,
while adjectives and prepositions probably are not. Moreover, there appears to
be variation in the inventory of functional categories a language can have. We
might want to include Gender as a functional category, a ‘satellite’ of N, in
Italian given the presence of grammatical gender in that language, but we would
probably not include it in English, since there is no grammatical gender. This
reasoning might lead us to exclude quite a few familiar functional elements,
determiners for example, from our grammar of Chinese. These questions remain
open and, as we said above, we will revisit them in Volume III, when we have
more facts and more theory to bring to bear on them. For now, we will assume
that English has the lexical categories N, V, Adj, Adv and P and at least the func
tional categories Aux(iliary), D(eterminer) and C(omplementiser). Furthermore,
our conception of the nature and inventory of the functional categories will be
successively revised as we go on.
Taking our cue from how we analysed (21d) in (7) of Chapter 1, we can take
each of these sentences to consist of at the very minimum a noun and a verb, and
assign them the labelled bracketings in (22):
D
a mouse
(Here NP1, N1, NP2 and N2 are numbered simply to keep them distinct; the num
bers have no theoretical significance.) It is important to see that tree diagrams
and labelled bracketings present exactly the same information in typograph
ically different ways. The choice between them is a matter of convenience.
However, most people find trees easier to work with, since the relations among
the elements of the tree are more readily visible than in labelled bracketings.
But this is just a matter of personal preference; nothing theoretical depends on
it. In what follows, I will mostly use trees to represent structural descriptions.
We can now introduce some terminology that relates to the parts of the tree
diagram. S, NP, VP, etc. are the nodes of the tree; the nodes are linked by
branches. Branches never cross and all emanate from S, which is often referred
to as the ‘root’ node (in this sense the tree is in fact upside down in relation to
botanical trees, since the root is depicted as being at the top; again this is purely a
matter of convention). The words are terminal nodes; the category symbols (S,
NP, VP, etc) are non-terminal nodes. Pursuing the tree comparison, one could
think of the terminal nodes as the leaves.
The most important relations in the tree are the vertical, hierarchical ones.
The two fundamental relations are dominance and constituency. A category A
dominates another category B just where A is both higher up in the tree than B
and connected to B. We can put this more precisely as follows:
(27) A node A dominates another node B just where there is a continuous path of
branches going down the tree from node A to node B.
Let us now apply these definitions to the tree diagram in (26), which I repeat
here for convenience:
D
a mouse
2.4 Conclusion
The central goal of the theory of syntax is to account for how words
are grouped into larger units, phrases and sentences. Categories and constituents
are the building blocks for analyses. In this chapter, we have seen how to test for
and isolate various categories of English. None of these tests is foolproof, espe
cially when taken alone, but taken together they do allow us to identify and dis
tinguish categories. The second thing we have seen here are the ways in which
we can present the structural description of a sentence: labelled bracketings and
Exercises
A. Categories
1. Assign the following words to a syntactic category (noun,
verb, adjective, preposition), using the tests discussed in
this chapter along with any others you know of. In your
answer, show not only the category, but also how you
apply the tests:
cat, moon, sing, fish, to, possibly, well, dispute, round
2. Now do the same with the following words: noble, enno
ble, near, asleep and around
B. Constituents
3. Give the labelled bracketings and tree diagrams for the
following examples:
a. Boris snorted.
b. The onx splooed the blarrg.
c. Fish fish fish fish.
d. Vurdy saw Goona.
IanRoberts_9781316519493_C002.indd 48 03-09-2022 06:53:58
Exercises 49
about
NP
Further Reading
All of the references below discuss the ways in which we can distinguish categories,
using tests broadly similar to those discussed here and, as here, concentrating largely (but
IanRoberts_9781316519493_C002.indd 49 03-09-2022 06:53:59
50 2 constituents and categories
not exclusively) on English. All of them therefore provide useful backup, and plentiful
further examples, to the ideas introduced here.
3.1 Introduction
In the previous chapter, we saw the basics of the structural descrip
tions of sentences, the essential features of syntactic representations. Structural
descriptions can be presented as labelled bracketings, as in (1a), or as tree dia
grams, as in (1b):
(1) a. [S [NP [N Clover ]] [VP [V ate ] [NP [D a ] [N mouse ]]]].
b. S
VP NP
NP N
V
Clover ate N
D
a mouse
We also introduced basic tree terminology: node, root node, terminal node,
non-terminal node, branch, etc., and the very important mutually definable rela
tions of (immediate) dominance and (immediate) constituency. In addition to
constituent structures, as represented in trees and labelled bracketings (i.e. struc
tural descriptions), categories are the other fundamental notion. We illustrated
the main categories of English, as well as the very important distinction
between lexical and functional categories, and showed how categories can be
fairly relia bly isolated by a battery of syntactic, semantic, morphological and
phonological tests (which vary somewhat from language to language).
The goals of this chapter are to start to describe (and explain) the difference
between grammatical and ungrammatical sequences. In order to do this, we
introduce the mechanisms that generate structural descriptions, i.e. the rules
that build trees and labelled bracketings. These are the Phrase-Structure
rules. The exact Phrase-Structure rules and the structural descriptions they
generate can be justified by independent tests designed to isolate and
distinguish the constituents in a tree or labelled bracketing; these are the
constituency tests. We look first at the Phrase-Structure rules and then at some
constituency tests which, for the most part, work quite well in English.
51
The question is: how do we, as speakers of English, know that these sen
tences are ungrammatical? Clearly we don’t merely store this information
for each individual sentence or we wouldn’t have grammaticality judgements
about novel sentences such as the fish sentences we saw in Chapter 1. What
is needed is some general schema determining which are the grammatical
sentences of English (that is, of the I-language of the typical native speaker
of English).
To put it more technically, we need a mechanism to generate the well
formed structural descriptions of English sentences. The mechanism in ques
tion is the set of Phrase-Structure rules, PS-rules for short. PS-rules are the
formal devices which generate constituent structure, by specifying all and
only the possible ways in which categories can combine. An example of a
basic PS-rule of English (and probably many other languages) is given in
(1), where S stands for Sentence;
(3) i. S → NP VP
The brackets indicate optional categories. Strictly speaking (4ii), for exam
ple, abbreviates VP → V, VP → V NP, VP → V PP and VP → V NP PP, but
we collapse these rules under (4ii) using the bracket notation. These rules can
generate the structural description for the sentence Clover ate a mouse, given
as a tree diagram in (1a) and a labelled bracketing in (1b), and similar ones.
Here is another example of a tree diagram (which we saw in the Exercises
sections in the previous chapter):
(5) S NP
VP
3.2 Phrase-Structure Rules 53
N V PP
Ron thought
P D N the problem
about
NP
This tree can be generated by the PS-rules in (6), which combines the rules in
(3) and (4):
(6) i. S → NP VP
ii. VP → V (NP) (PP)
iii. NP → (D) N (PP)
iv. PP → P NP
Each rule generates a small piece of tree, typically (but not always) a branch
ing node and the constituents which that branching node immediately domi
nates. As such, each rule specifies a set of immediate dominance/constituency
relations, such that the category to the left of the arrow is the label of the node
which immediately dominates the constituent(s) to the right of the arrow, and
the category or categories to the right of the arrow label the node(s) which
is/are immediate constituent(s) of the category to the left of the arrow. So,
PS-rule (6i) generates the structure in (7):
(7) NP VP
S
PS-rule (6ii) has two categories in brackets on the right of the arrow. The
brackets indicate that the categories in question are optionally immediate con
stituents of VP (another way to say this is to say that they are optionally part
of the ‘expansion of VP’, given that PS-rules generally expand the category
to the left of the arrow by specifying more than one category on the right).
Strictly speaking, (6ii) really collapses the four distinct PS-rules seen in (8):
(8) a. VP → V
b. VP → V NP
c. VP → V PP
d. VP → V NP PP
Like rule (6ii), rule (6iii) collapses four rules using the bracket notation. These
are given in (11):
(11) a. NP → N
b. NP → D N
c. NP → N PP
d. NP → D N PP
NP
NP DN
The pieces of structure in (7), (9), (10) and (12a,b) combine to form the tree in
(5). In (13), (5) is repeated with each part of the structure annotated giving the
PS-rule which generates it:
(13) NP rule (8c) (one
rule VP subcase of (6ii))
S
rule (6i)
(11a) V
N PP
Ron thought P
rule (11b) (one subcase
NP
of (6iii))
rule (6iv)
D N the problem
about
This illustrates how the structural description of the sentence is generated by
the PS-rules. The tree diagram in (13) gives the structural description of the
grammatical English sentence in (14):
In (2a) the verb and the subject are in the wrong order for English. PS-rule (6i),
S → NP VP states “Rewrite the symbol S as the sequence NP VP in that order.’
Whatever expansion of VP we choose from the options in (6ii) (see (8)), V is
always the first immediate constituent of VP. It follows from the combination of
rules (6i) and (6ii) that the subject NP (the NP in (6i)) will precede the verb. Hence
the sentences in (2) cannot be generated by our PS-rules where the immediately
postverbal NP is the subject; if it is interpreted as the object, then the subject is
missing and again the rules in (6) cannot generate the sentence. Our PS-rules are
examples of the PS-rules of English, specifying all and only the well-formed struc
tural descriptions of English sentences, and so the sentences in (2) are ill-formed.
It is important to see that the notion ‘ill-formed’ here means ‘not generated by
the syntactic rules of English’. Our aim is to make that notion coincide as far as
possible with native speakers’ intuitive judgements regarding the grammaticality
of English sentences, which we take to reflect their I-language competence. Here
it is important to remember what we saw in our discussion of centre-embedding in
Chapter 1: native-speaker judgements reflect performance, i.e. acceptability, rather
than competence, i.e. grammaticality. As we mentioned there, though, most of the
time grammaticality and acceptability coincide, and so most of the time matching
native-speaker judgements is a good benchmark for how well our theory is doing.
As we can see from our account of the ungrammaticality of (2), PS-rules
give information about the linear (left-to-right) order of nodes, including the
words at the terminal nodes. They also give information about hierarchical
structure, in that they specify immediate dominance and constituency rela
tions. Finally, since the symbols they use are category symbols (S, N, V, P,
etc.), they give information about the category labels of nodes. So PS-rules
specify three kinds of information about structural descriptions simultane
ously. We will see in Volume II that these three rather distinct kinds of
information can be teased apart.
3.3 Recursion
One of the most important things our discussion of the fish sentences
in Chapter 1 showed us was that the length of natural-language sentences is in
principle unbounded. If infinitely long sentences are grammatical (but maybe
not acceptable, as they can never be ‘performed’), they should be well-formed.
In other words, they should be generated by the PS-rules. Let us see how.
In fact, the rules we already have can generate infinite structures. Consider
again rules (6iii) and (6iv):
It is easy to see that NP appears on the left of the arrow in (6iii) and on the right
of the arrow in (6iv). Given the way PS-rules work, this means NPs can appear
inside other NPs, i.e. NPs can be constituents of other NPs, as illustrated in (15):
generated by (6iii)
(Again, the subscript numbers on the Ns and NPs serve merely to keep the two
occurrences of this category distinct; they have no theoretical significance.)
Since nothing requires us to apply rules (6iii) and (6iv) in that order, after
applying both rules so as to generate the structure in (15), we can go back and
, thereby intro
apply rule (6iii) again so as to expand NP2 identically to NP1 ducing a second PP,
which we can expand by rule (6iv) to give a third NP,
which we can expand by rule (6iii) to give a fourth NP, which we expand by
rule (6iv) to give a third PP, and so on. In principle, there is no limit to how
many times we can keep applying rules (6iii) and (6iv) to their own output.
The result of iterated application of these rules is complex recursive NPs of
the following kind:
(16) [NP [D the] height [PP of [NP [D the] lettering [PP on [NP [D the ] covers [PP of [NP [D the
] manuals [PP on [NP [D the ] table [PP in [NP [D the ] corner [PP of [NP [D the ] room …
The ability of rules to apply to their own output is known as recursion, a con
cept that originates in mathematics and logic. Rules (6iii) and (6iv) are not indi
vidually recursive, but together they form a recursive rule system. Recursion is
an extremely important concept, as it gives rise to the possibility of sentences
of unlimited length and underlies the fact that human languages are able to
make ‘infinite use of finite means’. As we have just seen, we can construct an
infinitely long NP using just rules (6iii) and (6iv). If our minds contain a recur
sive rule system of this kind for generating sentences in our native I-language,
then our finite brains have infinite capacity. This is clearly a very important
and interesting claim about human cognition, one which justifies looking at
language from the formal and cognitive perspective adopted here.
In fact, we already saw recursion in action in the fish sentences in Chapter 1.
Consider the structural description we gave there for the seven-fish example:
(17) [S [NP [N fish ] [RC [NP [N fish]] [VP [V fish] [NP [N fish ]]]]] [VP [V fish ] [NP [N fish ] [RC
[NP [N fish]] [VP [V fish] [NP [N fish ]]]]]]]
The relative clauses introduce NPs inside other NPs and the object relative
clause introduces a VP inside the main-clause VP. Since, as we said just after
introducing the representation in (17), relative clauses are really sentences,
i.e. of category S, we should substitute S for the RC labels in (17). Then
we see that relative clauses introduce Ss inside Ss. So the reason we can
have infinite fish sentences is that the PS-rules generating relative clauses
are recursive. They feature occurrences of the symbols S, NP and VP on both
sides of the arrow, just as rules (6iii) and (6iv) do for NP and PP.
Relative clauses are one kind of subordinate clause. I won’t give the PS-rules
that generate them here as certain details of their structure remain uncertain and
controversial. Another kind of subordinate clause whose structure appears to be
much more straightforward, however, are complement clauses. A simple com
plement clause is illustrated in (18):
(18) Goona hopes [ that Vurdy arrived].
(19) i. VP → V S′
ii. S′ → Comp S
Rule (19i) adds to the various possible expansions of VP seen in (8). We could add
‘(S′)’ to rule (6ii); this would subsume (19i) and predict that the sequence V NP S′
is possible (it is, as in persuade Mary that John arrived), the sequence V PP S′ is
possible (it is, as in say to Mary that John arrived) and the sequence V NP PP S′
is possible (this may not be correct, but I will leave this complication aside here).
In (19ii) ‘Comp’ abbreviates ‘complementiser’. Rule (19ii) states that Comp and S
are immediate constituents of the subordinate clause S′. S′ and S are not the same
category: S is a clause and constitutes the root node of a tree, while S′ is the label
of a subordinate clause introduced by rule (19i). Thanks to rule (19ii), we have a
further case of S-recursion: S appears on the left of the arrow in rule (6i) and to
the right of the arrow in rule (19ii). These rules will together generate structural
descriptions with S inside another occurrence of S.
We can expand S introduced by rule (19ii) as NP VP (in fact, this is the only
expansion of S our PS-rules allow as we have formulated them so far). We can
then expand VP using rule (19i) and introduce S again by rule (19ii) and expand
S again as NP VP, reapply rule (6i), reapply rules (19i) and (19ii), and so on
without limit. Once again we see rules applying to their own output, the basic
property of recursion. Recursive application of these rules in this way gives rise
to unlimited sequences of subordinate clauses embedded inside one another.
This is observed in English examples like (20):
(20) Mary hopes that John expects that Pete thinks that Dave said that …
As far as competence is concerned, just as with the fish sentences and the
complex NP in (16), there is no limit to the length, or to the depth of embed
ding, of sentences like (20). The usual performance restrictions mean that
infinite sentences cannot be uttered or written down, but three PS-rules ((6i),
(19i) and (19ii)) suffice to generate them.
The structural description of (18) in tree format is given in (21):
As indicated here, both subordinate clauses are S′s, with whether and for in C.
What follows whether in (22a) can stand alone as a complete sentence (Vurdy will
arrive on time), and so there is no difficulty with applying rule (19ii) here. The
difference with the that-complement in (18) is that the whole subordinate S′ stands
for the direct interrogative Will Vurdy arrive on time? What we observe in indirect
interrogatives is the absence of subject-aux inversion (briefly discussed in Chapter
2, see examples (3) and (4) there): in the S′ in (22a) Vurdy precedes will just as in
a declarative. We also observe that the presence of whether contributes what we
could call ‘interrogative force’, i.e. it makes the declarative into an interrogative.
In (22b), what follows the complementiser for cannot stand alone as a com
plete sentence: *Vurdy to arrive on time. We clearly have a subject (Vurdy) and
a predicate (arrive on time) as in (18) and (22), but what is missing is tense: there
is no finite verb or auxiliary. ‘Complete’ clauses, those able to stand alone, must
have a finite verb or auxiliary to mark tense (although the tense marker may be
silent in English, as in the following complete sentences: I/you/we/they/the boys
arrive). The particle to marks non-finiteness; it shows up in almost all infinitives
in English. It looks as though we cannot generate *Vurdy to arrive on time with
rule (6i), since there is no place for to. But it seems clear Vurdy is the subject NP
and arrive on time the VP. Assuming that to is not part of the subject NP, which
seems highly unlikely, there are two possible structural descriptions for *Vurdy
to arrive on time, which are shown in the labelled bracketings in (23):
(23) a. [S [NP Vurdy] [? to ] [VP arrive on time ]]
b. [S [NP Vurdy ] [VP to arrive on time ]]
Rule (6i) cannot generate (23a), but it could generate the NP VP structure in
(23b). But now rule (6ii) (with its various expansions) cannot generate the VP
there, which has to as its first constituent. In order to accommodate infiniti
val subordinate clauses, one of the two basic PS-rules in (6) will have to be
modified. I will leave this question open here, but we will come back to it in
Chapter 4 (see Section 4.4).
Combining (6) and (19), we now have the following set of PS-rules:
(24) i. S′ → Comp S
ii. S → NP VP
iii. VP → V (NP) (PP) (S′)
iv. NP → (D) N (PP)
v. PP → P NP
These rules can generate a large number, in fact, given their recursive nature,
an infinite number of English sentences. However, they do not generate all
the grammatical sentences English: we have seen that there is no place for the
infinitive to here, so we cannot generate (22b).
As we said at the end of the previous section, PS-rules give us three kinds of infor
mation: (i) information about hierarchical structure (‘vertical’ information, in terms
of tree diagrams), information about linear precedence (‘horizontal’ information)
and information about the category labels of nodes. These rules are very powerful
formal devices which are capable of generating structural descriptions in a precise
way. But how do we know that structural descriptions of the kind we have been
looking at in this chapter are the right ones for English? The descriptions make very
clear claims about constituent structure, but are they correct? We have also seen at
least one example where it is not clear where to place a constituent in a structure: the
question of where to put to in infinitival clauses like (22b). Which of the structures in
(23) is the correct one, and why is the possibility that to is a constituent of subject NP
‘highly unlikely’ as I said above? In the next section we will turn to questions of this
kind, by showing how there are various ways of testing constituent structure, just as
there are various diagnostics for categories as we saw in Section 2.2.
(6i) S → NP VP
This rule applies to give us the constituent structure (25b), rather than (25c),
for (25a):
(25) a. John ate the cake.
b. [S [NP John ] [VP ate the cake ]]
c. [S [VP John ate ] [NP the cake ]]
But why do we write the rule like this? What would be wrong with writing the
rule as (6i′), which would give the constituent structure we see in (25c) for (25a)?
(6i′) *S → VP NP
One obvious answer comes from the traditional idea that clauses consist of a
subject and a predicate. But this venerable idea does not really tell us anything
about phrase structure (despite what we said in Chapter 1): phrase structure rep
resents hierarchy, order and categories, as we have seen. It does not represent
grammatical functions or relations such as subject and predicate. From what we
have seen up to now, these notions have no place in our theory; we may wish to
build them in somehow, and we will do this in Section 5.2 (since these notions
seem to be so useful for informal discussion we might want to have a theoretical
way to understand them, but the fact remains that we have not actually said any
thing about this so far). We will look at the question of the theoretical status of
grammatical functions in full detail in Chapter 1 of Volume II.
So, coming back to choosing between (6i) and (6i′), what this really amounts
to is choosing between the structural description in (25b) and that in (25c) for
(25a). In (25b), which is the analysis we have been assuming up to now, the
verb and the direct object form a constituent that does not include the subject,
while in (25c) the subject and the verb form a constituent that does not include
the object. The question is: why should we prefer (25b) over (25c)? To put the
question another way, what is the evidence that ate the cake is a constituent and
John ate is not a constituent? The evidence is not directly audible on the basis
of what we hear as (25a); this is because constituent structure, like most of syn
tax, is silent. The linear order in (25a) is clearly compatible with either (25b)
or (25c). So we must find a way to determine what the hierarchical structure is.
The same question arises in relation to the NP P sequence in (26a) or the N A
sequence in (26b):
(26) a. They gave [NP the book ] [P to ] Mary.
b. They gave [NP Mary ] [AP sweet ] cookies.
Constituency tests of various kinds can show us to a large extent what the cor
rect constituent structures are. These tests are manipulations of sentences which
are sensitive to phrasal categories such as NP, VP, etc. There are several kinds
of constituency tests. These involve: (a) manipulating the order of elements in
such a way as to show that certain sequences of words must be manipulated
together, i.e. that they constitute a phrase (these are clefting, wh-movement
and fronting); (b) substituting a sequence of words with ‘pro-forms’ of various
kinds; (c) ellipsis, deleting a sequence of words in such a way that its interpreta
tion is recoverable from the linguistic context; (d) coordination, conjoining two
phrases of the same category with and, and (e) fragments, whether a sequence
can stand alone and be in an intuitive sense ‘complete’, even if it is not a com
plete sentence. We will now look at each kind of test in turn. These operations
are cases of a class of rules distinct from PS-rules, transformational rules, one
type of which (wh-movement) we will focus on in Chapter 5.
Let us begin with wh-movement. Fronting a phrase containing a wh-word (i.e.
the interrogative pronouns and determiners who, what, which, etc.; see Chapter
5) also targets constituents. This operation, known as wh-movement, is a very
important one for syntactic theory, and we will introduce it fully in Chapter 5.
Again, there is a gap in the sentence where the questioned phrase was, which we
mark with a ‘t’ (and which can also be taken as a silent copy):
(27) a. Which friends does Mary hope that John will like t ?
b. What did the Party Chairman send t to John?
c. Who did the Party Chairman send a book to t ?
d. To whom did the Party Chairman send a book t ?
These examples show us that the direct objects in (32a,b) are constituents, that
the NP following the P to in (32c) is a constituent, and that the PP to whom is a
constituent.
Non-constituents, i.e. strings of words that do not form an independent, unique
phrase, bolded in (28), cannot be fronted:
(28) a. *What to did the Party Chairman send t John?
b. *What to John did the Party Chairman send t ?
Strictly speaking the ungrammaticality of (28) does not give us information about
constituency; only the successful cases of wh-movement do this. Wh-movement
does not apply to VP: the various wh-phrases correspond to different grammatical
categories: who, what are NPs (respectively animate and inanimate), which is a
D, why an AdvP or PP (‘for what reason’), when a temporal NP or PP, where a
PP, how an AP or AdvP and how (many) a measure expression. But there is no
wh-word which questions a VP. Hence, we cannot use wh-movement as a diagnos
tic for a VP constituent. It does not follow from this that there is no VP constituent.
A further permutation we can apply to sentences in order to isolate constitu
ents is fronting. This operation ‘highlights’ phrasal constituents by placing them
at the beginning of the sentence. The fronted phrase often functions as a topic
which ‘is commented on’ by the rest of the sentence. So, from the neutral sen
tence (29a) we can derive (29b), where the direct object the new car is fronted:
(29) a. Mary hopes that John will like the new car.
b. The new car, Mary hopes that John will like t.
Again, we see a trace in the position where the direct object would normally be
in (29b). Fronting can apply to CP and VP, as in (30):
(30) a. That John will like her friends, Mary hopes t.
b. (Mary hoped that John would like her friends) … and [VP like her friends ]
he did t.
For (30b) to sound natural, it helps to give some context, as shown here.
Sentence (30a) may sound a little stilted, but it certainly seems acceptable. The
examples in (30) should be contrasted with (31) (where again non-constituents
are bolded):
(31) a. *A present to, the Party Chairman sent t John.
b. *A present to John, the Party Chairman sent t.
The sentence in (32a) ‘neutral’. In (32b), the sequence to John has been clefted;
in (32c) just the NP has been clefted, ‘stranding’ the preposition to. The general
schema for clefting is given in (33):
(33) S → It was XP that S.
In (34a), the NP the Party Chairman is clefted, while in (34b) it is the NP the
present. So the clefting test has isolated three constituents for us: [PP to John ],
[NP the Party Chairman ] and [NP a book ].
If we try to cleft non-constituents, bolded in (35), the result is ungrammatical:
(35) a. *It was the Party Chairman sent that t a book to John.
b. *It was a book to that the Party Chairman sent t John.
In (35a) the sequence the Party Chairman sent is clefted, and in (35b) a book
to. In both cases the result is ungrammatical. Since there could be independent
reasons why clefting some constituents is not good, the failure of a constituency
test does not really tell us anything. For example, clefting VP yields a rather odd
result (although probably not as bad as (35)):
(36) ??It was send a book to John that the Party Chairman did.
So (37) is interpreted to mean ‘John hopes that he, John, will win.’ Since John is
an NP, we should really call pronouns NPs too, since they stand for whole NPs
rather than just nouns. We can see this from the two cases of pronoun substitu
tion in (38b) and (38c):
(38) a. [NP The man who wears glasses ] hopes that he will win.
b. * The he who wears glasses hopes that he will win.
c. He hopes that he will win.
Other categories have pro-forms too. Most VPs can be replaced by (do) so,
as in (39):
(39) John has promised Mary a book, and Bill has done so too.
Here we see the VP promised Mary a book in the first conjunct, i.e. the verb
along with both the direct and the indirect object, are replaced in the second
conjunct by done so. As with pronoun (i.e. pro-NP) substitution in (38), do so
stands for the whole VP:
(40) a. * … , and Bill has done so Mary too.
b. *... , and Bill has done so a book too.
c. *... , and Bill has done so Mary a book too.
In (40a), just the verb and the direct object (promise a present) have been sub
stituted; in (40b), just the verb and the indirect object (promise Mary), and (40c)
just the verb promise. Again, the fact that these examples are all ungrammatical
doesn’t necessarily imply that the substituted categories are not constituents (in
fact, the verb on its own clearly is a consitituent); what it shows is that do so
substitutes an entire VP (and not, for example, V).
In the light of the conclusion that do so substitutes for a whole VP, consider
the following examples:
(41) a. John put his car in the garage on Tuesday, and Peter did so too.
b. *John put his car in the garage on Tuesday, and Peter did so on the driveway
on Wednesday.
c. *John put his car in the garage on Tuesday, and Peter did so his bike on
Wednesday.
d. John put his car in the garage on Tuesday, and Peter did so on Wednesday.
The grammaticality of (42a) also shows us that the temporal PPs are not required
for grammaticality. Such PPs are optional modifiers, or adjuncts, giving ‘extra’
information about the time the event took place, but their presence is not required
in order for the sentence to be intuitively ‘complete’. On the other hand, (42b–d)
feel ‘incomplete’: required information about what is being put (or) where is not
given. This is because the direct object and the locative PP are arguments of the
verb put; they must appear in the VP when put is V. More technically, put cat
egorially selects for a direct-object NP and a locative PP. But it does not select
for an adjunct temporal PP.
What the ungrammaticality of (41b,c) tells us then is that selected arguments
of V must form part of the VP with the verb, and hence must undergo do so
replacement. Then (41d) can be taken to indicate that adjuncts are outside the
VP, and so do not correspond to do so replacement. But what about (41a)? The
fact that the substituted VP is interpreted as put the car in the garage on Tuesday
indicates that the adjunct PP is part of the VP. So (41d) appears to be telling us
that the adjunct PP is outside VP, and (41a) appears to be telling us it is inside
VP. We conclude for now that adjuncts (of this kind, at least) are optional con
stituents of VP, while selected arguments are obligatory constituents of VP. This
conclusion accounts for the pattern seen in (41) as well as the examples in (42)
and is consistent with the idea that do so corresponds to VP. We will see a more
sophisticated treatment of adjuncts in the next chapter.
The do so test and the fronting test allow us to find VPs. We can now apply
these tests to our two structures for John ate the cake in (25), repeated here as (43):
With a slight tweak to the tense of the verb to create a natural context for
VP-fronting, (44a) is perfectly grammatical. Example (44b), on the other hand,
is strongly ungrammatical (bordering on what is sometimes called ‘word salad’,
a sequence of words so unintelligible that it is almost impossible to impose any
kind of structure or meaning on it).
Do so substitution, seen in (45), gives rise to a similar contrast:
(45) a. John ate the cake, and Mary did so too.
b. *John ate the cake, and did so the cake too.
Example (45b) is not word salad, but it is not acceptable. We can also notice that
Mary has completely disappeared from this example. This is a reflection of a
very general fact about substitution and ellipsis operations: substituted and elided
material is subject to a recoverability condition, i.e. the operations of substitu
tion and ellipsis (the latter of which we will look at in depth in Volume II) must
Example (46a) is ungrammatical for most English speakers, although there are
varieties in the north-west of England, around Liverpool and Manchester, where
it is accepted (the ‘%’ is used to indicate a form which is acceptable in one vari
ety of a language but not another; of course, this kind of variation simply reflects
the fact that the sociocultural E-language concept ‘English’ does not correspond
exactly to aggregates of I-languages – people from the relevant area of England
have slightly different I-languages from those from elsewhere, a fact reflected
in their differing grammaticality judgements of examples like (46a). However,
in these varieties it is interpreted as the direct object (a/the book); it cannot be
interpreted as substituting for a present to. Similarly, (46b) is grammatical, but it
can only be the direct object. An indirect object (i.e. to John) cannot be recovered
here. Again, then, a book to and a book to John fail the constituency test.
Ellipsis is the next kind of constituency test. Ellipsis elides, or deletes, mate
rial, subject to the recoverability requirement we have already seen in relation to
substitution. VP-ellipsis is quite natural in English, as in:
(47) a. John can speak Mandarin and Mary can speak Mandarin too.
b. John will leave tonight and Mary will leave tonight too.
c. John has passed the exam and Mary has passed the exam too.
d. John is writing a book and Mary is writing a book too.
(48) John ate a cake and Mary did eat a cake too.
This example illustrates an aspect of the English auxiliary system that will play
a major role in our analysis of clauses from the next chapter on. The struck-out
sequence in the second conjunct here consists of the verb and its direct object,
and so we can quite reasonably analyse this as VP. This would be consistent with
our other results for constituency tests for VP: fronting and do so substitution (see
(44a) and (45a)). Example (48) differs from (47) in that there is no auxiliary in the
first conjunct and the auxiliary do appears in the second one. There is a difference
between the two conjuncts here: in the first one, the verb bears the tense marking:
we have ate, not eat. In the second conjunct, the past-tense marking is carried by
the auxiliary do, in the form of did, and so the struck-out verb form lacks tense
marking (hence eat, not ate). We see from (47) and (48) that, generally speaking,
auxiliaries do not have to be deleted under VP-ellipsis. This suggests that finite
auxiliaries are not part of VP. VP-fronting confirms this:
IanRoberts_9781316519493_C003.indd 66 05-09-2022 19:49:23