Differential Geometry Notes
Differential Geometry Notes
DenisWerth
International Center for Fundamental Physics, 2020-2021
Preface
We have been doing physics for several years now. Since then, physics was mostly Fourier
transforms, solving dierential equations and linear algebra. The question behind most
of the physical problems was usually How to diagonalize the Hamiltonian? This is just
because we did a lot of Quantum Physics. We were (and still are) used to the Dirac
notations i.e. the bras and the kets, operators, commutators, etc. Even when it comes
to contracting indices and using the metric tensor, Special Relativity gave us some ease.
With General Relativity, the paradigm has changed, and so the mathematical formalism.
Everything is continuous and formula are impossible to remember1 . A lot of questions are
raised. What is the tangent vector tangent to? What does the connection connect? Does
a one-form act on a function? How to physically interpret the Riemann tensor? What
is the link between the Lie derivative, Killing vectors and symmetries? Does a change of
coordinates act at the same point? If you are just like me and all these questions prevent
you from sleeping, these lecture notes are made for you.
What are these Lecture Notes? This handout presents the mathematics of Gen-
eral Relativity. It aims at oering the reader a mathematical toolbox and ease behind
the concepts of Riemannian geometry. We will introduce the useful basic concepts of
dierential and Riemannian geometry, discuss their interpretations and give useful tips to
easily apply these abstract concepts to physics. I will try to write these lecture notes so
that you can have dierent reading levels. For example, you can entirely read the handout
or just the essential at the end of each section. Several examples in violet also illustrate
the abstract concepts throughout the script. We will not prove most of the results to
keep these lecture notes relatively short. Finally, these lecture notes are mainly based on
David Tong's lectures2 and the heavy "Big Black Book"3 .
What These Lecture Notes are Not? Personally in General Relativity, I know
how to apply formula, compute with several indices, integrate, change frames, etc. But
I feel that I do not deeply understand the mathematical objects I am using. To help me
understand the mathematics of General Relativity and to make things clear in my mind,
I started to write these lecture notes. And I learned a lot! Obviously, the aim was not
to write another lecture notes on General Relativity because many already exist on the
web. Instead, I wanted to focus on mathematics.
First, I need to say that this is not a physics lecture. We will not discuss Einstein
equations, present some solutions, linearize the theory to recover gravitational waves,
etc... It means basically that we will not make a clear and deep link between geometry
and gravity. But wait... Gravity is geometry! This rather simple statement gives me the
opportunity to focus on mathematics and geometry instead of physics, hoping that physics
will become more intuitive afterward. The rest is no longer my responsibility because (i)
I am just a beginner i.e. I know pretty much nothing about General Relativity, and (ii)
we have a wonderful teacher.
Finally, I would like to thank our teachers Marios Petropoulos and Philippe Grand-
clément for making me understand that I do not fully understand General Relativity.
1 Can you write the torsion tensor in terms of components without looking at the formula?
2 David Tong, General Relativity , University of Cambridge Part III Mathematical Tripos
3 Charles W. Misner, Kip S. Thorne and John Archibald Wheeler, Gravitation
2
Contents
1 Dierential Geometry 5
1.1 Dierential Manifolds . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . 5
1.2 Vectors Redened into Tangent Vectors . . . . . . . . . . . . . . . . . . . . 6
1.3 Integral Curves . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . 8
1.4 Transformation Law for Tangent Vectors . . . . . . . . . . . . . . . . . . . 10
1.5 One-Forms . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . 12
1.6 Tensors . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . 13
1.7 Operations on Tensor Fields . . . . . . . . . . . . . . . . . . . . . . . . . . 14
1.8 Commutators and Lie Derivative . . . . . . . . . . . . . . . . . . . . . . . 16
2 Riemannian Geometry 19
2.1 Abstract and Component Geometry . . . . . . . . . . . . . . . . . . . . . . 19
2.2 Metric . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . 20
2.3 Covariant Derivative and Connection . . . . . . . . . . . . . . . . . . . . . 22
2.3.1 Denition of the Connection . . . . . . . . . . . . . . . . . . . . . . 23
2.3.2 Levi-Civita Connection and Christoel Symbols . . . . . . . . . . . 25
2.3.3 Useful Properties of the Levi-Civita Connection . . . . . . . . . . . 26
2.4 Torsion and Curvature . . . . . . . . . . . . . . . . . . . . . . . . . . . . . 28
2.4.1 Abstract Denition and Components . . . . . . . . . . . . . . . . . 28
2.4.2 Identities and Properties of the Riemann Tensor . . . . . . . . . . . 31
2.5 Ricci and Einstein Tensors . . . . . . . . . . . . . . . . . . . . . . . . . . . 33
2.6 Parallel Transport . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . 34
2.7 Geodesics and auto-parallels . . . . . . . . . . . . . . . . . . . . . . . . . . 36
2.7.1 Ane and Non-ane Parametrisations . . . . . . . . . . . . . . . . 37
2.7.2 Geodesics as Extremal Line Elements . . . . . . . . . . . . . . . . . 38
2.7.3 Computing Christoel Symbols from the Action . . . . . . . . . . . 40
3 Advanced Topics 43
3.1 Symmetries . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . 43
3.1.1 Killing Vectors . . . . . . . . . . . . . . . . . . . . . . . . . . . . . 43
3.1.2 Conserved Charges . . . . . . . . . . . . . . . . . . . . . . . . . . . 44
3.1.3 Useful Identity Relating Curvature and Killing Vectors . . . . . . . 46
3.1.4 Maximal Symmetry and Constant Curvature . . . . . . . . . . . . . 47
3.2 Variational Calculus . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . 48
3.2.1 Covariant Volume Element . . . . . . . . . . . . . . . . . . . . . . . 48
3.2.2 Divergence Theorem . . . . . . . . . . . . . . . . . . . . . . . . . . 49
3.2.3 Einstein-Hilbert Action . . . . . . . . . . . . . . . . . . . . . . . . . 51
3.2.4 Action Variation . . . . . . . . . . . . . . . . . . . . . . . . . . . . 52
The Essential in a Scheme 55
3
4
Chapter 1
Dierential Geometry
The aim of this section is to fully understand the objects we manipulate in General
Relativity. Most of the denitions are usually very formal and one loses intuition. Here,
we give a meaningful sense of how dierent objects should be seen or at least imagined.
Even if our discussion of dierential geometry is not particularly rigorous, we will be
careful about building up the mathematical objects in the right logical order.
Figure 1.1. Example of a dierential manifold in two dimensions which locally looks
like R2 .
If one wants to describe physics with manifolds, one needs to make manifolds dif-
ferential. In simple words, a dierential manifold M means that we can continuously
5
(smoothly) go from one point p ∈ M of the manifold to another point q ∈ M, and that
M locally looks like RD . An illustration is provided above.
Example. Some simple examples in mathematics include Euclidean space RD , the sphere SD and the
torus TD = S1 × ... × S1 . In statistical physics, the phase space of N particles is a 6N -dimensional
manifold.
The last property is of course the Leibniz rule. Intuitively, tangent vectors tell us how
things change in a given direction. They do this by dierentiating. It is easy to check
that the objects
∂
∂µ = , (1.1)
p ∂xµ p
6
which act on functions obeys all the requirements of a tangent vector1 . However, they are
much more than simple tangent vectors.
Indeed, the set of all tangent vectors at a point p ∈ M forms an D-dimensional vector
space called the tangent space Tp (M). The tangent vectors ∂µ provide a basis for
p
Tp (M). This means that we can rewrite any tangent vector as
Xp = X µ ∂ µ , (1.2)
p
with X µ the components of the tangent vector in this basis. Note that to be fully rigorous
regarding the notation, one needs to write ∂µ instead of ∂µ because this object is a tangent
vector.
What is it tangent to? So far, we have not really explained where the name "tangent
vector" comes from. Consider a smooth curve in M that passes through the point p ∈ M.
We describe the curve by some coordinates xµ (λ) where λ parameterizes the curve such
that λ = 0 at p ∈ M. Before we learned any dierential geometry, we would say that the
tangent vector to the curve at λ = 0 is
dxµ (λ)
Xµ = . (1.3)
dλ λ=0
But we can take these to be the components of the tangent vector Xp which we dene
as
dxµ (λ)
Xp = ∂µ . (1.4)
dλ λ=0 λ=0
The tangent vector tells us how fast any function f ∈ C ∞ (M) changes as we move
along the curve. This gives also meaning to the term "tangent space" for Tp (M). It is
literally the space of all possible tangents to curves passing through the point p ∈ M. An
illustration of a two dimensional manifold embedded in R3 is shown below. Note that the
tangent spaces Tp (M) and Tq (M) at dierent points p 6= q are dierent. There is no way
to compare vectors in Tp (M) and vectors in Tq (M).
So far we have only dened tangent vectors at a point p. It is useful to consider objects
in which there is a choice of tangent vector for every point p ∈ M. We call objects that
vary over space elds. A vector eld X is dened to be a smooth assignment of a
tangent vector Xp to each point p ∈ M. Given a coordinate basis, we can expand any
vector eld as
X = X µ ∂µ , (1.5)
where the X µ are now smooth functions on M. This tangent vector eld X , given a
certain curve parameterized by xµ (λ), acts on a function as a dierential operator
dxµ ∂f df (λ)
X(f ) = X µ ∂µ f = µ
= . (1.6)
dλ ∂x dλ
1 Note that the index µ is a subscript rather than superscript that we use for the coordinates xµ .
7
Figure 1.2. Illustration of a tangent vector space Tp (M) at the point p ∈ M with a
tangent vector X = X ∂µ in a two dimensional manifold.
µ
It means that a tangent vector acting on a function tells us about how this function
vary along the curve.
Xp = X µ ∂µ . (1.7)
p
X = X µ ∂µ . (1.8)
The action of a tangent vector eld X , given a certain curve parameterized by
xµ (λ), on a function f is
df (λ)
X(f ) = X µ ∂µ f = , (1.9)
dλ
which is a number when evaluated at a point p ∈ M.
8
We have seen previously that the components of a tangent vector are
dxµ (λ)
Xµ = . (1.10)
dλ
Example. Let us consider the curve C : {θ = π/2, ϕ = λ} on the two dimensional sphere S2 that we
parameterize with θ and ϕ. This curve clearly is the equator. To explicitly compute a tangent vector at
some point, one needs to specify a basis. We choose {∂θ , ∂ϕ }. Then our vector X is written
dθ
0
X= dλ
dϕ = = ∂ϕ . (1.11)
dλ
1
The previous example shows that one can explicitly write a tangent vector if a specic
curve is given. Indeed, a tangent vector at a given point can be oriented in dierent ways.
that illustrates how the tangent vector components give the variation along a curve.
Example. • Consider again the sphere S2 in polar coordinates with a vector X = ∂ϕ . The integral
curves solve the equation (1.10), which are
dθ dϕ
=0 and = 1, (1.13)
dλ dλ
that has the solution θ = θ0 and ϕ = ϕ0 + λ. The integral lines are shown below.
• Consider the vector eld on R2 with Cartesian coordinates X = (1, x2 ) = ∂x + x2 ∂y . The equation for
the integral curves is now
dx dy
=1 and = x2 , (1.14)
dλ dλ
that has the solution x(λ) = x0 + λ and y(λ) = y0 + 31 (x0 + λ)3 . The associated ow lines are shown
below.
9
(a) X = X µ ∂µ = ∂ϕ . (b) X = X µ ∂µ = ∂x + x2 ∂y .
Figure 1.3. Some integral curves corresponding to the two examples given above.
∂
eµ = = ∂µ , (1.17)
∂xµ
that must be seen as a directional derivative along a curve. A transformation from one
basis to another in the same tangent space Tp (M) at the event p ∈ M is produced by a
matrix
2 It is more appropriate to call a point an event.
10
∂ ∂ ∂xµ
e0ν = = = eµ Jνµ . (1.18)
∂x0ν ∂xµ ∂x0ν
The matrix J is the Jacobian of the transformation xµ → x0µ . One must see this
transformation as the chain rule. The components of a tangent vector must transform
by the inverse matrix
X 0ν = (J −1 )νµ X µ . (1.19)
Note that we have (J −1 )νµ Jτµ = δτν . This inverse transformation law guarantees com-
patibility between the expressions X = X 0ν e0ν and X = X µ eµ
X = X 0ν e0ν = ((J −1 )ντ X τ )(eµ Jνµ ) = X τ ((J −1 )ντ Jνµ )eµ = X τ δτµ eµ = X µ eµ . (1.20)
Example. Let us consider the two dimensional at plane3 . One can use polar coordinates xµ : {r, θ} or
Cartesian coordinates x0µ : {x = r cos θ, y = r sin θ} to dene an event, a curve, etc. The Jacobian of the
xµ → x0µ transformation is dened as
∂x ∂x
cos θ −r sin θ 1 r cos θ r sin θ
J= ∂r
∂y
∂θ
∂y , J −1 = . (1.21)
∂r ∂θ
sin θ r cos θ r − sin θ cos θ
Now if we consider a tangent vector X = (0, 1) = ∂θ in polar coordinates, one can rewrite4 it in Cartesian
coordinates
∂θ = Jθµ ∂µ = Jθx ∂x + Jθy ∂y = −r sin θ∂x + r cos θ∂y = −y∂x + x∂y . (1.22)
Our tangent vector X is written X = (−y, x) in Cartesian coordinates. Notice that in Cartesian co-
ordinates, it is not obvious that our tangent vector gives rise to integral curves rotating around the origin.
X 0µ = (J −1 )µν X ν . (1.25)
a Note that the equations give the components of J and J −1 .
b We remember this transformation as dierentiating the new coordinates with respect to the
old coordinates.
11
1.5 One-Forms
We have seen that tangent vectors at an event p ∈ M live in the tangent space Tp (M).
Since Tp (M) is a vector space (one can choose {∂µ } as a basis), there exists a dual vec-
tor space to Tp (M) denoted Tp? (M). This space is called the cotangent space and its
elements are linear functions from Tp (M) to R.
An element ω ∈ Tp? (M) of the cotangent space is called a one-form. A one-form acts
on a tangent vector to give a number. This procedure must be seen as a scalar product5
∂xµ
µ
hdx |∂ν i = ν
= δνµ . (1.28)
∂x
ω = ωµ dxµ , (1.29)
where ωµ are the components of ω . Given a tangent vector X , one can act on this tangent
vector with the one-form6
∂x0µ
ων0 = ωµ (1.31)
∂xν
5 We draw an analogy with quantum mechanics when we write hφ|ψi with |ψi ∈ H and hφ| ∈ H? , the
Hilbert space and its dual space respectively.
6 Note that the scalar product is dened between a tangent vector and a one-form and not between
two tangent vectors or two one-forms.
12
A one-form ω is an element of the dual vector space Tp? (M) that acts on a tangent
vector X to give a numbera hω|Xi = ωµ X µ . Any one-form ω can be decomposedb
on the basis {dxµ } as
ω = ωµ dxµ , (1.32)
where ωµ are the components of ω . The basis elements satisfy
∂xµ
µ
hdx |∂ν i = ν
= δνµ . (1.33)
∂x
a This action denes an inner product.
b Remember that one-form components have low indices whereas tangent vector components
have upper indices.
1.6 Tensors
A tensor of type (q, r) is a multilinear object which maps q elements of Tp? (M) and r
elements of Tp (M) to a number. Such an object, called a tensor of rank q + r, is written
in terms of the bases described earlier as
Example. A tangent vector is a tensor of type7 (1, 0) and a one-form is a tensor of type (0, 1).
T (ω1 , ..., ωq ; X1 , ..., Xr ) = T µ1 ...µq ν1 ...νr ωµ1 ... ωµq X ν1 ... X νr . (1.35)
As with vector elds and one-forms, we can ask how the components of a tensor
transform. We know that tangent vector basis elements ∂µ and one-form basis elements
dxµ transform as
13
for a xµ → x0µ coordinate transformation. Note that we usually do not write the prime
because there is no ambiguity regarding the change of coordinates. This transformation
is just the chain rule and acts at two dierent points.
Similarly to tangent vectors elds, we can dene tensor elds by making each com-
ponents of the tensor a function that can be smoothly evaluated on the manifold. Note
that on a manifold of dimension D, a tensor T of type (q, r) has Dq+r components. For
a tensor eld, each of these is a function over M.
can dene a tensor independently of the basis that we chose, a tensorial equation is valid
in any coordinate system i.e. in any frame. Note that, in particular, a tensor is zero (at
a point) in one coordinate system if and only if the tensor is zero (at the same point) in
another coordinate system.
T (ω1 , ..., ωq ; X1 , ..., Xr ) = T µ1 ...µq ν1 ...νr ωµ1 ... ωµq X ν1 ... X νr . (1.38)
Under a coordinate transformation, the lower components of a tensor transform by
multiplying by J and the upper components by multiplying by J −1 . Intuitively,
this is the chain rule. For example,
∂xµ ∂x0τ ∂x0λ σ
T 0µρν = T τ λ. (1.39)
∂x0σ ∂xρ ∂xν
The principle of general covariance tells us that every physical equation should be
written in a tensorial form such that it is valid in every coordinate system.
Linear combination. Given two (q, r) tensors A, B and two scalars α, β , their linear
combination T = αA + βB is also a (q, r) tensor. In terms of components, this reads
14
T µ1 ...µq ν1 ...νr → T µ1 ...µq−1
ν1 ...νr−1 = T
µ1 ...µq−1 λ
ν1 ...νr−1 λ . (1.42)
Note that contraction over dierent pairs of indices will in general give rise to dierent
tensors. For example, Tνµ
µ
and Tµν
µ
will be dierent in general.
1
S(X, Y ) = [T (X, Y ) + T (Y , X)]
T (X, Y ) = S(X, Y ) + A(X, Y ) with 2 (1.43)
1
A(X, Y ) = [T (X, Y ) − T (Y , X)]
2
In index notation, this becomes
1 1
Sµν = (Tµν + Tνµ ) and Aµν = (Tµν − Tνµ ), (1.44)
2 2
which is just like taking the symmetric and anti-symmetric part of a matrix. These
operations being frequently used, we introduce some new notation. We dene
1 1
T(µν) = (Tµν + Tνµ ) and T[µν] = (Tµν − Tνµ ). (1.45)
2 2
Note that a product of a symmetric and an anti-symmetric tensor is always zero.
Indeed, let us consider S a symmetric tensor (Sµν = Sνµ ) and A an anti-symmetric
tensor (Aµν = −Aνµ ). Then,
1
T(µνλ) = (Tµνλ + Tµλν + Tνµλ + Tνλµ + Tλµν + Tλνµ ), (1.48)
3!
where we considered all permutations of three elements (hence the 3!). For a totally anti-symmetric
tensor, one obtains the same but with a minus sign when the permutation cannot be obtain from the
ordered indices with a cyclic permutation
1
T[µνλ] = (Tµνλ − Tµλν − Tνµλ + Tνλµ + Tλµν − Tλνµ ). (1.49)
3!
• If a (0, 2) tensor T is totally anti-symmetric, then Tµν = −Tνµ .
• The product of a totally symmetric tensor with a totally anti-symmetric tensor is zero. Indeed, let us
consider Aµν symmetric and Bµν anti-symmetric. Then one obtains
15
because µ and ν are dummy variables.
How to prove that an object is a tensor? To prove that an object is a tensor, one
has to verify that the object (its components) transforms as a tensor. Schematically, if
one obtains
Given a tensor, one can dene the symmetric and the anti-symmetric part of the
tensor
1 1
T(µν) = (Tµν + Tνµ ) and T[µν] = (Tµν − Tνµ ). (1.52)
2 2
These denitions are generalized to other tensors of higher rank. A product of a
symmetric tensor with an anti-symmetric tensor is zero.
16
However, one can build a new tangent vector eld by taking the commutator [X, Y ],
which acts on function f as
LX Y = [X, Y ]. (1.59)
Note the analogy with quantum mechanics and for example the angular momentum.
Written in terms of components, it is
∂Y µ ν ∂X
µ
(LX Y )µ = LX Y µ = X ν ν
− Y ν
= X ν ∂ν Y µ − Y ν ∂ν X µ . (1.60)
∂x ∂x
We usually write (LX Y )µ the components of the new tangent vector as LX Y µ . Using
the Jacobi identity, one can show that
LX LY Z − LY LX Z = L[X,Y ] Z. (1.61)
The denition of the Lie derivative is extended to one-forms and tensors. One just has
to take into account all possible congurations to permute a certain number of objects
and remember that a normal sum "lower index with an upper index" creates a plus sign
and "an upper index with a lower index" creates a minus sign. For a (1, 1) tensor T , the
Lie derivative along X reads
17
LX T µν = X τ ∂τ T µν + T µτ ∂ν X τ − T τν ∂τ X µ . (1.62)
Example. • In the two dimensional plane in Cartesian coordinates, let us consider f (x, y) = x2 − sin y
and X = sin x∂y − y 2 ∂x . Then, the Lie derivative of f along X is
Given two tangent vector elds X and Y , one can dene the commutator [X, Y ] =
XY − Y X that can be written in terms of components in the following way
ν ν
µ ∂Y µ ∂X
ν
[X, Y ] = X −Y . (1.66)
∂xµ ∂xµ
The commutator satises the Jacobi identity
LX Y = [X, Y ]. (1.69)
The Lie derivative satises
LX LY Z − LY LX Z = L[X,Y ] Z, (1.70)
and can be extended to one-forms and tensors. For a (1, 1) tensor T , the Lie
derivative along X readsa
LX T µν = X τ ∂τ T µν + T µτ ∂ν X τ − T τν ∂τ X µ . (1.71)
a We should remember that a ()µ ()µ gives a "+" and ()µ ()µ gives a "-".
18
Chapter 2
Riemannian Geometry
Abstract dierential geometry treats a tangent vector as existing in its own right,
without necessity to give its breakdown into components
X = X µ ∂µ = X 0 ∂0 + X 1 ∂1 + ..., (2.1)
just as one is accustomed nowadays in electromagnetism to treat the electric eld E ,
without having to write out its components. The abstract approach is useful when one
wants to derive results in a simple way. For example in electromagnetism, considering
the gradient operator ∇ instead of its components ∂x , ∂y , etc is the quickest and simplest
mathematical scheme one knows to derive general results.
Example. Let us take the example of the Lie derivative of a tangent vector Y along X . One can either
dene it by
LX Y = [X, Y ], (2.2)
or by
ν ν
µ ∂Y µ ∂X
LX Y µ
= X −Y . (2.3)
∂xµ ∂xµ
However note that the Jacobi identity for the Lie derivative is easier to derive using abstract notation
than in terms of components.
19
2.2 Metric
We have yet to meet the star of the show. There is one object that we can place on a
manifold whose importance dwarfs all others, at least when it comes to understanding
gravity. This is the metric. The existence of a metric brings a whole host of new concepts
to the table which, collectively, are called Riemannian geometry.
We all know that the metric is a way to measure distances between points on a man-
ifold. It does, indeed, provide this service but it is not its initial purpose. Instead, the
metric is an inner product on each tangent vector space Tp (M).
(ii) Non-degenerate i.e. if for any p ∈ M, g(X, Y )|p = 0 for all Y ∈ Tp (M) then
Xp = 0.
20
an object's velocity along a curve xµ (λ) parameterized by λ, depending on the sign of
X µ Xµ = gµν X µ X ν , the tangent vector can be spacelike2 , null3 or timelike4
X µ Xµ > 0 spacelike
X µ Xµ = 0 null (2.7)
X µ Xµ < 0 timelike
Figure 2.1. Illustration of the light cone and the three types of tangent vectors.
We can use the metric to determine the length of a curve. Given a parametrisation
x (λ) of curve with tangent vector X µ = dx , its length between two points a and b is
µ
µ
dλ
given by
ˆ b
(2.8)
p
`a→b = dλ −gµν X µ X ν .
a
Note that for timelike tangent vectors describing most of the physical trajectories, we
have gµν X µ X ν < 0 so the minus sign ensures that we have a positive quantity inside the
square root.
Raising and Lowering of Indices. Given a tensor, one can raise and lower indices using
the metric. These operations can of course be combined in various ways. For example,
given a tangent vector eld X µ , we can associate its dual covector eld Xµ
Xµ = gµν X ν , (2.9)
and likwise for covectors
X µ = g µν Xν , (2.10)
2 Thesevectors describe the velocity of objects travelling faster than the speed of light i.e. tachyons.
3 Thesevectors describe the velocity of massless objects travelling at the speed of light i.e. mainly
photons (maybe also gravitons?).
4 These vectors describe the velocity of massive objects/particles travelling slower than the speed of
light.
21
Note that there are dierent ways of lowering the indices, and they will in general give
rise to dierent tensors. It is therefore important to keep track of this in the notation.
For example gµν T ντ = Tµτ and not Tτ µ . This is the reason why the lower indices of a
tensor are written more to the right that the upper indices.
Finally note that this notation of raising and lowering indices with the metric is con-
sistent with denoting the inverse metric by raised indices because
g µν = g µτ g νρ gτ ρ , (2.11)
and raising one index of the metric gives the Kronecker tensor,
There is a dierent way to take derivatives, one which ultimately will prove more
useful. The derivative is again associated to a tangent vector eld X . However, this time
we introduce a dierent object, known as connection to map the tangent vector spaces
at one point to another tangent vector space at another. The result is an object, distinct
from the Lie derivative, called the covariant derivative.
22
2.3.1 Denition of the Connection
First, we give an abstract denition of the covariant derivative, also called connection.
This denition will be useful to understand that there are many such connections and so
covariant derivatives. Ultimately, we will chose one connection, the Levi-Civita connec-
tion, to perform computations in the physical spacetime.
Denition. A connection is a map from one tangent vector space Tp (M) to another
Tq (M). We usually write this map as ∇(X, Y ) = ∇X Y and the object ∇X is called the
covariant derivative. Note that the connection takes as inputs two tangent vectors. It
satises the following properties for all tangent vectors elds X, Y and Z ,
(i) ∇X (Y + Z) = ∇X Y + ∇X Z .
(ii) ∇f X+gY Z = f ∇X Z + g∇Y Z for all functions f and g .
(iii) ∇X (f Y ) = f ∇X Y + (∇X f )Y where we dene ∇X f = X(f ).
The covariant derivative endows the manifold M with more structure. To elucidate
this, we can evaluate the connection in a basis ∂µ of the tangent vector elds. We can
always express this as
Mapping two dierent tangent vector spaces is what allows the connection to act as a
derivative. In what follows, we will use the notation
∇ µ = ∇ ∂µ . (2.17)
This makes the covariant derivative ∇µ look similar to a partial derivative. Using the
properties of the connection, we can write a general derivative of a tangent vector eld as
∇X Y = ∇X (Y µ ∂µ ) = X(Y µ )∂µ + Y µ ∇X ∂µ
(2.18)
= X ν ∂µ Y µ ∂µ + X ν Y µ ∇ν ∂µ = X ν ∂ν Y µ + Γµνρ Y ρ ∂µ .
The fact that we can strip of the overall factor of X ν means that it makes sense to
write the components of the covariant derivative as
23
to write ∇X = X ν ∇ν and think ∇µ as an operator in its own right. In contrast, there is
no way to write "LX = X µ Lµ ".
The Connection is Not a Tensor. We can see this immediately from the denition 5
However, let us illustrate this using components. We ask what the connection looks
like in a dierent basis
∂ ∂xν ∂
∂µ0 = = = Jµν ∂µ . (2.20)
∂x0µ ∂x0µ ∂xν
In the basis {∂µ0 }, the connection is written ∇∂ρ0 ∂ν0 = Γ0µ
ρν ∂µ . Substituting in the
0
Γ0µ 0 λ σ λ σ λ τ σ λ
ρν ∂µ = ∇Jρσ ∂σ (Jν ∂λ ) = Jρ ∇∂σ (Jν ∂λ ) = Jρ Jν Γσλ ∂τ + Jσρ ∂λ ∂σ Jν . (2.21)
Γ0µ 0
(2.22)
σ λ τ σ τ
σ λ τ σ τ
−1 µ 0
ρν ∂µ = Jρ Jν Γσλ + Jρ ∂σ Jν ∂τ = Jρ Jν Γσλ + Jρ ∂σ Jν (J )τ ∂µ .
Stripping o the basis tangent vectors ∂µ0 , we see that the components of the connection
transform as
Γ0µ
ρν = (J
−1 µ σ λ τ
)τ Jρ Jν Γσλ + (J −1 )µτ Jρσ ∂σ Jντ . (2.23)
The rst term coincides with the transformation of a tensor. But the second term,
which is independent of Γ, instead depends on ∂J , is novel. This is the characteristic
transformation property of a connection.
Dierentiating other Tensors. One can use the Leibnizarity of the covariant derivative
to extend its action to any tensor eld. Here, we give some examples. For a one-form ω ,
we obtain
∇µ ων = ∂µ ων − Γρµν ωρ . (2.24)
The pattern is clear; for every upper index we get a +ΓT term while for every lower
index we get a −ΓT term.
5 Wesee here in practice the power of using abstract dierential geometry rather than in terms of
components.
24
The connection ∇(X, Y ) = ∇X Y is a map from one tangent vector space X ∈
Tp (M) to another Y ∈ Tq (M). The connection is not a tensor. Evaluated in a
basis {∂µ }, it reads
∇µ X ν = ∂µ X ν + Γνµσ X σ . (2.27)
The covariant derivative coincides with the Lie derivative for scalar functions
∇X f = X µ ∇µ f = LX f = X(f ) = X µ ∂µ f . Using the Leibnitz rule, one can
extend the action of the covariant derivative to any tensor eldsc
∇µ ων = ∂µ ων − Γρµν ωρ . (2.28)
Theorem. There exists a unique connection that is compatible with the metric g in the
sense that
25
We subtract the second and third of these from the rst, then multiply by the inverse
of the metric to obtain7
1
Γσµν = g σρ (∂µ gνρ + ∂ν gµρ − ∂ρ gµν ) . (2.32)
2
This is one of the most important formulas in this subject; commit it to memory.
This connection we have derived from the metric is the one on which conventional Gen-
eral Relativity is based. It is known as the Levi-Civita connection. The associated
connection coecients are called Christoel symbols. We then found a link between
two independent objects: the metric and the connection. Both concepts can be related if
one assumes a particular connection: the Levi-Civita connection.
∇σ gµν = 0. (2.33)
This connection is called the Levi-Civita connection and its components are the
Christoel symbols
1
Γσµν = g σρ (∂µ gνρ + ∂ν gµρ − ∂ρ gµν ) . (2.34)
2
The metric and the connection, two a priori independent objects, can be linked by
choosing a specic connection: the Levi-Civita one, on which conventional General
Relativity is based.
a Note that we need also the torsion-free property to dene a unique metric compatible connec-
tion.
Using the Levi-Civita connection, it is easy to show that the inverse metric also has
zero covariant derivative
7 Notethat we obtain the result if we consider the connection coecients being symmetric with respect
to both lower indices. This is a property of a torsion-free manifold. This notion will be introduce later
when we dene the torsion tensor.
8 Meaning we also consider the torsion-free property.
26
∇σ gµν = 0 and ∇σ g µν = 0. (2.36)
This nice property allows us to raise and lower indices inside the covariant derivative
by passing the metric through the covariant derivative. For example, if Xµ is a obtained
by lowering an index of the tangent vector X µ by Xµ = gµν X ν , then
∇σ Tνµ = g µτ ∇σ Tτ ν . (2.38)
∇µ ∇ν f − ∇ν ∇µ f = ∇µ ∂ν f − ∇ν ∂µ f
(2.39)
= ∂µ ∂ν f − Γλµν ∂λ f − ∂ν ∂µ f − Γλνµ ∂λ f = 0,
because the usual partial derivatives commute. Note that the second covariant derivatives
on higher rank tensors do not commute. We will come back to this in our discussion of
the curvature tensor later on.
Another useful property is when the covariant derivative lies in between a dened
constant norm. Let us consider a tangent vector X with constant norm9 Xµ X µ = .
Then,
X µ ∇ν Xµ = X µ ∂ν Xµ − Γσνµ X µ Xσ
= ∂ν (Xµ X µ ) − Xµ ∂ν X µ − Γσνµ X µ Xσ
(2.40)
= −Xµ ∂ν X µ − Xσ (∇ν X σ − ∂ν X σ )
= −Xµ ∇ν X µ .
Using the fact that one can raise and lower indices inside the covariant derivative with
the metric and using the previous result, one obtains
9 Forexample, this is the case for velocity tangent vectors parametrized with the proper time. In this
case we have = 0, ±1 depending on whether the object/particle is massive or not.
27
The Lie derivative can be expressed in term of the covariant derivativea
LX Y µ = X ν ∇ ν Y µ − Y ν ∇ ν X µ . (2.42)
One can raise and lower indices inside the covariant derivative by passing the metric
through the covariant derivativeb . For example,
∇σ Xµ = gµν ∇σ X ν
(2.43)
∇σ Tνµ = g µτ ∇σ Tτ ν .
Covariant derivatives commutec on scalars f
∇µ ∇ν f − ∇ν ∇µ f = 0, (2.44)
but it is not the case for higher rank tensors. For tangent vectors X with constant
norm Xµ X µ = , one has
X µ ∇ν Xµ = 0. (2.45)
a This property comes from Γσµν = Γσνµ .
b This property comes from ∇σ gµν = 0 and ∇σ g µν = 0.
c This property comes from the fact that usual partial derivatives commute.
28
R(X, Y ) = ∇X ∇Y − ∇Y ∇X − ∇[X,Y ] . (2.49)
It is not obvious from the denition that these objects are tensors. To actually demon-
strate that T and R are tensors, we need to show that they are linear in all arguments.
This is a good example of using the abstract notation to derive expressions. Linearity
in ω is straightforward. For the others, there are some small calculations to do. For
example, we must show that T (ω; f X, Y ) = f T (ω; X, Y ) where f is a scalar function.
To see this, we just run through the denition of the various objects
Thus, both torsion and curvature dene new tensors on our manifold.
The next step is to evaluate these tensors in a coordinate basis {∂µ } with the dual
basis {dxµ }. The components of the torsion are
ρ
Tµν = T (dxρ ; ∂µ , ∂ν )
= dxρ (∇µ ∂ν − ∂ν ∂µ − [∂µ , ∂ν ])
(2.54)
= dxρ (Γσµν ∂σ − Γσνµ ∂σ )
= Γρµν − Γρνµ ,
where we used the fact that dxρ ∂σ = δσρ and [∂µ , ∂ν ] = 0. We learn that, even though Γρµν
is not a tensor, the anti-symmetric part Γρµν − Γρνµ does form a tensor. Clearly, the torsion
is anti-symmetric in the lower two indices
29
ρ
Tµν ρ
= Tνµ . (2.55)
Connections which are symmetric in the lower indices, so Γρµν = Γρνµ have zero torsion.
Such connections are said to be torsion-free. This is the case for the Levi-Civita con-
nection10 .
σ
Rρµν σ
= −Rρνµ . (2.57)
Note that it would be quite unpleasant to have to verify the tensorial nature of this
expression by explicitly checking its behaviour under coordinate transformations.
Using the connection, one can dene the torsion, a (1, 2) tensor T , by
10 Actually, the torsion-free property is one requirement to derive the Levi-Civita connection components
in terms of the metric. The fundamental theorem of Riemannian geometry states the existence of a unique,
torsion-free, connection that is metric compatible.
11 Note the slightly counter intuitive, but standard ordering of the indices.
30
2.4.2 Identities and Properties of the Riemann Tensor
There is a closely related calculation in which both the torsion and te Riemann tensors
appear. We look at the commutator of covariant derivatives acting on tangent vector
elds. One can check that
[∇µ , ∇ν ]Z σ = ∇µ ∇ν Z σ − ∇ν ∇µ Z σ = Rρµν
σ
Z ρ − Tµν
ρ
∇ρ Z σ . (2.62)
This expression is known as the Ricci identity. Note that in practice, we will always
work with the Levi-Civita connection so that the torsion is zero. One then obtains
[∇µ , ∇ν ]Z σ = ∇µ ∇ν Z σ − ∇ν ∇µ Z σ = Rρµν
σ
Z ρ. (2.63)
It is perhaps surprising that the commutator [∇µ , ∇ν ], which appears to be a dier-
ential operator, has an action on vector elds which (in the absence of torsion, at any
rate) is a simple multiplicative transformation. The Riemann tensor measures that part
of the commutator of covariant derivatives which is proportional to the vector eld, while
the torsion tensor measures the part which is proportional to the covariant derivative of
the vector eld; the second derivative does not enter at all. One must remember that the
Riemann tensor measures the failure of the covariant derivative to commute.
The extension of the above formula to any higher order tensors follows the usual
pattern, with one Riemann tensor contracted with every index. For example,
[∇µ , ∇ν ]T αβ = ∇µ ∇ν T αβ − ∇ν ∇µ T αβ = Rσµν
α
T σβ + Rσµν
β
T ασ . (2.64)
There are also a number of symmetric properties satised by the Riemann tensor when
we use the Levi-Civita connection. We will just state but not prove the following results.
If we lower an index on the Riemann tensor and write Rσρµν = gσλ Rρµν
λ
, then the
resulting tensor obeys the following symmetry identities
Rσρµν = −Rσρνµ
Rσρµν = −Rρσµν (2.65)
Rσρµν = Rµνσρ .
The Riemann tensor also satises the following cyclic permutation relation
Finally, the Riemann tensor also satises the so-called Bianchi identity
31
∇λ Rσρµν + ∇σ Rρλµν + ∇ρ Rλσµν = 0. (2.67)
Note that for a general connection there would be additional terms involving the
torsion tensor. This cyclic identity is closely related to the Jacobi identity for the covariant
derivative
Figure 2.2. Illustration of the Ricci identity with the Levi-Civita connection.
The Riemann tensor measures the failure the failure of the covariant derivative to
commutea
[∇µ , ∇ν ]Z σ = ∇µ ∇ν Z σ − ∇ν ∇µ Z σ = Rρµν
σ
Z ρ. (2.69)
The above formula is extended to any tensor. For example,
[∇µ , ∇ν ]T αβ = ∇µ ∇ν T αβ − ∇ν ∇µ T αβ = Rσµν
α
T σβ + Rσµν
β
T ασ . (2.70)
The Riemann tensor veries the following properties
Rσρµν = −Rσρνµ
Rσρµν = −Rρσµν
(2.71)
Rσρµν = Rµνσρ
Rσρµν + Rσµνρ + Rσνρµ = 0.
and the Bianchi identity
32
2.5 Ricci and Einstein Tensors
There are a number of further tensors that we can build from the Riemann tensor and that
are especially important in General Relativity. First, it is frequently useful to consider
contractions of the Riemann tensor. Even without the metric, we can form a contracted
Riemann tensor called the Ricci tensor
σ
Rµν = Rµσν . (2.73)
The Ricci tensor associated with the Levi-Civita connection, which is always used
in practice, is symmetric. It inherits its symmetry from the Riemann tensor. We write
Rµν = g σρ Rσµρν = g ρσ Rρνσµ , giving
∇µ Gµν = 0. (2.79)
The Riemann tensor contains all the information about the curvature of the manifold.
However, note that vanishing Christoel symbols does not mean the manifold is at. This
is because the connection is not a tensor. But the Riemann tensor is a tensor and so if
it vanishes in one coordinate system, it must vanish in all of them. Given some horrible
coordinate system, with non-vanishing Christoel symbols, we can always compute the
corresponding Riemann tensor and thus the scalar curvature R to see if the manifold is
actually at.
The elementary and universal applicable method for computing the components of the
Riemann tensor and thus the Ricci tensor and nally the curvature starts from the metric
components in a coordinate basis, and proceeds by the following scheme:
σ µ
Rµσν Rµ
(2.80)
Γ∼∂g R∼∂Γ+ΓΓ
gµν −−−→ Γσµν −−−−−−→ Rρµν
σ
−−−→ Rµν −→ R.
We see here that if one knows the metric, one knows everything about the manifold.
Note that the Christoel symbols can be obtained in a straightforward by varying the
33
(a) Positive curvature R > 0. (b) Negative curvature R < 0.
∇µ Gµν = 0. (2.84)
Starting from the metric, one can have a complete description of the manifold by
following this scheme
σ µ
Rµσν Rµ
(2.85)
Γ∼∂g R∼∂Γ+ΓΓ
gµν −−−→ Γσµν −−−−−−→ Rρµν
σ
−−−→ Rµν −→ R.
34
The answer is that the connection connects tangent vector spaces at two dierent
points of the the manifold by mean of a map called parallel transport. As we stressed
earlier, such a map is necessary to dene dierentiation. Note that it does not make any
sense to ask if two tangent vectors are parallel in a curved space. However, given a metric
and a curve connecting these two points, one can compare the two by dragging one along
the curve to the other using the covariant derivative.
Take a tangent vector eld X and consider some associated integral curves with co-
ordinates xµ (λ), such that
dxµ
Xµ = . (2.86)
dλ
We say that a tensor eld T is parallely transported along the dened curve if
∇X T = 0. (2.87)
To illustrate this, consider the parallel transport of a second tangent vector eld Y .
In terms of the components, the last condition reads
Parallel transport is path dependent. It depends on both the connection and the
underlying path wich, in this case, is characterised by the tangent vector eld X . To
parallely transport an object between two points, one must rst dene the parametrized
curve.
Example. On the two dimensional sphere in spherical coordinates (θ, ϕ), let us parallely transport the
vector Y = ∂ϕ = (0, 1) between the point p = (θ = α, ϕ = ϕ0 ) and q = (θ = α, ϕ = ϕ0 + δ) that is
along a curve parallel to the equator. First we need to dene a parametrized curve C . Here we take
C = {θ = α, ϕ = ϕ0 + δ λ} where λ is the parameter. We then dene the tangent vector corresponding
to this integral curve
dxµ
X = X µ ∂µ = ∂µ = 0 × ∂θ + δ∂ϕ = (0, δ), (2.90)
dλ
so that
X µ ∇µ Y ν = 0 leads to ∇ϕ Y µ = 0. (2.91)
Using the denition of the covariant derivative in terms of the components, the last equation is a set
of coupled rst order dierential equations
∂ϕ Y θ + Γθϕϕ Y ϕ = 0
(
(2.92)
∂ϕ Y ϕ + Γϕ θ
ϕθ Y = 0.
35
In practice, these equations are very hard to solve. Here, we know the Christoel symbols and the
fact that θ is constant. Solving these equations (with constant θ = α) where the constants of integration
are found with the initial tangent vector leads to
(
Y θ = sin α sin(cos α δ)
(2.93)
Y ϕ = cos(cos α δ).
Note that after a round trip (δ = 2π ), we do not recover the same tangent vector except along the
equator θ = α = π/2.
The connection maps to tangent vector spaces at two dierent points of the manifold
by the parallel transport. A tensor T is parallely transported along an integral curve
xµ (λ) dened by the tangent vector eld X which satises X µ = dx if
µ
dλ
∇X T = 0. (2.94)
For a tangent vector eld Y , it reads in terms of components
dY ν
+ Γνµρ X µ Y ρ = 0. (2.95)
dλ
These are a set of coupled, ordinary dierential equations that can be solved exactly
only in a very few cases.
∇X X = 0. (2.96)
Along a curve xµ (λ) parametrized by λ, we can write the above equation in terms of
components
36
d2 xµ ν
µ dx dx
ρ
+ Γ νρ = 0. (2.97)
dλ2 dλ dλ
This is precisely the geodesic equation. We can characterise geodesics by the property
that their tangent vectors are parallely transported (do not change) along the curve. For
this reason geodesics are also known as auto-parallels.
Note that for the Levi-Civita connection, we have ∇X g = 0. This ensures that for
any tangent vector eld Y parallely transported along a geodesic X , we have
d
g(X, Y ) = 0, (2.98)
dλ
which tells us that Y makes the same angle with the tangent vector X along each point
of the geodesic.
A geodesic, also called auto-parallel curve, is a curve that is always tangent to its
tangent vector X , namely
∇X X = 0, (2.99)
which in terms of components reads
d2 xµ ν
µ dx dx
ρ
+ Γνρ = 0. (2.100)
dλ2 dλ dλ
These are coupled dierential equations that can be solved to nd the curve
parametrized by xµ (λ).
Thus the geodesic equation retains its form only under ane changes λ̃ = aλ + b
so that ddλλ̃2 = 0. Parameters that make the right-hand side of (2.102) vanish are called
2
ane parameters and are related to each other by ane transformations. The geodesic
equation (2.97) written in this form is said to be anely parameterised.
37
d2 xµ dxν dxρ dxµ
+ Γµνρ = C(λ̃) , (2.103)
dλ̃2 dλ̃ dλ̃ dλ̃
for some function C(λ̃), we can deduce that this curve is the trajectory of a geodesic, but
that it is simply not parametrised by an ane parameter. Denoting f (λ̃) = λ, an ane
parameter λ is determined by
ˆ
f¨
dλ
C(λ̃) = − i.e. = exp ds C(s) , (2.104)
f˙2 dλ̃
where f˙ = ddλλ̃ . So for C(λ̃) = 0, we have an ane transformation between λ and λ̃. Note
that the proper time τ is an ane parameter for massive particle trajectories.
(2.106)
p
ds = −gµν dxµ dxν .
We then introduce the following action
ˆ ˆ
(2.107)
p
S= dλ L = dλ −gµν X µ X ν .
38
∂L 1 ∂gµν µ ν
σ
=− X X
∂x 2L ∂xσ (2.109)
∂L 1
σ
= − gσν X ν ,
∂X L
dX ν
d ∂L d 1 1 ∂gσν 1
= gσν Xν
+ X α α X ν + gσν . (2.110)
dλ ∂X σ dλ L L ∂x L dλ
Writing down the Euler-Lagrange equation with all the terms and rearranging the
expressions leads to
d 2 xν
µ ν
dxν
1 dx dx
gσν + ∂ µ gσν − ∂σ gµν = C(λ)gσν , (2.111)
dλ2 2 dλ dλ dλ
1
∂µ gσν = (∂µ gσν + ∂ν gσµ ) . (2.112)
2
Multiplying the above equation by the inverse metric g ασ and using the denition of
the Christoel symbols, one nally obtains
d 2 xα µ
α dx dx
ν
dxα
+ Γ µν = C(λ) . (2.113)
dλ2 dλ dλ dλ
We note that this is the geodesic equation using a non-ane parameter. However, if
one executes the same steps with the following Lagrangian
L = −gµν X µ X ν , (2.114)
d2 xα µ
α dx dx
ν
+ Γ µν = 0. (2.115)
dλ2 dλ dλ
39
The Euler-Lagrange equation associated with a Lagrangian L(xµ , dx ) is
µ
dλ
dxµ
∂L d ∂L
= with Xµ =
= ẋµ . (2.116)
∂xµ dλ ∂X µ dλ
q
Extremizing the element line with the Lagrangian L = −gµν dx leads to a
µ dxν
dλ dλ
non-anely parametrized geodesic
d 2 xα µ
α dx dx
ν
dxα
+ Γ µν = C(λ) , (2.117)
dλ2 dλ dλ dλ
where C(λ) = . However, one can simplify the derivation by considering
1 dL
L dλ
. Extremizing the action with this Lagrangian leads to an anely
µ dxν
L= −gµν dx
dλ dλ
parametrized geodesic
d2 xα µ
α dx dx
ν
+ Γ µν = 0. (2.118)
dλ2 dλ dλ
One can always choose an ane parameter (for example proper time for massive
particles) and apply the Euler-Lagrange equations without the square root.
One normally thinks that the connection coecients Γσµν must be known before one
can write the geodesic equation. However, when one anely parametrizes the geodesic
equation, it is simple to apply the Euler-Lagrange equation and explicitly nd the geodesic
equation. Thus, one can use this fact to compute the Christoel symbols.
To do this, one must explicitly write down the Lagrangian that leads to an ane
parametrised geodesic
dxµ dxν
L = −gµν = −gµν ẋµ ẋν , (2.119)
dλ dλ
and use the Euler-Lagrange equation. One then identies the resulting equations with
the anely parametrized geodesic equations and reads out the Christoel symbols.
Example. Let us consider the two dimensional sphere parametrized with (θ, ϕ). We want to compute
the Christoel symbols without using the formula involving the metric. Instead, we write the anely
parametrized Lagrangian
L = −gµν ẋµ ẋν = − θ̇2 + sin2 θϕ̇2 . (2.120)
The Euler-Lagrange for x0 = θ leads to
40
∂L ∂L d ∂L
= −2 sin θ cos θ ϕ̇2 , = −2θ̇ and = −2θ̈, (2.121)
∂θ ∂ θ̇ dλ ∂ θ̇
so that
One can computes with less eort the connection components Γσµν by writing down
the Lagrangian
dxµ dxν
L = −gµν = −gµν ẋµ ẋν , (2.126)
dλ dλ
and using the Euler-Lagrange equation. One then identies the resulting equations
with the anely parametrized geodesic equations and reads out the Christoel
symbols.
41
42
Chapter 3
Advanced Topics
3.1 Symmetries
We all know that symmetries are very important in physics because in most cases they
simplify the problem. We also know that symmetries are very important because a sym-
metry implies that something is conserved. This is of course the N÷ther's theorem. Here,
we discuss the symmetries of the spacetime metric. Intuitively, the notion of symmetry
is clear. If you hold up a round sphere, it looks the same no matter what way you rotate
it. We want a way to state this mathematically.
Lξ gµν = ∇µ ξν + ∇ν ξµ . (3.2)
To mathematically dene a symmetry, we need the concept of a family of integral
curves1 , also called a ow. A ow can then be identied with a tangent vector eld ξ
which points along the tangent vector to the ow at each point p ∈ M
dxµ
ξµ = . (3.3)
dλ
This ow is said to be an isometry if the metric looks the same at each point along
a given ow line. Mathematically, this means that an isometry satises
Lξ g = 0, (3.4)
which, according to (3.2) can be written
∇µ ξν + ∇ν ξµ = 0. (3.5)
1 The concept of integral curves was dened in Section 1.3.
43
This is the Killing equation and any tangent vector eld ξ satisfying this equation
is known as a Killing vector.
In practice, it is not always easy to nd Killing vectors except in some peculiar situa-
tions. Indeed, let us explicitly write that the Lie derivative along some tangent vector ξ
og the metric vanishes
ξ α ∂α gµν = 0. (3.7)
We then see that the tangent vector ξ = (0, ..., 1, ..., 0) where the 1 is at the nth posi-
tion is a Killing vector.
Moreover, it is easy to show that if ξ and χ are Killing vectors, then any linear com-
bination with non-varying constants aξ + bχ and [ξ, χ] are also Killing vectors.
Example. • For a two dimensional sphere described in (θ, ϕ) coordinates, the metric reads2
• For the Schwarzschild metris written in the usual spherical coordinates (t, r, θ, ϕ)
−1
2M 2M
2
ds = − 1 − 2
dt + 1 − dr2 + r2 dΩ2 , (3.9)
r r
one sees that it does not depend on t nor ϕ so that ξ = ∂t = (1, 0, 0, 0) and χ = ∂ϕ = (0, 0, 0, 1) are
Killing vectors.
Lξ g = 0 i.e. ∇µ ξν + ∇ν ξµ = 0. (3.10)
If the metric does not dependent on a coordinate xn , then
ξ µ = (0, ..., 0, 1
|{z} , 0, ..., 0), (3.11)
nth position
is a Killing vector.
44
are two important examples.
Let ξ µ be a Killing vector and xµ (λ) be a geodesic associated to the tangent vector
X µ , then the quantity
ξµ X µ , (3.12)
is a conserved quantity along the geodesic. Indeed,
X ν ∇ν (ξµ X µ ) = X ν (X µ ∇ν ξµ + ξµ ∇ν X µ )
µ ν
= X
| {zX } ∇ν ξµ +ξµ X ν ∇ν X µ
symmetric
| {z } | {z } (3.13)
anti-symmetric =0 geodesic
= 0.
Let ξ µ be a Killing vector and T µν the covariantly conserved symmetric energy-
momentum tensor, ∇µ T µν = 0, then the current
ξν T µν , (3.14)
is covariantly conserved. Indeed,
∇µ (ξν T µν ) = |{z}
T µν ∇µ ξν +ξν ∇µ T µν = 0. (3.15)
| {z } | {z }
symmetric
anti-symmetric =0
45
Let ξ µ be a Killing vector and xµ (λ) be a geodesic associated to the tangent vector
X µ i.e. X ν ∇ν X µ = 0, then the charge
ξµ X µ , (3.20)
is a conserved quantity along the geodesic.
ξν T µν , (3.21)
is covariantly conserved i.e. divergent-less.
a For example the energy-momentum tensor.
∇µ ∇ν ξ σ − ∇ν ∇µ ξ σ = Rρµν
σ
ξρ, (3.23)
and the Killing equation
∇µ ξν + ∇ν ξµ = 0, (3.24)
one obtains
σ
∇µ ∇ν ξρ = ∇ν ∇µ ξρ − Rρµν ξσ
σ
= −∇ν ∇ρ ξµ − Rρµν ξσ
σ σ
= −∇ρ ∇ν ξµ − Rρµν ξσ + Rµνρ ξσ
σ σ
(3.25)
= ∇ρ ∇µ ξν − Rρµν ξσ + Rµνρ ξσ
σ σ σ
= ∇µ ∇ρ ξν − Rρµν ξσ + Rµνρ ξσ − Rνρµ ξσ
σ σ σ
= −∇µ ∇ν ξρ − Rρµν ξσ + Rµνρ ξσ − Rνρµ ξσ
so that
σ
∇µ ∇ν ξρ = Rµνρ ξσ , (3.26)
where we used the Bianchi identity.
46
Let ξ µ be a Killing vector, then it must obey
σ
∇µ ∇ν ξρ = Rµνρ ξσ , (3.27)
where Rµνρ
σ
is the Riemann tensor.
σ
∇µ ∇ν ξρ (x) = Rµνρ (x)ξσ (x). (3.28)
In particular, this shows that the second derivatives of the Killing vector at a point x0
are again expressed in terms of the value of the Killing vector itself at that point. This
means that, remarkably, a Killing vector eld ξµ (x) is completely and uniquely determined
everywhere by the values of ξµ (x0 ) and ∇µ ξν (x0 ) at a single point x0 . Since, in an D-
dimensional spacetime there can be at most D linearly independent vectors (ξµ (x0 )) at
a point, and at most D(D − 1)/2 independent anti-symmetric matrices (∇µ ξν (x0 )), we
reach the conclusion that an D-dimensional spacetime can have at most
D(D − 1) D(D + 1)
D+ = , (3.29)
2 2
independent Killing vectors. A spacetime with this maximal number of Killing vectors is
called maximally symmetric.
Example. The D-dimensional Minkowski spacetime is maximally symmetric. Note that D(D + 1)/2
for D = 4 agrees with the dimension of the Poincaré group6 , the group of transformations that leave
the Minkowski metric invariant. We can also cite the de Sitter and the anti-de Sitter spacetimes being
maximally symmetric.
One can also show that maximally symmetric spacetimes have constant curvature
R = constant. For such spacetimes, one can show that the Riemann tensor takes the
following form
R
Rµνρσ = (gµρ gνσ − gµσ gνρ ). (3.30)
D(D − 1)
47
A D-dimensional spacetime is maximally symmetric is it has
D(D + 1)
, (3.31)
2
independent Killing vectors. Those spacetimes have constant curvature R =
constant and the Riemann tensor can be written
R
Rµνρσ = (gµρ gνσ − gµσ gνρ ). (3.32)
D(D − 1)
Let us recall the standard tensorial transformation behaviour of the metric under
coordinate transformations x → x0 ,
∂xµ ∂xν
0
gαβ (x0 ) = gµν (x). (3.33)
∂x0α ∂x0β
it follows that the absolute value of the metric determinant |g| = |det g| does not
transform like a scalar but instead transforms as
2
∂x0
∂x −2
g = det
0
g = det g. (3.34)
∂x0 ∂x
√
In particular its square rootg transforms as
0
∂x −1 √
g 0 = det (3.35)
p
g.
∂x
√
Therfore the combined expression dD x g is invariant under general coordinate trans-
formation
√
(3.36)
p
dD x0 g 0 = dD x g,
and can be used to dene integrals of scalars f (x) in a generally covariant way
48
ˆ ˆ
√
(3.37)
p
D 0 0
d x g f (x ) = dD x g f (x).
0
Note that this is also frequently the quickest way to determine the volume element in
non- Cartesian coordinates in Euclidean space.
Example. Let us derive the volume element of the three dimensional euclidean space in spherical
coordinates (r, θ, ϕ) denoted by {x0µ }. The usual Cartesian coordinates are denoted {xµ }. Instead of
laboriously determining the Jacobi matrix for the coordinate transformation, and then calculating its
determinant, all one needs to know is the metric
therefore we have
√
d3 x = g d3 x0 = r2 sin θ dr dθ dϕ, (3.39)
which is of course the standard result.
√
The covariant volume element for integration is dD x g which is invariant under
general coordinate transformation x → x0
√
(3.40)
p
dD x0 g 0 = dD x g,
and can be used to dene integrals of scalars f (x) in a generally covariant way
ˆ ˆ
√
(3.41)
p
D 0 0
d x g f (x ) = dD x g f (x).
0
49
This is clearly true for a diagonal matrix, since the determinant is the product of
eigenvalues while the trace is the sum. But both trace and determinant are invariant
under conjugation, so this is also true for diagonalisable matrices. Applying it to our
metric formula above, we have
1 1 1 1 1
Γµµν = Tr [∂ν log g] = ∂ν log det g = ∂ν detg = √ ∂ν detg, (3.45)
p
2 2 2 detg detg
which is the claimed result. With this at hand, we can now prove the following theorem.
√ √ √ 1 √ √
g ∇µ X µ = g(∂µ X µ + Γµµν X ν ) = g(∂µ X µ + X ν √ ∂ν g) = ∂ν ( g X µ ), (3.47)
g
50
The contraction of the Christoel symbols can be written
1 √
Γµµν = √ ∂ν g, (3.52)
g
wherea g = det g .
In General Relativity, the gravitational eld is identied with a metric gµν on a four
dimensional Lorentzian manifold that we call spacetime. To build the Einstein-Hilbert
action that governs the metric dynamics, we know, from a previous section, that we need
the covariant volume element to integrate over a manifold. Furthermore, given that we
only have the metric to play with, the simplest scalar function is the Ricci scalar R. This
motivates us to consider the wonderful concise action
ˆ
√
S= d4 x −g R, (3.54)
which is the famous Einstein-Hilbert action8 . Note that the minus sign under the square
root arises because we are in a Lorentzian spacetime. As a quick sanity check, recall that
the Ricci tensor takes the schematic form R ∼ ∂Γ + ΓΓ while the Levi-Civita connection
itself is Γ ∼ ∂g . This means that the Einstein-Hilbert action is second order in derivatives,
just like other actions we consider in physics.
Note that written in this way, the Einstein-Hilbert action is suitable to consider that
classical General Relativity is just an approximated theory of a more general theory with
the following action
ˆ
√
S= d4 x −g (R + σ2 R2 + σ3 R3 + ...), (3.55)
8 Note that a dimensional analysis requires that the physical action must be multiplied by c3 /(16πG).
51
where σi are constants. For example truncating the series expansion up to second order
i.e. keeping only the ∼ R term gives rise to the so-called Starobinsky potential which
2
In General Relativity, the gravitational eld is identied with a metric gµν on a four
dimensional Lorentzian manifold that we call spacetime. The metric dynamics is
governed by the Einstein-Hilbert action
ˆ
√
S= d4 x −g R. (3.56)
The aim is then to derive the variation of the inverse metric, the covariant volume
element and the Ricci tensor.
It turns out that it is slightly easier to think of the variation in terms of the inverse
metric δg µν . This is equivalent to the variation of the metric δgµν , the two are related by
gρµ g µν = δρν so that
First, we use the standard trick log detA = Tr[log A] for any diagonalisable matrix A
to write
1
δ(detA) = Tr[A−1 δA]. (3.59)
detA
Applying this result to the metric, we have
√ 1 1 1√
δ −g = √ (−g)g µν δgµν = −gg µν δgµν . (3.60)
2 −g 2
Using g µν δgµν = −gµν δg µν , one obtains the variation of the covariant volume element
9 However,this potential has been proved wrong by the latest Planck satellite data of the Cosmic
Microwave Background.
52
1√ √
δ −g gµν δg µν .
−g = − (3.61)
2
We now claim that that the nal term g µν δRµν is a total derivative (boundary term)
and can be written as10
ˆ
√ √ √
d4 x [δ −g]g µν Rµν + −g[δg µν ]Rµν + −g g µν [δRµν ] . (3.65)
δS =
Varying the covariant volume element and the Ricci tensor reads
√ 1√
δ −g = − −g gµν δg µν
2 (3.66)
g µν δRµν = ∇µ X µ with X µ = g ρν δΓµρν − g µν δΓρνρ .
The last variation turns out to be a boundary term. Using all the previous results,
the variation of the action is written
ˆ
√
1
δS = 4
dx µν µ
−g Rµν − R gµν δg + ∇µ X . (3.67)
2
Ignoring the boundary term and requiring the action to be extremised δS = 0, one
obtains the Einstein eld equation in vacuum
1
Gµν = Rµν − R gµν = 0. (3.68)
2
Note that they simplify somewhat; if we contract the last equation with g µν , we nd
that R = 0. Substituting this back in, the vacuum Einstein equation is simply the
requirement that the metric is Ricci at
Rµν = 0. (3.69)
10 We do not prove this result because it is very long. Note that one can easily derive this result in
normal coordinates but these are not introduced in these lecture notes. For interested people, I found a
Youtube video that explicitly makes this derivation.
11 In vaccum.
53
Note that we happily discarded the boundary term, a standard practice whenever we
invoke the variational principle. It turns out that there are some situations in General
Relativity where we should not be quite so cavalier. In such circumstances, one can be
more careful by invoking the so-called Gibbons-Hawking boundary term.
54
The Essential in a Scheme
Dierential Manifold M
Geodesics
´ by extremizing
Auto-parallel curves and
p
S= dλ −gµν X µ X ν .
parallel transport
Spacetime symmetries
Killing vectors ξ µ
Fundamental theorem
of Riemannian geormetry
Conserved quantities
Levi-Civita connexion
(torsion-free, ∇X g = 0)
´ √
Einstein-Hilbert action S = d4 x −g R
55