Maths BEC Notes
Maths BEC Notes
Marcin Napiórkowski
February 2022
Chapter 1
Introduction
1
not be an exaggeration to say that the basis of much of the subsequent work
has been laid in the seminal 2002 paper of Lieb and Seiringer [3] in which the
authors proved for the first time that the ground state of a dilute, trapped
gas exhibits Bose–Einstein condensation.
The goal of this lecture is to discuss some of the topics mentioned above
in a mathematically precise way. We will define objects and set up problems
in a rigorous way. Some proofs we will be done in details, others will only
be sketched. From a mathematical point of view, most of the results we
will discuss have been obtained very recently (up to 15 years). During the
course we will formulate open problems which could lead to research and/or
master/PhD theses.
2
Chapter 2
Quantum many-body
systems
and, since we consider bosons, are symmetric under the exchange of two
particles, i.e.
ψ(x1 , . . . , xi , . . . , xj , . . . , xN ) = ψ(x1 , . . . , xj , . . . , xi , . . . , xN )
3
is assumed to be positive, i.e. λ > 0. It represents a coupling constant that
could also be N dependent. We will soon come back to this.
Finally, the one-body term consists of the kinetic energy operator −δxj
of the j-th particle and a external potential V (xj ) which acts on the j-th
particle too. On should think of the V as being a trapping potential which
keeps the particles located in a certain region of space. One example would
be a trapping potential of the form V (x) = x2 which represents a harmonic
trap. Another example would be a potential that is of the form
0 if x ∈ Ω
V (x) =
∞ if x 6∈ Ω
In this case the particles would be confined in Ω ⊂ R3 .
4
2.3 Grand-canonical ensemble
It is often convenient not to fix the particle number N , but rather work in
the grand-canonical ensemble, where one takes a certain average over the
number of particles in the system. For simplicity, consider a system of just
one species of particles. The N -particle Hilbert space, HN , is then the set of
square-integrable functions that are totally symmetric under permutations.
In the grand-canonical ensemble, one has as Hilbert space the Fock space
∞
M
F= HN .
N =0
J = −T ln TrF e−β(H−µN )
∂
hN i = Tr N ρβ,µ = − J.
∂µ
5
2.4 BEC in ideal gas
Let us now consider an non-interacting (ideal) Bose gas, that is confined in
a box of side length L, i.e. Λ = [0, L]3 . The Hamiltonian is given by
N
X
n-in
HN = −∆xi
i=1
e0 ≤ e1 ≤ e2 ≤ . . . .
and also X
β(H − µN ) = εi a∗i ai
i≥0
The spectrum of i≥0 εi a∗i ai is of the form i≥0 εi ni , with ni ∈ {0, 1, 2, . . .}.
P P
Summing over all possible occupation numbers is the same as summing over
all eigenstates, hence we have
P ∗ 1
Tr e−
YX Y
i≥0 εi ai ai = e−εi ni = .
n
1 − e−εi
i i
Here we have to assume that εi > 0 for all i for the geometric series to
converge. Then
∗
P
ln Tr e− i≥0 εi ai ai =
X
− ln(1 − e−εi ).
i
6
Note that εi > 0 can be achieved by taking µ < 0. This is not really
a restriction, however, as any particle number can be achieved even for
negative µ. In fact, the average particle number equals
∂ X 1
hN i = − J = β(p 2 −µ) .
∂µ e
3|
− 1}
2π
p∈( Z) L
{z
ha∗p ap i
Here the summands are ha∗p ap i, the average occupation number of momen-
tum [Link] µ varies between (−∞, 0), clearly hN i varies between (0, ∞).
We now perform a thermodynamic limit L → ∞. The sum over p can
then be interpreted as a Riemann sum for the corresponding integral. In
fact, Z
1 X 1
−→ dp
L3 3
(2π)3 R3
2π
p∈( L Z)
7
since ρc (β) = ζ(3/2)(4π)−3/2 β −3/2 . Here ζ denotes the Riemann zeta func-
tion X 1
ζ(z) = .
kz
k≥1
8
(1)
Notice that we chose the normalization Tr γψN = N .
Analogously, for k = 2, 3, . . . , N , we can define the k-particle reduced
density associated with ψN by
(k) N
γψN := Trk+1,...,N |ψN ihψN |. (2.8)
k
N
(1) (1) (1) (1)
X
(ψN , J (1) ψN ) = (ψN , Ji ψN ) = N (ψN , J1 ψN ) = Tr J1 γψN .
i=1
The reduced densities can be lifted to the Fock space setting. Note that
using (B.3) we see that the integral kernel (2.7) can be written as
(1)
γψN (x; y) = (ψN , a∗y ax ψN ).
This leads to the general definition that the one-body reduced density matrix
is the one-body operator (acting on the one-body Hilbert space) γ (1) defined
through
(g, γ (1) f ) = ha∗ (f )a(g)i, ∀f, g ∈ L2 (R3 ).
Here h·i denotes the expectation value in any state. In particular this defini-
tion applies to any state on Fock space, not only thermal equilibrium states.
One can also consider states of definite particle number, and hence recover
the definition for the canonical ensemble.
9
with λi > 0. According to the (generally accepted) definition of Penrose and
Onsager, Bose–Einstein Condensation occurs when γ (1) has an eigenvalue of
order N (or hN i). Note that this definition is independent of the fact,
whether the system under consideration is interacting or not.
Let us now check, how this definition relates to the concept of BEC
derived for the ideal Bose gas. In (2.4) we have shown that above the
critical density
ha∗0 a0 i supkf k=1 ha∗ (f )a(f )iβ,µ
0 < ρ − ρc (β) = lim < lim .
L→∞ L3 L→∞ L3
In particular, supkf k=1 ha∗ (f )a(f )iβ,µ is the largest eigenvalue of γ (1) so in-
deed it has to be of order N (as when taking the thermodynamic limit we
assume ρ = N/L3 ).
Now, recall our analysis of BEC in the ideal case. It shows that when there
is BEC we have (in the thermodynamic limit)
hnp i = δ0 N0 + ñp
where the singular term arises from the macroscopic occupation of the zero
momentum state. Thus, taking the thermodynamics limit in (2.9), we obtain
Z
γ (1) (x − y) = N0 + dp ñp eip(x−y) .
10
2.6 The curse of dimensionality
We shall now argue that, due to its high dimensionality, the many-body
Schrödinger equation is impossible to solve numerically at a high precision
for most physical systems of interest.
To this end let us consider a much simpler, discrete model - a quantum
spin system. We place a quantum (say, 1/2) spin on each of the N sites of a
square lattice. Since each spin is described the Hilbert space C2 , the Hilbert
space of the system is given by the tensor product ⊗x C2 where x enumerates
all the N sites. Thus, the dimension of the Hilbert space is 2N . In other
words, to encode a general quantum state we need 2N coordinates. Assume
we want consider a square lattice with length 20 and thus with 400 sites
altogether. Now, let us assume very optimistically, that one number can be
encoded using one bit which is carried on one elementary particle. Thus,
we would need 2400 elementary particles just to encode on the computer
one state of our quantum spin system. How much is 2400 ? According to
recent estimates of astrophysicists, there are 1086 elementary particles in
the universe. Since 2400 > 1086 , in principle we are not able to encode a
quantum state on the computer without even doing any operations on it.
That was a crude estimate for a spin system. The models we want to
consider are continuous and consist of thousands of particles (like in modern
experiments involving Bose–Einstein Condensates). To put the problem on
a computer one needs to discretize it and the example discussed above shows
that the complexity of this problem will be exponential.
All this forces us to consider effective theories. An effective theory is
a proposal to describe a system using less degrees of freedom. Of course,
there is price to pay - one looses information about the system. Let us
illustrate this idea using an example from classical physics. In classical
mechanics, the microscopic theory is given by Newton’s equations of motion.
Solving these equations would in principle give us information about all
positions and momenta of all particles involved. If one wants to understand
the behaviour of air in a room, then solving Newton’s equations is not a good
idea to say the least. But do we really need to know all the positions and
velocities of all particles? Maybe it would be enough to know the position
and momentum of a typical particle - that is, the probability that the particle
occupies a given very small region of the phase space (mathematically the
volume element dxdp) at an instant of time. We would thus be looking
for a probability distribution f (t, x, p) which now depends on only three
variables. Ludwig Boltzmann derived an equation for f which is now called
the Boltzmann equation. A rigorous justification of this equation starting
from Newton’s microscopic theory is still an active area of research and this
field of mathematical physics is called kinetic theory.
In this course we will try to understand to most important properties of
quantum many-boson systems by rigorously justifying effective theories. We
11
will be interested in effective descriptions of the the ground state properties,
excitation spectrum, effective dynamics. An effective theory usually is some
kind of limiting theory in a certain regime. In the next section we will
discuss the scaling regimes we will be studying.
NN
defined on HN = 2 3
sym L (R ). We want to find an effective description
of the model. Keeping in mind that we are interested in thermodynamic
quantities, we have to remember that we we will be taking the N → ∞
limit. We observe that the kinetic energy (or, more generally, the one-body
operator) is of order N as it consists of N terms which are independent of N .
If we assume that the two-body interaction is N independent, i.e. wN ≡ w,
then the interaction term is of order N 2 (for large particle numbers) as there
are N (N − 1)/2 terms in the double sum. In the limit when N → ∞ the
interaction term would thus become dominant. But we want to keep the
kinetic term as, in some sense, it is more quantum than the pure interaction
(recall, BEC has been computed for the ideal gas). Thus, the idea is to
balance these two terms (at least, as far as the naive asymptotics in N
would suggest). This motivates setting λ = O(N −1 ) and this scaling is
called mean-field scaling. The simplest mean-field Hamiltonian is of the
form
N
X 1 X
HN = −∆xj + V (xj ) + w(xj − xk ).
N
j=1 16j<k6N
Note, that in this setup there are two length scales. The first one is given by
the trapping potential V and basically tells us where the particles can move.
It is important for the kinetic term, as it has implications in the spectral
gap of the one-body operator.
But for now let us focus on the length scale set by the interaction. When
wN ≡ w, then it is independent of N and therefore is of order O(1) (with
respect to N ). In fact, the length scale of the interaction is characterized
by a parameter called the scattering length. To define the scattering length,
we consider the two-body scattering problem. We assume the two-body
interaction is radial, positive and has a finite range. Consider the zero-
energy scattering equation
1
− ∆f + wf = 0 (2.11)
2
12
with the boundary condition f (x) → 1 as |x| → ∞. Under the assumption
of compact support of w one can show that for |x| sufficiently large, the
scattering solution satisfies
a0
f (x) = 1 − (2.12)
|x|
for a constant a0 > 0. This constant is called the scattering length. Equiv-
alently, the scattering length a0 can also be defined through the integral
Z
8πa0 = w(x)f (x)dx (2.13)
where f (x) is the solution of (2.11). From the point of view of physics, the
scattering length a0 measures the effective range of the interaction potential;
two quantum mechanical particles interacting through the potential w, when
they are far apart, feel the other particle as a hard sphere with radius a0 (in
particular, the scattering length of a hard sphere potential coincides with
the radius of the sphere).
Problem 2.7.1. Show that (2.13) holds true.
Problem 2.7.2. Consider the hard sphere potential. Show that the scatter-
ing length is in this case equal to the radius of the hard sphere.
Problem 2.7.3. Let f be the solution of the scattering equation for a posi-
tive interaction potential. Show that 0 ≤ f ≤ 1. Deduce that
R
w
a0 ≤ .
8π
Having defined the scattering length, let us come back to the mean-
field model. Assuming the trapping potential restricts the particles to a
volume of order O(1), we see that the density is of order N and the average
distance between the particles is N −1/3 . Thus it is much smaller that the
effective range of the interaction. In other words, each particle feels and
interacts with many other particles, but the strength of interactions is weak
(of order O(1)). This situation corresponds to a high density regime where
the particles meet very often but interact only a little bit each time.
To model a different situation, which will call the dilute regime (a typical
situation in experiments with ultracold gases), we can scale the interaction
and make it N dependent. Assume wN (x) = N 2 w(N x). By rescaling x =
N x̃ we see from (2.11) that
1 1
− 2 ∆x̃ + w(N x̃) f (N x̃) = 0
N 2
Thus f (N ·) solves the scattering equation
1
−∆ + N 2 w(N ·) f (N ·) = 0.
2
13
It follows that as |x| → ∞ we have
a0 a0 /N
f (N x) = 1 − =1−
N |x| |x|
aρ1/3 1, (2.14)
which means that the effective range of the interaction a is much smaller
than the average distance between the particles ρ−1/3 .
14
Chapter 3
We shall now discuss the properties of the ground state energy (i.e. T = 0)
of a system of interacting bosons, i.e.
E0 (N ) = inf (ψ, HN ψ)
ψ∈L2sym (R3N ),kψk=1
where
N
X 1 X
HN = −∆xj + V (xj ) + wN (xj − xk )
j=1
| {z } N − 1 16j<k6N
=:hxj
Recall, that the mean-field scaling correspond to a situation when the in-
teraction between the particles is frequent, but weak. We know that the
ground state of a non-interacting system (i.e. when w ≡ 0) is given by a
product state ũ⊗N (x1 , . . . , xN ) := ũ(x1 ) · · · ũ(xN ) where ũ is the (normal-
ized) ground state of the one-body operator. In the mean-field scaling the
interactions are supposed to be weak, so a simple minded argument suggests
that maybe the ground-state will not be changed dramatically and will be
close, in some sense, also to a product state.
15
Let us make this ansatz, i.e. ψ(x1 , . . . , xN ) = u⊗N (x1 , . . . , xN ) for some
normalized function u ∈ L2 (R3 ). Computing the energy leads to the follow-
ing formula
(u⊗N , HN u⊗N ) = N E(u)
where
ZZ
1
E(u) = (u, hu) + w(x − y)|u(x)|2 |u(y)|2 dxdy. (3.2)
2 R3 ×R3
Problem 3.1.2. Let us assume that the Hartree functional has a minimizer
and let us denote it by u0 . It satisfies the so-called Hartree equation
hu0 + |u0 |2 ∗ w u0 = ε0 u0
(3.4)
R
where (f ∗ g)(x) = R3 f (x − y)g(y)dy denotes the convolution and ε0 is a
Lagrange multiplier due to the normalization constraint.
Note that in general there is no reason to expect that minimizers will be
unique. Uniqueness can be broken because of the one-body Hamiltonian h
or due to the interaction if it has an attractive part.
Let us denote by eH the infimum of the Hartree functional, i.e.
eH := E(u0 ).
From (3.3) we obtain
E0 (N )
≤ eH .
N
A natural question whether a matching lower bound holds. This is indeed
true.
Theorem 3.1.3 (Convergence of ground-state energy). Under appropriate
assumptions on h and w one has
E0 (N )
lim = eH .
N →∞ N
We will not state the most general assumptions on h and w that were
used to prove the most general version of Theorem 3.1.3 given by Lewin–
Nam–Serfaty–Solovej in 2014. Let us just mention that the one-body oper-
ator h can be much more general than in (3.1).
We will give a proof in a simpler case. We follow the proof of Lewin
(2015). Let us start with the simplest case when the following two conditions
are satisfied:
16
• h is symmetric and positive preserving: (u, hu) ≥ (|u|, h|u|);
Proof. We know from the definition of the reduced one-body matrix that
N
(1)
X X X
(ψN , hxi ψN ) = Tr(hγψN ) = nk (uk , huk ) ≥ nk (|uk |, h|uk |)
i=1 k k
(1) P
where we used the decomposition γψN = k nk |uk ihuk | and the assumption
on h. For real functions u1 , u2 we have
q q
(u1 , hu1 ) + (u2 , hu2 ) = (u1 + iu2 , h(u1 + iu2 )) ≥ 2 2 2 2
u1 + u2 , h u1 + u2 .
(1)
nk |uk |2 = γψN (x, x) = N ρψN (x) we get the desired
P
Using the fact that k
result.
Proof. We use
ZZ Z
w(x − y)f (x)f (y)dxdy = (2π) 3/2
ŵ(k)|fˆ(k)|2 dk ≥ 0
R3 ×R3 R3
PN
with f = i=1 δxi − η and expanding all terms.
17
We are ready to give the proof of Theorem 3.1.3.
(k)
The normalization is then such that Tr γψN = 1.
Theorem 3.1.6 (Convergence of ground states). Assume that h and w
satisfy the inequality
−C1 (T1 + T2 ) − C ≤ w(x1 − x2 ) ≤ C2 (T1 + T2 ) + C
for some constants 0 ≤ C1 , C2 < 1. Assume V (x) → ∞ as |x| → ∞ and
that w is symmetric and smooth enough. Let ψN be the ground state for HN
(n)
and γψN its n-th reduced density matrix with the normalization that
Then there exists a subsequence and a probability measure µ on the set
of minimizers of the Hartree functional M such that
Z
(k)
γψN → |u⊗k ihu⊗k |dµ(u) (3.5)
j
M
18
Theorem 3.1.7 (quantum de Finetti). Let H be a separable Hilbert space
and let Γ(n) be a sequence of positive, trace-class operators satisfying
where Trn+1 is the partial trace with respect to the last variable in Hn+1 .
Then there exists a unique Borel probability measure µ on the sphere SH =
{u ∈ H : kuk = 1} of H, invariant with respect to multiplication by a phase,
such that Z
(n)
Γ = |u⊗n ihu⊗n |dµ(u)
SH
19
(1)
It follows that TrH [T γψN ] is uniformly bounded and thus so is TrH [(T +
(1)
C0 )γψN ] which by the cyclicity if the trace yields (possibly, up to a further
subsequence) that
(1)
(T + C0 )1/2 γψN (T + C0 )1/2 *∗ (T + C0 )1/2 γ (1) (T + C0 )1/2
is also uniformly bounded in N and that nj=1 Tj also has compact resolvent
P
which allows for similar argument.
Thus we have proven for any n ∈ N that
(n)
γψN → γ (n)
but also
(n+1) (n+1) (n)
TrHn+1 [γψN (Bn ⊗1)] = TrHn [(Trn+1 γψN )Bn ] = TrHn [γψN Bn ] → TrHn [γ (n) Bn ].
This implies
Trn+1 γ (n+1) = γ (n)
and thus we can use Theorem 3.1.7. Before that, using the fact that (by
assumptions) T1 + T2 + w is bounded below by, say, CT , we write
1 (2) 1 (2)
lim inf TrH2sym [(T1 + T2 + w)γψN ] = TrH2sym [(T1 + T2 + w − 2CT )γψN ] + CT
N →∞ 2 2
1
≥ TrH2sym [(T1 + T2 + w − 2CT )γ (2) ] + CT
2
1
= TrH2sym [(T1 + T2 + w)γ (2) ]
2
20
where we used Fatou’s lemma for positive operators. Using the quantum de
Finetti theorem we obtain
Z
E0 (N ) 1
eH ≥ lim inf ≥ TrH2sym [(T1 + T2 + w)|u⊗2 ihu⊗2 |]dµ(u)
N →∞ N 2
Zu∈SH
≥ EH (u)dµ(u) = eH .
u∈SH
The second part of the theorem follows from the fact that all inequalities
become equalities.
(ψN , HN ψN ) = N eH + o(N )
we have
(1)
γψN → |u0 ihu0 |.
Thus the ground state exhibits Bose–Einstein condensation.
with wN (x) = N 3β w(N x) and β ∈ (0, 1]. For strictly positive β, the function
wN provides an approximation to the identity (up to normalization), i.e.
Z
3β
wN (x) = N w(N x) →N →∞ w δ0
21
Consequently, in the scaling regime when β > 0 the Hartree equation (3.4)
has to be modified and the new effective equation reads
Z
hu0 + w |u0 |2 u0 = ε0 u0 . (3.7)
If one assumes the two-body interaction is positive, then Lieb and Seiringer
proved that statements analogous to Theorems 3.1.3 and 3.1.6 hold true
when β = 1 with the effective theory now being given by the Gross–
Pitaevskii functional. The presence of the scattering length is, from a physics
perspective, crucial as it is a parameter which experimentalists can manip-
ulate.
One can interpret these results as a universality result. The effective
theory does not depend on the details of the underlying microscopic system,
but rather on an quasi-macroscopic quantity that is accessible in the lab.
Finally, let us mention that the Gross–Pitaevski functional cannot be
derived using a product wave-function as an ansatz. This would to the NLS
functional. It follows that the trial wave-function must include correlations
between the particles. Those correlations appear on very short length scales
and are therefore difficult to capture. We will not discuss the related details
in this course and refer to the original papers.
22
Chapter 4
Bogoliubov theory
P 0 = P − M V,
1 1 (4.1)
E 0 = P 02 /(2M ) = |P − M V |2 = E − P V + M V 2
2M 2
where M is the total mass of the fluid.
We first consider a fluid at zero temperature, in which all particles are in
the ground state and flowing along a capillary at constant velocity v. If the
fluid is viscous, the motion will produce dissipation of energy via friction
with the capillary wall and decrease of the kinetic energy. We assume that
such dissipative processes take place through the creation of elementary
excitation. Let us first describe this process in the reference frame K which,
rather confusingly, moves with the same velocity v of the fluid. In this
reference frame, the fluid is at rest and its energy is the ground state energy
that we denote by E0 . If a single elementary excitation with a momentum
p appears in the fluid, the total energy of the fluid in the reference frame
K is E0 + (p), where (p) is the energy of the excitation with momentum
p. Let us move to the moving frame K 0 in which the fluid moves with a
velocity v but the capillary is at rest. In this moving frame K 0 which moves
with the velocity v with respect to the fluid, the energy and momentum of
23
the fluid are given by setting V = −v in (4.1). We obtain
P 0 = p + M v,
1
E 0 = E0 + (p) + pv + M v 2 .
2
The above results indicate that the changes in energy and momentum caused
by the appearance of one elementary excitation are (p) + pv and p, respec-
tively.
Spontaneous creation of elementary excitations, i.e. energy dissipation,
can occur if and only if such a process is energetically favorable. This re-
quires
(p) + pv < 0.
This is satisfied when |v| > (p)|p| and pv < 0, i.e. when the elementary
excitation has the momentum p opposite to the fluid velocity v and the fluid
velocity |v| exceeds the critical value
(p)
vcr = min . (4.2)
p |p|
Thus, superfluidity will occur only if the critical velocity is strictly positive.
In particular, the ideal Bose gas has (p) = p2 and thus vcr = 0 so it is
not superfluid. In particular, the particle-particle interaction is a crucial
requirement for the appearance of superfluidity in a bosonic system.
24
because there are (N − 1) particles with zero momentum and only 2 with
nonzero momentum. Applying w over and over again we will ultimately get
a finite fraction of triplets, quartets, etc. as well as pairs, but hopefully if
the interaction is weak enough we need consider explicitly only pairs in the
ground state wave function. This suggests that only terms with pairs should
be relevant in the ”effective” Hamiltonian.
These were the arguments that led N. N. Bogoliubov to drop all terms
in the Hamiltonian involving more than two creation/annihilation operators
of a non-zero mode. Doing this we obtain
X a∗0 a0
ŵ(0) ∗ ∗
H≈ a a a0 a0 + k + 3 (ŵ(k) + ŵ(0) a∗k ak
2
2L3 0 0 L
k6=0
X ŵ(k)
+ (a∗0 a∗0 ak a−k + a∗k a∗−k a0 a0 )
2L3
k6=0
ŵ(0)ρ
= (N − 1) + HBog + R
2
where
ρ = N/L3 ;
X 1X
HBog := (k 2 + ρŵ(k))a∗k ak + ρŵ(k)(a∗k a∗−k + ak a−k );
2
k6=0 k6=0
−ŵ(0) X ŵ(k)
(a∗0 a∗0 − N )ak a−k ) + a∗k a∗−k (a0 a0 − N ) .
R := 3
(N − N 0 )(N − N0 − 1) + 3
2L 2L
k6=0
We used
a∗0 a∗0 a0 a0 = N0 (N0 − 1) = N (N − 1) − 2N0 (N − N0 ) − (N − N0 )(N − N0 − 1).
We argue that R is small, because
a∗0 a∗0 ≈ a0 a0 ≈ N0 ≈ N.
The last step is called the c-number substitution. Thus the effective Hamil-
tonian is given by a quadratic operator (quadratic in creation and annihi-
lation operators) that does not conserve the number of particles. It can
however be diagonalized. To this end we use a Bogoliubov transformation
By introducing
bp = cp ap + sp a∗−p , b∗p = cp a∗p + sp a−p
with appropriate cp , sp (see exercises) such that c2p − s2p = 1. Then
X
HBog = EBog + (p)b∗p bp (4.3)
p6=0
with p
(p) = |p| p2 + 2ρŵ(p).
25
Problem 4.2.1. Derive (4.3)
Note that it follows that
(p) p 2
= p + 2ρŵ(p)
|p|
and thus the critical velocity vcr > 0. Thus Bogoliubov’s computations
shows that the interacting Bose gas satisfies Landau’s criterion for superflu-
idity.
Let us define
HN
0 := Span(u ⊗ . . . × u0 )
| 0 {z }
N times
and for k > 0
k
⊗(N −k)
O
HN
k := Span(u0 ⊗ . . . × u0 ) ⊗s H+ = u0 ⊗s Hk+
| {z }
N −k times sym
HN = HN N N
0 ⊕ H1 ⊕ · · · ⊕ HN .
26
with ψk ∈ Hk+ . It is easy to check that
⊗(N −k) ⊗(N −l)
hu0 ⊗s ψk , u0 ⊗s ψl iHN = δkl hψk , ψl iHl
given by
UN (Ψ) = ψ0 ⊕ ψ1 ⊕ ψ2 ⊕ . . . ⊕ ψN
≤N
is a unitary operator from HN onto the truncated, excited Fock space F+ .
The latter can always be seen as being embedded in the the full excited Fock
space given by
M∞
F+ := Hn+ .
n=0
The full Fock space of excited particles appears as the limit of the truncated
excited Fock space when N → ∞.
The operator UN has some important properties. In particular we have
that UN can be written as
N
!
N −j
M a
UN (Ψ) = Q⊗j p 0 Ψ
j=0
(N − j)!
for all Ψ ∈ HN . Here Q = 1 − |u0 ihu0 | is the projection onto the excited
space. Similarly
N N
∗
M X (a∗ )N −j
UN ψj =
p 0 φj
j=0 j=0
(N − j)!
≤N
for all φj ∈ Hj+ . These operators satisfy the following indentities on F+ :
UN a∗0 a0 UN
∗
= N − N+ ,
UN a∗ (f )a0 UN
∗
= a∗ (f ) N − N+ ,
p
(4.4)
UN a∗0 a(f )UN
∗
p
= N − N+ a(f ),
UN a∗ (f )a(g)UN
∗
= a∗ (f )a(g).
∗
P
Here N+ = m≥1 am am is the operator counting the number of excited
particles.
27
4.3.2 The Bogoliubov Hamiltonian
We define the Bogoliubov Hamiltonian in the following way:
X 1 1
H= hum , (h + K1 )un ia∗m an dxdy + hum ⊗ un , K2 ia∗m a∗n + hK2 , um ⊗ un iam an
2 2
m,n≥1
(4.5)
where K1 : H+ → H+ and K2 : H+ → H+ are operators defined by
ZZ
hu, K1 vi = u(x)v(y)u0 (x)u0 (y)w(x − y)dxdy,
Z ZΩ⊗Ω (4.6)
hu, K2 vi = u(x)v(y)u0 (x)u0 (y)w(x − y)dxdy
Ω⊗Ω
28
for some constants c, C > 0. We will now state the main result about
Bogoliubov approximation, which heuristically can be written as
∗
UN (HN − N eH )UN → H.
Here it is.
Theorem 4.3.1 (Validity of Bogoliubov approximation). Under the as-
sumptions (A1)-(A3) the following holds true:
1 - weak convergence to H. For any Φ and Φ0 in the (quadratic form)
domain of H we have
lim UN ΨL
N =Φ
L
N →∞
where Tmn = hum , (−∆ + V )un i and Wmnpq = hum ⊗ un , w(x − y)up ⊗ uq i.
A tedious but straightforward computation using the relations (4.4)
shows that
X4
∗
UN (HN − N eH )UN = Aj
j=0
29
where
1 N+ (N+ + 1)
A0 = W0000 ,
2 N −1
X N − N+ − 1 p
A1 = T0m + W000m N − N+ am + h.c.,
N −1
m≥1
X
A2 = Tmn a∗m an − (T00 + W0000 )N+
m,n≥1
X N − N+
+ hum , (|u0 |2 ∗ w + K1 )un ia∗m an
N −1
m,n≥1
p
1 X (N − N+ )(N − N+ − 1)
+ hum ⊗ un , K2 ia∗n a∗m + h.c. ,
2 N −1
m,n≥1
1 X
Wmnp0 a∗m a∗n ap
p
A3 = N − N+ + h.c.,
N −1
m,n,p≥1
1 X
A4 = Wmnpq a∗m a∗n ap aq .
2(N − 1)
m,n,p,q≥1
≤M
We will show that when evaluated on a state Φ ∈ F+ , then A2 − H can be
made small. To this end we will use to following lemma whose proof follows
basically from (4.7).
Lemma 4.3.2. Under the assumptions on the interaction we have the fol-
lowing estimates
1
dΓ(QT Q) ≤ H + CN+ + C,
1 − α1 (4.9)
α2
dΓ(Q(|u0 |2 ∗ w)Q ≤ H + CN+ + C
1 − α1
where 1 > α1 > 0 and α2 > 0.
≤M
Let Φ ∈ F+ . Then
D N+ − 1 E
hA2 iΦ − hHiΦ = − dΓ(Q(|u0 |2 ∗ w)Q + K1 )
N −1 Φ
X
∗ ∗
(4.10)
+< hum ⊗ un , K2 ihan am XiΦ
m,n≥1
30
√
(N −N )(N −N −1)
+ +
where X = N −1 − 1.
Let us rewrite w = w+ − w− where w± ≥ 0. By linearity, using the
triangle inequality, we have
D N+ − 1 E D N+ − 1 E
− dΓ(Q(|u0 |2 ∗ w)Q) ≤ dΓ(Q(|u0 |2 ∗ w+ )Q)
N −1 Φ N −1 Φ
D N+−1
E
+ dΓ(Q(|u0 |2 ∗ w− )Q) .
N −1 Φ
Since the convolution of two non-negative functions is a non-negative func-
tion and as the multiplication operator by a non-negative function is non-
negative operator, we have that
dΓ(Q(|u0 |2 ∗ w± )Q) := A± ≥ 0
is a positive operator that commutes with N+ . It follows that A± =
(A∗± A± )1/2 and we can write
N+ − 1 q ∗ N+ − 1 p
A± = A± A± .
N −1 N −1
√
Using the fact that A± (and therefore A± ) does not change the number
of particles of a given state we have
Dq N+ − 1 p E M
A∗± A± ≤ hA± iΦ
N −1 Φ N −1
≤M
for Φ ∈ F+ . Thus
D N+ − 1 E M D E
− dΓ(Q(|u0 |2 ∗ w)Q) ≤ dΓ(Q(|u0 |2 ∗ (w+ + w− ))Q)
N −1 Φ N −1 Φ
M α2
≤ h H + CN+ + CiΦ
N − 1 1 − α1
where in the last step we used (4.9).
Furthermore, since for any ψ with kψk = 1 we have by the Cauchy-
Schwarz inequality
|hψ, ±K1 ψi| ≤ kK1 ψk ≤ kK1 k
it follows that
kK1 k1 + K1 ≥ 0.
Thus, similarly as above we get
D N+ − 1 E D N+ − 1 E
− dΓ(QK1 Q) = dΓ(Q(kK1 k1 + K1 −kK1 k1)Q)
N −1 Φ | {z } N −1 Φ
≥0
M D E M D E
≤ dΓ(Q(kK1 k1 + K1 )Q) + dΓ(QkK1 k1Q)
N −1 Φ N −1 Φ
3M kK1 k
≤ hN+ iΦ .
N −1
31
It follows that for the first term on the RHS of (4.3.3) we get the bound
D N+ − 1 E CM
− dΓ(Q(|u0 |2 ∗ w)Q + K1 ) ≤ hH + N+ 1iΦ
N −1 Φ N −1
where the constant C > 0 is independent of N . √
+ (N −N )(N −N −1)
+
It remains to look at the second term in (4.3.3). Recall X = N −1 −
1. By the Cauchy-Schwarz inequality we have
X X X
hum ⊗ un , K2 iha∗n a∗m XiΦ ≤ ( |hum ⊗ un , K2 i|2 )1/2 ( |ha∗n a∗m XiΦ |2 )1/2
m,n≥1 m,n≥1 m,n≥1
ZZ
1/2
X
≤( |K2 (x, y)|2 dxdy)1/2 ( ha∗n a∗m am an iΦ )1/2 hX 2 iΦ .
m,n≥1
32
Appendix A
A.1 Operators
Recall that a Hilbert space H is a vector space endowed with a sesquilinear
map (·, ·) : H × H → C (i.e., a map which is conjugate
p linear in the first
variable and linear in the second) such that kφk = (φ, φ) defines a norm
on H which makes H into a complete metric space.
One basic property of Hilbert spaces that we will use is the fact that for
any closed subspace V ⊂ H there corresponds the orthogonal complement
V ⊥ such that V ⊕ V ⊥ = H.
Another one goes under the name of the Riesz representation theorem:
to any continuous linear functional f : H → C there is a unique ψ ∈ H such
that f (φ) = (ψ, φ) for all φ ∈ H.
We shall always assume that our Hilbert spaces are separable and there-
fore that they have countable orthonormal bases.
Definition A.1.1. By an operator (or more precisely densely defined oper-
ator) A on a Hilbert space H we mean a linear map A : D(A) → H defined
on a dense subspace D(A) ⊂ H. Dense refers to the fact that the norm
closure D(A) = H.
Definition A.1.2. If A and B are two operators such that D(A) ⊆ D(B)
and Aψ = Bψ for all ψ ∈ D(A) then we write A ⊂ B and say that B is an
extension of A.
Note that the domain is part of the definition of the operator. In defining
operators one often starts with a domain which turns out to be too small
and which one then later extends.
Definition A.1.3. We say that A is a symmetric operator if
33
It is a fact that (A.1) holds if and only if (ψ, Aψ) ∈ R for all ψ ∈ D(A).
This property is of great importance in quantum mechanics as expectations
values of observables need to be real.
Definition A.1.4. An operator A is said to be bounded on the Hilbert space
H if D(A) = H and A is continuous, which by linearity is equivalent to
kAk = sup kAφk < ∞.
φ,kφk=1
The number kAk is called the norm of the operator A. An operator is said
to be unbounded if it is not bounded.
Definition A.1.5. If A is an operator we define the adjoint A∗ of A to be
the linear map A∗ : D(A∗ ) → H defined on the space
D(A∗ ) = {φ ∈ H| sup |(φ, Aψ)| < ∞}
ψ∈D(A),kψk=1
for all φ ∈ H.
Definition A.1.7. An operator A defined on a subspace D(A) of H is said
to be positive (or positive definite) if (ψ, Aψ) > 0 for all non-zero ψ ∈ D(A).
It is said to be positive semi-definite if (ψ, Aψ) ≥ 0 for all ψ ∈ D(A). In
particular, such operators are symmetric.
Definition A.1.8. If A and B are two operators with D(A) = D(B) then
we say that A is (strictly) less than B and write A < B if the operator
B − A (which is defined on D(B − A) = D(A) = D(B) is a positive definite
operator. We write A ≤ B if B − A is positive semi-definite.
34
We conclude that if K is a positive, compact operator then
X
K= λi |ui ihui | (A.2)
i
where now λi ’s can be chosen positive.
Among compact operators two special classes are of particular impor-
tance: trace class operators and Hilbert-Schmidt operators.
Definition A.1.9. Let H be a separable Hilbert space and {φn } its orthonor-
mal basis. Then for any positive operator A we define its trace to be
X
Tr A = (φn , Aφn ).
n
Moreover, Z
kAk22 = |K(x, y)|2 dµ(x)dµ(y).
35
A.2 Tensor product of Hilbert spaces
Let H and K be two Hilbert spaces. Consider Z = H × K. We form a vector
space which has Z as a basis. To this end we consider the vector space
of functions from Z → C and identify (v, w) with the function that takes
the value 1 on (v, w) and 0 otherwise. Let Z0 be the subspace spanned by
elements of the form
(u, v1 + v2 ) − (u, v1 ) − (u, v2 ),
(u1 + u2 , v) − (u1 , v) − (u2 , v),
(λu, v) − λ(u, v),
(u, λv) − λ(u, v)
where λ ∈ C.
Definition A.2.1. The algebraic tensor product is defined by
H ⊗al K = Z/Z0
and the image of (v, w) in this quotient is denoted v ⊗ w and is called the
tensor multiplication.
We introduce the inner product that satisfies
(u1 ⊗ v1 , u2 ⊗ v2 )H⊗al K = (u1 , u2 )H (v1 , v2 )K .
We set
H ⊗ K := (H ⊗al K)cpl .
Thus, Span{u ⊗ v|u ∈ H, v ∈ K} is dense in H ⊗ K. We call the vectors of
the form u ⊗ v pure tensor products.
If we have an operator A on the Hilbert space H and an operator B on
the Hilbert space K, then we may form the tensor product operator A ⊗ B
on H ⊗ K with domain
D(A ⊗ B) = Span{φ ⊗ ψ|φ ∈ D(A), ψ ∈ D(B)}
and acting on pure tensor products as
A ⊗ B(φ ⊗ ψ) = (Aφ) ⊗ (Bψ).
The tensor product may in a natural way be extended to more than two
Hilbert spaces. N
In particular, we may for N = 1, 2, . . . consider the N -fold
N
tensor product H of a Hilbert space H with itself. On this space we
have a natural action of N N group SN . I.e., if σ ∈ SN , then we
the symmetric
have a unitary map Uσ : N H → N H defined uniquely by the following
action on the pure tensor products
Uσ u1 ⊗ . . . ⊗ uN = uσ−1 (1) ⊗ . . . ⊗ uσ−1 (N ) .
NN NN
We shall denote by Ex : H → H the unitary corresponding to a
simple interchange of the two tensor factors.
36
Definition A.2.2. Let P+ be the orthogonal projection given by
X
P+ = (N !)−1 Uσ .
σ∈SN
hj = 1 ⊗ . . . ⊗ 1 ⊗ hj ⊗1 ⊗ . . . ⊗ 1.
|{z}
jth slot
The operator
n-in
HN = h1 + . . . + hN
may then be defined on the domain
n-in
D(HN ) = Span{φ1 ⊗ . . . ⊗ φN |φ1 ∈ D(h1 ), . . . , φN ∈ D(hN )}
37
Appendix B
The Fock space allows to describe states of the system where the number
of particles is not fixed. A normalized vector Ψ = {ψ (n) }n≥0 ∈ F describes
a state that with probability kψ (n) k22 has n particles. In particular, states
with exactly N particles are embedded in the Fock space; they are described
by vectors of the form {0, . . . , 0, ψN , 0, . . .} ∈ F having only one non-zero
component. These vectors are eigenvectors of the particles number operator
N defined by
(N Ψ)(n) = nψ (n)
for any Ψ = {ψ (n) }n≥0 such that
X
n2 kψ (n) k22 < ∞.
n≥0
38
B.2 Creation and annihilation operators
For any one-particle wave-function f ∈ L2 (R3 ) we define the creation oper-
ator a∗ (f ) and the annihilation operator a(f ) by setting
n
∗ (n) 1 X
(a (f )Ψ) (x1 , . . . , xn ) = √ f (xj )ψ (n−1) (x1 , . . . , xj−1 , xj+1 , . . . , xn ),
n
j=1
√
Z
(a(f )Ψ)(n) (x1 , . . . , xn ) = n + 1 f (xn+1 )ψ (n+1) (x1 , . . . , xn , xn+1 )dxn+1
a(f )Ω = 0.
[a(f ), a(g)] = [a∗ (f ), a∗ (g)] = 0, [a(f ), a∗ (g)] = hf, gi, ∀f, g ∈ L2 (R3 ).
(B.1)
Problem B.2.2. Check that a∗ (f ) and a(f ) are formal adjoints, i.e.
39
Despite the bosonic creation and annihilation operators being unbounded,
usually domain questions are unproblematic since one can take the domain
of a sufficiently large power of the number operator to easily make sense
of most expressions - see (B.5). One has to be more careful in this respect
when working with the operator-valued distributions as formally ax is ob-
tained form a(f ) by taking f to be the Dirac δ-function which is not an
element of the one-body Hilbert space.
Notice that if Ψ = {ψ (n) }n≥0 , then
√
(ax Ψ)(n) (x1 , . . . , xn ) = n + 1ψ (n+1) (x, x1 , . . . , xn ). (B.3)
Hence
X Z (n)
(Ψ, N Φ) = n dx1 . . . dxn ψ (x1 , . . . , xn )φ(n) (x1 , . . . , xn )
n≥0
XZ (n−1)
= dxdx1 . . . dxn−1 (ax Ψ) (x1 , . . . , xn−1 )(ax Φ)(n−1) (x1 , . . . , xn−1 )
n≥1
Z
= dx(ax Ψ, ax Φ).
for any f ∈ L2 (R3 ). To prove the first bound, we observe that by the
Cauchy–Schwarz inequality
Z Z 1/2
2
ka(f )ΨkF ≤ dx|f (x)|kax ΨkF ≤ kf k dxkax ΨkF = kf kkN 1/2 Ψk.
The second estimate follows from the first one and the canonical commuta-
tion relations.
40
a∗x and ax . Let J (1) be an operator on the one-particle space L2 (R3 ). The
second quantization of J (1) is the operator dΓ(J (1) ) on F defined by the
requirement that
n
(1)
X
(dΓ(J (1) )Ψ)(n) = Ji ψ (n)
i=1
(1)
where Ji denotes the operator acting on L2 (R3n ) as J (1) on the i-th particle
and as the identity on the other (n−1) particles. If the one-particle operator
J (1) has the integral kernel J (1) (x; y), we can write
n
(1)
XX
(Φ,dΓ(J (1)) Ψ) = (φ(n) , Jj ψ (n) )
n≥1 j=1
X Z
= n dxdydx2 . . . dxn φ(n) (x, x2 , . . . , xn )J (1) (x; y)ψ (n) (y, x2 , . . . , xn )
n≥1
XZ
= dxdydx2 . . . dxn J (1) (x; y)(ax Φ)(n−1) (x2 , . . . , xn )(ay Ψ)(n−1) (x2 , . . . , xn )
n≥1
Z
= dxdyJ (1) (x; y)(ax Φ, ay Ψ).
where the sum runs over all sets {i1 , i2 } of two different indices in {1, . . . , n}
(2)
and where J{i1 ,i2 } denotes the operator on L2sym (R3n ) acting as J (2) on the
variables i1 , i2 and as identity on the other (n − 2) variables. Show that if
J (2) has the integral kernel J (2) (x1 , x2 ; y1 , y2 ) then
Z
dΓ(J ) = dx1 dx2 dy1 dy2 J (2) (x1 , x2 ; y1 , y2 )a∗x1 a∗x2 ay1 ay2 .
(2)
41
Problem B.3.3. Consider the many-body Hamiltonian of the form
N
X X
HN = Ti + Wij .
i=1 i<j
Let {un } be an orthonormal basis of the one-body Hilbert space. Show that
the second quantization of HN equals
X 1 X
H= (um , T un )a∗m an + (uk ⊗ ul , W um ⊗ un )a∗k a∗l am an
m,n
2
k,l,m,n
Problem B.3.4. Consider the operator on the Fock space of the form
X
A= ei a∗i ai .
i
42
Bibliography
[2] F. London (1938). ”The λ-Phenomenon of Liquid Helium and the Bose-
Einstein Degeneracy”. Nature. 141 (3571)
43