Physics IV: Light and Quantum Mechanics
Physics IV: Light and Quantum Mechanics
Lecture notes
written by
István Nándori and Zoltán Trócsányi
Debrecen, 2019
Contents
1 Light 3
1.1 Wave properties of light . . . . . . . . . . . . . . . . . 3
1.1.1 Young’s double-slit experiment . . . . . . . . . 4
1.1.2 Intensity in the interference pattern produced
in the double-slit experiment . . . . . . . . . . 6
1.1.3 Thin-film interference . . . . . . . . . . . . . . 7
1.1.4 Single-slit diffraction . . . . . . . . . . . . . . . 9
1.1.5 Intensity in single-slit diffraction . . . . . . . . 11
1.1.6 Diffraction (or precisely interference) gratings . 14
1.2 Blackbody radiation . . . . . . . . . . . . . . . . . . . 17
1.3 Particle properties of light . . . . . . . . . . . . . . . . 23
1.3.1 Photoelectric effect . . . . . . . . . . . . . . . . 23
1.3.2 Compton scattering . . . . . . . . . . . . . . . 26
2 Particles 31
2.1 Wave properties of matter . . . . . . . . . . . . . . . . 31
2.1.1 X-ray diffraction . . . . . . . . . . . . . . . . . 32
2.1.2 Diffraction of electron beam on crystals . . . . 34
2.1.3 Double-slit experiment with electrons . . . . . 34
3 Atoms 41
3.1 Rutherford’s experiment . . . . . . . . . . . . . . . . . 41
3.1.1 Setup and result . . . . . . . . . . . . . . . . . 41
3.1.2 Algebraic properties of a hyperbola . . . . . . . 43
3.1.3 Motion in an 1/r repulsive potential . . . . . . 45
3.1.4 Differential cross section of scattering on a point-
like scattering centre . . . . . . . . . . . . . . . 48
1
2 CONTENTS
Light
by the angle θ on the screen (see Fig. 1.1) as the linear superposi-
tion of two waves starting at the slits of the second barrier with the
same phase, but travelling different distances, therefore, arriving with
a phase difference. The latter is due to the different distance that
can be computed by simple geometry. We assume that the distance
between the two slits d is much smaller than the distance from the
barrier to the screen D, d << D and the position of observation on
the screen P is at a distance r << D from the closest position on the
screen O to the slits. We denote the angle arctan r/D by θ. The path
difference from the two slits to P is approximately d sin θ. If this path
difference is equal to integer number n times the wavelength, then the
resulting phase difference is zero and the two waves add constructively.
Thus at distances r, satisfying the condition
d sin θ = mλ , (1.1.1)
there will be lines of total cancellation between the two light waves,
(minimum intensity).
6 CHAPTER 1. LIGHT
∆`
φ = 2π (1.1.3)
λ
d
φ ≈ 2π sin θ . (1.1.4)
λ
Before
The interference
Interface depends Interfer
ion can, on the reflections and the
The color
nterface. After path lengths. n1 n2 n3
ge, using light wave
elatively The thick
(a)
waveleng
r2
g. 35-16a ence of th
c
nsmitted Before
r1 Figur
situation After θ b of refract
θ a
index of source. Fo
he wave (b)
i n1 ! n3 in
L
hat is, its perpendic
Figure 35-16 Phase changes when a pulse is or dark t
Figure 35-15 Light waves, represented with
g. 35-16b reflected at the interface between two
ray i, are
Figure incident on a thin film of thick-
1.3: brightly il
ransmit- stretched strings of different linear densi-
ness
ties. The wave speed is greater in the lighter
L and index of refraction n2. Rays r1 The i
entation and r represent light waves that have of the film
string. (a) The incident pulse is in the 2
nusoidal
denser string.
oil film (b) cotes
that The incident pulse been
the water is in reflected
very thinly. Its by thickness
the front and back sur-
is comparable reflected
velength. faces of the film, respectively. (All three
thetolighter string. Only here
the wavelength is therelight,
of visible a phase
i.e. several hundreds of nanometers. the film t
medium change, and only
The light in the
waves reflected
travel our eyes, but there is a path differencetobe-
[Link] are actually nearly perpendicular refraction
e that is the film.) The interference of the waves of
ength.
tween the waves that are reflected from the front and back surfaces. it undergo
r1 and r2 with each other depends on their
action of
For some wavelength the resulting interference is constructive, while
phase difference. The index of refraction
by ray r2, i
for others it is destructive, giving rise to a colorful picture. This phe-
n1 of the medium at the left can differ If the
nomenon can be used to coverfrom objects to reduce or enhance reflection
the index of refraction n3 of the an interfe
of a certain wavelength, thusmediumincrease, or right,
at the decrease thenow
but for transmitted
we are exactl
ones. For instance, thin film onassume that both media are air,reflectivity
a window can enhance the with n1 ! dark to th
in the infrared, thus reduce nthe heating effect of sun light, without
3 ! 1.0, which is less than n2. phase diff
affecting the transmittance for the visible light. The K
Every time light
reflectsoff of a
The physics of thin-film interference is based upon the change of tween the
slow substance,, phase of a reflected light wave. As shown in Fig. 1.3, there are two the path i
there is a pi shift
cases. The wave is reflected from a surface dividing two materials of to b, and t
different indices of refraction. When the wave is coming from the side
through t
1, with smaller index of refraction than on side 2, n1 < n2 , then the
ence betw
phase changes as in the case of reflection of a wave from a fixed end,
between
resulting in a phase shift of 180◦ . If the wave is coming from the side
fference equivalen
1, with greater index of refraction than on side 2, n1 > n2 , then the
phase changes as in the case of reflection of a wave from a free end,
for two re
resulting in zero phase shift. and (2) re
Let us now consider a thin film of width L and refraction index
n n > 1 in air (of refraction index ' 1) as shown in Fig. 1.3, right. The
There is a phase difference between the light reflected from the surface
r2 shown
. Let’s next
5-15. At
CHAPTER 1. LIGHT 9
of 1st incidence (light approaching from the air) and from the 2nd
surface (where the light is approaching from the film). The origin
of this phase difference is two-fold. On the one hand there is 180◦
phase shift between these two waves due to the different relation of
refraction indices: (i) on the first surface n1 = 1 < n2 = n, (ii) while
on the second surface n2 = n > n3 = 1. On the other hand, the wave
travels in the film a path of length of approximately 2L. (Precisely,
it is 2l/ cos β, where β is the angle between the direction of light in
the film and the normal to the surface of the film, but we assume
β is small.) The condition for constructive interference is that the
optical path difference between the two waves results in a phase shift
of 180◦ to compensate for the phase shift due to the different reflection
mechanisms. As the light waves travel with a speed of c/n in the film,
so its wavelength is λ/n in the film. So there will be constructive
interference (maxima) if
1 λ
2L = m + , (1.1.6)
2 n
with m being a positive integer. Likewise, the condition for destructive
interference (minima) is
λ
2L = m . (1.1.7)
n
These equations hold if the index of refraction of the film is greater,
or less than the indices of media on both sides of the film because only
in this case there will be a phase shift of 180◦ for reflections at the two
surfaces. In other circumstances, we have to modify the conditions by
taking into account the phase shifts at the different reflections at the
various surfaces.
If the film thickness is not uniform, the conditions for construc-
tive and destructive interference depend on the position and bands
of maxima and minima appear, called fringes of constant thickness.
If such a film is illuminated by white light, one can observe that the
fringes of constant thickness depend on the wavelength and the film
appear in colours.
the slit is comparable to the wavelength of the light, then light beams
flare not only far beyond the geometrical shadow of the slit, but also
show a series of alternating bright and dark bands, similar although
quantitatively different as in the case of the double-slit interference.
Let us consider light waves falling on a barrier with a single slit
on it. In order to describe quantitatively the diffraction pattern we
assume that the distance between the barrier and the screen D is
much larger than the width of the slit d, D >> d. In principle, it
is possible to obtain the diffraction pattern using Fresnel’s principle
without such an assumption, but the required mathematics is more
involved.
In the case D >> d, we can approximate all waves by plane waves,
which can be produced directly using laser beams. The diffraction
pattern of plane waves is called Fraunhofer diffraction. Let us consider
first the point P0 closest to the slit on the screen. No matter from
where in the slit a spherical wave leaves, it arrives on this point with
the same phase, so there will be a constructive interference and a
bright line in the middle point P0 . Let us now consider another point
P1 at distance r from P0 , where we find the first dark line. The
distance to this point from the middle of the slit defines the angle
θ = arctan r/D as shown in Fig. 1.4. If the path from the closest
edge of the slit to P1 differs from the path between the middle of
the slit to P1 by exactly a half wavelength, then the spherical waves
from the edge and the middle arrive at P1 with opposite phase and
interfere destructively. Consider now another pair of spherical waves,
one emerging from a distance x < d/2 from the edge considered before,
and one emerging from a point at the same distance x from the middle,
so the distance between these two points is again d/2. These two waves
arrive at P1 with the same phase difference as the previous two, so
interfere destructively. As x is arbitrary, we conclude that the waves
interfere destructively at P1 giving a dark line. The path difference
between any of these pairs of waves is d/2 sin θ, which has to equal
λ/2 as discussed above, so the condition for the first dark line is
d sin θ1 = λ . (1.1.8)
Note that in our reasoning, the centers of the pairs always lay in
different zones. These two zones called Fraunhofer zones.
We can use similar reasoning to find the condition for the sec-
ond dark line. The only difference compared to the first minimum is
that this time we have to consider four equal Fraunhofer zones (see
[Link]
ximately tocross
reach being
P1with
section longer
a circular edge. than the path traveled by the wavelet of r1. T
this path length difference, we find a point b on ray r2 such that the pa
plifying)
and then from b to P1 matches the path length of ray r1. Then the path length diffe
ncel each tween the [Link]
CHAPTER rays
each
pair of rays cancel
LIGHTis the distance from the center of the slit
other at P1. So 11to b.
point P1. When viewing screen
do all such pairings.C is near screen B, as in Fig. 36-4, the d
n we ex-
pattern on C is difficult to describe mathematically. However, we can sim
y r2 from D
o rays to mathematics considerably if we arrange for the screen separation D to
ays from larger than the Totally
slit width a. Then, as in Fig. 36-5, we can approximate ray
destructive
er of the interference
r2 are in r1 P1
r1
t passing
irst dark r2
a/2 b θ
ifference This path
θ
elet of r2 P0
o display Central axis
b r2 difference
a/2 θ one wave
th length a/2
ence be- Figure 36-5 For D ! a, we Viewing
can θ other, wh
approximate rays r1 and r2screen
as
determine
ffraction being parallel, at
B angle u to the
C
Path length
Incident
mplify the central axis. the interfe
wave difference
be much
Figure 36-4 Waves from the top points of two
r1 and r2 zones of width a/2 undergo fully destructive
FigureC.1.4:
interference at point P1 on viewing screen
Fig. 1.5), and the pairs of waves that cancel at P2 originate from cen-
ength ters separated by distance d/4 on either side of the middle of the slit.
shifts As a result the condition for the second minimum is
rom the
h d sin θ2 = 2λ . (1.1.9)
ence. Using the Fraunhofer zones, it is easy to convince ourselves that the
condition for the mth minimum is
d sin θm = mλ , (1.1.10)
CHAPTER 1. LIGHT 13
an electromagnetic
ctric field. Here, this
pattern) is propor-
o E u2. Thus,
α α
(36-11) φ
R
1
2 f, we are led to Eq.
e phase difference f Eθ
may be related to a
Em φ Em
Then we use Eq. (1.1.3) where for point P at angle θ, the path differ-
ence between wavelets from adjacent zones of size ∆y is ∆y sin θ, so
the phase difference ∆φ between wavelets from adjacent zones is
∆y
∆φ ≈ 2π sin θ .
λ
The phase difference between the last and the first phasor is the sum of
the phase differences between adjacent phasors over the whole width
of the slit:
d
φ ≈ 2π sin θ ,
λ
14 CHAPTER 1. LIGHT
which leads us to the usual form of I(θ) expressed with the maximum
intensity Imax as
2
sin α(θ) πd
I(θ) = Imax , α(θ) = sin θ . (1.1.14)
α(θ) λ
neon laser is shown in Fig. 36-19b. The maxima are now very narrow
ed lines); they are separated by relatively wide dark regions.
P Diffraction Gratings
To point P
on viewing
We use a familiar procedure to find the locations of the bright lines θ One
of the most useful tools in the study of li
screen
screen. We first assume that the screen is far enough from the grat- absorb light is the diffraction grating. This devic
rays reaching a particular point P on the screen are approximately
they leave the grating (Fig. 36-20). Then we apply to each pair of θ arrangement of Fig. 35-10 but has a much great
s the same reasoning we used for double-slit interference. The sep- rulings, perhaps This pathaslength
manydifference
as several thousand pe
een rulings is called the grating spacing. (If N rulings occupy a total between adjacent rays
θ consisting of only five slits is represented in F
d ! w/N.) The path length difference between adjacent rays is determines the interference.
Fig. 36-20), where u is the angle dfrom the central axis of the grating
light isPath sent through the slits, it forms narrow
length
θ analyzed to determine
difference the wavelength of the lig
fraction pattern) to point P. A line will be located at P if the path d
θ between adjacent rays
ce between adjacent rays is an integer number of wavelengths : be opaque surfaces with narrow parallel gro
C Fig. 36-18. Light then scatters back from the gro
sin u ! ml, for m ! 0, 1, 2,λ. . . (maxima — lines), (36-25)
Figure 36-18 An idealized diffraction grating, Figure rather
36-20 Thethan being
rays from transmitted
the rulings in a through open slits
wavelength of the light. Each integer m represents a different line; Pattern.
diffraction grating With
to a distant monochromatic
point
consisting of only five rulings, that produces approximately parallel. The path length dif-
P are light inciden
tegers can be used to label the lines, as in Fig. 36-19. The integers
an interference pattern on a distant viewing gradually increase the
ference between each two adjacent rays is number of slits from two t
the order numbers, and the lines are called the zeroth-order Figure
line 1.7:
ne, with m ! 0), the screen C. line (m ! 1), the second-order line
first-order plot
d sin u, changes
where from
u is measured the typical
as shown. (The double-slit plot of F
rulings extend into and out of the page.)
o on. cated one and then eventually to a simple graph
ng Wavelength. If we rewrite Eq. 36-25 as u ! sin"1(ml/d), we pattern you would see on a viewing screen usin
given diffraction grating, the angle from the central axis to any
third-order line) depends
Exerciseson the wavelength of the light being
Intensity
en light of an unknown wavelength is sent through a diffraction
urements of the angles to the higher-order lines can be Interference
used in ∆θ hw
etermine the wavelength. Even light of several unknown wave-
distinguished and identified in this way. We cannot do that with 3
1. A viewing screen is separated from a double-slit source by 1.2 m.
t arrangement of Module 35-2, even though the same equation
h dependence apply there. The distance
In double-slit between
interference, the two slits is 0.030 mm. The second-
the bright
different wavelengths overlap too much
order to befringe
bright distinguished.
(m = 2) is 4.5 cm from the center line.
θ De-0°
termine the wavelength of the Figure
light!Figure 36-19 (a) The intensity plot produced
nes 36-21 The half-width #uhw of the cen-
by a diffraction grating with a great many
tral line is measured from the center of that
lity to resolve (separate)[Link]
Aofpossible
different wavelengths
means for depends on an
making airplane invisible to radar ishere
to labeled
he lines. We shall here derive an expression for the half-width of line torulings consists
the adjacent of narrow
minimum on a plotpeaks,
of
coat the plane with
e (the line for which m ! 0) and then state an expression for the
an antireflective u polymer.
like Fig. 36-19a. If radar
with their order numbers m. (b) The
I versus waves
havethea half-width
the higher-order lines. We define wavelength = 3.00 cmcorresponding
of λ line
of the central and the index brightof fringes seen on the
refraction
ngle #uhw from the center of ofthethe linepolymer
at u ! 0 outward
is n = to 1.50,
where how thick screen (d) are called
would lines
youandmakeare here also
the 3
vely ends and darkness effectively begins with the first minimum labeled with order numbers m.
coating?
such a minimum, the N rays from the N slits of the grating cancel
The actual width of the central line is, of course, 2(#uhw), but line Top ray
To first
3. Solar cell devices that generate electricity when exposed
ally compared via half-widths.) minimum
to sun-
36-1 we were also concerned light
with theare often coated
cancellation of a great with
many a transparent, thin film of silicon
to diffraction through a single slit. We obtained
monoxide (SiO, Eq.n= 36-3, which,
1.45) to minimize ∆θ hwreflective Bottomlosses from the
e similarity of the two situations, we can use to find the first ray
surface. Suppose that a silicon solar Nd cell (n = 3.5) is coated with
. It tells us that the first minimum occurs where the path length
ween the top and bottom rays a thin
equalsfilm
l. Forof silicondiffraction,
single-slit monoxide for this purpose. Determine the
Path length
is a sin u. For a grating of N rulings, each separated from the next ∆θ hw difference
minimum film thickness
he distance between the top and bottom rulings is Nd (Fig. 36-22),
d that produces the least reflection at
a wavelength of 550
th length difference between the top and bottom rays here is nm, near the center of the visible spectrum.
hus, the first minimum occurs where Figure 36-22 The top and bottom rulings of
4. A thin film of oil (n = 1.25) isa diffractionlocatedgrating
on ofa Nsmooth
rulings are wet pave-
Nd sin #uhw ! l. (36-26) separated to
by Nd. The pavement,
top and bottom rays
ment. When viewed perpendicular the the film
is small, sin #uhw ! #uhw (in radian measure). Substituting this in passing through these rulings have a path
reflects
the half-width of the central line as
most strongly red light at λ = 640 nm and
length difference of Nd sin #uhw, where
r reflects no
blue light at λ = 512 nm. How is the angle
thick
#uhw is to theoil
the firstfilm?
minimum.
l b
#u hw ! (half-width of central line). (36-27) (The angle is here greatly exaggerated
Nd for clarity.)
16 CHAPTER 1. LIGHT
Diffraction
Polarization
hεi =
physics for the spectral radiancy, for a kT where T is the temperature inside the cavity. Then on
dimensional grounds we expect that the spectral radiancy is
sical radiation law), (38-13)
P = σAT 4 (1.2.3)
CHAPTER 1. LIGHT 19
4ε 1 dε
=
3T 3 dT
where we could replace the partial derivative with ordinary derivative
as it depends only on the ratio ε/T . Integration with respect to T
leads to the Stefan-Boltzmann law in Eq. (1.2.3).
In Eq. (1.2.3) A denotes the area of the radiator and
W
σ = 5.670373 · 10−8 (1.2.5)
m2 K4
is the Stefan-Boltzmann constant that can be measured, but we shall
see that it is related to fundamental constants of nature.
The measured spectral radiancy has a maximum at λ = λmax .
In 1893 Wilhelm Wien published a theoretical article where he ar-
gued that the product λmax T (more precisely, the ratio T /fmax ) is a
constant,
λmax T = 2.8977729(17) mm · K , (1.2.6)
which is called Wien’s displacement law, demonstrated on Fig. 1.8,
right.
20 CHAPTER 1. LIGHT
2c2 h 1
S(λ) = 5 hc/λkT
, (1.2.7)
λ e −1
whose origin was not known until in 1917 Einstein produced a deriva-
tion assuming energy quanta both for the atoms of the cavity wall
as in the Einstein model of solids and for the electromagnetic radia-
tion. For the former we already found in the Thermodynamics course
that the Boltzmann distribution with quantized energy levels of the
Discrete Energy Levels: Electrons in an atom, vibrational modes of a
oscillators predicts molecule, energy levels of a quantum harmonic oscillator, quantum
n classical statistical mechanics, the equipartition theorem states that each ε0
degree of freedom (kin+pot) contributes hεi = ε0 ,states of a particle in a potential well.(1.2.8)
kT/2 to the average energy of the system. This theorem works well for exp ( kT ) − 1
systems where energy levels are continuous, and classical physics is
applicable as average energy per oscillator instead of kT . Thus, instead of
Eq. (1.2.2) we should have Continuous Energy Levels: Kinetic energy of a classical gas particle,
kinetic energy of a macroscopic object, energy levels of a classical
c ε0 harmonic oscillator, classical states of a macroscopic system.
S(λ) ∝ 4 ε0 . (1.2.9)
λ exp ( kT ) − 1
Z 2π Z π/2 Z 1
dφ cos θ sin θdθ = 2π xdx = π .
0 0 0
(kT )4
P = 2πcA I (1.2.11) dA cosθ
(hc)3
dA
where the integral
∞
x3 dx
Z
I= Figure 1.9:
0 ex − 1
2π 5 k 4
σ=
15 c2 h3
which gives the numerical value quoted in Eq. (1.2.5).
From the analytic form of the spectral radiancy function S(λ), we
can easily find the position of its maximum λmax from the condition
dS
= 0.
dλ λ=λmax
d x5 5x6 x7 e x
x2 = − =
dx ex − 1 ex − 1 (ex − 1)2
5ex − xex − 5
= x6 = 0,
(ex − 1)2
hc
λmax T = .
4.96511k
As discussed above the spectral radiancy results from the electro-
magnetic radiation inside the cavity. The function S(f ) has SI unit
W s/m2 ,1 so the differential energy density (with unit J/m3 ) in the
frequency range [f, f + df ] of that radiation should have the same
form,
S(f )
dε(f, T ) = 4π df (1.2.12)
c
where the factor of 4π is due to the integration of the spectral radiancy
over the complete solid angle (imagine that the radiator is the Sun).
In order to find the total energy density inside the cavity (or Sun), it
hf
is again more convenient to use the dimensionless variable x = kT as
integration variable instead of f , so
(kT )4 x3 dx
dε(x, T ) = 8π . (1.2.13)
(hc)3 ex − 1
Exercises
Black-body radiation
(max)
Ek = eVstop (1.3.1)
where e is the unit charge. Varying the intensity of the incident light,
(max)
we find that the stopping voltage, hence Ek does not depend on I.
In a classical picture, the alternating electric field in the light makes
the electrons oscillate inside the target. If the amplitude, hence the
intensity of the electric field is large enough it makes the electron
oscillate with so large amplitude that it can break free. So classically
(max)
we would expect that with increasing intensity Ek increases, but
this is not what happens. The intensity of the light does not influence
the maximum kick it gives to the ejected electrons.
If we think of the incident light as flow of particles with definite
energy that depends only on the frequency, then the energy that can
be transfered to the electron from the light is that of a single light
particle, called photon. Increasing the intensity increases the number
of photons in the light beam, but not the energy of the photons as
that depends only on the frequency, E = hf . Thus, the frequency
38-2 TH E PHOTOE LECTR IC E FF
being fixed, the energy transferred to the electrons is also fixed.
In the second experiment we mea- Electrons can escape only The escaping electron’s
if the light frequency kinetic energy is greater
sure the stopping voltage as a function exceeds a certain value. for a greater light frequency.
a
2.0
Figure 38-2 The stopping po-
Vstop = a(f − fc )Θ(f − f ) , (1.3.2)
c Vstop as a function of
tential
the frequency f of the inci- 1.0 Cutoff
dent light for a sodium target frequency f 0 c b
T in the apparatus of Fig. 38-
independently of the intensity. (Θ(x) = 1. (Data reported by R. A.
0
2 4 6 8 10 12
Millikan in 1916.) Frequency of incident light f (10 Hz) 14
1 if x > 0 and 0 if x < 0, called the step
function.) Based on classical physics
The existence ofwe a cutoff frequency is, however, just what we should expect
if the energy is transferred via photons. The electrons within the target are held
expect that if the intensity is sufficiently
there by electric forces. (If they weren’t, they would drip out of the target due to
Figure
the gravitational force on them.) To just escape from the 1.11:
target, an electron must
pick up a certain minimum energy !, where ! is a property of the target material
called its work function. If the energy hf transferred to an electron by a photon
exceeds the work function of the material (if hf " !), the electron can escape
the target. If the energy transferred does not exceed the work function (that is,
if hf # !), the electron cannot escape. This is what Fig. 38-2 shows.
CHAPTER 1. LIGHT 25
We also find that the work function and the cutting frequency are
related by
Wmin = hfc . (1.3.5)
26 CHAPTER 1. LIGHT
38-3 PHOTONS, M OM E NTU M , COM PTON SCATTE R I NG, LIG HT I NTE R FE R E NCE 1159
1.3.2 Compton scattering
Ideas
hough it is massless,The explanation
a photon has momentum,of the photoelectric effect using h the concept of light
h is related to its energy E, frequency f, and "l ! (1 # cos f),
particles was very encouraging. To extend this mc hypothesis further, in
length by where m is the mass of the target electron and f is the angle
1916 Einstein proposed that photons also have momentum. As the
at which the photon is scattered from its initial travel direction.
hf h
unit !of the
p! . ratio energy/momentum
● Photons: When is light
the interacts
unit of withspeed and
matter, the in theis
interaction
c l
case of light the only quantity particle-like,
that occurring
can have at a point
unitandoftransferring
speed isenergyc, itandis
Compton scattering, x rays scatter as particles (as momentum.
ons) from loosely boundnatural
electronstoin aassume
target. that a photon
● Wave: of Whenenergy
a singleE = hf
photon has momentum
is emitted by a source, we
he scattering, an x-ray photon loses energy and interpret its travel as being that of a probability wave.
entum to the target electron. ● Wave:E When many
h photons are emitted or absorbed by
e resulting increase (Compton shift) in the photon matter, = the combined light as a classical(1.3.6)
p = we interpret electro-
length is magneticcwave. λ
Intensity
Intensity
∆λ
h ∆λ 2
pγ = hf, 0, 0, , pe = (mc , 0, 0, 0) (1.3.7)
70 75 70 75 λ 70 75 70 75
Wavelength (pm) Wavelength (pm) Wavelength (pm) Wavelength (pm)
Intensity
Intensity
Intensity
he Compton shift. The value
the scattered x rays are de-
Figure 38-4 Compton’s results for four values of the scattering angle f. Note that the
= 45° φ = 90° as the scattering angleφincreases.
Compton shift "l increases = 135°
Intensity
Intensity
∆λ
∆λ
5 70 75 70 75
m) Wavelength (pm) Wavelength (pm)
Exercises
Photons
1. Using the plots for the stopping voltage versus frequency for
Cesium, Potassium, Sodium and Lithium in Fig. 1.14, estimate
the work function for these metals.
hown in countless other experiments, light is in fact quantized as
nstein’s explanation of the photoelectric effect is not the best ar-
fact. CHAPTER 1. LIGHT 29
um
m,
m
um
m
Vstop
iu
siu
ssi
di
th
e targets
ta
Ce
So
Li
Po
ots accord-
5.0 5.2 5.4 5.6 5.8 6.0 6.2
14
f (10 Hz)
Figure 1.14:
Particles
a0 θ θ
θ θ
d θ θ θ θ
d
0 d sin θ d sin θ
d θ θ
(b) θ θ
d θ θ
for either measuring the wavelength of X-rays, if the interplanar dis-
tance is known, or more a0 importantly, to study the atomic structure
d θ θ
(a) of crystals. It was derived by W.L. (b)Bragg, who shared with his father
W.H. Bragg the Nobel prize in 1915 for their “services in the analysis
of crystal structure by means of X-rays”.
The actual
Ray 2
process is differ-
Ray 1
ent from the simplified picture
of reflection θ by θ planes, which
nevertheless gives θ θ the correct re-
sult.
d Even this simplified picture d
predicts d asin θcomplicated
d sin θ
diffrac- d θ
tion pattern. θ AsθFig. 2.4 shows, d θ θ
there are other crystal planes in θ θ
The extra distance of ray 2 θ
thedetermines
same crystal. Although the
the interference.
(c) (d)
crystal is oriented the same way,
these
Figure 36-28 crystal
(a) The planes have
cubic structure differ-
of NaCl, showing the sodium and chlorine ions and
ent interplanar distance d
a unit cell (shaded). (b) Incident x rays undergo anddiffraction by the structure of (a). The
x rays are diffracted
are orientedas ifdifferently.
they were reflected by a family of parallel planes, 2.4:
Figure with angles
measured relative to the planes (not relative to a normal as in optics). (c) The path length
difference between waves effectively reflected by two adjacent planes is 2d sin u. (d)
2.1.2
A different Diffraction
orientation of the incident of elec-
x rays relative to the structure. A different family
of parallel
tronplanesbeam
now effectively reflects the x rays.
on crystals
The Davisson-Germer experiment is an exact analogue of the X-ray
diffraction on crystals, with the only difference that the incident beam
consists of electrons. For target Thomson used metal film while Davis-
son used crystal grid. They observed the same diffraction pattern as
can be observed using X-rays, which proved de Broglie’s hypothesis
of matter waves. Thomson and Davisson shared the Nobel prize in
1937 for their “experimental discovery of the diffraction of electrons
by crystals”.
Assuming that electrons can be considered matter waves and us-
ing de Broglie’s formula, the wavelength can be controlled well by
using definite acceleration voltage. Thus one can make a precise mea-
surement of the interplanar distances in the crystal by measuring the
positions of the maxima at several different electron energies.
experiment was repeated such that there was only at most one electron
at any time between the source and the detector screen by P.G. Merli
and his collaborators in the 1970s and later by A. Tonomura and his
collaborators in 1989 (see [Link] oWRI-
LwyC4). The technical details are not in our interest here. The im-
portant message is that even single electrons pass through the two
slits such that if many electrons do so independently, nevertheless an
interference pattern appears. So clearly, not the ensemble of elec-
trons behave as classical waves, but each electron separately behaves
as some sort of a wave. We call those waves matter waves.
We usually think of particles as little projectiles that cannot be
divided further. Classically, such projectiles path through one or the
other slit. Let us denote with Pi (x) the probability that a particle
passes through slit i (i = 1, 2) such that the other slit is closed and
reaches the detector at a distance x from the central point on the
screen that lies on the line connecting the source and the slits. Then
the probability distribution of the position of arrival on the screen
when both slits are open is
Figure 1. Simplified setup. a, An electron beam passes through a wall with two
slits in it. A movable mask is positioned to block the electrons, only allowing the
ones traversing through Figure
slit 1 (P1 ),2.5:
slit 2 (P2 ), or both (P12 ) to reach the backstop
and detector. b,c, Probability distributions are shown, (Experimental in false-colour
intensity) for electrons that pass through a single slit (b), or the double-slit (c). Inset
1,2, Electron micrographs of the double-slit and mask are shown. The individual slits
for electronsare
that50 nmpass
wide ×through a single
4 µm tall with a 150 nmslit (b),structure
support or themidway
double-slit (c).
along it’s height,
and separated by 280 nm. The mask is 5 µm wide × 20 µm tall. Reprinted from The
Feynman Lectures on Physics, Volume III, by Richard P. Feyman, Robert B. Leighton,
The detailed results
and Matthew of Available
Sands. Bach’sfrom experiment
Basic Books, an areimprint
shown inPerseus
of The Fig. 2.6.
Books
Group. Copyright ⃝ c 2011
On the left we see that the resulting probability distributions on the
screen as the mask is moved over the double-slit. The mask allows
the blocking of one slit, both slits, or neither slit in a non destructive
way. On the right we see how the electron interference builds up
from individual electrons. The bright spots indicate the locations of
detected electrons. Shown are intermediate build-up patterns from
the central five orders of the diffraction pattern (P12 ) with 2, 7, 209,
1004, and 6235 electrons (a-e).
Let us now imagine the following thought experiment: instead
of covering one of the slits, we put a source of intense light behind
the slits. Electric charges are known to scatter light, so when an
electron passes through one of the slits, then we can observe a flash
of light coming from behind that slit. In this experiment whenever an
electron reaches the screen, we also obeserve a flash from one of the
slits, so we conclude that the electron passes through only one slit at
a time. However, if we observe the probability pattern on the screen
in this experiment, we find that the interference disappears and the
probabilities simply add as in Eq. (2.1.2). If we now switch off the
intense light behind the slits, the interference patterm on the screen
reappears.
The message of the thought experiment is that the observation of
CHAPTER 2. PARTICLES 37
2760 nm
2640 nm
2540 nm
P1
2500 nm
2240 nm
1620 nm
P12
−140 nm
−1440 nm
−2320 nm
P2
−2500 nm
−2540 nm
−2580 nm
−2680 nm
electrons changes their path. This should not be a big surprise as light
is electromagnetic wave and the electromagnetic field exerts a force
on the electrons. How can we reduce the effect of light on the motion
of the electrons? We already know that light consists of photons, with
energy and momentum proportional to the frequency of light. We can
reduce the effect of light on the electrons if we reduce its momentum,
i.e. its frequency, or increase its wavelength. Indeed, increasing the
wavelength of the light beyond a certain value the interference pattern
on the screen reappears. This value is comparable to the distance
between the slits, but this light cannot resolve the position of the
slits, so we cannot tell any longer where the electron passed through.
We can conclude from the real and
imaginary experiments that we can-
not observe the path of the electrons
without destroying the interference on
the screen. This non-trivial statement
was formulated by W. Heisenberg into
the principle of uncertainty. Accord-
ing to this principle, the laws of Nature
would be contradictory unless there is
a fundamental limitation in the preci-
sion of our measurements. In the case
of double-slit experiment with electrons
this means that it is not possible to de-
vise an experiment such that we can tell
which slit the electron passes through Figure 2.7:
without destroying the interference on
the screen. The original problem that lead Heisenberg to his con-
clusion was that the path of an electron can be followed in a cloud
chamber (see for instance the trace of the first ever observed positron
in Fig. 2.7), so it seems that the notion of classical path can be em-
ployed. However, the notion of the classical path of an electron in-
side an atom has disastrous consequences as that makes the atom
unstable (we shall give more details in the next section). In order
to resolve this paradox, Heisenberg suggested to use only observable
physical quantities to descibe the properties of the electrons inside the
atom. The classical path (i.e. position and velocity) is not so. It is
not observable because the uncertainties in the simultaneous measure-
ment of position and momentum cannot be made arbitrary small, but
∆x∆px ≥ ~ = h/2π. If we want to localize the electron to a fraction
CHAPTER 2. PARTICLES 39
of the size of the atom, then its momentum becomes so large that it
leaves the atom immediately, i.e. the atom disappears similarly as the
interference pattern when we observe which slit the electron passes
through.
The modern viewpoint is that both matter and radiation are phys-
ical fields such as the electromagnetic field, which takes into account
both the relativistic and the quantum effects. These fields can show
both wave and particle properties. The latter are called elementary
excitations of the field or particles for short. In these lectures we
shall concentrate only on the experiments that lead to the develop-
ment of non-relativistic quantum mechanics–the understanding of the
quantum effects. For that purpose first we explore the structure of
atoms.
Exercises
Matter waves
Table 2.1:
U/kV R1 /cm R2 /cm
3.0 1.46 2.55
3.5 1.38 2.48
4.0 1.26 2.21
4.5 1.15 2.13
5.0 1.13 1.93
Atoms
41
42 CHAPTER 3. ATOMS
Atomic nucleus and atom are not the same1234. An atom is the smallest unit of matter that has the properties of an element3. An atom
42-1
consists of two regions: the nucleus and the electron cloud24. DISCOVE
The nucleus RofI NG
is the center TH
the atom Econtains
and N UCLE US
protons 127
and neutrons
The actual experimental setup is shown
ew) used in Rutherford’son the leftlaboratory figure of Fig. in 1911 3.1.– 1913 In 1911 to it
was already known
by thin metal foils. The detector can be rotated to vari- that some elements,
Alpha source
f. The alpha source called wasradioactive,
radon gas, a decay produce product radiation.
of
p” apparatus, theThree atomic types nucleus of radiation
was discovered. were identified:
α, β and γ. The α rays were known to
have relatively large mass (about 7300
times more massive than the electron)
and positive electric charge (−2 times Gold foil
s known thatthe certain
charge elements,
of the electron). called radioactive, Today we φ
spontaneously,know emitting thatparticlesthese are in the the process.
nuclei ofOne the
mits alpha (a) helium particles [Link] have Usingana energy source of of α aboutrays
42-1 DISCOVE R I NG TH E N UCLE US 1277
Their alpha source was a thin-walled glass tube of radon gas. The experiment
f the particles arethescattered through rather small 105
prising
involves counting because
number of alpha the
particlesprevailing
that are view
deflected through atvari- 10 Theory
7
5
ngles, approaching son. 180°. He advocated In Rutherford’s
arithmic. We see that most of the particles are scattered through rather small
that the
angles, but — and this was the big surprise — a very small fraction of them are
words:
positive “It 10
10 10
3 4
event that
scattered evercharge
through happened inside
very large angles,to theme atom
approaching in180°.
my was
In life. It was
homoge-
Rutherford’s words: “It 10
2
3
was quite the most incredible event that ever happened to me in my life. It was
10 10
had firedalmost aas15-inch
neously
incredible asshell you hadatfired
if distributed a apiece through
15-inch of attissue
shell theofpaper
a piece entire
tissue paper
2
3Ze2
R< ' 71.5 fm ,
4πε0 Ek
which is much smaller than the radius of the gold atom (about
129 pm). Thus Rutherford decided to derive the differential cross
section of a particle of charge ze on another particle of charge Ze
(e being the unit charge and z, Z are positive integers), which we also
do next.
The repulsive force force between the two positively charged point-
like particles is given by Coulomb’s law,
1 zZe2
F = .
4πε0 r2
This is a conservative force so there exists a corresponding potential
energy that is inversely proportional to the distance r between the
two particles, Ep ∝ 1/r. In the case of this potential energy, the path
of the scattered particle is a hyperbola, which we prove first.
or
bp 2
y(x) = ± x − a2 (3.1.3)
a
where the positive solution corresponds to the line above and the negative one to
that below the x axis. The hyperbola is symmetric with respect to the x axis and
crosses it at |x| = a where y = 0. The foci lie in points (±c, 0) where
b2
c2 = a2 + b2 = a2 1 + 2 ≡ a2 ε2 . (3.1.4)
a
p
The parameter ε = 1 + b2 /a2 > 1 characterizes the eccentricity of the hyper-
bola, the smaller ε, the closer its asymptotes (lines y = ± ab x) to the x axis.
The parametric equations of the hyperbola are
where the parameter u has meaning related to area. Let us consider the part of
the hyperbola in the upper half plane where y > 0. The position vector to the
point (x, y(x)) and the [0, x] section form two sides of a rectangular triangle whose
area is t = xy(x)/2. This triangle is cut into two parts by the hyperbola. Let us
denote the area of the part below the hyperbola by t1 , while that of the other part
by t2 , so t = t1 + t2 . The area t1 is simply
Z x p Z x/a p
b
t1 = x2 − a2 dx = ab x2 − 1dx
a a 1
ab x y(x) x y(x)
= − ln +
2 a b a b
xy(x) ab
= − ln(cosh u + sinh u)
2 2
ab
=t− u. thay vi chon vi tri tu goc toa do den P thi
2 thuong trong bai toan vat chuyen dong trong
Then truong luc, ta chon vecto vi tri tu vat 2 (>> vat
ab 1) den P
t2 = t − t1 = u , (3.1.6)
2
so u is the measuring number of the area t2 in units of 21 ab.
The radius vector is the vector from the more distant focus to the hyperbola.
The area of the triangle enclosed by the radius-vector, the x axis and the position
vector is
1 1
t0 = cy(x) = aε b sinh u . (3.1.7)
2 2
Finally, size of the area enclosed by the radius-vector, the x axis and the hyperbola
is
ab
t0 + t2 = f (3.1.8)
2
where
f = u + ε sinh u (3.1.9)
1
is the measuring number of this area in units of 2
ab.
CHAPTER 3. ATOMS 45
so
r = a(1 + ε cosh u) . (3.1.12)
1 l2 1
Ep = m 2 − mv 2
2 b 2
ml2 ε cosh u − 1 ml2
1
= 2 1− = 2 (3.1.21)
2b ε cosh u + 1 b ε cosh u + 1
ml2 a
= 2
b r
where we used Eq. (3.1.12). Eq. (3.1.21) is our result: in case of a
central force, the motion on a hyperbola implies that the potential
CHAPTER 3. ATOMS 47
Figure 3.4:
a2 1
= (3.1.26)
4 sin(ϑ/2) cos(ϑ/2) tan(ϑ/2) sin2 (ϑ/2)
a2 1
= .
4 sin4 (ϑ/2)
zZe2 ml2 a 2 a
Ep = K = 2 = mv∞ (K −1 = 4πε0 ) , (3.1.27)
r b r r
so
zZe2
a=K 2
. (3.1.28)
mv∞
The quantity 21 mv∞ 2
is simply the kinetic energy of the α particles
emerging from the source. The resulting cross section is depicted
as the solid line in Fig. 3.2. We see that the prediction agrees with
CHAPTER 3. ATOMS 49
Here and below the integration extends over the whole space if it is
not stated otherwise explicitly. For a point-like target nucleus this
normalised charge distribution is a δ distribution, %(~r) = δ(~r), which
means that if ~r 6= 0, then % = 0.
The Fourier transformed of the normalized charge distribution,
Z
F (~q) = %(~r) ei~q·~r/~ d3 r
is called form factor. We state without proof that the differential cross
section on an extended target is related to that on a point-like target
according to the equation
dσ dσ
= |F (~q)|2
dΩ % dΩ R
where R refers to the Rutherford formula (scattering on point-like
target) and ~q = ~k − ~k 0 represents the momentum transfer between
the projectile and target during the scattering process. For point-like
target %(~r) = δ(~r) and
Z
F (~q) = δ(~r) ei~q·~r/~ d3 r = ei·0 = 1
50 CHAPTER 3. ATOMS
1 q 2 hr2 i
=1−
6 ~2
R∞
because 4π 0 r2 dr %(r) = 1 according to the normalization condition.
The square root of the quantity
Z ∞ , Z ∞ Z
2 2 4
rrms ≡ hr i = %(r) r dr %(r) r2 dr = %(r) r2 d3 r
0 0
Exercises
Rutherford scattering
1. An α particle of kinetic energy Ek = 5.5 MeV collides with a gold
nucleus head-on. What will be the smallest distance between the
52 CHAPTER 3. ATOMS
1 q2 2
P = a
6πε0 c3
1 Ze2
= ma . (3.2.3)
4πε0 r2
As stationary states are not allowed in a classical theory, and these are
related to the emission or absorption of photons, i.e. energy quanta,
we call them quantum states. The labels m and n are integers that
enumerate these quantum states, called quantum numbers. Multiply-
ing Eq. (3.2.2) with the universal constant hc = hf λ and comparing
the resulting equation to Eq. (3.2.6) we see that the energy levels
in the hydrogen atom are characterized uniquely by positive integer
principal quantum numbers n as
hcRZ 2
En = − (3.2.7)
n2
where we allowed for hydrogen-like ions with the inclusion of the factor
Z 2 , Ze being the charge of the nucleus.
In order to make connection between the electron moving on a
classical orbit far from the nucleus and the electron moving on the
stationary states inside the microscopic atom, Bohr also used the cor-
respondence principle: any generalization of a theory must agree with
an established theory of narrower validity in a well-defined limit. We
have already seen the application of this principle when we were look-
ing for a generalization of Galilean transformations to motions with
large speeds (as compared to the speed of light) such that the Lorentz
transformation formulae were to fall back to the Galilean ones in the
limit of small speeds. In the present case the correspondence principle
means that
the new (quantum) theory must agree with the classical
theory in the limit of large quantum numbers.
We now apply the correspondence principle to the electron jump-
ing between states with large quantum numbers, from state with quan-
tum number n to that with m = n−1 such that n → ∞. Bohr’s second
postulate in Eq. (3.2.6) together with Eq. (3.2.7) gives
2cRZ 2
2 1 1
fn−1,n = cRZ − ' ,
(n − 1)2 n2 n3
which should equal to the result of the classical computation for the
same frequency in Eq. (3.2.5), with energy taken from Eq. (3.2.7).
The resulting equation can be solved for Rydberg’s constant, yielding
a prediction for R in terms of fundamental constants:
me4
R= . (3.2.8)
8ε20 h3 c
CHAPTER 3. ATOMS 57
the anode and the cathode, with UGC UGA . In the experiment the
current through the anode is measured as a function of the voltage
The result of our measurement performed in the class is shown in Fig. 3.8.
The current increased with increasing voltage up to about 4.9 V. This behavior is
typical of any vacuum tubes that don’t contain mercury vapor: the larger voltage
the larger anode current. At about 4.9 V the current dropped sharply, almost
vanished. The current then again increased steadily with increasing voltage until
about 9.8 V (= 4.9 V + 4.9 V) was reached where again a sharp drop was observed.
attached itself to the filament under highly repeatable This series of dips in current at about 4.9 V increments continued to voltage of
conditions.
This work was conducted during Advanced Lab
(PHY243W) at the University of Rochester.
⇤
Electronic address: [Link]@[Link]
†
Electronic address: [Link]@[Link]
[1] UR Advanced Laboratory Manual, The Franck-Hertz
Experiment [Online] [Link]
~AdvLab/2-Frank-Hertz/Lab02%[Link]
[2] P. Nicoletopoulos, Phys. Rev. E 78, 026403 (2008)
[3] R.E. Robson, [Link]., J. Phys. B 33, 507 (2000).
[4] G.F. Hanne, Am. J. Phys. 56 (8) (1988).
[5] D.R.A. McMahon, Am. J. Phys. 51 (12) (1983).
[6] E.B. Saloman, J. Phys. Chem. Ref. Data 35, 4 (2006)
[7] Yu. Ralchenko, Kramida, A.E., Reader, J., and
UGC between the grid and the cathode.
70 V.
Vonset .
58
Vonset was observed to be dependent on both Vf and
the temperature of the tube. While this phenomenon
was not studied exhaustively, there seemed to be a nega-
tive, linear dependence on the temperature and a positive
linear dependence on Vf (which, in turn, was roughly re-
lated to the electron density in the tube) [Fig 6]. Further
study is warranted to characterize and understand this
phenomenon.
In summary, experimental results of the historic
increased from 0.0 to 45.0V in increments of 0.05V with 6.3V. Altering the temperature did not result in any sta-
a scan rate of 10 scans per second. For each run, Iag tistically significant change to the calculated excitation.
and Vacc were stored in a text file. Each file was then However, data analysis showed that, for higher temper-
analyzed using Microsoft Excel 2003, looking at the de- atures, the current peaks began at a slightly lower volt-
pendence of Iag on Vacc [Fig 4]. Electrons that reached age. The maximum o↵set observed was 0.8V. This phe-
the anode were registered as negative current. Because nomenon was a result of the initial kinetic energy dis-
electrons were decelerated between the grid and anode, tribution of the electrons. For higher temperature, the
CHAPTER 3. ATOMS
this current is a measure of the electron energy. We have kinetic energy distribution is greater than that in a low59
reason to believe that the observed -2 nA o↵set visible in temperature environment. An electron with a large ini-
Fig 4 and all other data sets is a nonphysical product of tial kinetic energy required slightly less acceleration to
our data acquisition system.
The excitation potential for the favored 1 S0 !3 P1
Franck and Hertz noted that
transition was determined by multiplying the elementary
electron charge with the measured voltage di↵erence (for
the E = 4.9 eV characteris-
Vacc ) between the current peaks in a run. Data points
tic energy of the electrons in
representing the current peaks were selected by visually
inspection of the graphs produced by Microsoft Excel.
their experiment corresponded
For each run, the voltage di↵erence between adjacent
current peaks was calculated. The mean and standard
to the wavelengths λ = hc/E '
deviation of these measurements were calculated on a
run-by-run basis for all runs, including runs with varying
254 nm of light emitted by mer-
oven temperatures and filament voltages. The final ex-
1
citation potential was calculated by taking the weighted
cury atoms in gas discharges.
average of these measured values.
In fact, they also observed this
The excitation potential was measured to be 4.93 ±
0.06 eV. This value agrees with the expected value of
emitted light radiated from the
4.86 eV corresponding to the 1 S0 !3 P1 transition [7].
This determination of the excitation potential of mer-
tube in all directions of a single FIG.
cury also gives us an estimate for Planck’s constant, h.
4: Data from a run taken at 180 C with V at 6.3V. For
f
this run, the observed excitation potential was calculated to
We know
wavelength at 254 nm, but only the excitation be 4.89 ± 0.26eV. The arrows indicate the calculated value for
(3) Figure 3.8:
potential Anode
from all 6 datacurrent as a
sets, 4.93 ± 0.06eV,
hc/ = eV
if the voltage on the grid was big- typical run.
0 showing qualitative matching of the calculated value with this
where is the wavelength of the emitted light, c is the function of the voltage UGC be-
ger than 4.9 V.
tween the grid and the cathode as
They interpreted their dis-
measured in the Franck-Hertz ex-
covery as follows. Most collisions
periment.
between the mercury atom and
the electron are elastic. How-
ever, when the kinetic energy of the electron is about 4.9 eV the
collision with the mercury atom becomes inelastic, showing that the
electron could lose only a well defined (kinetic) energy before flying
away, leaving behind an excited mercury atom. A short time later,
the excited mercury atom released the deposited energy in the form
of ultraviolet light that has a wavelength of precisely 254 nm. Follow-
ing light emission, the mercury atom returns to its original, unexcited
state. Thus Franck-Hertz experiment, first conducted in 1914, was a
historic experiment that showed quantized internal energy excitation
in atoms.
Exercises
Atomic spectra
ready know that the same formula was used by Einstein in his explanation of the
photoelectric effect in 1905.
60 CHAPTER 3. ATOMS
which predicts the angular momentum of the electron inside the atom
as a natural number times the natural unit ~. The question is whether
or not this prediction is supported by observations. The answer is a
clear ‘no’. For instance, in the ground state of the hydrogen atom
n = 1, so the predicted value for the angular momentum is L1 = ~,
while the measured value for the (orbital) angular momentum is 0.
Hence, we conclude that while Bohr’s postulates can predict the en-
ergy levels fairly precisely2 , the understanding of the quantized states
of the angular momentum requires a new theory. Also, the explana-
tion for the existence of stationary states requires clarification.
the famous sodium D-lines at 590 nm, the distance between the two
lines is about 0.6 nm, i.e. about 1 h. The derivation of the precise
formula
En 2 1 3
∆En = (Zα) − (3.3.1)
n j + 12 4n
where θ is the angle between the directions of the domain and field.
This potential energy has its minimum when θ = 0, so the field acts
with a torque on the magnetic dipole momenta such that it turns their
directions to line up with the direction of the field. Then it follows
that the angular momenta belonging to the magnetic momenta also
line up, but opposite to the direction of the external field. In or-
der that the total angular momentum of the bar remains unchanged
(in the absence of external torque acting on the bar), the bar must
turn such that it has macroscopic angular momentum in the direction
of the magnetic field. Thus the Einstein-de Haas exepriment demon-
strates in a macroscopic way that atoms carry magnetism and angular
momentum in a non-separable way.
As the unit of angular momentum is the same as the unit of
Planck’s constant, we can also write the magnetic moment of the
electron-loop current as µ ~
~ e = −µB L/~ where
e~ J µeV
µB = ' 9.3 · 10−24 ' 57.9
2me T T
is called Bohr magneton, the natural unit of the magnetic moment in
the microworld.
There are lines that split into more than three lines. In such cases
the number of Zeeman sub-levels is even and the phenomenon is called
anomalous.
Zeeman shared the Nobel prize with H.A. Lorentz “in recognition
of the extraordinary service they rendered by their researches into the
influence of magnetism upon radiation phenomena” in 1902. The cor-
rect interpretation of the splitting of spectral lines emitted by atoms
in magnetic field required a long time and became possible essentially
with the birth of quantum mechanics.
vertical component, so
∂B
F~ ≈ ~kµz .
∂z
They expected that the magnetic moments of the silver atoms were
pointing in all directions in space, hence they would be deflected into
a continuous line, depending on the size of their µz component.
The schematic setup of the
experiment is shown in Fig. 3.10
together with the expected and
measured result. They observed
two spots, which was a puzzle
not only for them, but for the
whole physics community. One
might suspect that the silver
atoms are complex, so in order
Figure 3.10: SternGerlach experi-
to exclude that the effect had
ment: a beam (2) of silver atoms
anything to do with that com-
emerging from an oven (1) and
plexity, in 1927 T.E. Phipps and
travelling through an inhomoge-
J.B. Taylor reproduced the effect
neous magnetic field (3) are de-
using hydrogen atoms in their
flected depending on the direc-
ground state. Stern and Gerlach
tion of their magnetic moment.
could measure the separation be-
The classically expected result is
tween the spots (about 0.2 mm)
a blurred line (4), but instead two
and thus deduce the force acting
well separated spots are observed
on the magnetic moments of the
(5).
silver atoms. Knowing the mag-
nitude of the magnetic field (it
was 0.1 T), they could conclude about the size of the magnetic mo-
ments belonging to the atoms at the two spots,
µz ≈ ±µB ~ , (3.3.4)
ball. Goudsmit and Uhlenbeck could explain both the fine structure
and the anomalous Zeeman effect assuming a g factor of 2 for the
spinning electron.
If the g factor for the electron is 2, then its gyromagnetic ratio is
γe = e/me . Furthermore, we should write the magnetic moment found
in the Stern-Gerlach experiment instead of the formula in Eq. (3.3.4)
as
~
µz ≈ ±2µB . (3.3.5)
2
As a consequence the corresponding angular momentum for the elec-
tron in the ground state of the hydrogen atom in magnetic field
is ±~/2. History has proven the correctness of the assumption of
Goudsmit and Uhlenbeck, which lead to the correct quantum me-
chanical description of the angular momentum of the electron inside
the atom.
Exercises
Atoms in magnetic field
Relative intensity
was firstrange,
n the kiloelectron-volt observed by D.G. Barkla
electromagnetic radia-in 1909 who was awarded the Nobel
prize in 1917 “for his discovery
d. Our concern here is what these rays can teach us of the characteristic Röntgen [X-ray]
b or emit them. Figure 40-13 shows the wavelength
radiation of the elements”. Continuous
duced when a beam of 35
Such keV electrons
a radiation can falls on a using an X-ray
be studied K β that has a
spectrumtube
a broad, continuous spectrum
schematic viewofdepicted
radiation in
on Fig.
which3.11, left. λElectrons emerge from a
min
of sharply defined wavelengths.
heated cathodeThe andcontinuous spec- by the electric field between the
are accelerated
different ways, which
anodeweand
nextcathode.
discuss
40-6 Xseparately.
RAYS AN D TH E OR DE R I NG 1237
30 OF40TH E50E LE60
M E NTS
70 80 90
Wavelength (pm)
um
nd the Ordering of the Elements Figure 40-13 The distribution by wavelength of
nuous x-ray spectrum of Fig. 40-13, ignoring for the Kα
d target, such as solid copper or tungsten, is bombarded with electrons the x rays produced when 35 keV electrons
ent peaks that rise from it. Consider an electron of
Relative intensity
tic energies are in the kiloelectron-volt range, electromagnetic radia- strike a molybdenum [Link] sharp peaks
collides (interacts)
x rays is emitted. with one
Our concern hereofisthe
whattarget
these atoms,
rays canasteach
in us and the continuous spectrum from which they
ytoms
losethat
an absorb
amount or of energy
emit !K, which
them. Figure will appear
40-13 shows as
the wavelength rise are produced by different mechanisms.
Continuous
on that
f the is radiated
x rays producedaway
when from
a beam theof site of the
35 keV collision.
electrons falls on a spectrum K β
m target.
erred toWe theseerecoiling
a broad, continuous
atom becausespectrumof ofthe
radiation on which
relatively λmin
posed two peaks of sharply defined [Link] continuous spec-
we neglect that transfer.)
e peaks arise in different ways, which we next discuss separately.
in Fig. 40-14, whose energy is now less than K0, may 30 40 50 60 70 80 90
Wavelength (pm)
h a X-Ray
ous target atom, generating a second photon, with a
Spectrum
Figure 40-13 The distribution by wavelength of
is electron-scattering
amine the continuous x-ray process
spectrum can continue
of Fig. until the
40-13, ignoring for the
the x rays produced when 35 keV electrons
tationary. All the photons generated by these colli-
the two prominent peaks that rise from it. Consider an electron of strike a molybdenum [Link] sharp peaks
c energy
uous K0 that
x-ray collides (interacts) with one of the targetFigure
spectrum. atoms, as3.11:
in and the continuous spectrum from which they
The electron may lose an amount of energy !K, which will appear as rise are produced by different
Electron-atom mechanisms.
scattering
f that spectrum in Fig. 40-13 is the sharply defined
of an x-ray photon that is radiated away from the site of the collision.
w which
energy the continuous
is transferred toThe spectrum
kineticatom
the recoiling does
energy not
Ek exist.
because the This
gained
of by the
relatively
ofsponds
the atom; tohere
a collision
neglect in
weelectron thatwhich it an
transfer.)
until incident
reaches theelectron
cathode can
ttered electron K ∆K
Ek0-–ΔE
nergy K0 in ainsingle
Fig. 40-14, whose energy
head-on collision is now
withlessathan K0, may
target
k
be controlled precisely by
nd collision with a target atom, generating a second photon, with a
the voltage Target
ergy appears as between the energy theof two
a single photon, whose
electrodes. Inuntil
thethe ex- atom
hoton energy. This electron-scattering process can continue
minimum possible x-ray wavelength — is found
approximately stationary. All the photons generated by these colli- from
periment we observe the X-rays emit-
part of the continuous x-ray spectrum. KEk0
hc from the anode. A typical X-ray
ted
minentKfeature
0 # hf #
of that spectrum
, in Fig. 40-13 is the sharply defined Incident
spectrum is shown in does
[Link]
3.11, (= ∆ K)
hf (=ΔE k)
l
length lmin, below which min the continuous spectrum [Link].
This electron
X-ray
wavelength corresponds It hasto aseveral
collision characteristics:
in which an incident (i)electron
there photon
K0 – ∆ K
initial kinetic energy K0 in a single head-on collision with a target
ntially all this hc
isappears
a continuous spectrum, above which Target
" min # energy (cutoff
(ii) there
as the energy of a single photon, whose
wavelength).
are several (40-23)
discrete linesfrom
and
atom
wavelength —K the
0
minimum possible x-ray wavelength — is found
(iii) the continuous
hc spectrum starts at K0
, wavelength
Figure 40-14 An electron of kinetic energy K0
aKof
ally independent 0 #
well hfdefined
the #
target
l min material. If we were
λmin ,tocalled
passing near
Incident
hf (= ∆ K)
an atom in the target
electron may gener-
16
X-ray
target to a copper target,
cutoff for example,
wavelength thatall
is features of
independent of
ate an x-ray photon, the electron losing part
Figure 3.12:photon
-13 would change excepthc the cutoff wavelength. of its energy in the process. The continuous
" min # (cutoff wavelength). (40-23)
K0 x-ray spectrum arises in this way.
Figure 40-14 An electron of kinetic energy K0
wavelength is totally independent of the target material. If we were to
K-shell vacancy jumps from the shell with n ! 2 (called the L shell), the emitted
15 radiation is the Ka line of Fig. 40-13; if it jumps from the shell with n ! 3 (called
the M shell), it produces the Kb line, and so on. The hole left in either the L or M
shell will be filled by an electron from still farther out in the atom.
In studying x rays, it is more convenient to keep track of where a hole is
Energy (keV)
created deep in the atom’s “electron cloud” than to record the changes in the
CHAPTER
10 3. ATOMS quantum state of the electrons that jump to fill that hole. Figure 69 40-15 does
exactly that; it is an energy-level diagram for molybdenum, the element to
which Fig. 40-13 refers. The baseline (E ! 0) represents the neutral atom in its
ground state. The level marked K (at E ! 20 keV) represents the energy of the
the 5metal of the anode, depends onlyatom
molybdenum on with a hole in its K shell, the level marked L (at E ! 2.7 keV)
represents the atom with a hole in its L shell, and so on.
Ek . The relation
Kα between Ek The and λ min
transitions marked Ka and Kb in Fig. 40-15 are the ones that produce the two
is very simple: L (n = 2) x-ray peaks in Fig. 40-13. The Ka spectral line, for example, originates when an elec-
Kβ Lβ Lα
M (n = 3) tron from the L shell fills a hole in the K [Link] state this transition in terms of what
N (n = 4) the arrows in Fig. 40-15 show, a hole originally in the K shell moves to the L shell.
0
Figure 40-15 A simplified energy-level Ordering hc
diagram for a molybdenum atom, showing λ the=Elements,
min (3.4.1)
In 1913, BritishEphysicist
the transitions (of holes rather than elec- k H. G. J. Moseley generated characteristic x rays for as
many elements as he could find — he found 38 — by using them as targets for
trons) that give rise to some of the charac-
electron bombardment in an evacuated tube of his own design. By means of a
teristic x rays of that element. Each
trolley manipulated by strings, Moseley was able to move the individual targets
horizontal line represents the energy of
the atom with a hole (a missing electron) in
which indicates that at the into the path
cutoff of an electron beam.
wavelength the He measured
total kineticthe wavelengths
energy of the emitted
the shell indicated. x rays by the crystal diffraction method described in Module 36-7.
of the electron is transfered into
Moseleythethenemitted
sought (and found) regularitiesEin
radiation, = hf
k these minas, he moved from
spectra
element to element
with fmin = c/λmin . We interpret the inemergence
the periodic table.
ofInthe
particular, he noted that if, for a given
continuous
spectral line such as Ka, he plotted for each element the square root of the frequency
spectrum such that the electron passing
f against the near
position of an atom
the element is decelerated
in the periodic table, a straight line resulted.
Figure 40-16 shows a portion of his extensive data. Moseley’s conclusion was this:
by the electric field of the target atom, and the energy lost in this
deceleration is emitted in the We formhave here a proof that there is in the atom a fundamental quantity, which
increases ofregular
by X-ray stepsradiation
as we pass from(see Fig. 3.12).
one element to the next. This quantity
Hence this radiation is called bremsstrahlung can only be the charge on the (the
centralGerman
nucleus. word for
“breaking radiation”). As a result of Moseley’s work, the characteristic x-ray spectrum became the uni-
versally accepted signature of an element, permitting the solution of a number of
Moseley plot
The origin of the discrete
lines in the X-ray spectrum is 2.5
Pd Ag
completely different. The po- Mo Ru
2.0 Zr
No, the atomic
number and the
sition of these lines depend on Y Nb
of the nucleus.
70 CHAPTER 3. ATOMS
Ex − E0
Nx
= exp − .
N0 kB T
The typical value for the energy difference Ex − E0 is order of 10 eV,
while at room temperature kB T is about 25 meV, so the ratio in the
exponent is large, about 400, and the ratio is very small. According to
of N0 by of atoms between the ground state E
excited state Ex accounted for by th
N x $ N 0e#(Ex #E0)/kT, (40-29)
itation. (b) An inverted population,
in which k is Boltzmann’s constant. This equation seems reasonable. The quantity by special methods. Such a populati
kT is the mean kinetic energy of an atom at temperature T. The higher the sion is essential for laser action.
temperature,72 the more atoms — on average — will have been “bumped up” by 3. ATOMS
CHAPTER
thermal agitation (that is, by atom – atom collisions) to the higher energy state Ex.
Also, because Ex ! E0, Eq. 40-29 requires that Nx " N0; that is, there will always
be fewer atoms in the excited
Einstein, state than in
the probability of the ground state.
absorption of aThis is what
photon ofweenergy Ex − E0
expect if the level populations N0 and Nx are determined only by the action of
on an atom in its ground state
thermal agitation. Figure 40-19a illustrates this situation.
is the same as that of a stimulated
If we now flood the atoms of Fig. 40-19a with photons of energy Ex # E0, pho- radiate such
emission on an atom in its excited state. Hence, if we
tons will disappear via absorption
an ensemble by ground-state
of atoms atoms and
with photons ofphotons
energywill x −
Ebe gen-
E0 , most of the
W Discharge tube W
erated largelyphotons
via stimulated
will emission of excited-state
be absorbed by [Link]
their showed
groundthat state and much
the probabilities per atom for these two processes are [Link], because there
less will stimulate emission, simply because there are much Mmore of M2
are more atoms in the ground state, the net effect will be the absorption of photons. 1
To producethelaser
former
light,atoms
we mustthanhave the
morelatter ones.
photons emitted than absorbed; + Vdc –
(leak
If we
that is, we must have want that
a situation halfstimulated
in which of the atoms
emissionare in the Thus,
dominates. excited
we state, so that
Figure 40-20 The elements of a heliu
need more atoms in the excitedphotons
the irradiating state thancause
in the absorptions
ground state, as in Fig.
and 40-19b. emissions
stimulated inAn applied potenti
neon gas laser.
However, because
equal such a population
numbers, inversion is not
the temperature consistent
should be with thermal sends electrons through a discharg
equilibrium, we must think up clever ways to set up and maintain one. containing a mixture of helium gas
neon gas. Electrons collide with he
Ex − E0
The Helium–Neon Gas Laser T = ' 2 · 105 K . atoms, which then collide with neo
kB ln 2 which emit light along the length o
Figure 40-20 shows a common type of laser developed in 1961 by Ali Javan and
tube. The light passes through tran
his coworkers. The glass discharge tube is filled with a 20 : 80 mixture of helium windows
To produce laser light, we need even more photons emitted thanWab- and reflects back and f
and neon gases, neon being the medium in which laser action occurs. through the tube from mirrors M1
sorbed, so more atoms in the excited state than in the
Figure 40-21 shows simplified energy-level diagrams for the two types of atoms. ground one,
to cause more neon atom emission
calledpassed
An electric current population inversion,
through the which
helium – neon is clearly
gas mixture impossible
serves — through thermally
of the lightbe-
leaks through mirror M
collisions between
causehelium
the atoms
atomsanddisintegrate
electrons of the
atcurrent—to
much lowerraisetemperatures.
many helium form the laser beam.
The first successful experi-
mental demonstration of laser Metastable
state
light was performed by A. Ja-
The current (electrons)
van using a He-Ne gas laser. The
excite the helium atoms E3 E2
20
simplified diagram of the(but
by collisions energy
not the He–Ne
collisions E1 Laser light
levels of the twomore atoms is neon
massive shown (632.8 nm)
side by side in atoms).
Fig. 3.14. If
15 Excitation
an electron beam passes through Rapid Then the helium a
via collisions decay
the mixture of the gases, the excite the neon at
to level E2 by coll
Energy (eV)
Exercises
X-rays and lasers
reason was found by Heisenberg, and his discovery opened the door
to a completely new description of physical reality, which will be the
subject of the course on quantum mechanics. In this section we shall
suffice with describing the characterization of the electrons according
to quantum mechanics without discussing the theory behind it.
Bohr’s second postulate lead naturally to the concept of quantized
energy levels of electrons inside the atom. We also mentioned that
it suggested the quantized nature of other kinematic and dynamical
quantities, such as angular momentum, but the predicted values were
not correct. We now present the correct values. The energy E is a
scalar quantity and single quantum number n is sufficient to charac-
terize it. Angular momentum L ~ is a vector, so it appears natural that
it requires more than one quantum number.
µ = −ml µB ,
with s = 12 for the electron. With the discovery of new particles, it was
found that they always have spin either half integer (most commonly
1
2 ) or integer values (most commonly 1). We call the particles falling
into the first class fermions and the particles with integer spin bosons.
The total angular momentum is the sum of the orbital one and
the spin. In the ground state of the electron the orbital angular mo-
mentum is l = 0, so its total momentum is equal to its spin. Just like
the components of L, ~ the component of spin in an arbitrary direction,
conventionally called z is
1
Sz = ms ~ with ms = ± (3.6.4)
2
where ms is called the spin quantum number. The magnetic moment
78 CHAPTER 3. ATOMS
j = l ± s.
3.7 Atoms
The state of the electron inside the atom is characterized uniquely
with four quantum numbers: n, l, ml and ms . Without external
electric or magnetic field, the energy is determined uniquely by n.
One energy level with principal quantum number n is called a shell.
In quantum mechanics shells are denoted according to their principal
quantum numbers. In spectroscopy and chemistry the capital letter
K, L, M etc. are used such that K corresponds to n = 1, L to n = 2
and so on.
A shell has n states with n different values of the orbital angu-
lar momentum quantum number l. Multi-electron atoms states with
different l have different energies. There are 2l + 1 states with differ-
ent values of the magnetic quantum number ml belonging to each l,
which are said to belong to the same subshell. Subshells are denoted
usually with letters. The correspondence between the letters and the
orbital quantum numbers is shown in Table 3.1. We denote a subshell
by writing its principle quantum number followed by a letter charac-
terizing the orbital quantum number. For instance, the subshell 3p
has n = 3 and l = 2.
Finally, the spin quantum number ms can take two values for any
electron with values n, l and ml fixed. Thus, the total number of
CHAPTER 3. ATOMS 79
Table 3.1:
l 0 1 2 3 4 5
subshell s p d f g h
the sodium atom is due to the spin of its valence electron (l = 0 on the
s state). The alkali metals in the first column are chemically active
because their valence electrons combine readily with atoms that have
a “vacancy” (of electron) in their outermost shell.
The prime examples for such vacancies are the elements in the
column VIIA, called halogens. For instance, chlorine has 17 electrons
in a configuration 1s2 2s2 2p5 , which means that it has a vacany or
“hole” on its 3p subshell. If a chlorine and a sodium atom come close,
the valence electron of the letter will occupy this hole, making a strong
bound between the towo atoms. Indeed, NaCl is a stable compound.
The electron configurations of the elements can be deduced sim-
ilarly. The elements in column IIA (alkali earth metals), such as
potassium, have two valence electrons, while those in column VIA
(gases and metalloids), such as oxygen have two holes, which results
in stable compounds like CaO. Of course, elements with one valence
electron and two holes can also form compounds, the most common
example being H2 O.
It is fairly simple to derive the electron configurations of the ele-
ments in the first three horizonthal periods. In the 4th and 5th period
the 3d subshell gets filled before the 4p, leading to the appearence of
ten transtion metals in each period. In the 6th and 7th periods after
the 6s subshell, the 14 states of the 4f subshell start to fill, leading to
the inner transition metals: the lanthanide series in the 6th and the
actinide series in the 7th period.
It is remarkable that we can understand so much about the atoms
based on a few assumptions (quantization of energy according to
Bohr’s postulates, quantization angular momentum and the three
principles stated above). To learn about the motion of electrons inside
the atom we need quantum mechanics.
Exercises
Atoms and the periodic table of elements
Basics of Quantum
Mechanics
2. The probability that the electron passes through the two slits
is equal to the sum of probabilities that it passes through the
single slits separately, P12 = P1 + P2 only if we observe which
83
84 CHAPTER 4. BASICS OF QUANTUM MECHANICS
Figure 4.1:
In the upper figure, the first S-G apparatus splits the beam ac-
cording to the spins in the vertical direction, and then the spins in
the |z, −i are blocked. If we let the remaining beam through another
S-G apparatus, separating also in the z direction, then we do not
observe any more separation. This shows that the states |z, +i and
|z, −i have zero overlap, so these are orthogonal states, which can be
expressed formally as
hz, −|z, +i = hz, +|z, −i = 0 .
The operation in this equation is an inner product defined on the
elements of the abstract vector space.
In the 2nd S-G apparatus the intensity of the |z, +i beam does not
change, so we have perfect overlap between states |z, +i and |z, +i,
hz, +|z, +i = 1 (and also hz, −|z, −i = 1) .
We conclude that the states |z, +i and |z, −i form an orthonormal
basis of a two dimensional complex vector space. It cannot be a real
space because there |z, +i and |z, −i are anti-parallel, not orthogonal,
which makes spin so strange to us. All other spin states (spins in any
other direction) can be expressed as linear combination of these basis
states.
The orthonormal basis can be written in a more compact form if
we introduce the following notation for the basis vectors:
|z, +i ≡ |1i , |z, −i ≡ |2i , hence hi|ji = δij .
CHAPTER 4. BASICS OF QUANTUM MECHANICS 87
hx, +|z, +i =
6 0, and also hx, −|z, +i =
6 0.
Hence the vectors |x, ±i are linear combinations of the basis vectors
|z, ±i,
(x,±) (x,±)
|x, ±i = c+ |z, +i + c− |z, −i .
88 CHAPTER 4. BASICS OF QUANTUM MECHANICS
oi :a coefficient
Then we can write the following sequence of equalities:
or an ith
∗ ∗ ∗ ∗
element
associated w. oi hej |ei i = hej |Ô|ei i = hei |Ô† |ej i = hei |Ô|ej i = o∗j hei |ej i = oj hei |ej i
In the context of matrix elements, the bra state is used to "sandwich" an op. O
inn. prod gives a = oj hej |ei i , to compute the matrix element of the op-O
complex number between 2 basis states ket state ej, ei
representing the
hence hej |ei i = 0, i.e. the eigenvectors of a self-adjoint operator that
overlap or similarity
between the two states.
belong to different eigenvalues are orthonormal (there norm is 1 ac-
cording to our convention). Then the multiplication of Eq. (4.2.4)
CHAPTER 4. BASICS OF QUANTUM MECHANICS 91
Multiplying this equation with hej | and using the orthogonality con-
important dition of the ket vectors gives a complex number,
consequences in
quantum mechanics.
For example: hej |Â|ei i = aj,i .
Real Eigenvalues:
Hermitian operators Those numbers can trivially put into a matrix form, so linear operators
have real eigenvalues.
The eigenvalues can be represented by matrices on a basis, which is called the O-
correspond to the A Hermitian
possible measurement representation of the operator Â. If the operator is self-adjoint, its matrix is a
outcomes of the
observable associated
matrix representation has to be Hermitian. complex square
matrix that is
with the operator. The z component of the spin can have two states, so Ŝz acts on a equal to its own
conjugate
Orthogonal two-dimensional Hilbert space. Such an operator can be represented transpose. In
Eigenvectors: The other words, if A
eigenvectors by a 2 × 2 complex matrix. A general 2 × 2 complex matrix has eight is a Hermitian
corresponding to matrix, then A is
real
distinct eigenvalues of a parameters (four elements, each having a real and an imaginary equal to the
orthogonal to each
part).
Hermitian operator are The condition of Hermiticity means that the elements in the complex
conjugate of its
diagonal have to be real, while those in the off diagonal are com-
other. This orthogonality transpose:
property is essential for
the orthogonal plex conjugates of each other. Hence the matrix that represents the A = A^†,
decomposition of
quantum states. operator Ŝz has four real parameters, that we can write in the form
Conservation of
d + c a − ib
Probability: The
conservation of Ŝz ↔ = aσ1 + bσ2 + cσ3 + d1 (4.2.5)
probability in quantum
a + ib d − c
mechanics relies on the
Hermiticity of operators,
where we have introduced the notations for the Pauli matrices
as it ensures that the
total probability of all
possible outcomes 0 1 0 −i 1 0
sums to one. σ1 = , σ2 = , σ3 =
1 0 i 0 0 −1
92 CHAPTER 4. BASICS OF QUANTUM MECHANICS
and 1 is the 2×2 unit matrix. Eq. (4.2.5) shows that these for matrices
form a basis in the vector space of 2 × 2 Hermitian matrices.
In order to find the correct representation of the operator Ŝz , we
need specifiy the real coefficients a, b, c and d in Eq. (4.2.5) such that
Eq. (4.2.3) is fulfilled. To do so first we work out the actions of the
Pauli matrices on the doublet representations of the |ii basis vectors
easily:
0 1 1 0
σ1 |1i = = ,
1 0 0 1
0 1 0 1
σ1 |2i = = ,
1 0 1 0
0 −i 1 0
σ2 |1i = = −i ,
i 0 0 1
0 −i 0 1
σ2 |2i = =i ,
i 0 1 0
1 0 1 1
σ3 |1i = = ,
0 −1 0 0
1 0 0 0
σ3 |2i = =− .
0 −1 1 1
We see that the matrix S3 = ~2 σ3 has exactly the required action
on the basis vectors (cf. with Eq. (4.2.3)), so it provides a correct
representation of the Ŝz operator, Ŝz ↔ S3 .
We have emphasized before that the direction of the z axis is ar-
bitrary, or more precisely, we define it by the direction of the external
magnetic field. Let us now first fix our coordinate system and choose
the direction of the external magnetic field in the direction described
by the polar and azimuthal angles (ϑ, ϕ). Then the spin operator is
a self-adjoint operator that depends on these two angles Ŝ = Ŝ(ϑ, ϕ).
In order not to carry the physical dimension of the spin operator, we
introduce its dimensionless, rescaled version by
~
Ŝ(ϑ, ϕ) ≡ σ̂(ϑ, ϕ) . (4.2.6)
2
We would like to find the representation of the σ̂(ϑ, ϕ) matirx. As it
is a Hermitian matrix, the decomposition given in Eq. (4.2.5) applies
also here and we have to find the real coefficients a − d. Before you go
on you might challenge your intuition and try to find out how those
numbers are related to the angles ϑ and ϕ.
CHAPTER 4. BASICS OF QUANTUM MECHANICS 93
The overall phase drops out from averages, hence it cannot be mea-
sured and can be chosen at wish. We set ψ = 0 for simplicity. We do
not yet know the meaning of χ.
The σ̂i (i = 1 or 2) operators in σ3 representation can be given by
the 2 × 2 Hermitian matrices
x11 x12 y11 y12
σ1 = , σ2 =
x∗12 x22 ∗
y12 y22
where x11 , y11 , x22 , y22 are real numbers, representing the average
values of measuring the x and y components of the spin on Ŝz eigen-
states pointing up and down. We know from the S-G apparatus anal-
yses that those mean values all vanish as we can measure both +~/2
and −~/2 values with equal probabilities. We also know from the
S-G measurements that the eigenvalues of the σ1,2 matrices, i.e. the
solutions of the characteristic equations λ2 − |x12 |2 and λ2 − |y12 |2 ,
are ±1, so |x12 | = |y12 | = 1. Let us assume that
eiξ eiη
0 0
σ1 = and σ 2 = .
e−iξ 0 e−iη 0
Then
ϑ ϑ −i(χ+ξ)
hσ1 i+ = cos sin e + ei(χ+ξ) = sin ϑ cos ϕ ,
2 2
hence χ = ϕ − ξ, and similarly
ϑ ϑ −i(χ+η)
hσ2 i+ = cos sin e + ei(χ+η) = sin ϑ sin ϕ ,
2 2
so cos(χ + η) = cos(ϕ − ξ + η) = sin ϕ = cos ϕ − π2 , and η − ξ = − π2 .
We can choose one of the phases freely, then the other two phases are
fixed. By convention we set ξ = 0. Then η = − π2 and χ = ϕ. The
complete solution that reflects the observations made with the S-G
apparatus with this convention reads
cos ϑ2 sin ϑ2
−iϕ
e
|s+ i = , |s− i =
eiϕ sin ϑ2 cos ϑ2
and
x12 = eiξ = 1 , y12 = eiη = −i .
Thus we see that the matrix representation of the operator σ̂x in the
σ3 representation is the Pauli matrix σ1 , while that of σ̂y is σ2 . We
CHAPTER 4. BASICS OF QUANTUM MECHANICS 95
ϑ ϑ
|s+ i = cos |1i + eiϕ sin |2i ,
2 2
ϑ −iϕ ϑ
|s− i = cos |2i − e sin |1i
2 2
we obtain
ϑ ϑ
|1i = cos |s+ i − eiϕ sin |s− i ,
2 2
ϑ ϑ
|2i = e−iϕ sin |s+ i + cos |s− i .
2 2
Then we can compute the matrix elements of the operator σ̂(ϑ, ϕ) as
ϑ ϑ
h1|σ̂(ϑ, ϕ)|1i = cos2 − sin2 = cos ϑ
2 2
ϑ ϑ −iϕ
h1|σ̂(ϑ, ϕ)|2i = 2 cos sin e = sin ϑe−iϕ
2 2
ϑ ϑ
h2|σ̂(ϑ, ϕ)|1i = 2 cos sin eiϕ = sin ϑeiϕ
2 2
2 ϑ ϑ
h2|σ̂(ϑ, ϕ)|2i = sin − cos2 = − cos ϑ ,
2 2
or in explicit matrix form
sin ϑe−iϕ
cos ϑ
σij (ϑ, ϕ) =
sin ϑeiϕ − cos ϑ
= σ1 sin ϑ cos ϕ + σ2 sin ϑ sin ϕ + σ3 cos ϑ .
Using the definitions of the Pauli matrices, you can derive easily
the following general decomposition of their products:
σi σj = δij 1 + i
X
εijk σk , (4.2.7)
k
and with the help of Eq. (4.2.7), we find the following commutation
and anti-commutation relations
(ÂB̂ − B̂ Â)|ai bj i = 0 ,
If Ô is self-adjoint, then
D E D E D E 2
∆O2 = hx| Ô† − Ô Ô − Ô |xi = Ô − Ô |xi .
x x x
CHAPTER 4. BASICS OF QUANTUM MECHANICS 97
where we used that Ôi are both self-adjoint operators. Then the right
hand side of the inequality (4.3.1) can be written as
D E D E 2
Ô1 − Ô1 Ô2 − Ô2
x x
1 D E2 1 D E 2 cross terms that
= Ŝ1 + Ŝ2 +
4 4 add to zero
1 Dh iE 2
≥ i Ô2 , Ô1 .
4 x