1 Relativistic Quantum Mechanics
In order to describe the dynamics of particles involved in high-energy collisions we must be able to combine the theory of phenomena occurring at the smallest scales, i.e. quantum mechanics (QM), with the description of particles moving close to the speed of light, i.e. special relativity. To do this we must develop wave equations which are relativistically invariant (i.e. invariant under Lorentz transformations). In this section we will derive relativistic equations of motion for scalar particles (spin-0) and particles with spin-.
The Standard Model Lagrangian in popular form reads
| (1) | ||||
where the first line accounts for the self-interactions of the pure gauge fields, the second the free Dirac fields and their interactions with the gauge fields, the third line the interaction between the fermions and the Higgs field (i.e. the generation of masses to the fermions) and finally the fourth line the Higgs boson field and its interaction with the gauge fields. The focus of this brief course is to explain the notation contained in the two middle lines, and explore some of the complexity of the fields and the operator .
1.1 The Klein-Gordon Equation
Consider first the Hamiltonian for a particle in classical (non-relativistic) mechanics
| (2) |
To convert this into a wave equation for QM, we use the identifications and , so that a plane-wave solution
| (3) |
has the energy-momentum relation given in (2). Applied to a general wavefunction , a linear superposition of plane waves, this gives
| (4) |
where is the so-called Hamiltonian operator. We recognise this as the Schrödinger Equation, the cornerstone of QM. (4) cannot be relativistically invariant because time appears only through a first-order derivative on the left-hand side while space appears as a second-order derivative on the right-hand side. Yet we know that if we make a Lorentz transformation, this would mix the and components and therefore their derivatives will have to arise with the same orders.
The problem with the Schrödinger Equation arose because we started from a non-relativistic energy-momentum relation. Let us then start from the relativistic equation for energy. For a particle with 4-momentum and mass ,
| (5) |
Again we convert this to an operator equation by setting so that the corresponding wave equation for an arbitrary scalar wavefunction gives
| (6) |
where we have introduced the four-vector . This is the Klein-Gordon (KG) equation which is the equation of motion for a free scalar field. We can explicitly check that this is indeed Lorentz invariant.
Under a Lorentz transformation the contravariant four-vector and covariant transform as
| (7) |
The Lorentz transformations have the special property that
| (8) |
This equation says that the inverse of transformation is given by the transformation .
The field is a scalar, i.e. it has the transformation property
| (9) |
Therefore, in the primed system one obtains, using the aforementioned properties of the Lorentz transformations
| (10) |
and the equation still holds.
1.2 The Dirac Equation
The KG equation admits negative-energy solutions, because the energy appearing in the plane-wave in (3) can have the two values . This originates from the energy-momentum relation , (5). Dirac sought to find an alternative relativistic equation which was linear in like the Schrödinger equation since looks like it would remove the negative energy solution. If the equation is linear in , it must also be linear in if it is to be invariant under Lorentz transformations.
1.2.1 An alternative wave equation
We therefore start with the general form
| (11) |
We also want the solution to this equation to follow the energy relation (5). To implement this constraint, we square (11)
| (12a) | |||||
| (12b) | |||||
| (12c) | |||||
The KG equations, which is just the energy-momentum relation, requires the equality marked with the exclamation mark. Therefore, for and
| (13a) | ||||
| (13b) | ||||
| (13c) | ||||
Here we have defined the anti-commutator which is similar to the normal commutator . It is clearly not possible to satisfy these equations for numbers and . Instead, Dirac defined for and and to be a column vector. One can show that (13) requires that
| (14) |
and that the eigenvalues of all matrices need to be . Therefore, needs to be even. The simplest case of has only three independent matrices, the Pauli matrices
| (15) |
which is not enough. Therefore, the simplest solution is , for example,
| (16) |
Note that the and the here are both matrices and that this block notation is just a shorthand notation. In the future, I will drop the subscript. It is customary to collect these matrices as
| (17) |
and to define as a new Lorentz vector where each component is a matrix. The condition (13) then becomes
| (18) |
With these matrices we can now write down the Dirac equation by multiplying (11) with
| (19) |
In momentum space, i.e. setting via Fourier transformation, this becomes
| (20) |
These matrices are an example of a Clifford algebra. In fact, any set of matrices that fulfils (18) can be used to construct the Dirac equation. The representation in (16) is just an example, known as the Dirac representation. It is possible to find another representation by transforming
| (21) |
where is a unitary matrix.
We mentioned in passing that is a column vector rather than a scalar. This means that it contains more than one degree of freedom. Dirac exploited this property to interpret his equation as the wave equation for spin-1/2 particles, fermions, which can be either spin-up or spin-down. The column vector is known as a Dirac spinor, and the matrices operate on these Dirac spinors.
1.2.2 Negative energy solutions
Compare the Schrödinger equation (4) to the Dirac equation. This gives the Hamiltonian for a free fermion as
| (22) |
The trace of the Hamiltonian gives us the sum of the energy eigenvalues for all internal degrees of freedoms. Since and assumed to be traceless, the eigenvalues of must sum to zero. Therefore, we still have negative energy solutions, just like the KG equation!
Dirac’s solution to this problem is called the Dirac sea, cf. Figure 1. The negative-energy states are real but the vacuum is defines as the having all of those states already filled. This solves the practical problems because observations rely on energy differences but leaves a vacuum with infinite negative charge and energy.
Since it can be shown that fermions follow Pauli’s exclusion principle (spin-statistics theorem: particles that use anti-commutators rather than commutators have spin ), a positive-energy electron cannot fall into the negative-energy sea. It is however possible to excite on the negative-energy states into a positive-energy state, leaving behind a hole. This hole can be interpreted as a state with positive energy and positive charge, called a positron. Dirac predicted the existence of these particles in 1927 and they were experimentally confirmed in 1932.
A more consistent solution to this problem that also covers the KG equation was proposed later by Feynman and Stückelberg, based on considering the plane wave solution (3) which has a term . Their idea is to consider a particle that has to propagate forwards in time and a particle with to propagate backwards in time.