PHAS0067: Advanced Physical Cosmology

8 Fields

()

This maths in this Section is non-examinable and is included just to fill in the details for those who want it. The physical concepts about scalar fields in cosmology, which are examinable and will be discussed in class, are recapped in the summary section at the end.

When we make the transition from SR to GR, the Minkowski metric ημ⁢ν is promoted to a dynamical tensor field, gμ⁢ν⁢(x). In this sense, GR is an example of a classical field theory. Working in field theoretic language is helpful for theoretical derivations -- it allows a powerful sort of meta-analysis of the properties of equations2121 21 See, for example, the astounding Noether’s theorem — one of the most profound results in all of physics — which we previously mentioned in the context of why energy conservation does not apply to photons in an expanding universe; https://en.wikipedia.org/wiki/Noether%27s_theorem.. But even more importantly, understanding classical field theory is an essential pre-requisite for a working understanding of the mechanism thought to be responsible for cosmic inflation.

So this Section is a crash course in classical field theory. For a more complete overview of the underlying theory try the courses MATH6202 or PHAS3424. We will not discuss quantum fields in this chapter, though we will touch on them in the second part of the course when the origin of primordial perturbations is discussed.

8.1 Euler-Lagrange equations in flat-space field theory

In Section 7 we recapped the Euler-Lagrange equations to see how they simplify calculations for geodesics. To go further, we need just a bit more formal development to see how the equations work for fields, as opposed to single particles.

We first replace the single coordinate q⁢(t) by a set of spacetime-dependent fields Φi⁢(xμ), and the action becomes a functional of these fields. Such fields could be, for example, electric or magnetic fields, or gravitational fields, or scalar fields (like the Higgs field) and so on… in any case, i labels individual fields. A functional is a function of an infinite number of variables – in this case, the values of a field at each point in spacetime.

The Lagrangian can now be expressed as an integral over the space of a Lagrangian density, ℒ, which is a function of the fields Φi and their spacetime derivatives ∂μ⁡Φi:

L=∫d3⁢x⁢ℒ⁢(Φi,∂μ⁡Φi). (217)

Then the action becomes,

S=∫𝑑t⁢L=∫d4⁢x⁢ℒ⁢(Φi,∂μ⁡Φi). (218)

We should note in passing that, if we’re to have a hope of building a coordinate-independent theory, we need to make sure that S is independent of coordinate system. We will adopt this as a requirement.

The Euler-Lagrange equations, just like when we considered them for individual particles, come from requiring that S be invariant under small variations – but this time around, variations of the field:

Φi→Φi+δ⁢Φi,∂μ⁡Φi→∂μ⁡Φi+δ⁢(∂μ⁡Φi)=∂μ⁡Φi+∂μ⁡(δ⁢Φi). (219)

The expression for variation in ∂μ⁡Φi is the derivative of the variation of Φi, just as we saw in Exercise 7. Since δ⁢Φi is assumed to be small, we can Taylor-expand the Lagrangian under this variation,

ℒ⁢(Φi,∂μ⁡Φi) → ℒ⁢(Φi+δ⁢Φi,∂μ⁡Φi+∂μ⁡δ⁢Φi), (220)
= ℒ⁢(Φi,∂μ⁡Φi)+∂⁡ℒ∂⁡Φi⁢δ⁢Φi+∂⁡ℒ∂⁡(∂μ⁡Φi)⁢∂μ⁡(δ⁢Φi).

Correspondingly, the action goes to S→S+δ⁢S, with

δ⁢S=∫d4⁢x⁢[∂⁡ℒ∂⁡Φi⁢δ⁢Φi+∂⁡ℒ∂⁡(∂μ⁡Φi)⁢∂μ⁡(δ⁢Φi)]. (221)
😇 Exercise 8A Factor out the δ⁢Φi term from the integrand, by integrating the second term by parts. Hence show that δ⁢S=∫d4⁢x⁢[∂⁡ℒ∂⁡Φi-∂μ⁡(∂⁡ℒ∂⁡(∂μ⁡Φi))]⁢δ⁢Φi. (222) Hint: You will obtain one term which is a total derivative – the integral of something of the form ∂μ⁡Vμ – that can be converted to a surface term by the four-dimensional version of Stokes’ Theorem. As when we derived Euler-Lagrange equations for particles, we choose to consider variations that vanish at the boundary, along with their derivatives.

As before, we now argue that since this is to be true for any possible change δ⁢Φi, the term multiplying δ⁢Φi must be zero everywhere:

∂⁡ℒ∂⁡Φi-∂μ⁡(∂⁡ℒ∂⁡(∂μ⁡Φi))=0. (223)

Before moving on, let’s introduce a slightly more compact notation for this kind of argument. Namely, let δ⁢S/δ⁢Φi be the functional derivative of a functional S with respect to a function Φi, defined to satisfy

δ⁢S=∫d4⁢x⁢δ⁢Sδ⁢Φi⁢δ⁢Φi (224)

The actual function denoted by δ⁢S/δ⁢Φi has to be extracted by the exact same series of manipulations that we’ve explored above, so there’s nothing actually new here. It’s just that, for example, saying “δ⁢S/δ⁢Φi=0” will be a helpful short-hand for saying “δ⁢S=0 for all possible small changes to Φi”.

In addition to all the benefits of the Lagrangian formulation for particle mechanics that we discussed above, in field theory we also get a unique definition for the energy-momentum tensor. This is essential to have in hand as we move from familiar content for the universe (matter, radiation) on to more abstract content required for study of the early Universe and inflation.

8.2 Scalar fields (still in flat space)

A scalar field is simply an association of a single number with every point in spacetime. To turn this into a physical theory, we need to specify an equation of motion for the field – or, equivalently, a Lagrangian. We’ll go right ahead and start by writing down a Lagrangian:

Sφ=∫d4⁢x⁢(12⁢ημ⁢ν⁢(∂μ⁡φ)⁢(∂ν⁡φ)-V⁢(φ)) (225)

The equations of motion for this field are

□⁢φ+d⁢Vd⁢φ=0⁢, (226)

where the box operator (also known as the d’Alembert operator) is defined by □⁢φ≡ημ⁢ν⁢∂μ⁡∂ν⁡φ=φ¨-∇2⁡φ.

😇 Exercise 8B Derive equation (226) from equation (225) using expression (223).

What on earth does it mean? First, set V=0. Then φ obeys a wave equation. Recalling that c=1 in this course, the waves will propagate at the speed of light. You can see this explicitly by substituting the trial wave solution φ∝exp⁡(i⁢k⁢(x-t)).

So, localized features in the field travel along and start to look rather like a quantum particle (more in the exercise below).

Alternatively, imagine the field starts out completely homogeneous. Then, ∇2⁡φ=0 (and it remains zero for all time). All we have, at each point, is φ¨+d⁢V/d⁢φ=0. This equation is also familiar; for example, if we make the potential V⁢(φ)∝φ2 we just have a ‘mass on a spring’ which oscillates back and forth forever. The only difference is that φ is an abstract number, not an actual position. This is a useful analogy: quite often when discussing inflation we will talk about things like “balls rolling down a hill”. The thing to remember is that in reality there is no ball, no rolling, and no hill! There is just an abstract mean field value represented by the ball’s height, an abstract rate of change of that value represented by the motion of the ball, and an abstract potential represented by the hill.

☞ Exercise 8C What if we combine the two scenarios? Let’s say we have V⁢(φ)=m2⁢φ2/2 and wish to find a travelling wave solution. Substitute φ∝exp⁡(i⁢(k⁢x-ω⁢t)) and show that ω2=k2+m2⁢. (227) Recall from your quantum mechanics courses that ⟨E⟩=ℏ⁢ω, and that ⟨p⟩=ℏ⁢k. The relation you have derived then becomes the relativistic equation E2=m2+p2 (with c=ℏ=1 as always).

The result from the exercise above is not a coincidence – we’re seeing that relativistic particles can be understood as oscillations in quantum fields. We’ll touch on this later but the detail is beyond the scope of this course (see e.g. PHAS0073 instead). Note that in the limit k→0 – when the fluctuations in the field are on very long wavelengths – we get ω2≈m2, i.e. we automatically revert to the homogeneous, coherent oscillator picture outlined just above the exercise. This sort of coherent behaviour of a scalar field (suitably modified for expanding spacetime) will be our main focus for cosmology.

8.3 Into curved space

Before looking at cosmological implications, we need to generalise the field theory framework into curved space. One might imagine that we can use the equivalence principle to promote the fixed metric of Minkowski space ημ⁢ν into a general metric gμ⁢ν:

Sφfirst⁢guess=∫d4⁢x⁢(12⁢gμ⁢ν⁢(∇μ⁡φ)⁢(∇ν⁡φ)-V⁢(φ)) (228)

Sadly, it’s not true. The problem is that d4⁢x is dependent on our choice of coordinates. For example, suppose we transform coordinates such that x1→x1⁣′=2⁢x1. This isn’t permitted in special relativity, but in general relativity we have to allow it – remember the whole construction is about keeping physics invariant even in different coordinate systems. If Sφ is to have an actual physical meaning, it has to be invariant under this sort of transformation.

The integrand is correctly invariant under our trial transformation, because the g11⁣′ component of the metric picks up a factor 4 that cancels against a 1/2 attached to each of the ∂/∂⁡x1⁣′s. But the integral still changes because d⁢x1→d⁢x1⁣′=2⁢d⁢x1. This needs to be cancelled off, and it turns out the way to do it in general is to write:

Sφ=∫d4⁢x⁢-g⁢(12⁢gμ⁢ν⁢(∇μ⁡φ)⁢(∇ν⁡φ)-V⁢(φ)) (229)

where g≡det⁡gμ⁢ν is the determinant of gμ⁢ν. The point is, this determinant exactly cancels the effect of coordinate transformations on the d4⁢x measure, as we’ll now see…

😇 Exercise 8D Check that this works in our specific test case of x1→x′⁣1=2⁢x1. Start by verifying that, for the original metric gμ⁢ν=ημ⁢ν, the action reduces back to (225). Then, calculate det⁡gμ⁢ν and so write the action in the primed coordinates. Finally transform your new expression back into the unprimed coordinate system to once again recover (225).

Of course we’ve just shown it works for a simple, specific coordinate transformation but it’s not too much extra work to show that it actually works for any coordinate transformation (see Carroll for help with this if you want to give it a go)2222 22 It also turns out that measure theory gives you a more fundamental justification for why d4⁢x⁢-g has to be the correct way to integrate in a curved space; see https://en.wikipedia.org/wiki/Volume_form..

By Einstein’s equivalence principle, since the action (229) is coordinate-invariant and boils down to (225) in locally Minkowski coordinates, it’s the correct extension of scalar field theory into curved space.

8.4 Euler-Lagrange equations in curved space

Following the arguments given above for the specific case, it’s correct to generalise and say that actions in curved space look like this

S=∫d4⁢x⁢-g⁢ℒ^⁢(Φi,∇μ⁡Φi), (230)

where ℒ^ is a scalar (i.e. invariant under coordinate transformations). Here, we don’t consider gμ⁢ν to be one of the fields because we have to treat it on a different footing (later). So in the scalar field example, there is just Φ1=φ and no other fields to consider.

What do the Euler-Lagrange equations now look like? Well, there’s nothing special about this new expression, so all the old arguments go through and the correct equations of motion for the field Φi are:

∂⁡(-g⁢ℒ^)∂⁡Φi-∂μ⁡[∂⁡(-g⁢ℒ^)∂⁡(∇μ⁡Φi)]=0. (231)

We’ve derived this equation from a coordinate-invariant expression so it’s got to be coordinate-invariant itself. Knowing that, we can just throw caution to the wind and have a guess at what the above statement really means:

∂⁡ℒ^∂⁡Φi-∇μ(∂⁡ℒ^∂⁡(∇μ⁡Φi))=0. (232)

This is a correct guess. The technical details of why it works out are unnecessary for our future purposes, but if you’re interested, the next exercise guides you through it.

Exercise 8E Show that equation (232) is correct. 1. First you take the -g’s out side their immediate derivatives just by noting that gμ⁢ν is not one of the Φi field(s) under consideration; 2. Next, you prove that ∂μ⁡(-g⁢Aμ)=-g⁢∇μ⁡Aμ for any vector field Aμ. If you have trouble with this step, start by trying to show the following preliminary result: ∂⁡-g∂⁡gα⁢β = -g2⁢gα⁢β, (233) which takes a fair bit of work. It arises most naturally from the variation of the identity ln⁡(det⁡𝐌)=Tr⁢(ln⁡𝐌) for 𝐌 a square matrix. It’s equally worth remembering that exp⁡(ln⁡𝐌)=𝐌. 3. Then convince yourself that ∂⁡ℒ^/∂⁡(∇μ⁡Φi) is indeed a vector field Aμ, and so transform the second term appropriately; 4. Finally you cancel the -g that is now outside all the derivatives in both terms.

Now we can apply the Euler-Lagrange equations of motion in curved space, equation (232), to the action (229) to show that

□⁢φ+d⁢Vd⁢φ=0, (234)

where now

□⁢φ=∇ν⁡∇ν⁡φ=gμ⁢ν⁢∇μ⁡∇ν⁡φ. (235)

It’s all just the same as in flat space, but with partial derivatives replaced by covariant derivatives – a direct consequence of the equivalence principle, of course. It really couldn’t have worked out any other way.

8.5 An action for curvature itself

We have developed the machinary to work out how fields behave within a curved spacetime – but now we have to recognise the fact that the metric, gμ⁢ν, is itself a field (i.e. a function of spacetime position). What scalars can we make out of the metric to serve as a Lagrangian? We know that the metric can be set equal to the Minkowski metric (gμ⁢ν=ημ⁢ν) and its derivatives set to zero at any one point (remember Exercise 4). Consequently any non-trivial, coordinate-invariant scalar must involve at least second derivatives of the metric.

We have already encountered the Ricci scalar. It turns out to be the only independent scalar constructed from the metric, which is no higher than second order in its derivatives. The mathematician David Hilbert figured that this was the simplest possible choice for a Lagrangian, and proposed that the following action might be appropriate for general relativity:

SH=∫d4⁢x⁢-g⁢R (Einstein–Hilbert Action). (236)

This turns out to be a correct guess (as we will sketch out below). In some ways it gives a very natural explanation for why Einstein’s equations look the way they do – because their action is as simple as it possibly could be.

How can we check that the Einstein-Hilbert action correctly describes GR? Since R is a function of g, the approach that we so carefully developed above – working with ℒ^ and just applying the “covariantised” Euler-Lagrange equations (232) – does not work. We have to work with the “raw” Lagrangian (231) and replace Φi by gμ⁢ν (i.e. we treat each component of the metric tensor as a separate field when deriving the equations of motion).

This becomes a bit of a dreary exercise in algebra, so (although I certainly encourage you to go through the derivation once in your life; see Exercise 5) let’s for now just state the result:

δ⁢SH =-∫d4⁢x⁢-g⁢[Rμ⁢ν-12⁢R⁢gμ⁢ν]⁢δ⁢gμ⁢ν,
=+∫d4⁢x⁢-g⁢[Rμ⁢ν-12⁢R⁢gμ⁢ν]⁢δ⁢gμ⁢ν, (237)

meaning that at stationary points of SH with respect to variations in gμ⁢ν we have:

1-g⁢δ⁢SHδ⁢gμ⁢ν=Rμ⁢ν-12⁢R⁢gμ⁢ν=0, (238)

which has recovered Einstein’s equation in a vacuum (Tμ⁢ν=0).

To generalise to Einstein’s equation coupled to matter, we just need to include the relevant matter terms in the action. Consider

S=-116⁢π⁢G⁢SH+SM, (239)

where SM is the matter action, and we have pre-normalized the gravitational action to get the right answer. (Note our normalization includes a minus sign in front of SH, which is not present in all textbooks – it depends on the metric signature.) The calculation for the revised action goes through just by linearity:

1-g⁢δ⁢Sδ⁢gμ⁢ν=-116⁢π⁢G⁢[Rμ⁢ν-12⁢R⁢gμ⁢ν]+1-g⁢δ⁢SMδ⁢gμ⁢ν=0⁢, (240)

from which one recovers the Einstein equation provided that the energy-momentum tensor can be written

Tμ⁢ν=2-g⁢δ⁢SMδ⁢gμ⁢ν. (241)

(Note again the different sign to some textbooks, again due to our choice of metric signature.)

Consider again the action (229) for a scalar field. Now vary this action with respect to gμ⁢ν, not φ:

δ⁢Sφ=∫d4⁢x⁢-g⁢δ⁢gμ⁢ν⁢[12⁢∇μ⁡φ⁢∇ν⁡φ-gμ⁢ν2⁢(gρ⁢σ2⁢∇ρ⁡φ⁢∇σ⁡φ-V⁢(φ))], (242)

we obtain from (241) the energy-momentum tensor for a scalar field,

Tμ⁢ν(φ)=∇μ⁡φ⁢∇ν⁡φ-gμ⁢ν⁢[12⁢gρ⁢σ⁢∇ρ⁡φ⁢∇σ⁡φ-V⁢(φ)]. (243)

Here, it is worth pointing out that you will find different sign conventions to (243) in the cosmology literature. This can be traced to the fact that most of the cosmology literature uses the opposite metric signature to ours: (-,+,+,+).

Exercise 8F If you want to check that the variation of the Hilbert action really gives rise to (237), start by noting that δ⁢(-g⁢R)=R⁢δ⁢-g+-g⁢δ⁢R, then tackle the two terms separately. For the first term, remind yourself of result (233). For the second term, further split δ⁢R=Rα⁢β⁢δ⁢gα⁢β+δ⁢Rα⁢β⁢gα⁢β and use the expression (105) to show that δ⁢Rα⁢β=∇μ⁡δ⁢Γα⁢βμ-∇β⁡δ⁢Γα⁢μμ⁢. (244) Hence argue the term ∝δ⁢Rμ⁢ν leads to an integral with respect to the covariant divergence of a vector; by Stokes’ Theorem, this is equal to a boundary term that you can discard. Putting everything together, you should obtain the required result.
Exercise 8G We asserted that the energy-momentum tensor must be given by Equation (241); otherwise we don’t recover Einstein’s equations. This is certainly a valid argument, but we earlier remarked that conservation of the energy-momentum tensor (∇μ⁡Tνμ=0) is the result of a symmetry, and so can be derived from Noether’s theorem. If you’re curious to see how this works out, start in flat Minkowski space (the generalisation to curved spaces doesn’t change the main thrust of the argument). Consider coordinates xμ with metric gμ⁢ν=ημ⁢ν. Now change the coordinate system by some infinitessimal amount xμ→x~μ=xμ+ϵμ. Allow all field content to follow the transformed coordinates, so that for scalar fields φ→φ~ with φ~⁢(𝒙~)=φ⁢(𝒙). Show that g~μ⁢ν⁢(x~)=ημ⁢ν-ϵμ,ν-ϵν,μ. Hence argue that, if the action is invariant under the coordinate transformation, T,νμ⁢ν=0 where Tμ⁢ν=2⁢δ⁢S/δ⁢gμ⁢ν. How and why does this derivation differ from the derivation of the canonical energy-momentum tensor in quantum field theory books?

8.6 The practical consequence: energy-momentum of a scalar field

The above is all pretty technical, so let’s step back and see the practical consequences. So far we’ve used the Euler-Lagrange formalism to define what is meant by the energy-momentum tensor for a scalar field. Because it has been derived from a coherent, self-consistent framework, we can be assured that (for example) Tν;μμ=0 for our definition, making it properly compatible with Einstein’s theory.

The single most important result for our purposes is Eq. (243) – the correct energy-momentum tensor for a scalar field. Let’s simplify by writing it out in component form in the case that the field is homogeneous but time-varying (i.e. ∇i⁡ϕ=0, but ∇0⁡ϕ=ϕ˙≠0):

T(φ)νμ=(V⁢(φ)+φ˙2/20000V⁢(φ)-φ˙2/20000V⁢(φ)-φ˙2/20000V⁢(φ)-φ˙2/2). (245)

Compare this with Eq. (91), and we can read off the density and pressure of a scalar field in a homogeneous universe:

ρφ =12⁢φ˙2⏟“kinetic” energy+V⁢(φ)⏟“potential” energy (246)
Pφ =12⁢φ˙2-V⁢(φ)⁢. (247)

The ‘kinetic energy’ term is named by analogy with the corresponding term if we had just a single particle with position φ. But let’s be clear: φ is not a position, it’s a field value. It’s important to realise that the existence of kinetic energy in the field does not imply the field configuration represents a moving particle. In that light, the terminology is actually quite unfortunate. But it’s the convention, so we’re stuck with it.

The resulting equation of state

w=12⁢φ˙2-V12⁢φ˙2+V (248)

shows that a scalar field can lead to negative pressure (w<0) and accelerated expansion (w<-1/3) if the potential energy V dominates over the kinetic energy 12⁢φ˙2. In the next chapter, we will look at this phenomenology in detail.

8.7 Summary

As stated at the start of the chapter, the mathematical derivations are non-examinable. You may be asked to briefly explain the physics, as we will apply it to inflation in the next chapter. This can be summarised as follows:

  • •

    A field is just an assignment of a number, vector, or other quantity to every point in spacetime; often by ‘field’ we also have in mind that there are some equations of motion that predict its evolution, i.e. the future values of the field can be predicted from the past values of the fields.

  • •

    The equations governing classical evolution of fields can be derived from the Euler-Lagrange equations applied to suitable actions.

  • •

    Deriving things that way, rather than just writing down an evolution equation, helps us analyse the classical behaviour of the field – and is also essential for quantising the field (Chapter 14).

  • •

    In cosmology, we are usually most interested in scalar fields φ⁢(𝒙) (i.e. fields which assign a single number to each point in spacetime), because they may be responsible for powering the accelerated expansion of inflation and/or dark energy.

  • •

    Even in flat space, with the simplest possible non-trivial Lagrangian for a scalar field, the field can exhibit a range of behaviours from coherent oscillations to particle-like excitations. Within cosmology, it is normally coherent behaviour of the field, where the field takes one value throughout space, that is of most interest.

  • •

    When describing how this field changes over time, we quite often refer to analogies like “a ball rolling down a hill”. But there is no physical ball or hill! The horizontal position of the ball represents the field value φ (the same throughout space) and the hill is an abstract potential V⁢(φ) which you can think of as the ball’s height.

  • •

    The metric gμ⁢ν⁢(𝒙) can also be thought of as a tensor field. The action for Einstein’s relativity is then remarkably simple, Eq. (236).

  • •

    Using this action we can deduce the correct energy-momentum tensor for the fields within the Universe;

  • •

    This reveals the density ρ and pressure P of a scalar field are given by equations (246) and (247) respectively, in the limit that the field is completely homogeneous.

  • •

    If the potential V⁢(φ) dominates over the kinetic φ˙2/2 term, we find a negative equation of state P≃-ρ which we previously concluded (Section 2) will give rise to accelerated expansion.