Part III · General Relativity — Chapter 3.6
The Einstein Field Equations
The law of gravity, derived three times: cornered by a conservation requirement, extremised from an action, and its one free constant fixed by demanding that apples fall.
Everything Part III has built so far describes how matter moves in a geometry that is handed to us. Chapter 3.3 §8.3 obtained Newton's law of falling from the geodesic equation. Chapter 3.4 §4 identified the tide with curvature. Both chapters closed by naming the same missing piece: nothing yet says which geometry a given lump of matter produces. This chapter says it.
The field equations are not guessed here, and they were not guessed historically either, though the route taken below is cleaner than the route taken in 1915. The argument is a cornering.
Chapter 2.6 built the object that carries energy and momentum, and it proved that the object has vanishing divergence. Anything set equal to that object must therefore have vanishing divergence too, identically, for every geometry whatever. Chapter 3.4 §7 constructed the only combination of the curvature that does. There is essentially nothing else available. What is left is Einstein's equation with one undetermined constant in it.
Here is the route. Section 1 establishes what is on the right-hand side, and it is not mass. Section 2 lists the constraints the equation must satisfy, and says where each one comes from. Section 3 is the cornering itself.
Section 4 derives the same equation a second way, by writing down the simplest scalar that can be built from a metric and extremising it. That scalar is the entry Chapter 1.2 §8.1 put in its table of actions six chapters early. Section 5 fixes the constant by demanding Newton's limit, and it is the section where a reader decides whether to believe the whole edifice. Section 6 examines the one term the argument of §3 cannot exclude. Section 7 counts equations against unknowns, and finds that the count works out for a reason.
Conventions. The signature is . The Riemann and Ricci signs are as stated loudly in Chapter 3.4. Both and are written out. Three sign traps are met below and each is flagged where it occurs: the sign of in the weak-field metric, and the sign of the cosmological term.
Tools you'll need — Chapter 3.4 §7 above all: the second Bianchi identity, its double contraction (3.4.57), and the Einstein tensor (3.4.58) with . Also 3.4 §6 (Ricci and the scalar), §5.3 (locally inertial coordinates) and §4.5's , equation (3.4.33). Chapter 2.6 §10 for the energy–momentum tensor, its components, and . Chapter 3.5 §6 for the volume element, Jacobi's formula (3.5.47) and the divergence identity (3.5.51), and §5 for Stokes' theorem. Chapter 3.3 §7 (the Christoffel formula and metric compatibility) and §8.3 (the Newtonian limit, and its , equation (3.3.71)). Chapter 3.1 §6.5, the derived weak-field component , equation (3.1.39). Chapter 1.2 §3 (varying an action) and §8.1 (the table this chapter's fourth line comes from). Chapter 0.7 §7.5 for Poisson's equation.
1 · What sits on the right-hand side, and why it is not mass
Here is where this section is going. We identify the source of gravity as the energy–momentum tensor of Chapter 2.6. We promote its conservation law to a curved manifold, and we write down its form for the two kinds of matter this book needs. The important consequence is that pressure is a source of gravity, and no Newtonian intuition supplies that.
1.1 · The candidate, and why nothing simpler will do
In Newtonian gravity the source is the mass density , a single function. That cannot survive into a relativistic theory, and Chapter 2.5 §9 already established the reason. Mass is not additive, and it is not separately conserved.
Energy is both of those things, so energy is the natural replacement. But Chapter 2.5 §3 showed that energy is one component of a four-vector. A theory that sources gravity with energy alone would therefore source it with a frame-dependent quantity, and two observers in relative motion would compute different geometries for the same physical situation. That is not a difficulty to be worked around. It is a contradiction.
Chapter 2.6 §10 built the object that fixes this. The energy–momentum tensor arose there as the Noether current of spacetime translations. Its components were computed there too, and they are worth having in front of us again:
| Component | What it is |
|---|---|
| energy density | |
| momentum density, equivalently energy flux | |
| the stress: flux of -momentum across a surface of constant |
The tensor is symmetric, . Chapter 2.6 §10.1 checked that explicitly for the electromagnetic field, after fixing the Noether current with an improvement term. The same chapter, in §10.3, went on to prove
in flat spacetime for a closed system. That is four equations, one for each value of . Together they express local conservation of energy and of each component of momentum.
1.2 · The same statement on a curved manifold
Equation (3.6.1) is not yet a statement about a manifold. The trouble is that of a tensor is not itself a tensor, as Chapter 3.3 §4 showed. We need a replacement that is, and it is this:
and the reason for the replacement is not an appeal to elegance. It is Chapter 3.4 §5.3's locally inertial coordinates.
At any chosen point there is a chart in which and . In that chart and coincide, so (3.6.2) reduces to (3.6.1), and the equivalence principle says the flat law must hold there. Now use the second half of the argument. Equation (3.6.2) is an equation between tensors, so by Chapter 2.4 §6, holding in one chart at a point means holding in every chart at that point. And the point was arbitrary.
The recipe just used was to take the flat-space law and replace by . It is not always well defined. Suppose the flat law contains two derivatives. The two orderings and agree in flat space, and on a curved manifold they differ by a term containing the Riemann tensor, exactly as Chapter 3.4 §2 computed. So the recipe leaves the coefficient of a possible curvature term undetermined, and two laws that look different can reduce to the same flat one.
Here there is no ambiguity, because (3.6.1) has exactly one derivative and there is nothing to reorder. The ambiguity does bite elsewhere, and the standard example is the wave equation for a field in curved spacetime. There the missing coefficient has to be fixed by some other principle. This book meets the question again when quantum fields are put on a curved background.
1.3 · Dust, and then a fluid with pressure
Dust. Take non-interacting particles all moving together, with rest-mass density measured in their own rest frame and four-velocity field . In the rest frame the only thing present is energy density , so and every other component vanishes. There is exactly one symmetric tensor built from and with that property. At rest , so the tensor we want is this one:
It gives and nothing else, which is what was required, and it is built from tensors, so it is one. Its trace is worth recording while we are here. Using from Chapter 2.5, .
Perfect fluid. Now allow isotropic pressure . In the rest frame the stress block is . That is what the word "pressure" means: a flux of -momentum in the direction, the same in every direction. The energy density is still , so in the rest frame . We want a tensor that reproduces that, and only two objects are available to build it out of, and :
Let's check it component by component, in the rest frame, where and has only a entry, equal to . The component is ✓. The component is ✓, the minus sign coming from , and likewise for and . Off-diagonal entries vanish on both sides. ✓ Setting returns (3.6.3).
Its trace is worth computing now, because §5 will need it. Use together with in four dimensions, and the trace comes out as
1.4 · Three consequences worth naming before we start
(i) Pressure gravitates. contains , so pressure appears in the source. This has no Newtonian counterpart at all. In the pressure of a gas is invisible. Section 5 will find exactly how it enters, and the answer is that the combination doing the gravitating is .
(ii) A hot body weighs more than a cold one. Not because its atoms have gained mass, but because counts every form of energy, including the kinetic energy of thermal motion and the energy stored in the fields binding it. Chapter 2.5 §9 showed the same arithmetic from the other side, in the mass deficit of a bound system.
(iii) A gas of radiation has a fixed pressure. Chapter 2.6 §10 built the electromagnetic energy–momentum tensor, and its closing summary recorded that the tensor is traceless, . So set (3.6.5) to zero and solve for the pressure, which gives
and that result is derived rather than quoted. The equation of state of light follows from the tracelessness of its stress tensor and from nothing else. Section 5 and Chapter 3.9 both use it, and Chapter 7.3 builds a symmetry principle on the tracelessness itself.
Mass cannot be what gravity responds to, and the reason is not subtle once the previous part is in hand. Mass is not additive, is not separately conserved, and picking out energy alone would mean picking out one entry of a four-part object, so two observers in relative motion would compute different geometries for the same physical arrangement. What is needed is an object packaging energy, momentum and the flow of both, and the earlier part built exactly that when it asked what is conserved because the laws are the same here as there.
That object has ten independent entries and every one of them sources gravity. Energy density is only the corner entry. The diagonal spatial entries are pressure, the off-diagonal ones are shear and momentum flow, and all are on the same footing. The statement that the whole thing has no divergence is the local claim that energy and momentum are never created or destroyed, only moved, and it carries over to a curved arena unchanged because it contains only one derivative and so leaves nothing to be ambiguous about.
One consequence deserves flagging before it arrives. Pressure gravitates. Squeeze a gas without adding any material to it and the source term grows, which nothing in Newtonian gravity would lead anyone to expect and which matters enormously for dying stars and for the expansion history of the universe.
2 · What the equation is allowed to look like
Before writing anything down, let's list what the answer has to satisfy. Each item below is a constraint with a stated origin rather than a matter of taste. Section 3 will find that together they leave almost nothing to choose.
(C1) It must be an equation between tensors. The authority is Chapter 3.2 §7 together with Chapter 2.4 §6. A tensor equation true in one chart is true in all of them, and an equation that is not tensorial holds only in some charts. On a manifold there are no preferred charts for it to hold in. So a non-tensorial gravitational law would be a law about coordinates rather than about the world.
(C2) The source is a symmetric rank-2 tensor, so the geometry side must be one too. That is the content of §1. Since is symmetric with ten independent components, whatever equals it has the same type. This immediately rules out the Riemann tensor itself, which has four indices, and the Ricci scalar alone, which has none.
(C3) It must contain no derivatives of the metric beyond the second. There are two independent reasons for this, and they are worth keeping separate.
The first is the Newtonian limit, and it is the honest one. Newtonian gravity is , which is second order in the potential. Chapter 3.1 §6.5 derived , so the potential lives inside the metric. An equation that reproduces Poisson must therefore be second order in . Third or fourth derivatives would produce extra terms that Newtonian gravity does not have and that experiment does not see.
The second is the structure of every field equation so far in this book. Chapter 1.2 §8 showed that an action built from a field and its first derivatives yields an equation of motion with second derivatives, and every entry of that chapter's table of actions has that form. Section 4 below finds that the gravitational action is an exception in one specific and interesting way. The exception costs exactly one boundary term.
There is a deeper argument against higher derivatives that this book cannot yet make. A theory whose equations of motion contain more than two time derivatives generically carries extra degrees of freedom whose energy is unbounded below. The system can then lower its energy without limit by exciting them. This is Ostrogradsky's theorem, and we quote it rather than proving it, because the proof is Chapter 1.3's Hamiltonian machinery applied to a case that chapter did not cover.
It is flagged here because Chapter 7.1 leans on it. The reason one cannot repair quantum gravity by adding higher-derivative terms to the action is that the repair buys a ghost. For now, take (C3) on the Newtonian-limit argument alone, which is self-contained.
(C4) Its divergence must vanish identically. This is the constraint that does the work. Equation (3.6.2) says for every physical matter configuration. Suppose the geometry side of the field equation had a divergence that vanished only sometimes. Then setting the two sides equal would impose a further condition on the matter, and that is a condition no experiment supports and no principle supplies.
So the geometrical side must have vanishing divergence as an identity. That means automatically, for every metric anybody could write down, whether or not it solves anything.
(C5) It must reduce to Newton where Newton works. Chapter 3.3 §8.3 established the correspondence limit for the motion of a body. The same limit for the field equation is §5's business, and it is what fixes the one remaining constant.
Not one of (C1) to (C5) is an aesthetic preference. (C1) comes from the absence of preferred charts, (C2) from what matter is, (C3) from the requirement of reproducing an experimentally established law, (C4) from conservation of energy, (C5) from the same source as (C3).
The book's recurring claim is that most of the objects in it were forced rather than invented, and this chapter is the strongest case of it. Section 3 shows that (C1) to (C4) leave a two-parameter family of possible equations. Constraint (C5) then fixes one of the two parameters, and the remaining one is the subject of §6.
Before anything is designed there is a specification, and this one is short enough to hold in the head. The law must relate objects of the same kind on both sides, since otherwise it would hold only in some descriptions and be a statement about the description rather than the world. The matter side is a symmetric array with two slots, so the geometry side must be one too, which alone eliminates most of what the previous chapters built. The law must involve the shape of space and its first two rates of change and nothing beyond, because the old law it reproduces has two derivatives of the potential and the potential sits inside the shape.
Then the demanding one. Matter's bookkeeping object has no divergence, the statement that nothing is created or destroyed. If it is set equal to something built out of geometry, that something must have no divergence either, and not merely for the geometries that solve the equation. It must have none for every conceivable geometry, as an identity, or the law would quietly demand of matter something no experiment has seen matter obey.
Nothing on that list is a preference. Each traces back to the absence of preferred coordinates, or to what matter demonstrably is, or to the need to reproduce a law three centuries of astronomy confirmed. What remains after the list is applied is very nearly unique.
3 · The cornering — why and nothing else
Here is where this section is going. We list every symmetric rank-2 tensor that can be built from the metric and its first two derivatives. We impose (C4), and we find that a one-parameter family survives, up to overall scale. That family is the Einstein tensor plus a multiple of the metric, and there is nothing else.
3.1 · The list of candidates
Start with what is available. By (C3) the ingredients are , its inverse, and its derivatives up to second order. Chapter 3.4 §2 showed that the only tensor that can be built from the metric and its first two derivatives is the Riemann tensor. The first derivatives alone give the connection, which is not a tensor, and the second derivatives enter tensorially only in the combination (3.4.13). So the ingredients are and .
Now impose (C2), which asks for something symmetric with two indices. From the only contraction producing a rank-2 tensor is the Ricci tensor. Chapter 3.4 §6.1 showed that it is essentially unique, and §6 of that chapter showed that it is symmetric. Contracting once more gives the Ricci scalar , which can be multiplied by to make a symmetric rank-2 object. And itself qualifies, since it needs no derivatives at all. The complete list is therefore three terms, and the most general candidate we can write down is
with , , and constants. One caveat about that list should be said out loud now. If terms quadratic in the curvature were allowed, such as or , the list would be longer. Those terms contain squares of second derivatives, and (C3) as it was stated asked for the geometry side to be linear in the second derivatives, so they are out. Section 3.3 says what happens if that restriction is relaxed.
3.2 · Imposing (C4), one term at a time
We want to know what (C4) costs, so take the divergence of the left-hand side of (3.6.7), one term at a time. The index is raised with the metric, and the metric passes through untouched by metric compatibility (Chapter 3.3 §7.1).
Third term. , again by metric compatibility. So this term is divergence-free whatever is, and (C4) says nothing about it at all. Remember that. It is §6.
First term. Chapter 3.4 §7.2 contracted the second Bianchi identity twice and obtained (3.4.57):
So the first term contributes to the divergence.
Second term. , the metric having again passed through the derivative and then lowered the free index on .
Add them. Putting the three contributions together, the total divergence of the left-hand side is
Now the crucial step. It is a logical one rather than an algebraic one. Constraint (C4) demands that (3.6.9) vanish identically, meaning for every metric and not merely for solutions.
But is not identically zero. The sphere of Chapter 3.4 has constant , and a generic metric does not, and one has only to write down a metric with a position-dependent Ricci scalar to see it. So the only way to make the right-hand side vanish for every metric is for the coefficient in front of it to vanish:
That single condition ties to , so the next move is to put it back where it came from. Substituting the relation into (3.6.7) and factoring out the overall ,
The bracket is exactly the Einstein tensor of Chapter 3.4 (3.4.58). All that is left is to name the constants, so divide through by , absorb it into , and write :
Two constants remain. Section 5 fixes by (C5). Section 6 examines , which (C5) constrains only weakly and which the argument of this section cannot exclude at all.
In: the requirement that the equation be tensorial and symmetric, that it use the metric and its first two derivatives linearly, and that its geometry side be divergence-free identically. Plus one input from three chapters ago: the twice-contracted Bianchi identity.
Out: the form of Einstein's equation, with two undetermined constants. Notice that the factor of in the Einstein tensor was not chosen and was not fitted to data. It is the number that makes (3.6.9) vanish. It comes from the in the twice-contracted Bianchi identity, which in turn came from adding two identical dummy terms in Chapter 3.4 §7.2.
What it cost: the assumption that no terms quadratic in the curvature appear. That is the one hypothesis of this section which the earlier chapters do not force, and §3.3 says what is known about relaxing it.
3.3 · What happens if the hypotheses are weakened
Section 3.1 assumed that the geometry side is linear in the second derivatives of the metric, which excluded terms like . That assumption can be dropped.
Lovelock's theorem (1971). In four spacetime dimensions, the only symmetric, rank-2, identically divergence-free tensor that can be constructed from the metric and its first and second derivatives is for constants and . We quote this and do not prove it. The proof is a classification argument of a kind this book does not develop.
Two of the hypotheses deserve emphasis, because they are where the theorem's content lies.
The dimension matters. In five or more dimensions there are additional terms that are also identically divergence-free and also give second-order equations. They are built from the square of the curvature in a particular combination, and the first of them is called the Gauss–Bonnet term. In four dimensions that combination contributes nothing to the equations of motion at all, which is a fact about four dimensions rather than a fact about gravity. Chapter 7.8 meets these terms where they are live.
And the derivative order matters. If third derivatives are permitted, the list grows again, and Ostrogradsky's theorem is the reason not to permit them.
Subject to the flagged theorem, then, (3.6.12) is not a possible law of gravity in four dimensions. It is the possible law, up to the two constants.
This is the section where the law of gravity gets cornered rather than proposed. Start by listing everything with the right shape that can be assembled out of the geometry: there are three items, the boiled-down curvature array, the single curvature number multiplied by the distance rule, and the distance rule by itself. Write the most general mixture of the three, with unknown coefficients, and then impose the requirement that its divergence vanish for every conceivable geometry.
Two of the three items have divergences that are automatically nothing, so they impose no condition. The other two have divergences proportional to the same quantity, which is not generally zero, so the only escape is for their coefficients to cancel. That single demand fixes the ratio of the first two coefficients, and the combination it produces is exactly the one built three chapters earlier from a differentiated identity. The famous factor of one half was never chosen; it is what makes the cancellation work.
What survives is one equation with two unfixed numbers in it, and the rest of the chapter is about those two numbers. A quoted theorem strengthens this considerably: even if the restriction to the mildest possible dependence on second rates of change is dropped, nothing new appears in four dimensions, though it does in five. The law is not one option among many. It is what is left.
4 · The same answer from an action
Chapter 1.2 §8.1 printed a table of eight actions and promised that each would be constructed in its own chapter. Line four read Gravity (Einstein–Hilbert), Chapter 3.6. This section constructs that action and varies it. The equation that comes out is (3.6.12), obtained a second time by a route that shares no step with §3.
Let's lay out the structure before starting, because this is the longest computation in Part III after Chapter 3.4 §2. The integrand is , and the thing being varied is the metric. Write , and the product rule splits the variation into exactly three pieces:
Piece 2 needs no work at all. It is already in the form (something) times , which is exactly what a variational calculation wants. Piece 1 is a determinant identity, and grind box A does it.
Piece 3 looks the worst and is the most interesting. It is a total derivative, so it contributes nothing to the equations of motion and everything to a discussion of boundaries, and it takes two grind boxes, B and C. The reasoning stays here. Only the algebra is folded away.
4.1 · Why , and the action
An action is a number, so what we need this time is a scalar rather than a tensor. Constraint (C3) still says at most two derivatives of the metric, so the question is what scalars can be built under that restriction.
A scalar has no free indices, so every index must be contracted. With one factor of the Riemann tensor there is essentially one way to do that. Contract the first and third indices to get the Ricci tensor, and Chapter 3.4 §6.1 showed that the other choices give zero or the same thing with a sign. Then contract the remaining two indices with the inverse metric to get .
With no factor of Riemann there is only a constant. With two factors of Riemann, meaning or or , one has products of second derivatives, and by the same counting as §3.3 those are excluded by (C3) as stated. So the complete list of admissible scalars is a constant and , and the action is
with and constants, and with whatever action describes the matter. The factor in front of is pure convention. It is chosen so that the symbol matches the of (3.6.12), and we shall see it do so.
The volume element is Chapter 3.5 §6.2's. It is there because without it the integral would depend on the chart, which by (C1) is not allowed. Chapter 0.6 §8.3 said this six chapters ago and named this action as the reason.
Set until §6, to keep the derivation clean. The constant is restored in three lines at the end of §4.4.
4.2 · Piece 1: the variation of the volume element
Piece 1 asks how the volume element responds to a change in the metric. The claim is
Everything needed for it is already in hand. Chapter 3.5 §6.3 derived Jacobi's formula (3.5.47), which gives the change in a determinant produced by a change in its entries, and it derived the specialisation to the metric, (3.5.48). The only additional ingredient is the relation between varying and varying , which comes from differentiating the statement that the two are inverse. Grind box A does it in five lines.
Grind box A — , from Jacobi's formula
Line 1. The two variations are related. The metric and its inverse satisfy . The right-hand side is a constant array, so varying both sides gives
Multiply by and sum over , which turns into in the first term:
Line 2. Contract it. Multiply the previous display by and sum. On the right, and then , so
The two contracted variations are negatives of each other. This is the step that is easiest to get backwards, and getting it backwards costs a sign in the field equations.
Line 3. Jacobi's formula. Chapter 3.5's (3.5.48) applied to the metric reads , where . Using line 2 to swap which variation appears,
Line 4. The square root. . Substituting line 3,
Line 5. Tidy the determinant. Since is negative, , so and therefore . Hence
which is (3.6.15). (This was checked symbolically against a direct differentiation of for a four-dimensional metric family with off-diagonal entries. The two agreed exactly, as did the equivalent form .)
Now combine that result with piece 2, which was already in the shape a variational calculation wants. Together the two of them produce this:
The Einstein tensor has appeared without being sought. In §3 the factor of came from the contracted Bianchi identity. Here it comes from the derivative of a determinant. The two routes have nothing in common, and they produce the same coefficient, which is the strongest kind of check a derivation can have.
4.3 · Piece 3: the Palatini identity, derived
What remains is , which is piece 3. The strategy is to show that it is a total derivative, and it takes two steps.
Step one is the Palatini identity,
and the reason this identity can be derived rather than quoted is a fact worth stating on its own. Although the connection is not a tensor, the difference of two connections is.
Chapter 3.3 §5.2 derived the transformation law (3.3.26) and pointed out that its offending inhomogeneous term depends only on the change of chart, not on the metric. So take the difference of two connections belonging to two nearby metrics. That term cancels, and what is left transforms as a tensor. A variation is exactly such a difference. Grind box B uses that fact together with Chapter 3.4 §5.3's locally inertial coordinates.
Grind box B — the Palatini identity, in four lines
Line 1. The definition. Contracting Chapter 3.4's (3.4.13) on the first and third indices gives the Ricci tensor written out in terms of the connection:
Line 2. Vary it. The first two terms give . The quadratic terms give four contributions by the product rule, and every one of them carries an undifferentiated as a factor.
Line 3. Evaluate at a point in locally inertial coordinates. Chapter 3.4 §5.3 constructed, around any chosen point , a chart in which while . In that chart, at that point, all four quadratic contributions vanish, and
Note what is not being claimed: is not zero at , only is, which is exactly the point that made Chapter 3.4 §7.1's proof of the Bianchi identity work.
Line 4. Promote to covariant derivatives, and then to every chart. Since is a tensor, its covariant derivative is defined, and at in this chart every correction term in that covariant derivative carries a factor of . So the partial derivatives above may be replaced by covariant ones without changing anything, giving (3.6.17) at in this chart. But (3.6.17) is an equation between tensors, so by Chapter 2.4 §6 it holds in every chart. And the point was arbitrary.
(This was checked symbolically. For a four-dimensional metric family with off-diagonal entries, all sixteen components of the two sides of (3.6.17) agreed exactly at randomly chosen points.)
Step two contracts the Palatini identity with the inverse metric and recognises a divergence. The move that makes it work is metric compatibility. Since passes through untouched, it can be taken inside the derivative, and then the whole expression is the divergence of something. Grind box C does the index bookkeeping.
Grind box C — the contraction, and the vector whose divergence it is
Contract (3.6.17) with :
Take the metric inside. By metric compatibility (Chapter 3.3 §7.1), , so for anything . Applying that to both terms:
Give the two terms the same derivative index. The first differentiates with respect to and the second with respect to . Both are summed, so both names are private. Rename in the second term, which forces its other to become as well, and then rename its internal dummy and its to avoid a clash. The result is that both terms are of something, and
where we have used the symmetry of in its two lower indices to write as . That symmetry is inherited from the symmetry of itself, Chapter 3.3 §7.2.
Then the volume factor. Chapter 3.5 §6.4's identity (3.5.51) says , an ordinary partial derivative of an ordinary product. So piece 3 of (3.6.13) is exactly
(The whole identity was checked symbolically. For four-dimensional and three-dimensional metric families, computed by direct differentiation agreed exactly, at random points, with using the displayed above.)
4.4 · Assembling, and the field equations
All three pieces are now in hand, so put them together. Using (3.6.16) for pieces 1 and 2, and grind box C for piece 3,
The second term is an ordinary divergence, so by Chapter 3.5 §5's theorem it equals an integral over the boundary of the region. Adopt the same boundary condition Chapter 1.2 §3.3 adopted for the pendulum: the variation is required to vanish, along with its first derivatives, outside some bounded region. Then vanishes on the boundary and the term contributes nothing. Section 4.5 asks what that condition costs, because it is not free.
The matter side. Define the energy–momentum tensor by how the matter action responds to a change in the geometry:
This is a definition, so by itself it proves nothing. What makes it legitimate is that it reproduces the tensor Chapter 2.6 built by an entirely different argument. Worked example 1 checks exactly that, by feeding Chapter 2.6's electromagnetic Lagrangian (2.6.72) into (3.6.19) and recovering (2.6.80), sign included. The definition also delivers a symmetric tensor automatically, because is symmetric. Chapter 2.6 had to arrange that by hand with an improvement term.
Stationarity. Demanding for every allowed , and applying the fundamental lemma of the calculus of variations (Chapter 1.2 §3.4) to strip off the arbitrary variation,
That is (3.6.12) with , obtained a second time. Note that is still undetermined. The action principle fixes the form of the law and cannot fix its coupling, exactly as Chapter 2.6 §9 could not fix from the shape of the electromagnetic Lagrangian alone.
Restoring , in three lines. Put the constant back in (3.6.14). Its variation involves only , so by (3.6.15),
which adds to the left of (3.6.20), and hence to the left of the field equation. That reproduces (3.6.12) in full, and it justifies the in the action. The cosmological term is the constant that can be added to any Lagrangian. Section 6 takes that sentence seriously.
4.5 · The boundary term is not free
Two things were quietly discarded above and both deserve to be named.
Why the equations are second order at all. Constraint (C3) said the field equations must not contain derivatives of beyond the second, and Chapter 1.2 §8 established that an action built from a field and its first derivatives delivers exactly that. But contains second derivatives of the metric, so (3.6.14) is not of that form, and one would naively expect fourth-order equations.
It does not happen, and grind box C says why. Every second derivative of the metric in sits inside , which is a total derivative and drops out of the equations of motion entirely. The Einstein–Hilbert action is second order in disguise, and that is a special property of rather than a general feature of curvature scalars. It is precisely why adding to the action does give fourth-order equations.
What the discarded term costs. Because the second derivatives live in the boundary term, making vanish required more than fixing the metric on the boundary. It required fixing the metric's normal derivative there too, and that is one condition too many. The standard fix is to add to (3.6.14) a boundary integral whose variation cancels the unwanted piece. What is left is a variational principle in which only itself is held fixed on the boundary.
The required addition is , where is the metric induced on the boundary and is the trace of its extrinsic curvature. That trace measures the rate at which the boundary's normal direction turns as one moves along it. We are quoting the form and not deriving it, because extrinsic curvature is machinery this book has deliberately avoided. Chapter 3.2 §1 forbade any reference to an ambient space, and extrinsic curvature is precisely what a surface looks like from outside.
Why it is worth flagging rather than ignoring. For the problems of Chapters 3.7 and 3.8 the boundary term contributes nothing and can be forgotten. It stops being ignorable exactly when the value of the action itself is the thing one wants, rather than the equations it produces. The outstanding case is the thermodynamics of black holes, where the numerical value of the gravitational action supplies the entropy. Chapter 3.9 §7 quotes that entropy, in appropriate units, and says the derivation is beyond this book. This boundary term is one of the places where the derivation happens, and Chapter 7.9 returns to it.
The same law arrives again by a road sharing no step with the first. Attach a single number to each possible shape of spacetime, namely the total curvature added up with the correct volume weighting, and ask which shape makes that number stationary. There is essentially one number available to attach, because insisting on no more than two rates of change leaves exactly one scalar and a constant.
Varying it splits into three parts. The first asks how the volume weighting responds, and the rule for differentiating a determinant answers it. The second needs no work. The third looks worst and is a total derivative, contributing nothing inside the region and everything on its edge. Combining the first two, the object cornered in the previous section appears unbidden, and the notorious factor of one half arrives from the derivative of a determinant rather than from a differentiated identity. Two unrelated routes, one coefficient.
The discarded edge term is not tidy-up. It is where the second rates of change hide, which is why the equations come out second order though the number attached is not. Setting it aside also demands more of the boundary than one is entitled to, and repairing that costs an extra term whose value matters in one place above all: the entropy of a black hole is what the gravitational number evaluates to, and that story waits for the last part.
5 · The constant, fixed by demanding that apples fall
This is the section where the whole construction is put on trial. Everything so far has been structural. The equation has the form it has because of conservation, because of symmetry, and because of the number of derivatives allowed. None of that touches the world. Here it does.
The route has five steps and each one is small, so here they are before we start. We take a weak, static gravitational field and slow-moving matter. We rearrange the field equation into a form where the Ricci tensor stands alone. We evaluate its component on the left from the metric, and then the same component on the right from the matter. Finally we compare the result with Poisson's equation. Out comes .
Everything below rests on the weak-field metric component. In this book's signature it is
with a plus sign, and with the Newtonian potential, which is negative near a mass. This is not a convention adopted here for convenience. Chapter 3.1 §6.5 derived it, as equation (3.1.39), from the gravitational redshift of a signal climbing inside an accelerating cabin, using special relativity and the equivalence principle and no general relativity whatever.
The sense of it can be checked directly. A clock lower in a well runs slow, so the coefficient relating proper time to coordinate time must be smaller there, and is more negative there. The signs agree.
Books using write . That is the same physics in the other signature, and copying it into this chapter would flip the sign of the coupling constant derived below. This is the second of the two sign traps announced at the start of the chapter. The first was the Riemann convention of Chapter 3.4, and the third is in §6.
5.1 · Step 1 — trace-reverse the field equation
The field equation (3.6.20) has the Ricci scalar buried inside , and that is inconvenient. We want the scalar out where we can see it, so contract the equation with . Using and ,
with . That is an expression for in terms of the matter, so we can now eliminate the scalar entirely. Substitute back into , and the result is
This is the trace-reversed form, and it is exactly equivalent to (3.6.20). Contract it and the previous step runs backwards. It is the form to use whenever the Ricci tensor is easier to compute than the Einstein tensor, which is most of the time, and Chapter 3.7 uses it from its first line.
One consequence comes free and matters a great deal. In vacuum, where , the field equations reduce to . Chapter 3.4 §6.3 already showed that this does not mean flat, because the Weyl part of the curvature survives. Chapter 3.7 solves .
5.2 · Step 2 — the assumptions, each named as it is made
Assumption 1, weak field. We write with every , and we keep only terms linear in . In particular .
Assumption 2, static. Nothing depends on time, so .
Assumption 3, slow matter. The matter's four-velocity is , so in (3.6.4) we keep the leading term only. Pressure is taken to be small compared with rest-energy density, . Section 5.5 removes this one.
5.3 · Step 3 — the left-hand side,
We want the component of the left-hand side, so write out the Ricci tensor in terms of the connection, exactly as grind box B did, and then set :
Now take those four terms one at a time. Three of them turn out to be nothing.
The two quadratic terms are second order. Each is built from one derivative of the metric, so by Assumption 1 each one is first order in . A product of two of them is second order, and it is dropped.
The second term vanishes. It carries , and by Assumption 2 nothing depends on .
The first term is the whole answer. Split the sum over into its time and space parts. The time part is , which vanishes by Assumption 2 again. What is left is the spatial sum, and Chapter 3.3 §8.3 has already computed the connection coefficient that sum needs. It is equation (3.3.71), obtained there from exactly these three assumptions:
Now put that coefficient into the surviving term and carry out the remaining differentiation, which finishes this side of the equation:
where the last step recognises the Laplacian of Chapter 0.7 §7.5. At this order indices are raised and lowered with , so the distinction between and costs only a sign that appears twice and cancels.
Equation (3.6.26) is Chapter 3.4's (3.4.33). That chapter got it by a completely different argument. It derived the geodesic deviation equation, set that beside Chapter 3.1's Newtonian tidal equation, read off , and then took the trace. Here the same result comes straight from the definition of the Ricci tensor and the Christoffel symbols of Chapter 3.3.
The agreement is not a coincidence, and it is not circular either. The two routes share only the weak-field metric component, which Chapter 3.1 obtained from the redshift with no general relativity in it at all.
(This was confirmed symbolically as well. Computing from the full metric , and expanding to first order gives for every value of . So the spatial part of the metric, which none of these arguments has pinned down, does not affect this component at this order.)
5.4 · Step 4 — the right-hand side, and the comparison
Now for the right-hand side. Use the trace-reversed form (3.6.23) with . By Assumption 3 the matter is dust at rest, so from (3.6.3) and ,
Those are the two ingredients the right-hand side asks for, so assemble the bracket that appears in (3.6.23):
where the last step drops , because it is smaller than by the factor , and Assumption 1 declared that negligible.
Let's look at what that line is actually saying. Note the arithmetic: . The full energy density enters, the trace-reversal removes half of it, and that surviving factor of one half is exactly what turns the of Poisson's equation into the of the final answer.
Both sides of the equation are now in hand, so set them equal. That means (3.6.26) on the left and times (3.6.28) on the right:
That is a Poisson equation with an undetermined coefficient in it, so compare it against the real Poisson equation, which Chapter 0.7 §7.5 derived from Newton's inverse-square law and the divergence theorem:
The two agree for every density only if the coefficients of agree, which requires . Solving that for ,
Put that value of back into the boxed result of §3, and there, at last, are the field equations of general relativity:
And by (3.6.20), , so the action is fixed too:
Chapter 1.2 §8.1 previewed this action as , quoted forward and flagged there as quoted. Equation (3.6.33) has the opposite overall sign, and the discrepancy is a convention rather than a disagreement.
Here is why. Replace everywhere by , which is what changing signature does. The Christoffel symbols are unchanged, because the formula (3.3.50) contains one inverse metric and one derivative of the metric, and the two sign changes cancel. Hence and are unchanged as well.
But flips sign, since the inverse metric does. In the second term picks up two sign flips, which cancel, so is unchanged. In there is only one flip, so that quantity changes sign. The field equations are signature-independent. The Lagrangian producing them is not.
Landau and Lifshitz, who use as this book does, write the gravitational action with the minus sign, in agreement with (3.6.33). Books using write it with the plus, and Chapter 1.2's table quoted that form. So check the signature before importing a gravitational Lagrangian from anywhere, exactly as Chapter 3.4 said to check both Riemann conventions before importing a curvature formula.
5.5 · What the number means, and one generalisation
The size of . Numerically . The reciprocal is more telling: .
Read (3.6.32) as (curvature) (stress), and becomes the stiffness of spacetime, meaning the stress needed to produce unit curvature. It is about newtons. That is why an object as massive as the Earth bends spacetime by only the that Chapter 3.4 §4.5 computed, and why gravity looked for three centuries like the weakest thing in physics rather than like geometry.
Restoring the pressure. Assumption 3 dropped , and we can now afford to put it back. Use the perfect fluid (3.6.4) instead of dust, and take its trace from (3.6.5), which is . Then, still at rest and to leading order, the bracket becomes
Nothing else in the calculation changes, so run steps 3 and 4 again with this bracket in place of . What comes out is
This is the promise of §1.4 kept. Pressure gravitates, and it does so three times over. Squeeze a gas without adding anything to it and the source of its gravitational field grows.
The effect is invisible in ordinary matter. There is smaller than by the square of the ratio of thermal speeds to , which for the Sun's centre is about . But it is not invisible everywhere. In a neutron star the pressure term is a sizeable fraction of the source, and it works against the star. Adding pressure to resist collapse also adds to the gravity doing the collapsing. That is part of the reason a sufficiently massive star has nothing left that can hold it up.
And for radiation, where (3.6.6) gives , the bracket becomes . A gas of light gravitates twice as strongly as the same energy density in cold dust. Chapter 3.9 needs that.
The same expression also contains the seed of §6, and it takes one line to see. Suppose a substance could have . Then the bracket would be negative, would have the wrong sign, and the gravity of that substance would push rather than pull. Nothing in (3.6.35) forbids it.
Everything so far has been shape without scale, since the constant tying geometry to matter was carried along unfilled, and filling it in is where the construction becomes a theory of the world rather than bookkeeping. Take a weak, unchanging field and slow matter, and rearrange the law so the boiled-down curvature stands alone on the left. Its timekeeping entry works out to be the ordinary second-derivative operator applied to the potential, divided by the square of the speed of light, which the curvature chapter had already reached by a different argument about drifting dust.
The matter side gives the energy density less half its own trace, which for slow cold matter is half the energy density. Comparing with the three-centuries-old equation relating potential to density fixes it at eight pi times the gravitational constant over the fourth power of the speed of light. Its reciprocal is more eloquent: about ten to the forty-two newtons, the stress needed to bend spacetime usefully, and the reason gravity looks like the feeblest thing in physics.
One term survives that Newton had no way to see. Because the source is the whole bookkeeping object and not merely its topmost entry, pressure appears in it, tripled. Compressing a gas increases the gravity it makes, light pulls twice as hard as cold matter of the same energy, and a substance with sufficiently negative pressure would push rather than pull.
6 · The one term the argument cannot exclude
Now return to . Section 3 found it and could say nothing about it, because constraint (C4) was silent: automatically. Section 4 found it again from the other direction, as the constant that may be added to any Lagrangian. This section asks what it is.
6.1 · It is a substance with negative pressure
The way to find out what a term means is usually to move it to the other side and read it as matter. So move it from the geometry side of (3.6.32), which costs a sign:
Now read that bracket as a total energy–momentum tensor. Whatever turns out to be, its contribution to the source is
We want to know what substance would produce that, so compare it with the perfect fluid (3.6.4), with both indices lowered: . For (3.6.37) to have this form, two conditions must hold, and we take them one at a time. The term is absent from (3.6.37), so its coefficient must vanish:
That is the first condition. Matching the remaining term, , gives , and then (3.6.38) converts that into a density:
Three things follow immediately, and not one of them was put in by hand.
(i) The equation of state is fixed, not chosen. A cosmological term is a fluid with , which is in the figure's notation. It is the far-left point of that plot. By (3.6.35) its source term is . So with a positive energy density it repels, and it repels twice as hard as the same energy density in dust attracts.
(ii) It is the only fluid that looks the same to everybody. Equation (3.6.37) is proportional to , and the metric is the one tensor whose components are the same in every local inertial frame. Every other fluid singles out a rest frame, through . This one does not. That is exactly what one would demand of the energy of empty space, since there is no such thing as moving relative to the vacuum.
(iii) Its density is constant. That is not an assumption. It follows because is a constant and is a constant. Expand the universe and ordinary matter thins out while this does not, which is why a term negligible in the early universe can come to dominate the late one. Chapter 3.9 works that out.
Equation (3.6.39) carries a minus sign. With the term written as on the left, a positive vacuum energy density corresponds to a negative . That is a consequence of our signature, and it would be a permanent nuisance, so we do what everyone does and absorb it into the definition. Write , so that
With this definition means positive vacuum energy and gravitational repulsion, and has dimensions of one over length squared, as a curvature should.
Books using write and mean the same . The flip is the one traced in §5.4. Under the tensor is unchanged, is unchanged, and is not. Since is a measured number rather than a convention, it is the placement of the sign that has to move. Chapter 3.4's preview wrote the term as "for constants and ". That is this same term with the naming still open, and it is settled here.
6.2 · Why nothing forbids it
The quickest way to see that nothing rules the term out is to go back through the chapter and check every constraint in §2 against in turn.
| Constraint | Does satisfy it? |
|---|---|
| (C1) tensorial | yes, since is a tensor |
| (C2) symmetric, rank 2 | yes |
| (C3) at most two derivatives of | yes, it has none |
| (C4) identically divergence-free | yes, by metric compatibility, |
| (C5) Newtonian limit | yes, provided is small enough, as below |
Every row passes. And §4 makes the point from the other side. In the action, the cosmological term is a constant added to the Lagrangian, and nothing forbids adding a constant to a Lagrangian.
In ordinary mechanics, adding a constant to the Lagrangian changes nothing, because the extra contribution to the action is itself a fixed number, and a fixed number has no variation. Here it is not fixed. The constant is multiplied by , which depends on the very thing being varied. So gravity, alone among the theories in this book, notices a constant in its Lagrangian. What it notices is this term.
What (C5) says is a bound rather than a prohibition. Redo §5 keeping . The extra term contributes to the left of the field equation, so (3.6.29) becomes . For that to be indistinguishable from Poisson's equation in the Solar System, must be tiny compared with there, which it is, by an enormous margin. The constant is not excluded. It is bounded.
6.3 · What it is measured to be, and the problem that leaves
Observations of distant supernovae and of the cosmic microwave background give , corresponding by (3.6.39) to . That is about four hydrogen atoms per cubic metre, and roughly of the total energy density of the present universe. These are quoted as measurements. Chapter 3.9 derives what a universe with this term does. It does not derive the number.
Two things about that value are worth stating now. It is fantastically small in any natural unit. Expressed as a length, is about metres, which is comparable to the size of the observable universe and about times the Planck length.
And it is not zero, which is worse. A vanishing constant might be explained by a symmetry forbidding it. A tiny non-zero one has to be explained by something that gets the number right. Nothing in this book explains it, and nothing anywhere else does either. Chapter 7.9 is where the failure is accounted for honestly.
One historical note, because it is usually told wrongly. Einstein introduced this term in 1917 to permit a static universe, since with the equations have no static solution containing matter. The balance he found is unstable, as Chapter 3.9 §3 shows by integrating the equations. Once the expansion of the universe was established, the motivation evaporated.
What did not evaporate is the term. It was never inserted. It was always allowed. Removing it requires an extra assumption, and that assumption turned out to be false.
One extra piece slipped through the cornering untouched, namely the distance rule itself multiplied by a constant. It slipped through because the requirement doing all the work, that the geometry side have no divergence, is automatically satisfied by the distance rule. Seen from the action, the same piece is a constant added to the quantity being extremised, and no principle in this book forbids that.
Moved across to the matter side, the term describes a substance, and its properties are forced rather than assumed. Its pressure must be exactly the negative of its energy density, which is the unique choice that looks identical to every observer, as the energy of empty space ought to. Because the source of gravity contains three times the pressure, such a substance has a negative source and therefore pushes rather than pulls. And its density cannot dilute as the universe grows, since it is built from constants, so a contribution negligible early on can come to dominate later.
The observed value is small beyond ordinary description, and small is worse than zero. A quantity forced to vanish can be explained by a principle forbidding it; a quantity very small and not zero needs an explanation producing the actual number, and none exists. This is the single most embarrassing number in physics and the last part of the book returns to it without pretending to fix it.
7 · Ten equations, four identities, and why that is exactly right
The field equations are written. Before spending them, let's count them, because the count reveals something the equations do not say out loud.
7.1 · The count
Unknowns. The metric is symmetric with two indices in four dimensions, so it has independent components. That is ten functions of four variables.
Equations. Both sides of (3.6.32) are symmetric rank-2 tensors, so it is also ten equations. Ten equations for ten unknowns looks like a well-posed problem, and looks determined.
But four of the ten are not independent. Chapter 3.4 §7.3 proved identically, and as well, so the divergence of the left-hand side of (3.6.32) is zero whatever the metric. That is four differential relations among the ten equations, one for each value of . Only six of the ten carry independent information about how the metric evolves.
And that shortfall is exactly right. A solution can be relabelled by any smooth change of chart , which is four arbitrary functions. The relabelled metric describes the same geometry, so it must also be a solution.
So the equations cannot determine all ten components. If they did, they would forbid the relabelling, and by (C1) the relabelling is not physical. Four functions' worth of freedom has to remain undetermined, and four is exactly what the Bianchi identities leave undetermined.
This is not a peculiarity of gravity. Chapter 2.6 §3.1 wrote Maxwell's sourced equations as , which is four equations for the four components of . Taking of both sides gives automatically, since is antisymmetric and the two derivatives commute. So one of the four equations is not independent. Three carry information, and the missing one matches the one function of gauge freedom, , that Chapter 3.5 §10.3 showed is always available.
The dictionary is exact: one identity and one gauge function in electromagnetism, four identities and four coordinate functions in gravity. In both cases the identity is forced by the structure of the left-hand side, and in both cases it exists to make room for a freedom that no measurement can see. Chapter 6.3 makes this correspondence into a principle.
7.2 · Which four are the constraints
Let's look at where the four identities bite. Write out . It contains and , so it expresses the time derivative of in terms of quantities involving one fewer time derivative. Since contains second derivatives of the metric, the consequence is that the four equations contain no second time derivatives at all.
So they are not evolution equations. They are conditions that the initial data must satisfy, and the identities then guarantee that if the conditions hold at one time they hold at every later time.
Once again electromagnetism did this first. In Chapter 2.6, Gauss's law contains no time derivative of . It is a constraint on initial data, preserved by the other equations because charge is conserved. Gravity has six evolution equations and four constraints. Electromagnetism has three and one.
7.3 · The equations are nonlinear, and that is physics
One last structural fact, and it is the one that makes Part III hard and Chapter 7.1 harder. is built from and its derivatives, and contains . So the Einstein tensor contains the inverse metric multiplied by derivatives of the metric, over and over. It is nonlinear in , badly so.
What the nonlinearity means physically. Superposition fails, so the field of two masses is not the sum of their separate fields. The reason is not technical. Gravitational fields carry energy, and by §1 everything carrying energy sources gravity, so gravity gravitates. There is no way to write a linear theory with that property, because linearity means the source is independent of the field.
Contrast electromagnetism, which is linear. Chapter 2.6's is linear in , and two solutions superpose. The reason is that the electromagnetic field is not itself charged. Chapter 6.4 builds a theory where the field does carry the charge it responds to, and finds equations that look startlingly like these.
Two consequences are worth filing. Exact solutions are rare and precious, which is why Chapter 3.7's Schwarzschild solution is a landmark rather than an exercise. And the standard technique of physics, which is to expand about a simple solution and keep the first correction, becomes the only technique available. That is why Chapter 3.1's warning that nearly everything is an approximation applies here with particular force.
Counting is worth doing before solving. The unknown is a symmetric array with ten entries and the law supplies ten equations, a matched set until one notices that the identity driving the chapter makes four of them redundant. Only six carry information about how the geometry develops.
That shortfall is not a defect and could not have been otherwise. Any solution can be repainted with different coordinate labels, which takes four arbitrary functions and changes nothing measurable, so the law is obliged to leave four functions' worth undetermined. The identity exists to make room for the freedom. The four leftover equations are not useless; they are conditions the starting data must satisfy, automatically preserved thereafter, exactly as the law relating electric field to charge constrains starting data rather than governing its development.
The last observation is the expensive one. The geometry side is nonlinear in the geometry, so the field of two masses is not the field of one added to the field of the other. There is no way to avoid this, because everything carrying energy is a source, and the gravitational field carries energy, so gravity is a source of itself. Electromagnetism escapes because its field carries no charge. That difference is why exact solutions are rare, why approximation is the normal state of affairs, and why the last part of the book finds gravity so much harder than everything else.
8 · Worked examples
Feed Chapter 2.6's electromagnetic Lagrangian into (3.6.19) and check that the result is (2.6.80), sign included. Without this check, the definition (3.6.19) would be no more than an assertion.
The action. Chapter 2.6 §9 gave the electromagnetic Lagrangian density as (2.6.72), and its source-free part is . Written out on a manifold, the action is therefore
The indices have been written out explicitly because the metric dependence is the whole point. Note that carries no metric dependence at all, by Chapter 3.5 §2.2. It is , and the exterior derivative needs no connection. So only the two inverse metrics and the volume element vary.
Vary the contraction. Two inverse metrics, so the product rule gives two terms:
The two terms are equal after relabelling dummies and using the antisymmetry of twice, which supplies two minus signs that cancel.
Vary the volume element with (3.6.15). Collecting:
Read off by comparing with (3.6.19), which says the bracket times equals divided by the prefactor:
Compare with Chapter 2.6. Equation (2.6.80) reads . Lower both free indices, and then handle the first term: . The first step uses the antisymmetry of , and the second raises and lowers the summed index, which is free of charge. So Chapter 2.6's tensor is . Identical. ✓
The sign, checked against a physical number. For a pure electric field along , , so , and by Chapter 2.6 §7.1 . Substituting, , which is positive, and which is Chapter 2.6's (2.6.82). So the definition (3.6.19) has the sign that makes energy density positive in this book's signature. (Books using define with the opposite sign, for the same reason the action carries the opposite sign there.)
Use the trace of the field equations to compute the Ricci scalar inside a body of uniform density, and evaluate it for air, water, the Earth and a neutron star.
The formula. From (3.6.22) we have , and for slow cold matter (3.6.3) gives . Putting the two together,
Note that this is a local statement. The Ricci scalar at a point depends only on the density at that point, not on how much matter is elsewhere. The mass of the Earth as a whole does not appear. Everything non-local about gravity lives in the Weyl part of the curvature that Chapter 3.4 §6.2 named and set aside, which is exactly why vacuum is not flat.
Numbers. Curvature has dimensions of one over length squared, so the quantity worth putting beside it is the radius of curvature .
| Material | |||
|---|---|---|---|
| Air at sea level | |||
| Water | |||
| Earth, mean | |||
| Neutron star, core |
What to take from the table. Inside the Earth, the radius of curvature of spacetime is about two-thirds of an astronomical unit. The geometry departs from flatness on a scale comparable to the Earth's distance from the Sun. That is why nobody noticed for three hundred years, and it is the same conclusion Chapter 3.4 §4.5 reached from the tidal side with the closely similar number .
The last row is the interesting one. For a neutron star the radius of curvature drops to about ten kilometres, which is the size of the object itself. When the radius of curvature becomes comparable to the body producing it, no expansion in is available, and the full nonlinear equations are needed. That is the boundary of the regime §5 assumed, stated as a length rather than as a small parameter.
9 · Your turn
Problem 1 — three forms of the same equation
(a) Derive the trace-reversed form (3.6.23) in the presence of , showing that . (b) Deduce the vacuum equations with and without , and give in each case. (c) Chapter 3.7 solves the case , . Say in one sentence why that is a sensible thing to do for the space outside the Sun even though in the universe. (d) In two dimensions identically (Chapter 3.4 Problem 2). What does that say about the field equations there?
Solution
(a) Contract with : , so . Substitute into , and collect: the terms give , leaving the stated result.
(b) With and we get and hence . With and we get , and contracting gives . Note that the second is not flat and not even Ricci-flat. Empty space with a cosmological constant is curved, which is the whole of Chapter 3.9's late-time behaviour.
(c) Because , while the curvature scales relevant to planetary orbits are set by , which at Mercury's orbit is about . That is larger by twenty-two orders of magnitude. Neglecting inside the Solar System is not an approximation anyone can detect.
(d) The left-hand side vanishes identically, so the equations read . That is not a field equation but a prohibition, since it says there can be no matter. Gravity has no dynamics in two dimensions. Chapter 7.2, which works entirely in two dimensions, is therefore not doing gravity, and that turns out to be a feature.
Problem 2 — geometry tells matter how to move, and it is not a separate law
The field equations imply , since the left-hand side is identically divergence-free. Take dust, . (a) Expand by the product rule. (b) Contract the result with and use to show . (c) Substitute back and deduce the geodesic equation. (d) Say what has just been proved about the logical structure of general relativity.
Solution
(a) .
(b) Contract with . The first term gives . The second gives , where the metric was taken through the derivative to combine the two factors, and the middle step is the product rule read backwards. So , which is the continuity equation for the dust, saying that no particles are created or destroyed.
(c) With the first term of (a) gone, , and dividing by where it is non-zero gives . That is the geodesic equation in the form Chapter 3.3 §8.1 derived it: the tangent parallel-transports itself.
(d) The equation of motion for matter is a consequence of the field equations, not an additional postulate. In Newtonian gravity, "the field satisfies Poisson's equation" and "a particle accelerates according to " are two independent laws. Here the second follows from the first, through the Bianchi identity. Geometry does not merely tell matter how to move. It is not permitted to say anything else.
Problem 3 — the spatial metric, and a grievance from Chapter 3.1 settled
Chapter 3.1 §6.5 obtained from the redshift and then complained: "this argument has said absolutely nothing about the spatial components of the geometry." Settle it. Take the static weak-field ansatz , with far away, and a dust source at rest. (a) Write down the components of (3.6.23) for and for . (b) Given that to first order , deduce . (c) Write the metric. (d) Say which later result this is needed for.
Solution
(a) For dust at rest, and , so to leading order, using . So the off-diagonal components must vanish and the three diagonal ones must be equal.
(b) Write . Off-diagonal, : the stated expression gives . Diagonal: the three components are equal to each other, so . Call the common value , so that . The diagonal equation itself reads , while gives for the same right-hand side. Subtracting, , that is , so . Every second derivative of therefore vanishes, which makes a linear function of position. Requiring both potentials to vanish at infinity then leaves only . Hence .
(c) Note the signs: time is stretched by and space by , so with near a mass, clocks run slow and rulers are shortened.
(d) Chapter 3.1 §7.3 computed the deflection of light past the Sun using only the equivalence principle and got exactly half the observed value, promising that the missing half would be identified. It is the spatial part of (c). A light ray is affected by the geometry of space as much as by the geometry of time, and the cabin argument saw only the latter. Chapter 3.8 §§3–4 do the integral and collect the factor of two.
Problem 4 — the vacuum as a fluid, and Einstein's static universe
(a) Using (3.6.39) and , compute in kilograms per cubic metre and in hydrogen atoms per cubic metre. (b) At what distance from the Sun does the repulsive source match the Sun's own averaged over a sphere of that radius? Put another way, where does start to matter? (c) Show, using (3.6.35) with restored, that a static uniform universe of density requires . (d) Argue that this balance is unstable.
Solution
(a) . Dividing by the hydrogen mass gives about atoms per cubic metre.
(b) The Sun's mass spread over a sphere of radius has mean density . Setting that equal to and solving, . With this is about , or roughly parsecs. So is irrelevant on any scale smaller than a substantial piece of a galaxy, and dominant on scales much larger. That is what §6.2's bound said in other words.
(c) With restored, §6.2 gave . A static uniform universe has no preferred point and hence a potential with no second derivative anywhere, so , giving .
(d) Compress the universe slightly. Ordinary matter's density rises, since the same matter now occupies less volume, while is a constant and does not. So the source becomes positive, gravity wins, and the compression accelerates. Expand slightly and the reverse happens. The equilibrium is a pencil balanced on its point. Chapter 3.9 §3 shows the same thing by integrating the equations rather than arguing from them.
You have the law of gravity, and it was not guessed. Matter's energy, momentum, pressure and stress are packaged in one symmetric tensor whose divergence vanishes. Anything set equal to that tensor must be divergence-free identically, for every metric. The only symmetric rank-2 objects available from a metric and its first two derivatives are , and . Imposing the divergence condition on the most general combination of those three fixes the ratio of the first two and leaves the third free. That is (3.6.12), and ⚑ Lovelock's theorem says there is nothing else in four dimensions.
The same equation, from an action. There is essentially one scalar that can be built from a metric with at most two derivatives, so the action writes itself. Varying it splits into three pieces, and the Einstein tensor appears with its factor of coming this time from the derivative of a determinant instead of from a contracted Bianchi identity. The third piece is a total derivative, which is why the equations are second order even though the action is not. The term thrown away is the one ⚑ Gibbons–Hawking–York repairs, which Chapter 7.9 needs when the value of the action becomes an entropy.
The constant, fixed. Weak field, static, slow matter. works out to using , with the plus sign belonging to this book's signature and derived in Chapter 3.1 from the redshift with no general relativity in it. The trace-reversed source is , and matching against Poisson's equation gives . Restoring the pressure gives , so pressure gravitates, radiation gravitates twice as hard as dust of the same energy density, and a substance with would push.
And the term nothing excludes. passes every constraint in the chapter, it is the constant one may add to any Lagrangian, and it behaves as a fluid with . That is the unique equation of state which looks the same to every observer, and the one that does not dilute as space grows. It is measured to be about , tiny and not zero, and that is a problem nobody has solved.
Where this gets spent. Chapter 3.7 sets and solves outside a spherical mass, using Chapter 3.5's Killing vectors to make the geodesics tractable, and pays the factor-of-two debt from Chapter 3.1 using the spatial metric of Problem 3. Chapter 3.9 puts a perfect fluid on the right and a homogeneous, isotropic metric on the left, and there the pressure term and the cosmological term both do real work. And Chapter 7.1 asks what happens when the same equations are quantised. At that point the nonlinearity of §7.3, which is gravity gravitating, stops being an inconvenience and becomes the obstruction.