Part III · General Relativity — Chapter 3.8
Light, Redshift, and What a Horizon Is
Collect the factor of two Chapter 3.1 confessed. Then take the horizon apart, and find that the only thing broken was the chart.
Chapter 3.7 solved Einstein's equations. Outside a spherical mass the geometry is fixed completely by , spherical symmetry and one boundary condition. The answer contains a single length that nothing in the field equations knew about until the Newtonian limit put it there. That chapter then spent the solution on matter: two Killing charges, an effective potential, an innermost stable circular orbit at , and Mercury's per century. This chapter spends the same solution on light, and then on the length itself.
Two debts are payable here, and both were written down in advance. The first is Chapter 3.1 §7.3. That section computed the bending of starlight from the equivalence principle alone, obtained exactly half the observed value, and refused to fudge it. It entered the shortfall in what it called the register of debts, with this chapter's §4 named as the place of payment. The second debt is older and larger. Chapter 3.2 insisted that a manifold is not its coordinates, and nothing since has forced you to care. At the metric components misbehave. Whether that is a fact about spacetime or a fact about labels is the question §§6 and 7 exist to settle. It is the most misunderstood question in this part of physics, and you now own every tool needed to answer it.
Here is the route. Section 1 asks what changes when the interval along a path is zero, and finds that light is described by one number where a planet needed two. Section 2 finds a circular orbit for light and a critical aim that separates capture from escape. Section 3 does the deflection integral and produces . Section 4 takes the metric apart and identifies which half of it Chapter 3.1 was missing, then shows why a planet collects one half and a ray collects both. Section 5 derives the gravitational redshift three separate ways, and puts the satellite-navigation arithmetic on one metric instead of two effects glued together. Section 6 asks the horizon question and answers it with three calculations, the last of which is an invariant and therefore cannot be argued with. Section 7 builds a better chart, watches the light cones tip over, and compares the result with the horizon Chapter 2.3 found in flat spacetime. Section 8 says what is genuinely singular, and what general relativity declines to predict.
Conventions. Signature , and . Riemann and Ricci signs are as stated in Chapter 3.4. We keep and explicit throughout, and we write out rather than abbreviating it to a mass. The Schwarzschild metric is used in the form Chapter 3.7 derived it,
and throughout this chapter is written for the repeated factor. Remember what is: Chapter 3.7 §2 defined it as the areal radius, the label for which the sphere carrying it has area . It is not the distance to the centre, and §6 shows that reading it as one is most of what makes the horizon look catastrophic.
Tools you'll need — Chapter 3.7 above all: its §3 solution of , its §5 Killing charges and , and the radial equation of its §6, which is the line this chapter's §1 takes a limit of. Chapter 3.5 §9, where a Killing vector conserved quantity, used three times below. Chapter 2.4 §§5.1 and 6 for the fact that a full contraction is a scalar, which is the instrument §6 leans on. Chapter 3.4 §5 for the symmetries of the Riemann tensor and §5.3 for locally inertial coordinates, and §4 for geodesic deviation. Chapter 3.3 §8.1 for the geodesic equation and the meaning of an affine parameter, and §2, the flat plane in polar coordinates, which is the whole of §6 in miniature. Chapter 3.2 §2 on charts, atlases and transition maps: changing chart is the operation §7 performs. Chapter 3.1 §6 (the cabin, the redshift, and ), §7 (the half-sized deflection and the register of debts) and §5 (how big a freely falling laboratory may be). Chapter 3.6 Problem 3, the spatial part of the weak-field metric, which is the missing half in weak-field clothing. Chapter 2.3 Problem 3, the Rindler horizon, compared in §7.5, and §6.2 on conjugate points. Chapter 2.5 §7 for the Doppler formula. Chapter 1.2 §5.1, the conjugate point, which §3.6 collects. Chapter 0.8 §3.1 for the superposition that makes a linear second-order equation tractable and §6 for particular solutions, both used below, and Chapter 0.2 §5 for the improper integrals of §6.
1 · Null geodesics, and what changes
Let's announce the destination before we start. We take Chapter 3.7's radial equation for a massive particle, replace the timelike normalisation by the null one, and find that the two conserved quantities collapse into a single length. That length is the impact parameter, and everything light does outside a spherical mass is a function of it alone.
1.1 · The normalisation, and the parameter that is no longer available
A massive particle's worldline was parametrised by its proper time. Chapter 3.3 §8.1 showed this is legitimate because is constant along any geodesic, so it may be set to once and the parameter is thereby fixed. For light that route is closed. The interval along a light ray is zero by construction, which is what "moves at " became in Part II. So the normalisation we have to work with reads
and there is no proper time to divide by. Between the emission of a flash at one end of the galaxy and its absorption at the other, the flash's own clock records nothing at all.
Something weaker does survive, and Chapter 3.3 §8.1 proved it as well: the geodesic equation holds for an affine parameter, and affine parameters are related to one another by
Look at what that freedom means here. For a massive particle, fixing the constant in was what removed the freedom . Here the constant is zero, and setting zero to zero fixes nothing, so remains free. That leftover freedom is the whole content of this section.
1.2 · Two charges, one number
Chapter 3.7 §5 applied Chapter 3.5 §9, which says that a Killing vector supplies a quantity conserved along every geodesic. It applied that result to the Schwarzschild metric's two obvious symmetries: the metric components do not depend on , and they do not depend on . The derivation there did not use the timelike normalisation anywhere, so it applies verbatim to a null geodesic, and the two charges are
(The reduction to the equatorial plane is legitimate for the same reason Chapter 3.7 §5 gave. Spherical symmetry lets you rotate any single geodesic into that plane, and a geodesic starting there with stays there, since makes the only source term in the equation vanish.)
Now apply (3.8.2). Rescaling divides both and by . Neither is therefore a property of the ray. You can make either one anything you like by relabelling the parameter, and no measurement can distinguish the labellings. Precisely one combination is immune, and we name it
which is unchanged by the rescaling because numerator and denominator scale together. Section 3.2 will show that is the perpendicular distance from the centre at which the ray would have passed had it gone straight, which is the impact parameter. Note the shape of the argument, because it is one the book has made before: two quantities, individually meaningless because of a normalisation nobody can fix, and one ratio that is physical. That is the same move as the impact parameter in a scattering problem, where the incoming flux and the beam normalisation are conventions and the ratio is what the detector sees.
It is tempting to read as "the photon's energy" and as "its angular momentum", and to conclude that light gains energy falling into a well. Resist it. and as defined in (3.8.3) are constants of the motion attached to a curve, and their numerical values depend on a choice of parameter that nothing physical fixes. Multiply the affine parameter by seven and the "energy" divides by seven with no change to the ray.
What is physical is the frequency a specified observer measures with their own apparatus at their own location, and that is built from together with that observer's four-velocity. Section 5.1 constructs it. The observer-dependence is not a defect but the whole content of the gravitational redshift. Until then, treat and as bookkeeping and as the physics.
1.3 · The radial equation, as a limit
Chapter 3.7 §6 obtained, for a massive particle, from the normalisation together with (3.8.3),
We want the same line for light, and the good news is that we do not have to redo the work. Every step of that derivation used the normalisation exactly once, in the single place where the constant appears on the right. Swapping (3.8.1) in for (3.8.5)'s normalisation therefore amounts to setting that one to zero, and to writing for , which is no longer available:
Now spend the leftover freedom, with the aim of getting and to appear only in the combination . Divide (3.8.6) by and define the rescaled affine parameter , so that . Using from (3.8.4), that gives
One equation, one parameter , and no trace of or separately. The same division turns the angular charge into , which is the form §3 needs.
Let's look at what that last line is actually saying. Compare (3.8.7) with (3.8.5). The Newtonian-looking that sat inside has gone, taking with it every trace of the inverse-square law. Whatever bends light here, it is not the term that makes planets orbit.
Light carries no wristwatch, and that is the first thing to give way. Every path in the last few chapters was labelled by the reading of a clock carried along it, and for a path on which the interval vanishes there is no such reading: between the emission of a flash and its absorption a billion years later, the flash records nothing. What stands in for the missing clock is a parameter chosen so that the equation of motion keeps its shape, and that choice is fixed only up to a change of scale and a shift of origin.
The freedom is not a nuisance. It is why the two conserved quantities the previous chapter extracted from the two directions in which the geometry does not change stop being separately meaningful for light. Rescale the parameter and the one playing the part of energy and the one playing the part of angular momentum rescale together. Only their ratio is untouched, and once the units are arranged that ratio is a length: the perpendicular distance from the centre at which the ray would have passed had nothing deflected it. Anyone who has watched a beam scatter off a target has met the same quantity doing the same job.
So a planet needs two numbers and a ray needs one. Everything light does outside a round mass is a function of that single length.
2 · The photon sphere
Again, the destination first. Equation (3.8.7) has the form (kinetic term) (constant) (potential), so it can be read the way Chapter 3.7 §6 read its timelike counterpart. We locate the maximum of that potential, find a circular orbit for light at , show it is unstable, and extract the aim that separates capture from escape.
2.1 · The null effective potential
We want (3.8.7) in the shape of an energy balance, so move the -dependent piece to the left:
The left-hand term is a square and so is never negative. A ray with impact parameter can therefore only reach radii where . Think of the graph of as a barrier, with the height at which the ray is launched. Everything in this section is read off the shape of one function.
The barrier's top is what decides who gets through, so differentiate and look for it:
The bracket vanishes at exactly one radius, and rises for smaller and falls for larger, so that radius is a maximum:
Now let's check that this really is a circular orbit. A circular orbit needs and . Differentiating (3.8.8) once with respect to gives , so the second condition is . There is a circular orbit for light, and it is at . Because has a maximum there, with negative, the orbit is unstable. Nudge a ray inward and it falls in. Nudge it outward and it leaves. Nothing sits on the photon sphere for long, which is why it behaves as a boundary rather than as a place.
Newtonian gravity says nothing about light, but a corpuscular reading of it does. Treat a light particle as an ordinary projectile moving at speed and ask for a circular orbit, . That gives , which is inside the radius §6 is about, by a factor of two.
The real answer, (3.8.10), is at : three times further out, and comfortably in the region where the solution is valid and nothing is strange. The discrepancy is not a small correction to be expanded in , because at these radii that quantity is of order one. It is the same failure that made Chapter 3.1's cabin calculation half the right answer, wearing different clothes. A corpuscle argument uses only the pull, and light samples more of the geometry than the pull.
2.2 · The critical impact parameter
A ray coming in from far away has at large , since . Whether it turns around or is captured depends on whether clears the top of the barrier. So we need the height of the peak. Evaluate at :
The critical aim is the one whose launch height sits exactly level with that peak. So set , take the square root, and read off the impact parameter:
Three cases follow from that one number.
- A ray aimed with is turned back.
- A ray with goes over the barrier and never returns.
- A ray aimed at exactly spirals in and asymptotes onto the photon sphere, taking infinitely many turns to arrive. Problem 1 derives the logarithm that says so.
So the object presents a capture disc of radius . To a distant observer the region from which no light returns is therefore an apparent disc of that radius, larger than by a factor of , and larger than the photon sphere too. The reason is that a ray arriving from far away is bent inward on the way in, and so gets closer than its aim suggests.
Numbers, for two objects whose masses are known from the orbits of things around them.
| Object | distance | ||||
|---|---|---|---|---|---|
| Galactic centre | |||||
| M87 centre |
Forty microarcseconds is radians. ⚑ An aperture resolves an angle , a result quoted here rather than derived, since this book has not built diffraction anywhere. Resolving that angle at a wavelength of needs , comparable to the radius of the Earth. No single instrument will do, and the aperture has to be synthesised from stations spread across the planet. So (3.8.12) is not a decorative result. It is the number that says what instrument is required, and it came from maximising one function of one variable.
Ask where light can be made to go round in a circle and the older theory is nearly silent, because nothing in it makes light respond to gravity at all. Treat a light corpuscle as an ordinary projectile moving at the limiting speed and Newtonian arithmetic does answer, and it is wrong by a factor of three: it puts the circular orbit inside the radius this chapter spends its second half worrying about, where the real one sits comfortably outside. The correct surface is where the barrier keeping light out of the central region reaches its highest point.
That the barrier has a highest point is the whole content of the section, and it says two things. A ray aimed to pass just outside the peak climbs, slows, turns and comes back out, having wrapped part of the way round. A ray aimed just inside goes over the top and never returns. The dividing aim is a single number, and it is not the radius of the circular orbit but something larger, because a ray arriving from far away is bent inward on the way in.
The perch itself is unstable, in the way the top of a hill is unstable. Nothing settles there, which is why the surface is a boundary rather than a place, and why the dark patch a distant instrument would record is larger than the object casting it.
3 · Deflection: the integral, done
Here is the destination, and the route with it, because there are four steps and each is small. We convert (3.8.7) into an equation for the shape of the orbit. We notice that its unperturbed solution is a straight line. We solve it to first order in the mass, read the deflection off the asymptotes, and put in numbers for the Sun.
3.1 · The orbit equation
We do not want the radius as a function of the parameter. We want the shape of the path, so we should trade the parameter for the angle. That is Chapter 3.7 §8's trick, used there for planets, and it works best in the variable rather than . From §1.3, , so the chain rule gives
the second from the chain rule applied to , since . Multiplying these two gives , so squaring kills the sign and (3.8.7) becomes
That is a first-order equation with a square in it, and squares are awkward to solve. We would rather have a linear second-order equation, so differentiate with respect to . Every term on the right carries out by the chain rule, and the left gives :
The common factor is now begging to be cancelled. Divide by , which is legitimate wherever the ray is not at a turning point, and the resulting equation then holds everywhere by continuity:
Compare this with the planetary case. Chapter 3.7 §8 obtained . There the constant source term is Newton's inverse-square law in disguise and gives closed ellipses, and the small correction is what precesses them. For light the constant term is absent altogether. The entire deflection of starlight is the term that for a planet is a correction of one part in .
3.2 · The unperturbed solution is a straight line
Before solving the full equation, let's see what it gives when the mass is switched off. Set in (3.8.16). What is left is , the equation Chapter 0.8 §4 solved in its sleep, with general solution . Choose the origin of so that and the amplitude so that the closest approach is at :
Now let's check what that is. In the plane, and , so (3.8.17) says . That is a straight line at perpendicular distance from the origin, traversed from (far away on one side) through (closest approach) to (far away on the other). It confirms the promise of §1.2, namely that is the impact parameter. It confirms something else worth naming too. With no mass present, light goes straight. The metric is then flat, and Chapter 3.3 §2 warned that a flat geometry in curvilinear coordinates has non-constant metric components and non-zero connection coefficients anyway. Here that warning is cashed as a formula that looks curved and is not.
3.3 · First order in the mass
Name the approximation before making it. For a ray grazing the Sun, , so the term on the right of (3.8.16) is a millionth of the terms on the left. Write with of first order in , substitute, and keep only first-order terms. Since satisfies the equation exactly with the right-hand side dropped, what is left is
and the discarded pieces are , second order in the mass. This is first-order perturbation theory in exactly the sense Chapter 3.1 §7.2 used it. You evaluate the small term along the uncorrected path, so that the neglected piece is the correction to the correction. Grind box A solves (3.8.18) in four lines, and the answer is
and substituting it back into (3.8.18) gives residual exactly zero, not approximately zero. The solution is exact for the equation it solves.
Grind box A — solving the first-order equation, and the one trigonometric identity used
Line 1. Get rid of the square. The source in (3.8.18) is , and a squared sine is not a natural right-hand side for a linear equation with resonant frequency . Use the double-angle identity of Chapter 0.3, , rearranged:
Line 2. The source has two pieces and the equation is linear, so solve for each and add (Chapter 0.8 §3.1). Write for brevity.
Line 3. The constant piece. For , try constant . Then and the equation reads . So .
Line 4. The oscillating piece. For , try . Then , so the left-hand side is . Matching, , so . Note why this works and would not have worked for a source: the driving frequency is and the natural frequency is , so there is no resonance and no secular term. Chapter 0.8 §6 is the general statement.
Adding. , which is (3.8.19).
The homogeneous solution has been set to zero, and that is a choice of boundary condition rather than an omission. Adding to amounts to shifting the impact parameter and rotating the axis, which is a redefinition of and of where sits, not a different ray.
(Checked symbolically: substituting (3.8.19) into (3.8.18) returns exactly , and substituting into the full equation (3.8.16) leaves a residual whose coefficient of vanishes identically, so the error really is second order.)
3.4 · Reading off the deflection
We have the shape of the path. The deflection is the difference between where the ray comes in and where it goes out, so we need the two angles at which the ray is infinitely far away. The ray is at infinity when , that is when . Without the mass that happened at and , and the total swing was exactly , which is a straight line. With the mass, solve near each end. Put with small, expand and , and keep first order in both small quantities:
That is one end. By the symmetry of (3.8.19) under the other end is displaced by the same , so the total swing is instead of and the ray has been turned through
In: the Schwarzschild metric, the null normalisation, two Killing charges collapsed to one ratio, one change of variable to , one differentiation, one linear second-order equation with a source, and one expansion about a straight line.
Out: a deflection , exactly twice Chapter 3.1's .
What it cost: first order in . That is not a limitation anybody meets in practice. For the Sun the neglected term is one part in of an angle already too small to see with the eye. It is a genuine restriction all the same, and it is why in (3.8.21) may be read interchangeably as the impact parameter or the distance of closest approach: at this order they agree.
3.5 · The number
For a ray grazing the Sun, and , with . Every conversion on the page:
Radians are not the unit astronomers quote, so convert to arcseconds. One radian is degrees and one degree is arcseconds, so one radian is arcseconds:
Two observational results are quoted here and neither is derived. In May 1919 two eclipse expeditions organised from Britain measured the displacement of stars near the eclipsed solar limb and reported deflections consistent with (3.8.23) rather than with Chapter 3.1's half of it, at a precision of some tens of per cent. Since the 1970s the same quantity has been measured with very-long-baseline radio interferometry, using quasars occulted by the Sun rather than stars, and the general-relativistic value is confirmed to about one part in .
The reason the second is so much better is worth a sentence, because it is not about better telescopes. Radio interferometry does not need an eclipse, since the Sun is faint at centimetre wavelengths. The measurement can therefore be repeated at leisure, at many impact parameters, with the dependence of (3.8.21) itself checked rather than a single number.
3.6 · Lensing, and two chapters' promise collected
Equation (3.8.21) has a consequence that Chapters 1.2 and 2.3 both pointed at and neither could develop. Put a source far behind a mass and consider all the null geodesics running from source to observer. If the source is exactly behind, every ray leaving it in a cone of the right opening angle is bent by the same amount and arrives, so the image is a ring. If it is slightly off-axis the ring breaks into two or more separate images. There is more than one geodesic between the same two events.
That is precisely the phenomenon Chapter 1.2 §5.1 named a conjugate point and illustrated with the two arcs of a great circle joining two cities, both of them stationary points of arc length and only one of them the shortest. Chapter 2.3 §§6.1 and 6.4 met it again from the other side. In flat spacetime the straight worldline is not merely locally but globally the longest, and that chapter noted that in a curved spacetime the global statement can fail. Here is the failure, and it is observed. The several geodesics are the several images. The surface on which their number changes is the caustic, and the caustic is where the second variation of Chapter 1.2 stops being definite.
Turning (3.8.21) into image positions requires the geometry of a thin lens, plus bookkeeping about three distances: observer to lens , observer to source , and lens to source . We quote the result rather than deriving it, because a careful derivation needs the propagation of a bundle of neighbouring null geodesics, which is Chapter 3.4 §4's geodesic deviation applied to null curves and is machinery this book does not build. With the true angular position of the source and the observed position of an image,
the second expression being (3.8.21) with the impact parameter written as . What follows from it is derived, and it is worth doing. Set , meaning the source sits exactly behind the lens. That gives , so the ring has angular radius
Put numbers in: a galaxy of , halfway to a source at , so that and . Then and radians, which is arcseconds. That is a comfortably resolvable angle, which is why lensed images are catalogued in their thousands. And for a general the lens equation is a quadratic in with two real roots of opposite sign: two images, on opposite sides. The count is the derived part. The equation it is counted from is the quoted part.
Solving the equation for a ray exactly is not on offer, and the honest move is the one this book has made in every part: expand about something already understood and keep the first correction. What is already understood here is startling in its plainness. Take the mass out and the equation says light travels in a straight line, and that line, written in the polar language the problem forces on us, is the tidiest formula in the chapter. The whole deflection is a small disturbance of it.
The correction is found by feeding the straight line back into the piece of the equation that was dropped and asking what must be added to absorb it. This is legitimate because the disturbance is small: for a ray grazing the Sun the trajectory shifts by a few parts in a million, so evaluating a small term along the uncorrected path makes an error that is the correction to the correction. Where the ray runs off to infinity in either direction, the disturbed solution has tilted the two asymptotes toward each other by a definite angle.
The result is the exact double of the estimate made from the accelerating cabin, and the arithmetic producing one and three-quarter seconds of arc is worth doing on paper once. That angle, and the geometry of a source lying behind a mass, is the reason astronomers wait for eclipses.
4 · Where the missing half was
This is the section Chapter 3.1 §7.3 named in writing. Entry (a) of its register of debts reads: "The factor of two in — Chapter 3.8 §4, which derives from the Schwarzschild solution and names the missing half as spatial curvature." Half of that is now done, because §3 derived the . What remains is to identify the missing half rather than assert it, and then to explain why light collects both halves and a planet does not.
Here is the destination. We mutilate the metric in two different ways. The first keeps only the distortion of time, the second only the distortion of space. We run §3's calculation on each, and get exactly from each. Then we run the same pair of calculations for a particle moving at arbitrary speed and find that the space half is independent of while the time half grows as . Their ratio is , which is the sentence Chapter 3.1 promised.
4.1 · Why the two halves may be computed separately
Before splitting anything, we should check that splitting is allowed. The deflection was obtained to first order in . At that order the first-order equation (3.8.18) is linear in the perturbation. Splitting the metric perturbation into two pieces gives a source that is the sum of two sources, and a linear equation with a summed source has the sum of the two solutions. So the contributions add, and computing them separately is not a heuristic. It is a consequence of linearity, and §4.3 checks it explicitly by showing that the two first-order solutions sum to (3.8.19) term by term.
Name the two mutilations precisely. Write the Schwarzschild line element as
in the equatorial plane, with as before. The angular term carries no mass and belongs to neither. Mutilation one keeps in the time slot and replaces by . That is Newtonian clocks-only gravity, which is precisely what Chapter 3.1's cabin argument had access to. Mutilation two does the reverse, with in the time slot and in the radial slot.
4.2 · Mutilation one: keep only the distortion of time
Repeat §1.3 with . The Killing charges are unchanged in form, and , since neither used . The null condition now gives, after substituting and rescaling to as before,
The factor has moved. It now sits under the instead of multiplying . We want the orbit equation again, so convert to as in §3.1 and expand , keeping first order in the mass:
the second expression obtained by differentiating and dividing by exactly as in (3.8.15). Now look at what the right-hand side has become. The source is a constant, not a . Its particular solution is that same constant, , so . Putting and as in (3.8.20) gives the deflection:
That is Chapter 3.1 §7.2's exactly, obtained here from a metric rather than from a cabin, and it is at the solar limb. It is also the prediction Einstein published in 1911, with the solar constants of his day, and superseded in 1915. The equivalence principle sees the time part of the geometry and nothing else, and this calculation is what that sentence means.
4.3 · Mutilation two: keep only the distortion of space
Now and . The time charge becomes with no in it. The null condition rearranges to , and rescaling as before,
We want this in the same second-order shape as §3.1's orbit equation, so differentiate with respect to and divide by once more:
Two source terms this time. Substituting on the right and using as in grind box A, the source becomes , whose particular solution by the same two-line method is
At this equals , the same value as the time part's constant, so the asymptote calculation of (3.8.20) runs identically and delivers
Now the check §4.1 promised. If linearity really lets us compute the halves separately, adding the two first-order solutions must reproduce §3's answer. Add (3.8.26)'s constant to (3.8.30):
which is (3.8.19) exactly, not merely at the asymptotes but at every angle. The debt is now identified rather than asserted: the missing half of Chapter 3.1's deflection is the in (3.8.24), the curvature of space itself. Chapter 3.6 Problem 3 met the same object in weak-field clothing, deriving and predicting that its spatial factor would be the missing half. It is.
Mutilation one has an exact classical counterpart. A light ray in a medium of refractive index follows the path that extremises . That is Fermat's principle, which is Chapter 1.2's action principle with in place of the Lagrangian. A ray crossing a region where varies is bent toward the higher index at a rate . Take and that bending rate is transverse component, which integrated along the ray gives (3.8.27). The time part of the metric is a graded-index medium, and a mirage over hot tarmac bends light by mathematics identical to Einstein's 1911 calculation.
Where it breaks, and it matters. A real medium slows light. The locally measured speed in glass is , and that is a measurable fact about glass. Here the locally measured speed is always exactly , in every local inertial frame, which is what Chapter 3.1 §3's equivalence principle demands. The index describes coordinate speed, a bookkeeping quantity. And the analogy does not survive at all into §4.3, because an isotropic index cannot reproduce the spatial curvature term with the correct dependence on direction. The analogy delivers exactly the half Chapter 3.1 already had, and not one part of the half it was missing.
4.4 · Why light collects both halves and a planet collects one
The two halves are equal for light. They are not equal for anything else, and computing the inequality is what turns the slogan into an explanation. So run §§4.2 and 4.3 again for a massive particle arriving from far away with speed and impact parameter . At infinity the metric is flat, so Chapter 2.5's kinematics applies there and the two charges are
the first because and far away. The second holds because angular momentum per unit mass is , evaluated on the incoming asymptote, where the speed is and the perpendicular distance from the centre to that asymptote is what means. (Not at closest approach: there the speed has been increased by the fall inward, and the distance is not .) These give using , so the unperturbed orbit is again and is again the impact parameter. Everything is set up as before.
Now run the two mutilations through the same three steps, which were normalisation, change of variable and differentiation. They give
and the second of these is identical to (3.8.29). The space half does not know how fast the particle is going. Reading off both deflections as before,
Three checks on that. Setting gives , recovering (3.8.21). Letting gives , which is the Newtonian deflection of a fast projectile by an inverse-square force and can be checked against Chapter 1.4's Kepler machinery. And the ratio of the two contributions is
That is the sentence Chapter 3.1 promised, and it is now a formula. Let's say it in words. Both parts of the geometry are distorted by the same fractional amount, of order . What differs is how much of each a moving body samples. A body's motion through spacetime is divided between the time direction and the space directions, and for a slow body the division is grotesquely lopsided: in one second of its own time it covers metres of the time direction and metres of space. The spatial distortion therefore enters its trajectory suppressed by .
Light divides its motion equally. It covers exactly of space in every of time, which is what moving at means, and so it samples both distortions with the same weight. Two equal contributions, and hence a factor of two.
| Body | space share of the deflection | ||
|---|---|---|---|
| Mercury at perihelion | |||
| Earth in orbit | |||
| A electron | |||
| Light |
The last column is , that is . Read the table downward. The entire content of the factor of two is that this column runs from zero to a half. Newtonian gravity is the top row, and it is right there because that is where the planets are.
Four rows is what a page holds. The figure below runs the two mutilated calculations at every speed between them, by integrating the orbit equation rather than evaluating (3.8.35). So the boxed formula is something the picture checks rather than something it draws.
Chapter 3.1 §6 obtained from the redshift and complained that it had learned nothing about the spatial components. Chapter 3.6 Problem 3 supplied them from the field equations. Chapter 3.7 supplied them exactly, with no weak-field assumption. This section separated their effects and found them equal for light. Four chapters, one debt, and the reason it took four chapters is that identifying the missing half required the spatial metric, which required the field equations, which required the whole of Part III's machinery. Chapter 3.1 could not have paid this debt with any amount of cleverness. It could only record it.
Here is the debt from the beginning of this part, and the payment is more interesting than the number. The cabin argument used only the part of the geometry governing clocks, and it could not have done otherwise: every step in it was about how long something took. So repeat this chapter's calculation on a mutilated geometry: keep the distortion of timekeeping, throw away the distortion of distances, and see what survives. Exactly half survives, and it is the value published in nineteen eleven and withdrawn four years later. Do it the other way round, keeping only the distortion of distances, and the other half appears.
The two halves add, and they add exactly, because at this accuracy the disturbance is small enough for its effects to superpose. That much is arithmetic. The interesting part is why a planet does not collect both.
Run the same calculation for a body moving at any speed and the bending contributed by the distortion of distances is fixed, while the bending contributed by the distortion of timekeeping grows without limit as the body slows. Their ratio is the square of the speed divided by the square of the speed of light. For Mercury that is a part in twenty-six million, which is why three centuries of astronomy got away with only the clock half. For light the ratio is one, and nothing can be got away with.
5 · Gravitational redshift, three ways
The destination first. A signal climbing out of a gravitational well arrives with its frequency reduced by the factor . We derive this three times: from the conserved Killing charge of §1.2, from the relation between a static clock and coordinate time, and from Chapter 3.1's cabin, recovered as the weak-field limit. Then we do the satellite-navigation arithmetic on one metric instead of two effects added by hand.
5.1 · Route (a): the conserved charge, against a local clock
This is where §1.2's warning is cashed. A photon has a four-momentum proportional to its tangent . Chapter 3.5 §9 says is constant along the ray for every Killing vector , and for the time-translation Killing vector in coordinates that constant is
That is the bookkeeping quantity. What we want is what somebody actually measures, so put an observer in. An observer sitting still at radius has four-velocity with only a time component, and normalisation gives , so . The energy they measure is the projection of the photon's momentum on their own four-velocity, which is Chapter 2.5 §4's rule, unchanged:
The bracket is the constant of (3.8.37). So the locally measured frequency is a fixed number divided by , and comparing an emitter at with a receiver at infinity where ,
which is less than one, so the signal is reddened. Notice what did the work here. The conserved quantity did not change, and the local clock did. That is why §1.2 insisted that is bookkeeping and the observer's measurement is the physics.
5.2 · Route (b): counting crests, with the metric only
Forget photons. Chapter 3.3 §3 wrote proper time along a worldline at rest as , which here is
Now repeat Chapter 3.1 §6.3's crest-counting argument verbatim, because every word of it still applies. The geometry is static, meaning the metric components do not depend on . A crest leaving the emitter at coordinate time arrives at the receiver at coordinate time , and because nothing in the geometry depends on , the same applies to the next crest. So the coordinate-time interval between crests is the same at both ends, and no crests are created or destroyed in between. Converting each end to its own proper time with (3.8.40),
which reduces to (3.8.39) when the receiver is at infinity. The two routes share not one step. Route (a) used a Killing vector and a four-momentum. Route (b) used only the component of the metric and the fact that nothing depends on .
5.3 · Route (c): the cabin, recovered
The third route is to check that the exact answer reduces to the one Chapter 3.1 already had. So expand (3.8.39) for , using :
with the Newtonian potential. Chapter 3.1 §6.4 obtained from an accelerating cabin, the Doppler formula and the equivalence principle, with no general relativity in it whatever. With the receiver at infinity, and that formula gives , which is (3.8.42). Three derivations, one answer. The third is the one that matters historically, because it shows that the whole effect was available in 1907 from special relativity plus one principle, and that the exact solution merely resums it.
Numbers, from (3.8.39) directly. The fourth column converts the fractional shift into the speed a source would need in order to produce the same shift by ordinary Doppler, which is how spectroscopists usually report it.
| Surface | equivalent speed | |||
|---|---|---|---|---|
| Earth | ||||
| Sun | ||||
| White dwarf, , | ||||
| Neutron star, , | — |
The last row has no equivalent-speed entry because the shift is and the expansion that produced (3.8.42) has stopped being useful. That is the row where the cabin argument dies and (3.8.39) keeps working.
Both derivations above put an observer at rest at radius and used . That requires , that is . As the factor and (3.8.39) says the received frequency goes to zero, which is an infinite redshift. Below the expression under the square root is negative and there is no such observer at all. The four-velocity of a stationary body would have negative norm, which by Chapter 2.3 §4 means no body can be stationary there.
Do not read this as a proof that something dreadful happens at . It is a proof that hovering is impossible there, which is a statement about a family of observers and not about the geometry. Sections 6 and 7 separate the two, and the distinction turns out to be the whole of what a horizon is.
5.4 · Satellite navigation, on one metric
Chapter 3.1's Worked example 2 computed the clock-rate difference between a navigation satellite and the ground, and was explicit that it was gluing together two effects from two different chapters: "adding is legitimate here only because both effects are of order and their product is of order … the place where Chapter 3.8 §5 does it properly with a single metric rather than two separate small corrections." Here is the single metric.
A satellite in a circular equatorial orbit of radius has and . We want its proper time against coordinate time, so divide the line element by :
with . One square root, containing both terms. Nothing has been added by hand. The height dependence and the speed dependence are two entries of the same metric evaluated on the same worldline. The ground station, at rest at , has with the rotation of the Earth neglected. What we want is the ratio of the two rates, so expand both square roots to first order and divide:
The two small terms are the two lines of Chapter 3.1's calculation, and now they are visibly two terms of one expansion rather than two competing physical effects. With , , and :
Both of those are fractional rates, so to get an accumulated time we multiply each by the in a day. That gives and , netting
And now the justification for dropping the cross terms is a calculation rather than a promise. The product of the two small quantities is , which over a day amounts to seconds. That is four femtoseconds, about of the effect being kept, and far below the stability of any clock that has been flown. Note also that the gravitational term is six times the kinematic one, so a working navigation system is predominantly a test of this chapter and only secondarily of Part II.
Clocks are the subject again, and the fact about them arrives three times from three directions, with the agreement worth more than any one derivation. The first route uses the direction in which the geometry does not change: a quantity built from it is the same at every point of the ray, while the frequency an observer actually measures involves that quantity divided by the rate of their own clock, and the two differ by exactly the factor this section is about. The second counts wave crests, as the equivalence-principle chapter did, and needs only the relation between a stationary clock and the coordinate label. The third takes the answer far from the mass, where it collapses onto the accelerating-cabin formula obtained with no general relativity in it at all. Three routes, one answer, and the third closes a loop opened eight chapters ago.
Then satellite navigation, usually presented as two competing corrections added by hand. It is not two corrections. There is one geometry, one square root and one expansion of it; the term involving height and the term involving speed drop out of the same line. They may be added because each is around a ten-billionth while their product is around a hundred-billion-billionth, which over a day comes to a few femtoseconds. Neglecting the whole thing costs eleven kilometres a day.
6 · The horizon is not where the metric blows up
Look at the Schwarzschild metric again and read off what happens at :
One component vanishes and another diverges. Stop here and decide what you think that means, before reading on. You have the tools to settle it and the question is worth answering for yourself first: is (3.8.47) a statement about spacetime, or a statement about the labels somebody chose for it?
Chapter 3.3 §2 is the reason to hesitate. There, the flat plane in polar coordinates was worked early and deliberately. Its metric is , the components are not constant, one of them vanishes at , the inverse metric component diverges there, and the plane is the plane. Nothing is wrong with the paper. The moral was stated in that chapter and has not been used since: badly behaved components are not evidence of badly behaved geometry.
Here is a first hint that the same thing is happening. Compute the determinant of the Schwarzschild metric, in the very coordinates where (3.8.47) misbehaves:
Look at what happened to the two offending factors. They cancel exactly. The determinant is the same function of and as it is in flat spherical coordinates, it is finite and non-zero at , and its only zeros are at and on the polar axis. That last one is the familiar artefact of spherical coordinates that Chapter 3.2 §2 discussed. This is suggestive, not conclusive, because a determinant is not an invariant. It changes under a change of chart by the square of a Jacobian. Three real calculations follow, and the third is decisive.
6.1 · Calculation one: the distance down to the surface is finite
"" is often read as "the horizon is infinitely far away". Let's test it. The proper radial distance between two radii at one instant of is obtained by setting in the line element and integrating :
Grind box — the proper radial distance, done
Substitute , so that and . The two hyperbolic sines cancel and what is left is elementary:
using from Chapter 0.3 and . Undo the substitution with and , so that . At the lower limit we have and both terms vanish, which is why only the upper limit survives.
The integrand diverges at the lower limit like , and Chapter 0.2 §5 classified exactly this case. An inverse power is integrable, so the improper integral converges. For the answer is , a perfectly ordinary length, slightly larger than the coordinate difference . Being larger is the correct behaviour, and it is the reason is called the areal radius rather than the distance to the centre. A diverging metric component has produced a convergent distance.
6.2 · Calculation two: two clocks disagree about the crossing
Drop something in radially from rest a long way out and follow it. Radial means . Released from rest at infinity means there, so by (3.8.3). Substituting both into (3.8.5),
That is remarkable on its own account. It is exactly the Newtonian free-fall speed , written in relativistic clothing, and it is perfectly finite and smooth at , where it equals . Nothing in (3.8.50) notices that anything special is happening. What we want next is the time this takes on the faller's own watch, so integrate, as grind box B does:
which is finite for every , including and including . The faller's own watch records an ordinary number of seconds and keeps running.
Now the same journey in coordinate time. From (3.8.3), , so and
The extra factor is the whole difference between the two clocks. Near the integrand behaves as , an inverse first power. Chapter 0.2 §5 classified that as the borderline case that diverges, logarithmically. Grind box B does the integral exactly and extracts
Grind box B — both integrals, done exactly
Proper time. Invert (3.8.50) to get and integrate from down to , flipping the limits to absorb the minus sign:
Nothing in the integrand is singular anywhere in , so there is nothing to discuss. Setting gives for the whole journey to the centre, and setting , gives the celebrated from horizon to centre.
Coordinate time. Integrate (3.8.52). Substitute , so and :
Split the integrand. Add and subtract in the numerator:
The second piece. Put , so and :
the middle step being polynomial division, , and the last using the partial fractions of Chapter 0.8 §2.1. Collecting, and writing ,
(Differentiating returns exactly; checked symbolically.)
The divergence, isolated. Only the logarithm misbehaves as . There and , so , and since it enters with a minus sign, . Solving for gives , which is (3.8.53). The approach is exponential, with the length setting the scale.
The contrast, stated as plainly as possible. The faller crosses after a finite, unremarkable interval on their own watch, computed by (3.8.51), and carries on. The coordinate label assigned to the crossing is . Both statements are correct, and they are statements about different things. One is about the traveller. The other is about a labelling scheme anchored to observers who are at rest far away, and §5's warning already established that no such observer exists at at all. Only one of these two numbers is about anybody's experience.
6.3 · What a distant observer actually sees, and how briefly
"The infalling object appears frozen at the horizon" is the usual gloss on (3.8.53), and it is misleading in a way worth correcting with two lines of arithmetic, because the object does not hang there visibly: it winks out.
Take the falling body to emit a steady signal. Its four-velocity is , the first entry from and the second from (3.8.50). An outgoing radial photon has , the second entry forced by the null condition with the sign chosen so that . The frequency measured by the emitter is , and the frequency at infinity is , so writing ,
using . That is a tidy exact result, and it vanishes at . Near the horizon it is , which is proportional to the very quantity (3.8.53) says decays exponentially.
One correction is needed before the exponential can be quoted, and skipping it costs a factor of two. The signal emitted at coordinate time takes time to climb out, and that climb time also diverges. Section 7.1 constructs the function whose difference gives the travel time, and logarithmically at the horizon in exactly the same way. Adding the two logarithms doubles the coefficient, so in terms of the time at which the light arrives,
The characteristic time is , which is for a hole and for the one at the galactic centre. After a few multiples of that the signal is redshifted out of any detector, and the last photon has been emitted. The image does not linger. It fades on that timescale and is gone.
Equation (3.8.55) has a form you use every day. A quantity whose rate of decrease is proportional to itself obeys , whose solution is . The constant is the mean lifetime, is the half-life, and after five or so multiples of the quantity is gone for practical purposes. That is first-order elimination kinetics, and it is also radioactive decay, and it is also (3.8.55). The mathematics is not analogous. It is the same equation, because (3.8.53) came from an integrand behaving as and that is what always produces an exponential.
Where it breaks, and the break is the point of the section. A drug concentration falling exponentially is a real physical quantity in a real patient, and asymptotic approach means the drug is genuinely still there in ever-diminishing amounts. Here the exponential describes a coordinate label and the brightness one particular family of observers records. The faller's own clock, by (3.8.51), records nothing exponential at all: a finite interval, an ordinary crossing, and a continuation. If you take the exponential as a fact about the object rather than about the observation, you will conclude that nothing ever falls in, which is exactly the error §7 exists to prevent.
6.4 · Calculation three: the invariant, which cannot be argued with
Distances and clock readings are physical but chart-dependent in their bookkeeping, and the argument above therefore relies on interpretation. Chapter 2.4 §§5.1 and 6 supplied the instrument that does not. The Riemann tensor is a tensor, so a full contraction of it with itself is a scalar: every observer, in every chart, computes the same number at the same event. The simplest such quantity quadratic in the curvature is the Kretschmann scalar
Why not something simpler? The Ricci scalar is useless here. Chapter 3.7 solved , so and vanish identically outside the mass and tell us nothing. What survives is the Weyl part that Chapter 3.4 §6 named and set aside, and is how you get at it. Grind box C computes it, and the answer is
Grind box C — the Kretschmann scalar for Schwarzschild, in three steps
The computation is mechanical and long, so here is the route with every intermediate result. The metric is in coordinates with , and .
Step 1. The connection, from Chapter 3.3's formula (3.3.50). Nine non-zero components, counting each symmetric pair once, and these are the same nine Chapter 3.7 §3 listed while solving the field equations:
Step 2. The Riemann tensor, from Chapter 3.4 (3.4.13), then lowered with the metric. After the symmetries of Chapter 3.4 §5 are used, six independent components survive and every other one is either zero or one of these up to a sign:
Two things follow without further work. Every component is linear in , so , being quadratic, carries . And has dimensions of (length), so with only and available it must be a pure number times . Step 3 only fixes the pure number.
Step 3. The contraction, and the counting. All the surviving components have the same pair of index pairs on both sides, with , and every component with vanishes here. Each such component appears in four index assignments, namely , , and , and the two sign flips cancel in the square, so
the last because the metric is diagonal, so raising an index is multiplication by one entry of . Take the term as the model. There and , so
Every factor of has cancelled between the lowered component and the raising factors, which is the structural reason the answer contains no and therefore nothing that could misbehave where . Doing the same for the other five and multiplying each by four, in units of :
(Verified symbolically: the connection, the full four-index Riemann tensor and the complete contraction were computed from the metric above with nothing assumed, returning exactly, and in all sixteen components, which is the check that the metric being differentiated really is the vacuum solution Chapter 3.7 derived.)
Now evaluate (3.8.57) at the two places in question. At the horizon, substituting ,
a perfectly finite number. And at exactly one radius, namely .
That settles it, and it settles it in a way no change of chart can touch. is a scalar. If it is finite at an event in one chart it is finite at that event in every chart, by Chapter 2.4 §6. So there is no chart in which the curvature at is infinite, because there is nothing infinite there to display. Whatever (3.8.47) is telling us, it is not that the geometry has failed. Since 2.3 the word invariant has meant exactly this: a number all observers agree on, which therefore cannot be an artefact of any of them.
6.5 · How gentle the edge is, and why bigger is gentler
A number being finite is not the same as its being small, so put a scale on it. Chapter 3.4 §4 derived geodesic deviation and identified it with Chapter 3.1's tidal acceleration: two freely falling bodies separated radially by accelerate apart at . Evaluate that at the horizon, using :
The same scaling is visible in (3.8.58), where . Indeed identically, so the tidal coefficient is the square root of the invariant up to a fixed factor. Chapter 3.1's Worked example 1 remarked on this scaling and pointed here for the reason. Here is the reason, and here are the numbers for a person of height :
| Mass | head-to-toe stretch at | as a multiple of | ||
|---|---|---|---|---|
A bigger hole is gentler at its edge. That is counter-intuitive and it follows from elementary scaling: the tidal field goes as and the horizon radius as , so the tidal field at the horizon goes as . It also joins up with Chapter 3.1 §5, which bounded the size of a freely falling laboratory by requiring tidal effects to stay below the measurement precision. Small means a large local inertial frame, and for the last row of the table the local inertial frame at the horizon is enormous. An observer crossing it in a sealed cabin has no local experiment that will tell them anything has happened. That is precisely the equivalence principle doing its job, and precisely why §7 is needed to say what has happened.
Shown. The geometry at is regular: an invariant built from the curvature is finite there, the proper distance to it is finite, and a falling body crosses it in finite proper time with nothing locally remarkable occurring.
Not shown, and not true. That the surface is therefore unimportant. Something real does happen at , and §7 computes what: it is a one-way surface. The point of §6 is only that whatever is special about it is not that the curvature blows up, because it does not.
Also not shown. That the Schwarzschild chart is bad everywhere. It is an excellent chart for , which is where every measurement in §§2–5 was made, and Chapter 3.7's orbits and this chapter's deflection and redshift are all computed in it without difficulty. Chapter 3.2 §2 made exactly this point about the sphere: a chart failing on part of a manifold is normal, and the repair is a second chart, not a different manifold.
A distinction prepared when the flat grid became a patchwork, and flagged again one chapter ago, carries everything here: whether a failure is a fact about the labels rather than a fact about the thing described. At one particular radius two of the numbers describing the geometry misbehave, one dropping to zero and the other running away to infinity. The temptation to read that as a catastrophe is enormous and should be resisted, because the same thing happens at the origin of a flat page in circular labels, and nothing is wrong with the page.
Three calculations settle it, and none is an argument from taste. Somebody falling inward crosses the offending radius after a perfectly ordinary and finite interval on their own watch, while the coordinate label assigned to the crossing runs away without limit, so only one of them is about the traveller. The distance down to the surface, measured with the geometry's own ruler, is finite as well, despite the ruler component being the one that diverges.
And then the number that cannot be argued with. There is a quantity built from the curvature that every observer computes the same, whatever the labels, and at the offending radius it is finite. It becomes infinite in one place only, the centre. Bigger holes are gentler at the edge, so gentle that for the largest nothing detectable happens at crossing.
7 · Eddington–Finkelstein, and crossing
Section 6 showed that the geometry is fine at and that the chart is not. Chapter 3.2 §2 said what to do about a chart that fails on part of a manifold: find another chart, check that the two agree on the overlap, and read the geometry in whichever one is well behaved where you are working. That is the entire operation performed here. It is worth noticing that no new physics enters. All that changes is the labels, and by Chapter 2.4 §6 a change of labels cannot change anything a tensor says.
Here is the destination. We build a new radial coordinate by integrating the condition for a radial light ray, use the arrival label of an ingoing ray as a new time coordinate, rewrite the metric, check its determinant, and then compute the two radial null directions at every radius. Inside both of them have . That computation is what a horizon is.
7.1 · Following the light: the tortoise coordinate
The construction is dictated by the problem rather than chosen. What misbehaves in the Schwarzschild chart is the description of light. As , radial rays take infinite coordinate time to get anywhere, which is why is a bad label. So build a coordinate for which they do not.
A radial null curve has and , so from the line element
The right-hand side is a function of alone, so it can be integrated once and for all. Define the tortoise coordinate by :
where the third step is the polynomial division and the constant of integration has been chosen to make the logarithm's argument dimensionless. Differentiate the answer back to check: ✓. Notice what does: it is far away, and it runs to as , stretching the last kilometre above the horizon into an infinite coordinate range. That is exactly the divergence §6.2 found, isolated into one function of where it can be dealt with.
Now we can use it to label the rays themselves. With (3.8.61) in hand, radial null curves are , so consider the two constants
which are respectively constant along ingoing and outgoing radial rays. Use , the label of the ingoing ray that passes through an event, as the new time coordinate, keeping , and as they were. This is the ingoing Eddington–Finkelstein chart, and both and have dimensions of length.
7.2 · The metric in the new chart
From (3.8.62), , so . Substitute into the line element and expand the square:
and now watch the two terms, which are the only place appeared. They cancel outright:
Every component is finite at , and the metric is perfectly ordinary there. The component vanishes at , but that on its own is not a failure. The component of a flat plane's metric vanishes at the origin too. What would be a failure is degeneracy, meaning the metric becoming non-invertible, so that raising an index is impossible. Let's test it. In the ordering the matrix is block-diagonal with a block , whose determinant is , so
which is non-zero for every off the polar axis, and in particular is at the horizon. The determinant does not merely fail to vanish. It never notices that exists. And the block's determinant is independently of , which is the algebraic reason: the off-diagonal carries the invertibility, and it is there precisely because the new time coordinate was built out of the light rays rather than out of the static observers.
Done: a change of chart, with , which is smooth and invertible on . That is exactly a transition map in the sense of Chapter 3.2 §2. On that overlap the two charts describe the same geometry, and any tensor computed in one may be transported to the other by Chapter 3.2 §7's rule.
Not done: anything physical. No field equation was re-solved, no assumption added, no matter introduced. Equation (3.8.64) is Chapter 3.7's solution, written down in different letters. (Computing directly from (3.8.64) returns zero in all sixteen components, as it must.)
What was gained: the new chart covers , not merely . It is defined on more of the manifold than the old one, which is the whole reason for building it. The Schwarzschild chart was never wrong. It was incomplete, in the same way that a single chart on the sphere is incomplete.
7.3 · The two radial null directions, computed
Now the calculation that makes a horizon a horizon. Set with in (3.8.64):
A product is zero when a factor is, so there are exactly two families, as there must be, because a light cone has two radial generators:
Branch (i) is the ingoing family, by construction: was defined to be constant along ingoing rays. Branch (ii) is everything else. To see what it does, plot against a time-like label rather than against . Put , which is the natural choice because it makes the ingoing rays run at as they do in flat spacetime. Then and
One thing has to be fixed before those signs mean anything: which way is the future. It is settled once and for all outside the horizon, where plainly increases with . A time orientation is a continuous choice on a connected spacetime, so it carries inward unchanged rather than being re-chosen at . Check it survives: an infalling particle has and , so . Increasing is the future everywhere, inside as much as outside.
Both slopes are now explicit functions of and there is nothing left to interpret. Read branch (ii) off:
| Radius | for branch (ii) | What that means |
|---|---|---|
| outgoing at the speed of light, as in flat spacetime | ||
| still outgoing, but slowed in these coordinates | ||
| the ray stays at forever | ||
| the "outgoing" ray moves inward | ||
| both branches fall at the same rate, and the cone has closed |
The middle row is the horizon, arrived at by evaluating a formula and not by drawing anything. The row below it is the content of the whole chapter's second half:
Inside every future-directed light ray moves to smaller . A material particle's worldline lies strictly inside the light cone, by Chapter 2.3 §4, and that statement is unchanged by the move to a manifold because it is a statement about the local light cone. So every material particle inside also moves to smaller . Not because anything pushes it. Because the only directions available in which its own future lies all point inward. That is what the surface is: a one-way membrane, and the word horizon means precisely this and nothing more.
One more fact falls out of (3.8.68) for free. Both branches can be integrated in closed form. Branch (i) gives . For branch (ii), integrating by the same polynomial division as (3.8.61),
Both families are therefore drawable exactly, with no numerical integration, and that is what the figure below draws.
7.4 · What has and has not been shown about escape
It is worth separating two statements that are easy to run together. (3.8.69) is a local fact: at any event with , the future light cone lies entirely in the direction of decreasing . What follows from it is that a particle inside cannot stop, cannot turn round and cannot hover, since all three would need a future direction with and there is none. How long it has left is then §6.2's integral. For a body falling from rest far away, (3.8.51) with and gives a proper time of from the horizon to the centre. That is for a hole and for the one at the galactic centre.
What has not been shown is that the surface is a horizon in the general sense that would apply to a spacetime with no symmetry at all. That statement needs a definition, and the definition needs machinery beyond this book.
The general definition is this. Consider the set of events from which a signal can eventually reach an observer who remains arbitrarily far away for arbitrarily long. Formally, that is the causal past of future null infinity. The event horizon is the boundary of that set. We quote this rather than deriving it, because "arbitrarily far away for arbitrarily long" has to be made precise, and doing so requires attaching a boundary to the spacetime by conformal compactification, which is machinery this book does not build.
Two consequences of that definition are worth knowing even though we cannot prove them here. It is global and teleological. Whether an event lies on the horizon depends on the entire future of the spacetime, so a local observer crossing it cannot in principle detect the crossing. That is consistent with everything §6.5 computed about the tidal field being unremarkable there. And it is why qualifies. Section 7.3 showed that no future-directed causal curve from ever reaches larger , so the boundary of the region that can signal to infinity is exactly , for this spacetime. The general definition and the calculation agree here. In a spacetime that is still collapsing they need not coincide with any locally computable surface at all.
7.5 · The Rindler horizon, compared honestly
Chapter 2.3 Problem 3 met a horizon already, in flat spacetime with no matter in it anywhere. A rocket with constant proper acceleration follows the hyperbola , whose asymptote is the null line . That problem showed by explicit calculation that a flash emitted from beyond that line never catches the rocket, however long it chases. The dividing surface is the Rindler horizon, and Chapter 2.3 promised that this chapter would meet the same object again.
What is the same. Both are one-way surfaces with no local marker on them: nothing measurable happens there, and an observer crossing either can perform no local experiment that detects the crossing. Both are an artefact of a chart in one specific sense. A coordinate system adapted to a family of observers, the static ones here and the uniformly accelerated ones there, assigns infinite coordinate time to the crossing, while the crossing takes finite proper time. Both produce an exponentially growing redshift with a characteristic time set by the surface gravity. For Rindler that is , which is about a year for . Chapter 3.1 §6.1 noted the companion length , about a light-year, which is where the horizon sits. For Schwarzschild (3.8.55) gives . And in both cases the repair is a change of chart, which was the point of §7.2.
And now the honest difference, which matters more than the similarity. The Rindler horizon belongs to an observer. Different accelerations give different horizons. An inertial observer sails across it noticing nothing and has no horizon at all. If the rocket cuts its engines its horizon dissolves and every delayed signal arrives. The Schwarzschild horizon belongs to nobody. It is at for every observer, however they move and whatever chart they use, because is the areal radius and the area of a sphere is a geometric fact. Stop accelerating and it does not go away. The reason for the difference is exactly the one §6.4 established. Flat spacetime has everywhere and there is nothing for a horizon to attach to, while here is a definite non-zero number that marks the surface unambiguously even though it is finite.
The correct summary is that the two horizons have the same local character and different global status, and the vocabulary for saying so precisely is the flagged definition of §7.4. Chapter 2.3's claim that they are "the same kind of object" is right about the local part and overstated about the rest. This is the sharper statement it pointed forward to.
If the trouble is with the labels, change the labels. The recipe is the one the manifold chapter set out and never used in anger: find a better chart on the overlap and check that the geometry written in it behaves. The better chart is built by following the light. Ask what path an inward radial flash takes, integrate that condition to get a new radial ruler, and use the arrival label of the flash itself as the new time. Written in those coordinates the geometry contains no bad number at that radius at all, and the quantity measuring whether a description has collapsed is as healthy there as anywhere else.
What the new chart makes visible is the thing genuinely there. At each radius the two directions a flash can take are computable, and following them inward one watches them lean over together. Far out they lean opposite ways, which is what being able to leave means. At the special radius the outward one stops leaning outward and stands still. Inside it, both lean inward. None of this is drawn; it is a pair of slopes read off an equation.
An accelerating rocket in empty space has a boundary of the same local kind, worked out in full earlier. The honest difference is that the rocket's boundary belongs to the rocket and dissolves when it stops accelerating, and this one belongs to nobody.
8 · What is actually singular
Everything so far has been about a place where the geometry is fine and the chart is not. This section is about the other place, and the whole point of the previous two sections is that the distinction is now a computation rather than a matter of taste.
8.1 · The centre, and why no chart can repair it
Return to the Kretschmann scalar (3.8.57), , and read it in the other direction:
Set the two statements side by side, because they are the chapter in miniature. At an invariant is finite, so no observer computes anything infinite there, so §7 could and did repair the chart. At an invariant is infinite, and there is nothing to repair. A change of chart is a relabelling. A scalar's value at an event does not depend on labels, by Chapter 2.4 §6. So no relabelling can make finite at . Every clever coordinate system anybody could invent has already been ruled out, in one line, without inventing any of them.
This is what the word curvature singularity means and it is all it means: an invariant built from the curvature grows without bound as an event is approached. It is not a claim about infinite density. That would be a statement about , and there is no matter here, because Chapter 3.7 solved the vacuum equations. It is a statement about the geometry alone.
8.2 · What general relativity says, and where it stops
Take the falling observer of §6.2 and follow them all the way. Equation (3.8.51) with gives a finite proper time. Section 7.3 showed that once inside they have no choice about going. And diverges on arrival, so the tidal forces that were gentle at the horizon of a large hole grow without bound as . Then the worldline ends. Not "reaches a point of infinite density", because there is no point there to reach. is not part of the manifold: a manifold is a set on which the metric is defined and smooth, and the metric is defined nowhere on .
The technical name for what has happened is geodesic incompleteness: there is a geodesic which, parametrised by its own proper time, cannot be extended past a finite parameter value. And the honest statement of the situation is short.
General relativity does not predict what happens at . It predicts that it does not know. The theory's own equations produce a place where its own equations stop making sense, after a finite and computable interval of somebody's watch, and they do so from perfectly innocuous initial data, namely a star, sitting there. That is not a paradox and not a scandal. It is a theory reporting the boundary of its own domain of validity, which is the most useful thing an incomplete theory can do, and one should trust the report precisely because the equations making it are the ones being invalidated.
What is expected to matter there is quantum mechanics. The curvature radius falls through the Planck length at a radius that is still finite, and at that scale a classical smooth metric is not a description anybody has reason to believe. Chapter 7.1 is where this book asks what replaces it, and Chapter 7.9 is where the honest accounting of the answer is done.
8.3 · Is it the symmetry's fault?
One reasonable objection remains, and it deserves an answer rather than a reassurance. Everything above was computed for a perfectly spherical vacuum solution. Perfect spheres do not occur. Perhaps the singularity is an artefact of the idealisation, in the way that the infinite field at the centre of a point charge is an artefact of pretending a charge is a point. Give the star a little rotation or a little lumpiness and perhaps the collapse misses the centre and comes back out.
It is not an artefact, and the results establishing that are quoted here rather than derived, because their proofs are global differential geometry of a kind this book does not build. The Penrose theorem of 1965 and its successors say, in outline: if a spacetime contains a trapped surface, and satisfies a causality condition, and its matter satisfies an energy condition, then it is geodesically incomplete. No symmetry is assumed anywhere.
Each hypothesis, since a theorem is its hypotheses. A trapped surface is a closed two-surface both of whose families of outgoing light rays are converging. That is exactly what §7.3 computed inside , where both radial null directions have , so the calculation you have done is the local content of the hypothesis. The causality condition forbids closed timelike curves, which is the assumption that cause precedes effect. The energy condition is the one this book has leaned on silently everywhere and never stated, so it is stated now.
The energy conditions. These are inequalities on formalising "matter has positive energy and gravity attracts". The one Penrose's theorem uses is the null energy condition: for every null vector . The related strong energy condition, used in the Hawking theorems about the past, is for every timelike , which in the weak-field limit is exactly Chapter 3.6's . That is the combination that chapter derived as the source of . So the strong energy condition is the demand that gravity be attractive, and Chapter 3.6 §6 already showed that a cosmological constant violates it, since there . That is not a technicality: it is why the singularity theorems do not forbid the accelerating universe of Chapter 3.9, and it is why the hypotheses have to be named rather than waved at. `GAPS.md` records that these conditions had gone unmentioned in this book until here.
What the theorems do and do not say. They say a singularity in the sense of geodesic incompleteness is unavoidable once collapse has passed a certain point. They do not say that diverges, that the singularity is a point, that it is spacelike, or what the geometry near it looks like. Those are all extra questions, and for the rotating solution the answers differ from the spherical case.
So the situation is this. A star more massive than any pressure can support collapses. A trapped surface forms. The theorems then guarantee that some observer's worldline ends after finite proper time. And general relativity, which predicted all of that from its own equations, has nothing to say about the ending. That is where this chapter stops, and it is where the theory stops. The next chapter takes the same equations and the same machinery and puts the largest available source on the right-hand side, which turns out to raise the same question about the beginning.
The geometry does fail in exactly one place, and by now the criterion for saying so is not a matter of opinion. The invariant that stayed finite at the surface everyone calls the edge of a black hole runs away to infinity at the centre, and being an invariant it does so in every description at once. No cleverer chart is available, because the failure is not in the labels.
What should be said about it is less than people expect. The theory does not describe what happens there. It stops, in the specific sense that the paths of falling matter end after a finite interval of their own time with no continuation the equations can supply. That is not a paradox and not a scandal; it is a theory reporting the edge of its own domain, which is the most useful thing an incomplete theory can do, and its own equations are what report it.
Nor is it a peculiarity of the perfectly round solution read here. Theorems exist showing such an ending is forced under conditions assuming no symmetry whatever, and those conditions include an assumption about matter this book has leaned on silently everywhere and never named. Naming it is the last thing done, because knowing which assumption carries the weight is the difference between a result and a slogan.
9 · Worked examples
Somebody falls radially from rest far away into (a) a hole and (b) the hole at the galactic centre. For each, compute the proper time from crossing the horizon to reaching , the head-to-toe tidal stretch at the moment of crossing, and the radius at which that stretch reaches across a body. Then say which of the two is survivable at the horizon and why.
The tools. Three formulas, all derived above: (3.8.51) for the proper time, (3.8.59) for the tidal stretch at the horizon, and Chapter 3.4 §4's radial geodesic deviation for a general radius. Use .
(a) Ten solar masses. and
Proper time from to , from (3.8.51) with and , which collapses to :
Tidal stretch at crossing, from (3.8.59) with :
And the radius at which :
which is . You are destroyed a hundred and twenty-four horizon radii out, long before the horizon exists as an issue.
(b) Four-point-three million solar masses. , so , which is about a fifth of Mercury's orbital radius. Then
which is . That is a stretch of a tenth of a millimetre per second squared across a human body, which is to say nothing at all. The radius is , which is inside the horizon by a factor of .
The comparison, and why. For (a) the tidal limit is far outside the horizon, and for (b) it is far inside. The reason is (3.8.59)'s . Multiplying the mass by divides the tidal field at the horizon by , and the two numbers in the table of §6.5 differ by exactly that. So an observer falling into (b) crosses the horizon with no local indication whatsoever, has just under half a minute of ordinary experience, and is destroyed only in the last fraction of a second, when has fallen to . By that time, from (3.8.51), has left to run.
What this example is really about. Nothing in it required §7. The whole calculation is §6.2 plus §6.5, and it demonstrates the chapter's central claim numerically: the horizon is not where anything happens to you. What is special about it is a fact about the light cones, which no local measurement of tidal force can detect.
A radio source is observed with a ray grazing the Earth's limb. (a) Compute the deflection from (3.8.21). (b) Compute what Chapter 3.1's cabin argument would have given, and confirm the ratio is . (c) A spacecraft passes at with impact parameter . Compute its deflection from (3.8.35), identify how much of it is the spatial curvature, and say why the impact parameter had to be enlarged. Use and .
(a) Straight into (3.8.21):
which is arcseconds, or . Small, and about ten times the shadow of the galactic-centre hole computed in §2.2. So it is measurable with the same instruments, and in fact this deflection has to be modelled in any radio measurement made from the ground.
(b) Chapter 3.1's answer is (3.8.27), which is half: . The ratio is , as (3.8.36) requires with .
(c) First the reason for enlarging , because it is the more instructive half of the question. Everything in §§3 and 4 was first order in the perturbation, and inspecting (3.8.35) shows the expansion parameter is , not . At and that parameter is , which is not small at all, so the formula does not apply. A slow body grazing a mass is strongly deflected, as the orbit of any comet shows. Taking brings the parameter to and the expansion is safe. Then
A light ray with the same impact parameter would be deflected by , so the spacecraft is turned about times more sharply, because it moves slowly and so lingers in the field. Meanwhile the fraction of its deflection contributed by the curvature of space is
The moral, in one line. The same geometry deflects both objects and the same two halves of it are present in both calculations. The ray gets equal contributions from the two. The spacecraft gets one part in from the spatial half. There is no separate "Newtonian" and "relativistic" bending. There is one metric, and how much of each part of it a body samples depends entirely on how fast it is going.
10 · Your turn
Problem 1 — the photon sphere, and a light ray that goes round twice
(a) Show from (3.8.8) that a ray with slightly greater than has its turning point slightly outside , and find the turning radius to first order in . (b) Near the peak, write and show that obeys for small , so that grows exponentially in . (c) Deduce that the number of times such a ray circles the hole grows like , and find the needed for two complete circuits. (d) Say in one sentence what this implies about the appearance of the bright ring around a black hole.
Solution
(a) A turning point has , so . Expand about its maximum: , since there. And . Equating, , and with and from §2,
so . Outside, as claimed, and the square root is the signature of a quadratic maximum.
(b) Use the orbit form. From (3.8.14), . Differentiating gives (3.8.16), . Put with , and note that , confirming that solves it. Keeping first order in ,
Since to first order, as well. The solutions are , which is exponential growth in the angle. That is the analytic statement that the circular photon orbit is unstable, matching the sign of .
(c) A ray comes in with of order one, shrinks to its minimum , and grows back out. Since changes by a factor per radian, each leg takes an angle set by , that is . There are two legs, inward and outward, so the total angle swept near the peak is plus a constant. The number of circuits is that divided by ,
so two circuits needs , that is of order . The additive constant is not fixed by this argument, so read the number as an order of magnitude. The logarithm is the exact part, and it is the part that matters.
(d) Because the required falls exponentially with the number of circuits, the images formed by rays that go round once, twice, three times… pile up on top of one another within an exponentially thin annulus just outside : the ring is a stack of infinitely many increasingly faint images of the whole sky, and its outer edge is at to a precision far beyond anything measurable.
Problem 2 — Shapiro delay, from the same metric
A radar pulse is sent from Earth past the Sun to a reflector on another planet and back. The extra round-trip time compared with flat spacetime is a third classical test, and it comes out of §7.1's tortoise coordinate with almost no new work. (a) For a ray moving radially, show from (3.8.60) that the coordinate time to travel from to is . (b) Using (3.8.61), show that the excess over is , and simplify for . (c) The full non-radial calculation replaces the logarithm's argument by for a ray passing at impact parameter . Take that as given and compute the round-trip excess for a signal from Earth to a spacecraft on the far side of the Sun, with , and . (d) Why is this a test of the same half of the metric as §4.2, or of both halves?
Solution
(a) Equation (3.8.60) gives for an outgoing ray, and by definition, so and integrating gives .
(b) From (3.8.61), , using so that the in the denominators cancels in the difference. Subtracting the flat-space time leaves for .
(c) One way, . With , :
whose logarithm is . So one way, , and the round trip is seconds. That is about , which is of light travel and enormously larger than the timing precision of a radar system.
(d) Both halves. The tortoise coordinate (3.8.61) was built from , and that single factor came from (3.8.60), which used and together, since the null condition sets one against the other. Running the exercise on §4.2's mutilated metric gives and hence half the coefficient, exactly as for the deflection. The Shapiro delay is the same factor of two, measured with a clock instead of a protractor.
Problem 3 — a chart that is worse, and one that is better
(a) Show that the outgoing coordinate of (3.8.62) gives the metric , and compute its determinant. (b) Compute the two radial null slopes with , and show that inside both are positive. What kind of region does this chart describe? (c) Show that the Schwarzschild chart's determinant, (3.8.48), is the same as the Eddington–Finkelstein one, (3.8.65), and explain why that had to be so given that the transformation has unit Jacobian. (d) Using (c), say precisely why the finiteness of in the Schwarzschild chart was suggestive but not decisive in §6, whereas the finiteness of was decisive.
Solution
(a) , so and . Subtracting leaves . The block is with determinant , so again, which is also regular at .
(b) Null: , so or . With , , the first branch gives and the second gives . Inside that is positive, so both slopes are positive: every future-directed ray moves outward. This chart covers a region from which everything must escape and into which nothing can fall, which is the time reverse of a black hole. It is a legitimate solution of the field equations, which are time-symmetric, and it is not what forms when a star collapses.
(c) Both are . Under a change of chart, with the Jacobian determinant of the transformation. Here , , so the Jacobian matrix is , triangular with unit diagonal, so and the determinants must agree.
(d) Because is not a scalar. It transforms with , and a transformation with singular Jacobian can make it vanish or diverge without anything happening to the geometry, exactly as happens on the polar axis. Its finiteness in one chart therefore proves nothing on its own, and it was worth noting in §6 only as a hint that the zero of and the pole of were conspiring. is a scalar: its value at an event is chart-independent outright, so finiteness in any one chart is finiteness in all.
Problem 4 — how deep a well before the redshift is not a small correction
(a) A spectral line emitted at the surface of a static star of mass and radius is observed far away. Using (3.8.39), define the redshift and express it in terms of . (b) At what does reach ? ? ? (c) A static star cannot be arbitrarily compact: ⚑ take as given the Buchdahl bound , which follows from requiring the pressure at the centre to be finite for any fluid whose density does not increase outward. What is the largest surface redshift a static star can show? (d) The neutron star of §5.3's table has . Compute and compare with (c). What would it mean to observe a line with from a stellar surface?
Solution
(a) Wavelength is inversely proportional to frequency, so , and
(b) Invert: . For , so . For , so . For , so . Note how fast it runs: a factor of ten in costs less than a factor of nine in the first time and only a factor of four the second, because the square root is turning over.
(c) gives , so . The maximum surface redshift of any static star is .
(d) , so . Comfortably below the bound, as it must be. A line at would require , that is . That is still above the Buchdahl bound of , so it is permitted, but only just, and it would place extremely tight constraints on the equation of state of matter at that density. Anything with from a static surface would mean either that the object is not static or that one of the bound's hypotheses fails.
Light needs one number. The null normalisation removes proper time as a parameter and leaves an affine parameter fixed only up to scale, so the two Killing charges of Chapter 3.7 are individually meaningless and only the ratio survives. One equation, (3.8.7), then contains everything: its effective potential has a maximum at , giving an unstable circular orbit for light and a critical aim that separates capture from escape and sets the apparent size of the dark region.
The deflection, and the debt. The orbit equation is , with no Newtonian source term at all. Its unperturbed solution is a straight line and its first correction gives , that is at the solar limb. Then §4 did what Chapter 3.1 could not: keeping only gives exactly , keeping only gives exactly , and the two first-order solutions sum term by term to the full one. Repeating for a body at speed gives , which is why a planet samples one half and a ray samples both. The register of debts is closed.
Clocks. , derived from a conserved Killing charge against a local clock, from crest-counting with , and as the weak-field limit of Chapter 3.1's cabin. The satellite-navigation numbers, which are , , per day and , come out of one square root rather than two effects glued together, with the neglected cross term computed at four femtoseconds a day.
And the horizon. At the components misbehave and the geometry does not. The proper distance down is finite. The proper time to cross is finite while the coordinate time is logarithmically infinite. And the invariant is finite there, which no change of chart can alter. Eddington and Finkelstein's chart, built by integrating the radial null condition, is manifestly regular with , and in it the two radial null slopes are and : the second passes through zero at and is negative inside, so both future directions point inward. That is a one-way surface, computed rather than drawn. Chapter 2.3's Rindler horizon has the same local character and a different global status, since it belongs to an observer and this one belongs to nobody. And diverges at exactly one place, , where general relativity stops, ⚑ irremovably, by theorems whose hypotheses include an energy condition this book had never named.
Where this gets spent. Chapter 3.9 puts the largest available source on the right-hand side and finds that the universe has no timelike Killing vector. So §5's whole method, which was to get a conserved charge from a symmetry, has nothing to work with, and energy is not conserved. Its redshift is computed from null geodesics exactly as §1 set them up, but with the scale factor in place of . The horizons it finds are integrals of the same null condition §7.1 integrated here. And Chapter 7.9 returns to with the one number this book leaves deliberately unpaid: the entropy of the surface §7 has just built, which is proportional to its area and which nothing in Part III can explain.