An error budget added in quadrature
Assumes Seeing and Refraction.
A telescope’s resolution is a property of its aperture, and the atmosphere destroys it. Adaptive optics restores some of it by measuring the incoming wavefront and deforming a mirror to cancel the distortion — but it never restores all of it, and the natural question is how much.
The answer is not a resolution. A corrected image has a diffraction-limited core sitting on a broad halo, and the useful figure of merit is what fraction of the light is in the core. That fraction is the Strehl ratio, and it is set by a sum of independent variances.
That framing is unfamiliar to anybody who has met the seeing as a limit on resolution, where the answer is an angle. The difference is real rather than presentational: an uncorrected image is a blur with a width, and a corrected one is two things at once — a sharp core and a wide halo — so a single width describes neither.
Why a sum of squares, and why an exponential
Two facts do all the work.
Wavefront errors from independent causes add in quadrature. The residual phase after correction is a sum of contributions — what the mirror could not fit, what the servo could not keep up with, what the sensor could not see for lack of photons — and these are statistically independent, so their variances add. That is why an error budget is a list of squares rather than a list of magnitudes, and why the largest term dominates so completely: a term half the size of another contributes a quarter as much.
The Strehl ratio is the exponential of minus that variance. For small residuals the fraction of light remaining in the core is , with measured in radians of phase. The approximation is Maréchal’s, it is good to a few per cent up to , and it is what turns a wavefront specification into an image quality.
The second fact carries the whole of the wavelength dependence, and it is worth stating slowly because it is the single most important thing about the subject. A wavefront error is a path difference, measured in nanometres, and it does not depend on the observing wavelength at all. The phase error is that path divided by the wavelength. So the same physical mirror error is four times smaller in phase at two microns than at half a micron, and the Strehl — the exponential of the square — is transformed out of all recognition.
A system with two hundred nanometres of residual delivers a Strehl of a few per cent at 550 nanometres and about seventy per cent at 2.2 microns. That is the reason adaptive optics was an infrared technique for twenty years, and it is not a statement about detectors or about the atmosphere.
The five terms, and what each one is a failure of
Fitting error is the part of the wavefront the deformable mirror cannot represent. A mirror with actuators spaced apart cannot correct structure finer than , and the atmosphere has structure at every scale. The residual scales as , where is the atmospheric coherence length, and since scales as the fitting error in nanometres is nearly wavelength-independent — the mirror does not know what colour it is reflecting.
Servo lag is the part that changed between measuring and correcting. The air moves across the aperture at the wind speed, so the wavefront decorrelates on a timescale , which is a few milliseconds. A loop running at a kilohertz has a delay of a millisecond or two, and the residual scales as .
Sensor noise is the part that could not be measured because there were not enough photons. It rises as the guide star gets fainter and is the term that decides how much of the sky a system can work on.
Anisoplanatism is the part that is wrong because the guide star is not the target. The atmosphere is a stack of layers at different heights, so two lines of sight separated by an angle pass through different air above some altitude, and the correction measured on one does not apply to the other. The residual scales as , with the isoplanatic angle.
Calibration and non-common path is the part that is wrong because the wavefront sensor and the science camera do not see the same optics. The sensor is corrected to flat, and flat at the sensor is not flat at the detector. This term is static, it does not average away, and it is the one that limits high-contrast work.
It is worth noticing which of the five are properties of the atmosphere and which are properties of the instrument, because the distinction decides what money can buy. The fitting error, the servo lag and the anisoplanatism all scale with atmospheric parameters and can be reduced by building faster, denser and more expensive hardware — they are engineering. The sensor noise depends on the guide star and is therefore a property of the sky rather than of the system. And the calibration term is neither: it is a static error in the optics that could in principle be removed entirely and never quite is, because measuring it requires knowing the wavefront at the science detector, where there is no wavefront sensor. That last term is the one that limits the direct imaging of faint companions, where the enemy is not the width of the core but the speckles in the halo, and a static speckle is indistinguishable from a planet.
One more consequence of the five-thirds powers deserves attention, because it is what makes the subject so unforgiving. Three of the terms scale as a ratio raised to five-thirds, which is steeper than linear: halving the actuator spacing reduces the fitting error by a factor of three, and doubling the guide-star separation raises the anisoplanatic error by the same factor. Steep scalings cut both ways. They are why modest engineering improvements produce large gains when a system is close to working, and why a system that is a factor of two away from working is much further away than it looks. The requirement that the correction outrun the air is the same exponent seen in the time domain rather than the spatial one.
What the laser fixes and what it costs
The sky-coverage problem is solved by making a guide star: a laser tuned to the sodium D lines excites a layer of meteoric sodium ninety kilometres up, producing an artificial source wherever it is pointed.
That removes the requirement to find a bright star nearby, and it introduces two new terms.
The cone effect. A beacon at ninety kilometres is not at infinity, so the light returning from it samples a cone rather than a cylinder of atmosphere. The edges of the telescope’s aperture see air the beacon’s light never passed through, and the mismatch grows with aperture diameter. For an eight-metre telescope in the visible the residual is comparable to everything else in the budget; for a thirty-metre telescope a single beacon is useless, and the correction requires several lasers whose returns are combined tomographically.
Tip-tilt indeterminacy. The laser beam is refracted on the way up by the same atmosphere that refracts it on the way down, so the beacon appears fixed relative to the outgoing beam and carries no information about the overall tilt of the wavefront — which is the largest single component of the distortion. A natural star is still required for the tilt, but it can be much fainter, and it can be much further away, because tilt is coherent over a larger angle.
The two new terms illustrate a pattern worth naming: a fix for one term in a budget almost always adds terms of its own, and whether it is a gain depends on the arithmetic rather than on the idea. A laser guide star trades a sky-coverage limit for two wavefront errors, which is a clear gain for an eight-metre telescope in the infrared and a marginal one in the visible. The same trade run on a thirty-metre telescope is a loss with one laser and a gain with five, which is why the number of lasers on the coming generation of telescopes is a computed quantity rather than a choice.
It also explains why the technique arrived when it did. The idea was published in 1953 and the first astronomical systems ran in the late 1980s, and what changed in between was not the optics but the ability to read a detector, compute a correction and drive a mirror inside a millisecond. The same constraint appears wherever a correction has to outrun the thing it is correcting, and the atmosphere is an unusually demanding case because its timescale is set by wind speed rather than by anything that can be slowed down.
What was actually measured
The budget is not a theory; every term in it is measured, and the measurements are what the systems are commissioned against.
The total is measured from the image. Photometry of the corrected point spread function against a model gives the Strehl directly, and inverting Maréchal’s approximation gives the residual wavefront in nanometres. The number agrees with the sum of the individually measured terms, which is the check that the budget is complete — an unexplained excess means a term nobody listed.
Individual terms are measured by switching them off. Running the loop faster and extrapolating to zero delay isolates the servo lag; observing a guide star at a sequence of separations from a target measures the anisoplanatism directly and returns the predicted five-thirds power; observing progressively fainter guide stars measures the noise term.
And the wavelength scaling is verified across the same image. An instrument observing simultaneously in two bands measures two Strehls with one wavefront, and the ratio has to be what the exponential requires. It is, to a few per cent, across the range where the Maréchal approximation holds.
Where the picture stops
There are three, and the second is the reason very large telescopes need a different architecture.
Maréchal’s approximation fails when the correction is poor. At greater than about one the exponential understates the Strehl, and at very low correction the concept stops being useful — there is no core to speak of, and the image is characterised by its width instead. The budget is a good tool exactly in the regime where the system is working.
The terms are not all independent. Quadrature addition assumes independence, and the servo lag and the sensor noise are anti-correlated in a specific way: running the loop faster reduces the lag and increases the noise, because there are fewer photons per frame. Optimising a system means finding the minimum of the sum, which sits at a loop speed that depends on the guide star’s brightness — so the optimum configuration is different for every target.
And the whole framework describes correction on one axis. For a thirty-metre telescope the isoplanatic patch is a small fraction of the field, and single-beacon correction is not the technique; the architectures under construction measure the atmosphere in three dimensions with several beacons and correct with several mirrors conjugated to different altitudes. The budget becomes a different list, and the dominant term becomes the tomographic error — the part of the atmosphere the beacons did not sample.
A fourth limit is worth adding because it is about what the number means to a user rather than to a builder. A Strehl ratio describes the fraction of light in the core, and two systems with the same Strehl can have quite different halos — one smooth, one full of static speckles — and be worth completely different amounts for a given science case. For photometry the Strehl is nearly the whole story; for astrometry what matters is the stability of the core’s position; for high contrast what matters is the structure of the halo at a few tenths of an arcsecond, which the Strehl does not describe at all. A single-number summary of an image is a compression, and which information it discards depends on what the image is for.
Why a budget rather than a specification
There is a general point about instruments hiding in this, and it is worth extracting.
A budget of independent variances has a very particular character: it is dominated by its largest term, it is insensitive to its smallest ones, and it therefore rewards effort applied in exactly one place. An engineering programme that improves four terms by twenty per cent each and leaves the largest untouched achieves almost nothing. One that halves the largest term achieves a great deal, until the second-largest becomes the largest and the programme has to move.
That structure is why the first question about any such system is which term dominates, and why the answer changes with the observing conditions, the target’s brightness, the wavelength and the airmass. There is no single figure of merit for an adaptive-optics system, and a Strehl ratio quoted without the conditions is not a number.
The same shape appears wherever independent errors combine. The residual wavefront that limits an interferometer is a budget of this kind; so is the noise in a photometric measurement, where the sky, the source and the detector each contribute a variance and the largest decides the exposure time. Recognising the shape is worth more than remembering any particular budget, because it says immediately where to look.
Close on the number that makes the whole subject possible, because it is not obvious that it should. The atmosphere’s coherence length in the visible at a good site is about ten centimetres and its coherence time a few milliseconds, and the number of independent patches across an eight-metre aperture is a few thousand. A system that measures and corrects a few thousand numbers a thousand times a second is a substantial machine and is not an absurd one. Had the coherence length been a centimetre — which it is at a poor site — the count would have risen a hundredfold and the technique would have been impossible for a generation. The parameters of the Earth’s atmosphere sit, by no design, just inside what electronics can keep up with. Every other correction in this collection is applied from a model; this one is measured and applied a thousand times a second, and that is only feasible because of an accident of scale.
There is one property of a quadrature budget worth stating plainly because it governs where effort should go, and it is routinely violated. When one term dominates, halving it is worth far more than eliminating any other — and when several terms are comparable, halving one buys almost nothing. A budget with terms of 100, 90 and 80 nanometres has a total of 156; removing the largest entirely leaves 120, a gain of a quarter for the elimination of a third of the budget. So the correct question is never “which term can be improved” but “is any term dominant”, and a well-balanced budget is a sign that the system has been optimised and that no cheap improvement remains. A badly balanced one is an opportunity. Read a published budget with that in mind and it says immediately whether the instrument was designed or assembled.
A last practical note. Because the delivered fraction is an exponential of the budget, the same system’s performance varies enormously with the observing wavelength: a wavefront error that is a small fraction of a wave in the infrared is a large fraction in the visible, and the exponent turns that into the difference between a working instrument and a useless one. The hardware does not change; the ratio of the error to the wavelength does.
Where the ladder goes next
The natural next rung is tomography: how several beacons at different positions are combined into a three-dimensional estimate of the atmosphere, and why that estimate degrades with the number of layers rather than with the number of beacons. The rung after it is the regime beyond correction — what happens when the residual is used rather than removed, which is what speckle and interferometric methods do with the same broken wavefront.
What links here
Essays that link to this one from their own argument.
The objects this essay names
Each one links to every other essay that touches it.
Adaptive opticsAnisoplanatismCone effectDiffraction limitError budgetFitting errorLaser guide starServo lagStrehl ratioWavefront error