The direction a photon count throws away
Assumes Extinction, Photometric systems and Photon noise.
Light arriving from a star carries four independent quantities. A photometer measures one of them.
The other three describe the orientation of the electric field’s oscillation — whether it prefers one plane, whether it rotates, and how strongly. They are usually small, they are difficult to measure, and they survive the journey untouched by anything that merely attenuates. Where a brightness is degraded by distance, dust and atmosphere, a polarisation is not: it is a ratio, and the things that dim the light dim both components of the ratio together.
That makes polarimetry the only channel in optical astronomy that carries a geometry rather than a scalar, and there are objects for which it is the only channel at all.
What the four numbers are
Stokes’ 1852 parameterisation is the one that survived, because its components add when incoherent beams are combined. Written as intensities:
The linear polarisation is and its position angle is . The factor of a half is not a convention: a rotation of the analyser by 90° exchanges ’s two intensities, which is a rotation by 180° in the plane. Polarisation has no sign and no direction, only an orientation, and everything odd about handling it follows from that.
The circular component is almost always negligible in astronomy and almost always the interesting one where it is not — it appears in the Zeeman-split lines of magnetic stars and in the synchrotron emission of a few sources, and measuring it well enough to be believed is harder than everything else in this essay put together.
Where the polarisation comes from
Two mechanisms produce nearly all the linear polarisation seen in optical astronomy, and they have opposite geometries.
Scattering polarises the light perpendicular to the plane containing the source, the scatterer and the observer. For a particle small compared with the wavelength the degree of polarisation is
exactly zero in the forward and backward directions and exactly one at a right angle. Neither of those is a fitted number: they are what a dipole radiator looks like end-on and side-on.
Dichroic absorption by aligned grains polarises the light along the direction the grains are aligned by, which is the magnetic field. Interstellar grains are elongated, they spin, and they align with their long axes perpendicular to the field — so they absorb preferentially the component of the electric field along their long axis, and what gets through is preferentially the component along the field.
The mechanism of the alignment took fifty years to settle and is still not entirely closed. Davis and Greenstein proposed paramagnetic relaxation in 1951; the modern account is radiative torques, in which an irregular grain in an anisotropic radiation field is spun up to suprathermal rotation and its angular momentum then aligns with the field. What matters for the observation is that the alignment happens, that it happens with the long axis perpendicular to B, and that the transmitted polarisation is therefore parallel to it.
A map of a field nobody can see
The Galactic magnetic field is a few microgauss. It threads the whole disc, it confines the cosmic rays, it resists the collapse of molecular clouds, and it is invisible to every direct measurement optical astronomy has.
Polarimetry of background stars maps it. Each star gives one number — the position angle of the polarisation its light acquired crossing the intervening dust — and that angle is the field direction projected on the sky, averaged along the line of sight and weighted by the dust. Thousands of such measurements make a vector field.
The picture that came out, first from Hiltner and Hall in 1949 and in modern form from Planck’s submillimetre polarimetry, is a field running predominantly along the Galactic plane, with a coherent large-scale component and a turbulent one of comparable strength. The ratio between those two is a measurement of how magnetically dominated the interstellar medium is, and it is obtained by dispersion of the position angles rather than by measuring any field. The peak of the Serkowski curve is the one number in it that carries physics, and moving it is the whole of what the law is used for.
A shape from an unresolved point
The second use of polarimetry is the one that produces information no telescope can otherwise obtain.
A spherically symmetric source that scatters its own light produces no net polarisation, however strongly each individual photon is polarised: the contributions from opposite sides of the source cancel exactly. Break the symmetry and they no longer cancel, and the residual is a measure of the departure from spherical.
For a supernova, that is the whole of what is known about its geometry. A supernova at ten megaparsecs subtends a few microarcseconds; nothing resolves it, and nothing will. What is measured is a continuum polarisation of a few tenths of a per cent, and the modelling that turns it into an axis ratio is straightforward enough to be believed: a per cent of continuum polarisation corresponds to an asphericity of about ten per cent.
The results have been decisive. Core-collapse supernovae are systematically aspherical, and the asphericity increases as the ejecta thin and deeper layers are exposed — so the explosion mechanism itself is aspherical rather than the envelope being disturbed on the way out. Type Ia supernovae are nearly spherical in the continuum, at a few tenths of a per cent, which is a constraint on their explosion models that no other observation supplies. There is a third use, smaller in scope and unusually direct. A planet reflecting its star’s light is polarised by the scattering in its atmosphere, and the polarisation varies through the orbit exactly as the scattering angle does — zero at full phase, maximum near quadrature. That is a signal whose shape in time is fixed by geometry alone, and it can be recovered from a light curve in which the planet is never separated from the star.
Why it is rare
Every polarimetric measurement is a difference of two nearly equal counts, and that is the whole of the difficulty.
To measure a polarisation of one per cent to a tenth of its own value requires distinguishing two intensities that differ by one part in a hundred, to one part in a thousand. From counting statistics alone that needs photons; in practice it needs far more, because the systematic errors are larger than the statistical ones and they do not average down. The instrumental answer is to make the measurement differential in as many ways as possible. A dual-beam polarimeter splits the incoming light into two orthogonal polarisations and records both simultaneously on the same detector, so a fluctuation in transparency affects both equally and cancels in the ratio. Rotating a half-wave plate through four positions and combining the eight resulting beams in the right order cancels, to first order, both the difference in the two channels’ efficiencies and the difference between the two detector regions. The instrument is designed so that the quantity of interest survives four cancellations and the systematics do not.
That is why polarimetry is done on a minority of telescopes by a minority of observers, and why a polarimetric result is generally presented with a longer description of its calibration than of its interpretation.
The bias, and how it is handled
Return to the positive bias, because it is the part of polarimetry most often got wrong and it has a clean structure.
The measured is the length of a two-dimensional vector whose components carry Gaussian errors. Its distribution is the Rice distribution, and its mean exceeds the true value for every true value. At high signal-to-noise the excess tends to , which is negligible. At zero true polarisation the mean is , which is not.
The practical consequences are worth stating as rules, because they are the difference between a result and an artefact.
A measurement with is not a detection of polarisation. It is consistent with zero, and quoting the measured as a value overstates it.
Debiasing estimators exist — the simplest subtracts in quadrature — and they are useful in the middle range and misleading at the bottom, where they return zero for a genuine small polarisation as readily as for none.
And the position angle is well behaved where the polarisation is not: has no bias at all, because the two components’ errors are symmetric about the truth. A weak polarisation whose angle is consistent across four measurements is more convincing than a strong one measured once, and the angle is the quantity most of the astrophysics is carried by. Both applications share a difficulty that is worth naming once rather than twice: what is measured is an integral along a line of sight, so a polarisation is a weighted average over everything between here and the source, and untangling a foreground from a signal is the same operation in both cases.
Where the model stops
Three limits, and each is a live research question rather than a settled correction.
The alignment efficiency of grains is not known independently. The Serkowski law is empirical; the constant relating its peak to is a fit; and the fraction of grains aligned varies with environment, falling in dense cloud interiors where the radiation that drives the alignment cannot penetrate. So a polarisation map of a molecular cloud under-represents the field exactly where the field matters most.
The scattering geometry is degenerate in a way that a single measurement cannot lift. The same net polarisation can come from a moderately aspherical source seen edge-on or a strongly aspherical one seen obliquely, and separating them requires either a time series — as the geometry changes — or a wavelength dependence.
And the interstellar contribution has to be removed before anything intrinsic can be claimed. That is done by measuring nearby stars along the same sight line, which assumes the dust between here and them is the same dust as between here and the target. For a supernova in another galaxy that assumption spans two interstellar media and an intergalactic one, and it is the single largest uncertainty in most published supernova polarimetry.
Two applications are worth setting out before the ladder, because they are where the technique is currently doing the most work — one on a field that fills the Galaxy, and one on a signal from before there was a galaxy.
Two more settings are worth putting beside those, because each shows a different way the measurement fails.
Three measurements, three components
The essay has described one probe of the interstellar magnetic field, and it measures one thing: the direction of the field projected on the plane of the sky. A field is a vector, and getting the other components requires two further techniques that share nothing with this one.
Faraday rotation measures the line-of-sight component. A linearly polarised radio wave passing through a magnetised plasma has its plane of polarisation rotated, by an amount proportional to the wavelength squared and to the integral along the path of the electron density times the field component along the line of sight. Observing a background source at several wavelengths and fitting the rotation against wavelength squared gives that integral.
What it returns is a product of two unknowns — the field and the electron density — so converting it to a field strength requires a separate measurement of the density, usually from the dispersion of a pulsar’s pulses along the same sight line. Combining the two gives a density-weighted mean field along the path, with a sign: toward the observer or away.
The Zeeman effect measures the line-of-sight component again, and directly. A spectral line formed in a magnetic field is split, and the splitting is proportional to the field; for the twenty-one-centimetre line of neutral hydrogen the splitting is far smaller than the line width, so what is measured is not a split line but a small circular polarisation whose profile is the derivative of the line’s. It is one of the hardest measurements in radio astronomy and it is the only one that returns a field strength with no other quantity multiplied into it.
So the three techniques divide the vector between them: dust polarisation gives the plane-of-sky direction, Faraday rotation gives the line-of-sight component weighted by electrons, and Zeeman gives the line-of-sight component weighted by neutral hydrogen. Each is sensitive to a different phase of the medium, so combining them is not a matter of adding components of one field — it is a matter of comparing fields measured in different gas.
The three-dimensional field is assembled from measurements that do not overlap, and the assembly is the least secure part of every published field map.
The division above also explains why field strengths quoted for the interstellar medium vary so much between papers: they are frequently measurements of different phases by different techniques, presented as though they were the same quantity.
The same scattering, at recombination
The mechanism at the start of this essay — scattering polarises light perpendicular to the scattering plane — has a second application at a completely different scale, and it is the one the largest experiments in cosmology are built for.
At the moment the universe became transparent, the last photons to scatter off free electrons were scattered from a radiation field that was not perfectly isotropic. A quadrupole anisotropy in the incident radiation produces a net linear polarisation in the scattered light, for exactly the reason a scattering angle of ninety degrees produces one: the two orthogonal components of the incident field are not equal.
So the microwave background is polarised, at a level of a few per cent of its temperature anisotropies, and the pattern carries information the temperature alone does not.
The pattern is decomposed into two kinds. One has no curl and is produced by density perturbations, which is most of what is there. The other has a curl and cannot be produced by density perturbations at all, so its presence would be evidence for something else — gravitational waves from the very early universe being the sought-after source.
The difficulty is the foreground, and the foreground is the subject of this essay. Aligned interstellar grains in this galaxy emit polarised light in the far infrared and microwave, at exactly the frequencies the background is measured at, with a polarisation fraction of several per cent and a pattern that contains both kinds.
Separating them requires observing at several frequencies and exploiting the different spectra — the background’s is a blackbody and the dust’s is a modified one — which is why every experiment of this kind is multi-frequency by design.
A claimed detection of the curl component in 2014 turned out, on the release of better dust maps, to be consistent with Galactic dust alone. The measurement that would say something about the first fraction of a second is limited by the alignment of grains a few hundred parsecs away, and improving it means understanding the dust rather than the background.
Where this ladder goes next
This rung establishes what a polarisation is, the two mechanisms that produce one, and the two failure modes — instrumental and statistical — that make it hard.
Above it lies spectropolarimetry, where the polarisation becomes a function of wavelength and the loops traced in the Stokes plane across a spectral line carry the geometry of the line-forming region.
Beside it lies the circular component, which is nearly always the Zeeman effect and which measures the line-of-sight field strength directly rather than the plane-of-sky direction — the two together giving a three-dimensional field from a measurement made on a point.
And below it, as a caution worth its own rung, lies the whole business of what a small measured quantity means when its estimator cannot be negative. That is not a fact about polarisation; it is a fact about lengths, and it recurs wherever a modulus is measured — in a proper motion, in a lens magnification, in an amplitude of any kind.
What this makes readable
Essays that name this one as a prerequisite.
- A direction measured by something with no strength in it galaxies
- A field strength read off a line that will not split starlight
- A magnetic clock read off a butterfly stars
- A minimum that was mistaken for a principle starlight
- An instrument more polarised than the sky starlight
- A slope that needs no source starlight
- A spectrum with no temperature in it starlight
- Two instruments blind in opposite directions starlight
About the same objects
Not linked from either essay — found by the objects both name.
- The dust is not lost light, it is moved light grain alignment · interstellar dust
What links here
The 8 of 11 essays linking to this one that name the most of the same objects.
- A direction measured by something with no strength in it galaxies
- An instrument more polarised than the sky starlight
- A field strength read off a line that will not split starlight
- A minimum that was mistaken for a principle starlight
- A slope that needs no source starlight
- Two instruments blind in opposite directions starlight
- A one-per-cent distortion, and a million galaxies to see it galaxies
- A ratio that is an energy and a distance cosmology
The objects this essay names
Each one links to every other essay that touches it.
AsphericityGrain alignmentInterstellar dustMagnetic fieldPolarimetric biasPolarisationPosition angleReflection nebulaRice distributionScatteringSerkowski lawStokes parameters