An instrument more polarised than the sky
Assumes Polarimetry and Photometric systems.
Polarimetry measures a fraction of a per cent. Interstellar dust polarises starlight by a few per cent at most and usually much less; a scattering atmosphere gives tenths of a per cent; the alignment signal from a magnetic field in a molecular cloud is at the same level. These are small numbers by the standards of any other astronomical measurement, and they are not small compared with what the telescope contributes.
An oblique reflection from a metal surface reflects light polarised parallel to the plane of incidence differently from light polarised perpendicular to it. Every mirror at a non-zero angle of incidence therefore polarises, and a telescope with a tertiary mirror or a Nasmyth focus can add a per cent or more.
Polarisation is the piece of information a photon count throws away, and recovering it means measuring a difference between two intensities that are nearly equal. That structure is what makes the technique powerful and what makes it vulnerable: a difference of two large numbers inherits every systematic that affects the two unequally.
Why it is a vector and not a scale
The single most important structural fact about the problem is that instrumental polarisation adds rather than multiplies.
Polarisation is described by two numbers, conventionally and , which are differences of intensities measured in perpendicular directions. They combine linearly: light that is the sum of two beams has Stokes parameters that are the sum of theirs. So an instrument that adds its own polarised component adds a fixed vector to whatever arrived.
The degree of polarisation and the position angle are the modulus and the argument of that vector. Adding a vector changes both. An instrumental contribution of half a per cent applied to a source with half a per cent of its own can double the measured degree, halve it, or rotate the angle by ninety degrees, depending on the relative directions.
This is why a photometric-style calibration is useless here. Determining that the instrument transmits ninety per cent of the light corrects a scale; nothing about a scale correction touches an additive offset, and the offset is the whole of the problem.
There is a second structural fact that follows from the first and is worth stating separately. Because the offset is a vector, its effect on the measured degree of polarisation is not symmetric: a source whose own polarisation happens to point along the instrumental vector is measured too high, and one pointing against it is measured too low, and one at right angles has its angle rotated with its degree barely changed. So a population of sources with random position angles, measured through an uncorrected instrument, acquires a correlation between degree and angle that no physical mechanism would produce. That correlation is the diagnostic: a plot of measured degree against measured angle for a sample that should have no such relation is the cheapest test of whether the correction is working.
It is worth putting the numbers side by side once. A Cassegrain focus with a symmetric optical train contributes of order a hundredth of a per cent; a Nasmyth focus with a flat tertiary at forty-five degrees contributes several tenths of a per cent, and can exceed a per cent in the blue where aluminium’s reflectivity is most asymmetric. Interstellar polarisation of a star a kiloparsec away is a per cent or two; of a nearby star, a hundredth. Scattering in a circumstellar disc gives a few per cent in the disc and much less when averaged over an unresolved source. So the instrument is comparable to the strongest astrophysical signals and larger than most of them, and the ordering depends on which focus the instrument happens to sit at — a decision made for reasons of mechanical convenience decades earlier.
How it is measured
The calibration is conceptually simple and observationally tedious: observe objects known to have no polarisation, and whatever is measured is the instrument’s.
The standards are nearby stars, chosen because interstellar polarisation is produced by dust and there is little dust within a hundred parsecs. Their polarisations are known to be below a hundredth of a per cent from decades of measurement with many instruments, and the residual disagreement between those measurements is the floor on how well the zero point can be established.
There is a second calibration and it is equally necessary: the efficiency. An instrument does not merely add polarisation, it also fails to transmit all of what arrives, so a source with one per cent of polarisation may be measured as 0.95. That is measured by observing standards of known non-zero polarisation, or by inserting a polariser into the beam, and it is a scale correction — the multiplicative half that the additive half above is not.
Both calibrations have an awkward property in common: they are measured on bright stars and applied to faint ones, and nothing guarantees the instrument behaves identically. Detector non-linearity, charge transfer inefficiency and the different exposure times involved all enter, and the resulting brightness dependence of the polarimetric zero point is one of the least-characterised systematics in the technique. It is the same shape as a photometric calibration transferred from standards to targets, with the difference that a polarimetric measurement is a difference and therefore twice as sensitive to anything that scales.
The parts a zero point does not fix
If instrumental polarisation were a constant vector the story would end there. Three things stop it being constant.
It depends on the pointing. The angles of incidence on the mirrors of an alt-azimuth telescope with a Nasmyth focus change as the telescope moves, so the instrumental vector rotates and changes magnitude through the night. Standards therefore have to be observed at the same altitude and azimuth as the target, or the dependence has to be mapped and modelled.
Crosstalk. A real optical train does not merely add polarisation; it converts one kind into another. Linear polarisation becomes circular and back again on reflection from a metal surface, and the conversion is described by the off-diagonal elements of the instrument’s Mueller matrix. A source with strong circular polarisation and no linear can be measured as linearly polarised, which is a systematic that no unpolarised standard reveals.
Depolarisation. Averaging over a field of view, over a bandpass, or over time reduces the measured polarisation if the position angle varies across whatever is being averaged. That is not an instrumental error in the ordinary sense — the instrument reports the average correctly — but it means the measured value depends on the aperture, the filter and the exposure.
A fourth complication belongs with those three and it is astrophysical rather than instrumental: the sky itself is polarised. Moonlight scattered in the atmosphere is strongly polarised, with a degree that depends on the angle from the Moon and reaches tens of per cent, and it enters the aperture along with the source. For a faint target the sky contributes a large fraction of the light in the aperture, so the sky’s polarisation has to be subtracted as a Stokes vector rather than as an intensity — which requires measuring the sky’s polarisation nearby and at the same time, and which is why polarimetry of faint objects is not attempted near a bright Moon.
What was actually measured
Three results establish both the size of the problem and the level to which it can be beaten.
The standards themselves. Catalogues of unpolarised standards, built up over decades, agree at the level of about a hundredth of a per cent. That number is the floor of the technique from the ground, and it is set by the accumulated disagreement between instruments rather than by any single measurement.
The pointing dependence, mapped. Observing standards across the sky at a Nasmyth focus produces a smooth pattern of instrumental polarisation with altitude and azimuth, of amplitude up to a per cent and reproducible from run to run. Modelling it reduces the residual by an order of magnitude, and the residual after modelling is what limits the instrument.
The best achieved precision. Dedicated instruments using rapid modulation — switching between polarisation states hundreds of times a second, so that atmospheric and instrumental drifts are common to both states — reach a few parts per million on bright stars. That is four orders of magnitude below the raw instrumental polarisation, and it is achieved not by measuring the offset better but by arranging that it cancels.
There is a fourth result that is really a warning, and it comes from the history of the subject. Several claimed detections of polarisation in objects of great interest — the alignment of quasar polarisation vectors over large scales, the polarisation of the light from certain supernovae, the circular polarisation of some stellar sources — have been reduced or removed by later observations with better-characterised instruments. In each case the original measurement was at the level of a few tenths of a per cent, which is exactly where the instrumental contribution sits, and in each case the disagreement was traced to calibration rather than to variability. That history is why polarimetric results at that level are reported with the instrumental correction described in detail, and why a paper that does not describe it is not usable.
Where the picture stops
The picture stops in three places, and the second is the one that decides what can be attempted.
The correction is only as good as the standards’ distribution. Unpolarised standards are bright, nearby and unevenly distributed on the sky, so at some pointings there is no suitable standard within a reasonable slew. Interpolating the instrumental model across such gaps is where the residual systematic lives.
Rapid modulation beats calibration, and it costs an instrument. The reason the best polarimeters reach parts per million is that they switch states faster than anything drifts, so the two measurements being differenced share every systematic. That is a design decision made before the instrument is built, and it cannot be retrofitted to data taken with a slow rotating analyser.
And circular polarimetry is much harder than linear. Circular signals in astronomy are typically ten to a hundred times weaker than linear ones, and the crosstalk that converts linear into circular is often larger than the circular signal being sought. A measurement of circular polarisation in the presence of strong linear polarisation is therefore mostly a measurement of crosstalk, which is why the Zeeman signature that measures a magnetic field is one of the most demanding measurements in observational astronomy.
A fourth limit belongs here because it changes what is possible from space rather than from the ground. A space telescope has no sky polarisation, no atmospheric depolarisation and no pointing-dependent gravity flexure, and it has an optical train that cannot be recoated or realigned. So the instrumental polarisation is smaller and far more stable, and the limiting factor becomes the stability of the detector rather than of the optics. That is why the most precise polarimetry of faint objects has come from space and the most precise polarimetry of bright ones from the ground: the two regimes are limited by different terms, and neither instrument is simply better.
There is also a case worth mentioning where the instrumental contribution is deliberately made large. A polarimeter that modulates by rotating a waveplate introduces a known, large, rapidly varying polarisation on purpose, so that the small unknown one is measured as a modulation amplitude rather than as a difference between two exposures. That converts an additive systematic into a phase — and a phase is much easier to keep stable than a level, which is the same reason a frequency measurement beats an amplitude measurement nearly everywhere in this subject.
Why the shape of the signal is the real defence
That last point generalises and it is the most useful thing in this essay.
A constant instrumental offset can be subtracted if it is measured, and measuring it is limited by standards. But a signal with a distinctive structure — a wavelength dependence, a spatial pattern, an antisymmetric line profile — can be separated from an offset without measuring the offset at all, because the offset does not have that structure.
That is why the Zeeman analysis works: the signal is antisymmetric about the line centre and the instrumental contribution is not, so fitting for an antisymmetric component rejects the offset by construction. It is why dust polarisation in a cloud is measured from the pattern of directions rather than from the degree at any one point. And it is why polarimetric imaging of a disc uses the fact that scattered light has a position angle that circles the star — a pattern that an additive constant cannot produce.
The general prescription: when a systematic is additive and constant, look for a signal that varies. The variation need not be large. It has to be a variation the systematic cannot imitate, and identifying one is worth more than any amount of effort spent on the calibration.
The same argument runs through the whole collection — a contaminant with a known dependence is not a contaminant but a second observable, and the design of a measurement is largely the business of arranging for the two to differ in something.
A concluding observation about why this is a rung on the polarimetry ladder rather than an instrumental appendix. Polarimetry is the one common astronomical measurement whose systematic is larger than its signal for most targets, and that inverts the usual order of work. In photometry or spectroscopy the calibration is a correction to a measurement; here the measurement is a correction to the calibration. Everything about how the observations are planned follows from that — standards at the same pointing, modulation faster than any drift, and a preference for signals with structure over signals with amplitude. Two instruments blind in opposite directions is the same principle applied to the physics rather than to the hardware, and both are cases of designing a measurement so that its dominant error has nowhere to hide.
Say which measurements are unaffected by all of this, because the list is short and useful. Anything that depends on a change in polarisation — a variable source measured repeatedly through the same instrument at the same pointing — is nearly immune, because the offset is common to every epoch. Anything that depends on a spatial pattern within one image is similarly protected. What is vulnerable is the absolute polarisation of a single unresolved source measured once, which is unfortunately the most commonly wanted quantity.
One consequence of the vector character deserves stating separately, because it decides what an unpolarised standard is worth. Observing a star known to be unpolarised measures the instrument’s own contribution at that moment and at that pointing — which is exactly what is wanted, and which is only valid for that pointing, because the reflection angles change as the telescope tracks. On an alt-azimuth mount the field rotates with respect to the instrument through the night, so the instrumental vector rotates with respect to the sky while the source’s does not. That is a nuisance and it is also the escape: observing the same source across a large range of parallactic angle modulates the two contributions differently, and fitting for both separates them. The measurement therefore has to be spread across hours rather than concentrated, which is the opposite of what signal-to-noise alone would suggest.
Where the ladder goes next
The rung directly above is the Mueller matrix itself: the full sixteen-element description of what an optical train does to polarised light, how much of it can be measured on the sky, and which elements matter for which observation. The one above that is the modulation scheme — how switching between polarisation states faster than anything drifts converts a calibration problem into a differential measurement.
What links here
Essays that link to this one from their own argument.
The objects this essay names
Each one links to every other essay that touches it.
CalibrationCrosstalkDepolarisationInstrumental polarisationModulationMueller matrixPolarimetric standardPosition angleStokes parametersSystematic error