The craters that were not primary
Assumes Surface chronology and Collisional cascade.
Crater counting is the only method that dates a surface anywhere in the solar system without going there. It rests on one assumption that is close to unimpeachable — objects arrive from space at a rate that is known as a function of time — and one that is not: that the holes being counted were made by those objects.
A large impact does not merely excavate a hole. It throws out a mass comparable to the mass it excavated, at speeds of hundreds of metres a second, and that material comes down again. Where it lands it makes craters, and those craters are indistinguishable from small primaries in a photograph. They are also enormously more numerous.
The problem is not a curiosity at the edge of the method. It is the difference between a surface being one billion years old and four, on exactly the surfaces where the answer is most wanted — the young volcanic plains, the recent flows, the deposits that might record something that happened rather than something that has always been there.
Why the contamination has an edge
The saving grace of the problem is that it is confined, and the confinement follows from mechanics rather than from observation.
A block ejected from a crater on an airless body follows a ballistic arc. Its range is set by its launch speed, and its ability to make a crater when it lands is set by the same speed. A primary arrives at some tens of kilometres a second — the escape speed of the Sun at that distance, near enough — and carries kinetic energy per unit mass some four orders of magnitude larger than a block leaving at half a kilometre a second. So an impactor of a given mass makes a very much larger crater when it comes from space than when it comes from next door.
Run that backwards and it gives the edge. The largest secondary on a body is made by the largest block the largest primary threw out, at the fastest speed at which a block survives launch, and for the Moon that combination gives craters of about one kilometre. Above that diameter every crater is a primary, and a count restricted to large craters is clean by construction.
The trouble is that large craters are rare. The primary production function falls steeply with diameter, so restricting a count to craters above a kilometre on a small area leaves a handful of objects and a Poisson error of tens of per cent. The temptation to go smaller is enormous, and going smaller is exactly what invites the contamination.
The arithmetic of that temptation is worth spelling out because it is what drives practice. A cumulative production function of slope −2 means that halving the diameter threshold multiplies the number of countable craters by four. Going from one kilometre to a hundred metres multiplies it by a hundred, which turns a ten per cent counting error into a one per cent one. On a target the size of a single lava flow — a few hundred square kilometres — counting above a kilometre may yield three craters, which is not a measurement at all, while counting above fifty metres yields thousands. There is no way to date a small young unit without going small, and there is no way to go small without meeting this problem. The method’s precision and its accuracy pull in opposite directions, and the pull is steep.
A further wrinkle is that the secondary branch is not a fixed feature of a body. It is a feature of the neighbourhood: a patch of ground a few crater radii from a fresh ten-kilometre impact is drowned in secondaries, and a patch on the far side of the body from every recent large impact is comparatively clean. So the contamination varies from site to site by orders of magnitude, and two counts on terrain of genuinely identical age can differ by a factor of three in the age they return.
What it does to the answer
The measurement being made is a density of craters, converted to an age through a chronology function calibrated on the Apollo and Luna samples. Inflating the density inflates the age. The interesting question is by how much, and the answer has a shape worth knowing.
Two features of that curve matter more than its height.
It is monotonic, so averaging does not help. A count taken over a range of diameters and combined is a weighted average of a set of answers all of which are biased in the same direction. The formal error shrinks and the systematic does not.
And it saturates. Because the chronology function rises exponentially towards the early solar system, a badly contaminated count on young ground does not return an absurd number; it returns something around four billion years, which is a perfectly plausible age for a lunar surface. The failure mode is not a result that looks wrong. It is a result that looks old.
This is why the discovery of the problem was so disruptive. It was not that ages were wrong by a known factor that could be divided out. It was that the ages of young surfaces, which are the ones the method is most useful for, were the ones most vulnerable.
The direction of the bias is also worth stating on its own, because a systematic with a known sign is a different object from an unknown error. Secondaries only ever add craters. There is no mechanism by which they remove any, so a contaminated count is always too high and a contaminated age is always too old. That makes every crater age a kind of upper limit, and it makes any comparison between two counts taken at the same diameter on differently contaminated terrain unreliable in a way that a comparison at large diameters is not. The relative chronology of a planet’s surface units — which came first — is the part of the method most people actually use, and it is not immune.
There is one further consequence, and it is the one that makes the problem hard to route around by being careful. Because the contamination depends on the local history of large impacts, it is correlated with the very quantity being measured. Old terrain has had more large impacts on it, so it has more secondaries; young terrain has had fewer, so it has fewer. The bias therefore does not merely add a constant offset to every age — it stretches the age scale, making old surfaces look older than they are by more than young ones do. Any attempt to calibrate the effect out by measuring it on one surface and applying the correction elsewhere runs straight into that.
The statistic that does not use a size
The size argument for identifying secondaries has an uncomfortable circularity: it assumes the production function whose slope is exactly what is in dispute. A better instrument would be one that does not use the size at all.
There is one, and it uses the fact that secondaries are not independent events. Blocks come from a single primary in a single moment, and they land in rays, clusters and chains. Primaries arrive from space independently, so their positions are a Poisson process.
The distinction being exploited is a distinction between a process with a length scale and a process without one. A Poisson field has no scale: the distribution of nearest-neighbour distances depends only on the density, and the shape of the curve is fixed. Any clustering imprints the scale of the clump, and the imprint is at small separations where the Poisson field has almost nothing. That is why the test is sensitive even when the contaminated fraction is modest — it is comparing something with something close to nothing rather than comparing two comparable numbers.
The test is old and the application is not. Nearest-neighbour statistics, two-point correlation functions and simple counts of chains and clusters are all sensitive to the same thing, and they can be computed on a crater catalogue for nothing. What they cannot do is label an individual crater; a secondary that happens to have landed alone is unidentifiable by any means. So the statistic returns a fraction, and the fraction is what decides whether a count can be believed.
It is worth being explicit about what the statistic can and cannot deliver, because it is often quoted as though it settled the matter. It measures the excess of close pairs over what independence would give, which is a lower bound on the contaminated fraction: a secondary field that has been thoroughly stirred, or one whose parent crater is far enough away that the rays have spread, leaves no clustering signature and is counted as primary. So a clustering test that comes back clean does not certify a count. It only fails to condemn it. The asymmetry is annoying and it is the honest situation, and it is the reason the disagreement over background secondaries has stayed open for two decades — the instrument that identifies them is blind to precisely the population whose size is in dispute.
What was actually measured
The argument was made concrete on Mars, at a crater called Zunil, in 2005. Zunil is ten kilometres across, young, and surrounded by a field of small craters that can be traced to it directly: they are aligned in rays radiating from it, they are shallower than primaries of the same size, and many are in tight clusters or elongated chains. Counting them gave something like ten million secondaries above ten metres from that single impact.
There is a second observable that made the identification firm rather than suggestive: the depth-to-diameter ratio. A secondary arrives at a few hundred metres a second rather than at ten kilometres a second, so it excavates less deeply for its width, and it often arrives at an oblique angle, so it is elongated. Neither property is decisive for a single crater — primaries are sometimes shallow and sometimes oblique — but the two are measurable in bulk from a stereo image, and the population around Zunil is shallower and more elongated than the population far from it by amounts that no plausible primary distribution could produce.
Ten million is the number that mattered. The total number of primaries above ten metres expected on the whole of Mars in the time since Zunil formed is far smaller. So on the terrain around Zunil, essentially every small crater is a secondary, and a count of small craters there measures Zunil rather than the age of the ground.
The same argument was then run on the Moon and found the same thing. The most direct check is the one where the answer is independently known: several lunar surfaces have radiometric ages from returned samples, and the crater ages derived from large-diameter counts agree with them while the ages derived from small-diameter counts run systematically old.
The controversy is not fully settled and the disagreement is quantitative rather than conceptual. Nobody disputes that secondaries exist or that they dominate the small-diameter population near a fresh large crater. What is disputed is what fraction of the background small craters, far from any identifiable source, are secondaries from impacts long ago whose rays have faded. Estimates range from a small contamination to nearly all of them, and the difference between those two positions is a factor of a few in the ages of every young surface in the solar system.
The practical protocol that has emerged from all of this is unglamorous and worth stating, because it is what the argument is for. Count at the largest diameters the area allows and report the Poisson error honestly, even when it is large. Plot the answer against the counting diameter and show the plot rather than a single number. Compute a clustering statistic on the catalogue and report the contaminated fraction. And where a small-diameter count is unavoidable, treat the result as an upper limit on the age rather than as a measurement of it. None of that recovers the precision the small craters appeared to offer; all of it replaces a precise wrong answer with an imprecise defensible one.
Where the picture stops
Three limits stand out, and the second one is the reason the problem is worse on some bodies than on others.
Saturation is a different effect and it looks similar. On very old terrain the craters are packed so tightly that each new one destroys an old one, and the count stops rising with age. That also flattens a size-frequency distribution at the small end, and it also has to be diagnosed before a count is believed. The two can be separated — saturation depletes, contamination adds — but only if the production function is trusted, which returns to the same circularity.
The escape speed decides how bad it is. On a body with a low escape speed, most ejecta leaves entirely and comes back as sporadic impacts distributed over the whole surface, or does not come back at all. On a large body almost all of it lands nearby. So the Moon and Mars are heavily affected, and a small asteroid is affected differently: a rubble pile held together by almost nothing loses its ejecta and gains a global veneer instead of local clusters.
And an atmosphere removes the small end entirely. On Mars the smallest primaries burn up or are decelerated below crater-forming speed, so the primary production function turns over at a few tens of metres while the secondary branch does not. That makes the contamination at small diameters worse than on the Moon, not better, and it is the reason the Zunil work was done on Mars.
The general shape of the trouble
The defect this essay is about has a shape that recurs throughout the collection, and it is worth naming separately from the craters.
A measurement is made by counting things. The count is converted to a physical quantity by a function that was calibrated on a population. A second population, which the calibration did not include, contributes to the count in a way that is negligible at one end of the range and dominant at the other. Because the contamination is monotonic in the counting variable, it is invisible in any single count; it shows up only as a trend — the answer depending on the range over which the count was taken, when it should not.
The same shape appears in a survey’s luminosity function, where an unrecognised second population at the faint end changes the slope; in the sizes of a debris population, where the small end is dominated by fragments rather than by original bodies; and in a collision rate estimated from a flux, where the small end of the impactor distribution is the part that is least directly observed and most heavily extrapolated. In every case the honest report is not a number but a number together with the range it was measured over.
That trend is the diagnostic, and it is nearly always available for free. If a measurement is supposed to be independent of some choice, plotting the answer against that choice costs nothing and is the most reliable single test of a systematic that a counting experiment has. A collision dated from the spread of an asteroid family is checked the same way, and so is every luminosity function in this collection.
One closing thought about why this is filed as a rung on the chronology ladder rather than as a caveat. Crater counting is the only chronometer that works on a surface nobody has visited, and it is therefore the instrument that fixes the timeline of the entire solar system beyond the handful of sampled sites. A systematic in it does not corrupt one measurement; it propagates into the ages of the Martian valley networks, the resurfacing history of Venus, the ages of the outer planets’ moons, and the impact rate that is used, in turn, to argue about the early history of the Earth. That is a very large structure resting on a distinction between two kinds of hole.
A last comparison is worth drawing, because it says what kind of statistic a crater count is. It is a count of events accumulated over time, read as a rate — the same structure as any measurement that infers a flux from a sample, and it fails in the same places. It fails where the counted objects are not independent, which is what a secondary field is; it fails where the counting is incomplete at small sizes, which a count with a knee in it describes; and it fails where the population being counted changed during the interval. A collision rate inferred without watching a collision is the same inference made on a debris population, and the arguments about correlated events are identical.
Where the ladder goes next
The rung above this one is the production function itself: where the impactor size distribution comes from, and what it would take to measure it independently of the surfaces it is used to date. The rung after that is the calibration — the handful of radiometrically dated sites that the whole chronology hangs on, and what the ages of every other surface in the solar system would do if one of them turned out to be wrong.
What links here
Essays that link to this one from their own argument.
The objects this essay names
Each one links to every other essay that touches it.
Crater countingEjectaLunar chronologyProduction functionResurfacingSaturationSecondary craterSize frequency distributionSpatial clusteringSystematic error