A length nobody derived, fitted to one star
Assumes Energy transport and Hydrostatic equilibrium.
A star’s outer third is a boiling fluid. Energy arrives at the base of that region as radiation and leaves the top as radiation, and in between it is carried by the bulk motion of gas — hot material rising, cool material sinking, on scales from the size of a granule to the depth of the whole convection zone.
Computing that flow is a three-dimensional, compressible, radiative hydrodynamics problem with a Reynolds number of — a harder calculation than any of the fluid problems this collection has met, and one that has to be embedded inside a calculation of the whole star’s evolution. It cannot be done inside a stellar evolution code, which has to integrate a star’s structure through a million time steps.
So a quantity that decides the radius, the effective temperature and the evolutionary track of every star with a convective envelope — which is every star cooler than about 7,000 kelvin, and the outer layers of most others — is a single fitted number. A star is held up by its own weight is a statement that needs no free parameters; how the heat gets out is not.
What the parameter stands in for
The mixing-length theory is a piece of dimensional analysis dressed as a physical model. A blob of gas is displaced upward, it is hotter than its surroundings because the star’s temperature gradient is steeper than the adiabatic one, so it is buoyant and accelerates. After travelling a distance it dissolves and gives up its heat.
The convective flux follows from the excess temperature, the velocity and the density. The velocity follows from the buoyant acceleration acting over . And is set to , with the local pressure scale height and a number of order one that nothing in the theory determines.
There is no derivation of , there is no reason it should be the same at every depth in one star, and there is no reason it should be the same in two different stars. It is a fitting parameter with the dimensions of a length, and everything about the outer layers of every stellar model depends on it.
It is worth being clear about the scale of the approximation. The theory replaces a flow with structure across ten orders of magnitude in length — from the depth of the convection zone down to the viscous dissipation scale — with a single length and a single velocity, both local. It is the crudest kind of closure, it has no small parameter to expand in, and it is not obviously better than several alternatives that were proposed and abandoned. What recommends it is that it is one number and that it fits.
Why it controls the radius
The reason matters at all is that convection in a star’s deep interior is almost perfectly efficient. The gas carries so much heat that only a minute excess over the adiabatic gradient is needed, so the temperature profile there is adiabatic whatever is, and the answer is insensitive.
Near the surface it stops being efficient. The density falls, the heat capacity per unit volume with it, and the gas has to be substantially superadiabatic to carry the flux. How superadiabatic depends on — a longer mixing length carries the flux more easily and needs less of a gradient.
That superadiabatic layer is thin, and it sets the entropy of the whole adiabatic interior below it, which sets the radius at which the star’s structure matches its surface boundary condition. So a parameter describing a few hundred kilometres of a star’s outer skin fixes the radius of the entire object.
There is a corollary that makes the sensitivity easier to remember. The luminosity of a low-mass star is set by its core, which does not care about ; the radius is set by the envelope, which does. Since , fixing and changing changes , and the change appears as a horizontal shift of the whole evolutionary track in the temperature–luminosity plane. That is why an age read off the bend of a cluster’s main sequence depends on the parameter: the bend’s position in temperature moves with it, and matching an observed sequence to a model sequence is matching a temperature.
There is a second reason the outer layers dominate, and it is geometric rather than thermodynamic. The superadiabatic region is a few hundred kilometres thick in a star seven hundred thousand kilometres in radius — a part in two thousand. Its thermodynamic role is to set the entropy of everything below it, and entropy is conserved along an adiabat, so a change made in that thin skin propagates unchanged through the entire interior. A star is unusually vulnerable to its own surface in a way that most physical systems are not, and the mixing-length parameter is the lever that acts there. The same leverage makes a spectroscopic gravity so sensitive to the treatment of the outer layers, for the same reason.
The calibration, and what it assumes
The standard procedure is to build a model of one solar mass with the solar composition, evolve it for 4.57 billion years, and adjust — together with the initial helium abundance — until the model reproduces the Sun’s radius and luminosity.
That gives a value near 1.8 for most codes, and the value is code-dependent because it absorbs whatever else that code gets wrong about the outer layers: the equation of state, the low-temperature opacities, the treatment of the atmosphere.
Two assumptions are then made without much comment. That the calibrated applies to stars of other masses, other metallicities and other evolutionary states. And that the same applies at every depth within one star.
Neither is obviously true and both are now testable, because three-dimensional simulations of stellar surface convection can be run for a grid of stellar parameters and the effective read off by matching a one-dimensional model to the simulation’s entropy profile.
Notice that the second assumption is known to be false in a specific way. The pressure scale height varies by orders of magnitude through a convection zone, so varies with it — but the physical eddy size does not, near the top, where the granules have a size set by the depth at which they form rather than by the local scale height. So a constant is certainly wrong in the layer that matters most, and it is used because a depth-dependent would require knowing the depth dependence, which is the thing that was too hard to compute in the first place.
What the simulations found
The three-dimensional calculations give a definite answer and it is not a constant.
The effective varies systematically across the temperature–gravity plane, by about ten per cent from the Sun to a metal-poor turn-off star and by more towards the giants. The variation is smooth, it is reproducible between independent simulation codes, and it goes in the direction of lower for hotter and lower-gravity stars.
Ten per cent in is about twenty kelvin at the Sun and more elsewhere, which is small compared with the older uncertainties and large compared with modern measurements. Adopting a varying from the simulations changes stellar radii by a per cent or two and cluster ages by a few per cent — not a revolution, and not negligible either.
The simulations also say something the one-dimensional theory cannot express: the convective flux is carried by a small number of fast downdrafts rather than by a symmetric up-and-down exchange, which is a qualitatively different flow from the one the mixing-length picture describes. That the parameterisation works as well as it does, given that its physical picture is wrong in that specific way, is the genuinely surprising part.
The third of these is the one that connects the parameter to something a telescope sees directly rather than to a model’s output, and it is worth ending the sequence on it.
What was actually measured
Three classes of measurement bear on the parameter, and they disagree in a way that is itself informative.
Interferometric and eclipsing-binary radii. For stars with directly measured radii, comparing against models constrains directly. The results are consistent with the solar value for solar-type stars and require lower values for hotter ones, in the direction the simulations predict.
The solar convection zone’s depth. Helioseismology measures the depth of the Sun’s convective envelope to a fraction of a per cent, and it is one of the constraints a solar model must satisfy. It is not very sensitive to — the depth is set by the opacity and the composition — but it does constrain the combination of and helium abundance, which is why the two are calibrated together.
Red giant temperatures. The effective temperature of a red giant branch is almost entirely set by , because the envelope is convective and enormously extended. Requiring that models reproduce the observed giant branches of clusters gives values that differ from the solar one, and the differences are the most direct evidence that a single value is inadequate.
There is also a fourth result worth stating because it undermines the parameter’s interpretation rather than its value. Comparing simulations run with different numerical resolutions and different codes shows that the effective recovered depends slightly on how the matching to the one-dimensional model is done — on which quantity is required to agree, and at what depth. That is not a numerical error; it is a statement that “the mixing length” is not a well-defined property of the flow, only of the procedure used to summarise it. A parameter defined by a fitting procedure inherits the procedure, which is the same observation as the one about a Coulomb logarithm’s cut-off being a choice rather than a quantity.
One more consequence of the simulations deserves recording, because it is a genuine physical result rather than a calibration. The flow they show is not a symmetric exchange of rising and falling blobs. It is a broad, slow, warm upflow occupying most of the area and a set of narrow, fast, cold downdrafts occupying very little of it — a topology forced by the steep density stratification, since a rising parcel expands and a sinking one is compressed. The mixing-length picture has no way to express that asymmetry, and the asymmetry is what determines the overshooting at the base of the zone, the mixing of material below it, and the excitation of the oscillations the previous essay’s technique depends on.
Where the picture stops
The picture stops in three places, and the third is where the field is heading.
The parameter absorbs other errors. A calibration that adjusts to fit the Sun will compensate for an error in the opacities, the equation of state or the atmospheric boundary condition by shifting . So the calibrated value is not a measurement of a mixing length; it is a measurement of one number’s worth of everything the code gets wrong near the surface.
Rotation and magnetic fields are not in it. A magnetised, rotating convection zone carries flux differently, and the effect is in the direction of inflating a star. That is one of the leading explanations for the radius discrepancy in low-mass stars, and it is degenerate with by construction — both act by changing the efficiency of convection.
And the replacement is coming from a different direction. Rather than improving , the current work replaces the surface boundary condition of a one-dimensional model with a patched-on three-dimensional simulation. That removes the parameter’s most important role — setting the entropy of the adiabat — and leaves it doing much less. It is expensive, it is being done on grids rather than star by star, and it is the likely end of the parameter’s forty-year career as the largest single uncertainty in stellar structure.
A fourth deserves stating because it changes what can be claimed from a comparison. Because absorbs the code’s other surface errors, two codes with different opacities and different atmospheres will calibrate to different and then agree with each other about the Sun by construction. Their disagreement about a red giant is therefore a measurement of the difference between their absorbed errors, not of anything about convection — and quoting the spread between codes as an uncertainty on a stellar age counts that difference as if it were physical. The same objection applies to comparing two spectroscopic analyses that share a calibration: agreement between two methods that were tuned on the same object is not evidence.
Why one fitted number is worth an essay
The general shape is one this collection meets from several directions, and the repetition is the point.
A theory with derived structure and one fitted constant is enormously powerful and is not a measurement. The Coulomb logarithm in dynamical friction is one; the constants in the seismic scaling relations are another; the mixing length is a third. In each case the exponents and the functional form come from physics, the normalisation comes from a fit, and the result is used far outside the range the fit was made in.
Such a construction behaves well differentially and poorly absolutely, it hides other errors by absorbing them, and its failures show up as small systematic discrepancies rather than as visible breakdowns. That last property is the dangerous one: a model with a well-tuned fudge factor fits everything slightly and nothing exactly, and there is no single observation that refutes it.
What eventually resolves such a parameter is never a better fit. It is a calculation from below that does not need it — a simulation, a first-principles derivation, a measurement of the thing itself. That is what the three-dimensional convection calculations are, and it is worth noticing how long it took: the parameter was introduced in 1958 and the calculations that measure it became routine around 2010.
One remaining observation about what the parameter’s persistence says. It has been known to be inadequate since it was introduced, its replacement has been technically possible for over a decade, and it is still what nearly every published stellar model uses. The reason is not inertia: it is that a stellar evolution calculation has to run through millions of time steps and a three-dimensional simulation of one instant costs more than the whole track. So the parameterisation survives because of an arithmetic constraint on computation rather than because anybody believes it, and it will be replaced when the grids of simulations are dense enough to be interpolated rather than when the physics improves. That is a common shape in computational science and it is worth naming, because the resulting models carry an assumption that everybody involved knows to be wrong and that nobody can currently avoid.
Be precise about what the solar calibration does and does not transfer. Fixing the parameter on the Sun makes one model reproduce one star’s radius at one age, and that is a single constraint on a single number — it says nothing about whether the number should be the same for a star twice as massive, half as metal-rich, or in a completely different evolutionary state. The three-dimensional simulations that have been run precisely to answer that find a parameter that varies systematically across the temperature–gravity plane, by twenty per cent or so between the Sun and a red giant. Twenty per cent is small against the range the parameter is sometimes given and large against the precision that seismology now delivers, so the calibration’s constancy has stopped being a harmless simplification and become a measurable error.
Where the ladder goes next
The rung directly above is the surface boundary condition: how a three-dimensional simulation is attached to a one-dimensional interior model, what has to match at the join, and what is left of the mixing length afterwards. The rung after that is the radius discrepancy in low-mass stars — a well-measured disagreement with three candidate causes, all of which act through the efficiency of convection.
About the same objects
Not linked from either essay — found by the objects both name.
- A better measurement that made the model worse convection · systematic error
- A planet radius is a stellar radius stellar radius · systematic error
- A temperature that depends on where the observer stands effective temperature · stellar radius
- An angle of five hundredths of an arcsecond effective temperature · stellar radius
- The interior read from a comb of frequencies convection · stellar radius
- The star that swells because its centre shrank effective temperature · evolutionary track
What links here
Essays that link to this one from their own argument.
The objects this essay names
Each one links to every other essay that touches it.
ConvectionEffective temperatureEvolutionary trackFree parameterMixing lengthSolar calibrationStellar radiusSuperadiabatic gradientSystematic errorThree-dimensional simulation