The cheapest shape the walls allow
Worth reading first: The price of a gradient · Mass has nowhere to go.
Every course in fluid mechanics arrives at the parabola the same way. Write the Navier–Stokes equations, throw away everything that vanishes in a long straight channel, integrate twice, apply no slip at both walls. Out comes , and the reason it is a parabola is that the equation was second order and the source term was constant.
That is a correct derivation and it explains nothing. Here is a different route to the same shape, which explains a great deal.
Of all the velocity profiles that conserve mass, vanish at both walls and carry the required flow rate, the parabola is the one that destroys the least energy. Not one of several; the unique minimum. The equations produce it because it is the cheapest, and the cheapness is the reason rather than a coincidence.
The family, and where its minimum is
The quickest way to see it is to give the profile one adjustable parameter and differentiate.
Take . Every member has no slip at . Fixing the flux fixes in terms of : . The dissipation is , and carrying the algebra through gives
Differentiate the fraction and the numerator is . There is one turning point in the whole family and it is at .
Two things about that curve deserve separate attention, and the second is the reason the theorem is not better known.
The first is that is not approximate. It is the root of a linear factor, and it does not move if the flux changes, the viscosity changes, the gap changes, or the fluid is replaced with something else. It is a property of the geometry and of nothing else.
The second is that the minimum is flat. A profile with — visibly pointier than a parabola — costs four per cent more. One with costs seven. A cosine, which is what a smooth guess produces if asked to draw a plausible channel profile, costs 1.47 per cent more than the true answer, and no measurement anybody would make on a real channel would distinguish them.
What “least” is being taken over
The claim needs its constraints stated precisely, because a variational principle with sloppy constraints is worth nothing.
The comparison is over velocity fields that are divergence-free, that take the same values on the boundary, and — in a channel, where the boundary values are all zero — that carry the same flux. Within that class, the Stokes solution has the least dissipation. Anything outside the class is not being compared: a field that leaks mass is not a competitor, and neither is one that slips at the wall.
This is Helmholtz’s theorem, and its proof is three lines. Write any competitor as the Stokes solution plus a difference , where is divergence-free and vanishes on the boundary. The dissipation of the sum is the dissipation of the solution, plus twice a cross term, plus the dissipation of . The cross term integrates by parts into a boundary integral — which vanishes, because does — plus a volume integral against the Stokes equations, which vanishes because they are satisfied. What is left is
with equality only when is zero. The excess is exactly the dissipation of the difference field.
The consequence that can be measured
That last line is a sharper statement than “the parabola is the minimum”, and it is testable in a way that the minimum itself is not.
If the excess is , then scaling the disturbance by scales the excess by — with no linear term at all. A functional that had a linear term would have a direction in which the flow could be made cheaper, and there would be no minimum.
That is the rejection test this site’s habit asks for. The theorem is not being illustrated; it is being given something it could fail, which is a measured exponent, and the exponent comes out at two.
Where the expensive shapes are
The family above is well behaved because every member of it is smooth. The instructive part of the comparison is what happens when a profile is allowed to be a poor shape rather than a slightly wrong one.
The plug is the case worth understanding, because it is the shape a turbulent flow actually has.
A plug of speed joined to the wall by a ramp of thickness dissipates . As the layer thins, the gradient in it rises as and the volume falls as , so the product rises as — without limit. Concentrating the shear is expensive, and it is expensive by exactly the same arithmetic that makes the wall the only place a pipe’s heat is made.
The same statement in a bearing, where it is worth money
A channel is the clean case and it is not the case anybody is paid to think about. The version that matters commercially is a bearing, and there the theorem says something a designer can use.
A journal bearing carries its load because the film it runs on is converging, and the pressure it generates is a consequence of viscous flow through a narrowing gap. The load capacity is therefore inseparable from the dissipation: a bearing that destroyed no energy would carry nothing. What the theorem adds is that the relation between them is stationary — the film shape that a real bearing settles into is a minimiser, so a small error of form changes the load at second order rather than first.
That is the reason a plain bearing is as forgiving as it is. A pad machined a few per cent off its intended taper does not lose a few per cent of its load; it loses a fraction of a per cent, because the quantity being spoilt is at a minimum with respect to exactly that kind of change. The whole design tradition of running clearances quoted to one significant figure rests on it, and nobody states it.
What happens when the constraint is a velocity rather than a flux
One more variation is worth doing, because it changes the answer and is the case a great deal of machinery is actually in.
In the channel above the flux was held and the profile was free. In a Couette flow it is the wall speed that is held, and the flux is whatever it turns out to be. The minimiser is then the straight line, and the argument is the same: any competitor is the straight line plus a field vanishing at both walls, the cross term integrates away, and the excess is the disturbance’s own dissipation.
The straight line, unlike the parabola, dissipates uniformly. Every part of the gap is being sheared at the same rate, so every part is paying the same. That is a striking difference from the pressure-driven case, where nearly all of the bill is at the wall, and the two cases sit side by side in almost every real film — a bearing is driven by both a wall speed and a pressure gradient at once, and the dissipation it makes is not the sum of the two considered separately.
Which is why the theorem does not say what it seems to
Here is the trap, and it is worth setting out plainly because the theorem is frequently quoted with the constraints dropped.
Turbulent pipe flow has a plug-like profile with a thin wall layer, which is the shape the law of the wall describes. By the arithmetic above it must dissipate far more than the parabola at the same flow rate — and it does, by roughly a factor of three at Reynolds number , rising with Reynolds number because the two friction laws have different exponents.
So a real flow at that Reynolds number is emphatically not minimising its dissipation. It is doing several times worse than a solution that exists, satisfies the same boundary conditions, and carries the same flux.
There is no contradiction, and locating it precisely is the useful exercise. Helmholtz’s theorem is a statement about the Stokes equations, in which the nonlinear term is absent. Those equations have exactly one solution for given boundary data, so “the solution” and “the minimiser” are the same object and there is nothing for a flow to choose between. The Navier–Stokes equations have no such guarantee: the laminar solution is still a solution, and it is no longer the only one, and nothing in the equations says the cheapest available one is taken.
Minimum dissipation is a property of uniqueness, not a principle of selection. Every attempt to turn it into one — and there have been many, from Helmholtz onwards — founders on the same fact.
The variational idea, running the other way
Minimum dissipation is not a selection principle is the right conclusion and it is not the end of the variational programme, because the same style of argument survives if the inequality is turned round.
The failure above is that the laminar solution is a lower bound the flow declines to take. Ask instead for an upper bound — how much can a flow between these walls dissipate, whatever it is doing — and the question becomes answerable, and the answer is a theorem about the full Navier–Stokes equations with no closure, no model and no measurement in it.
The modern machinery is disarmingly simple in outline. Split the velocity into a steady background field that carries the boundary conditions, plus a fluctuation that vanishes on the walls. Substituting that split into the energy balance leaves a quadratic form in the fluctuation, and if the background is chosen so that the form is positive definite, everything the fluctuation could possibly be doing is bounded — so the dissipation is bounded, by a quantity depending only on the background. Optimising over backgrounds gives the best bound the method can produce.
What comes out is genuinely a theorem. For shear-driven turbulence it gives a drag coefficient bounded independently of Reynolds number, which is the rigorous half of the dissipation anomaly: the statement that the losses do not fall as the viscosity does, obtained without assuming anything about eddies. For a heated layer it gives a Nusselt number bounded by the square root of the Rayleigh number.
The bounds are not tight — typically a factor of several above what is measured — and that is the honest position. What the variational method cannot do is say what a turbulent flow does; what it can do is prove what no flow may exceed, which is the same relationship this collection’s control-volume field has to the machines it bounds.
What it is good for anyway
Three things, and all of them are about bounding rather than predicting.
It bounds a drag from above. Any admissible field’s dissipation is an upper bound on the true one, so guessing a plausible profile and integrating gives a number that is too large and is known to be too large. That is how the first estimates of the drag on awkward shapes were made, and it is why they were quoted as bounds rather than as answers — including the estimates that preceded the exact solutions this collection uses for a sphere.
It says perturbations are quadratic. A slightly deformed geometry has a dissipation that differs from the original at second order in the deformation, not first — which is why a bearing’s load capacity is insensitive to small errors of form and why the harmonic mean sets the pressure peak of a tapered pad so robustly.
It explains why a Stokes solve is forgiving. A numerical Stokes solve that is slightly wrong is wrong in its dissipation by the square of how wrong it is in its velocity. That is a genuine practical advantage and it has no counterpart at high Reynolds number: a Navier–Stokes solve that is one per cent wrong in its velocity field can be several per cent wrong in its drag, because there is no stationarity to protect it.
What the picture cannot show
The comparison is at fixed boundary values, and a real design changes them. Every profile here carries the same flux between the same walls. A designer who is allowed to change the gap, or to add a wall, is not in the class the theorem covers, and the theorem has nothing to say about that comparison.
The family is one-parameter and the theorem is not. Finding the minimum of shows that the parabola beats every member of one particular family. The theorem says it beats every admissible field, which is an infinite-dimensional statement, and the perturbation ladder is the closest this essay comes to testing it in that generality.
And nothing here is about stability. A minimum of the dissipation is not a stable state; the laminar profile is both the global minimiser and, above transition, unstable. Those are answers to different questions and the second is where the neutral curve is.
The number the flatness explains
The flatness of the minimum has one more consequence, and it is the reason this theorem is more often useful as an excuse than as a tool.
Almost every practical estimate of a viscous flow is made by assuming a profile. Lubrication theory assumes a parabola across the film; integral boundary-layer methods assume a one-parameter family; network models of a piping system assume fully developed flow in every branch. Each of those assumptions is wrong in detail, and each of them is wrong in a direction that raises the dissipation, because the true profile is the minimiser.
So the errors are all of one sign and all of second order. An assumed profile that is ten per cent wrong in shape gives a dissipation that is about one per cent too high — never too low, and never proportionally. That is why lubrication theory works as well as it does on films that are not quite thin, and why an integral method’s drag is usually good to a few per cent while its profile is visibly not the right shape.
The exception is the shape that is wrong in the expensive direction. A method that assumes a plug where the flow is really parabolic is not making a second-order error; it is on the steep part of the curve, and the further it goes the worse it gets without bound. That asymmetry is worth carrying: the penalty for guessing a smooth profile is negligible, and the penalty for guessing a flat one is not.
Who found it, and when
Helmholtz proved it in 1868, and Korteweg gave the converse — that the minimiser satisfies the Stokes equations — in 1883, which is the half that makes it a genuine variational principle rather than a property. Rayleigh restated it in 1913 in the form most often quoted, and spent some effort trying to extend it to flows with inertia, which cannot be done.
The surprising connection is with electrical networks. The dissipation of a Stokes flow is a positive-definite quadratic form in the velocity field, minimised subject to a linear constraint — which is the same mathematical object as the power dissipated in a resistor network, minimised subject to Kirchhoff’s current law. Thomson’s principle for currents and Helmholtz’s for Stokes flows are the same theorem about the same kind of functional, discovered independently fifteen years apart, and both fail for exactly the same reason when the medium stops being linear.
Where the ladder goes next
Beside this rung is the reciprocal theorem, which is the other thing linearity buys: a force obtained without solving for the flow that makes it. Both are properties of the Stokes equations and both stop at the first appearance of inertia.
Below it is the price of a gradient, which supplies the functional being minimised, and mass has nowhere to go, which supplies the constraint. Above it, in a sense, is what it costs to go turbulent — the measurement of how far a real flow sits from the cheapest one available to it.
What links here
Computed from the collection rather than written here: the essays that point at this one.
Reads more easily once this is understood
Essays that name this one as worth reading first.
Shares its objects with
Essays naming at least two of the same things, that neither author linked.
- A swimmer that cannot go backwards — both name boundary condition, dissipation, stokes flow
- The discontinuity that has a thickness — both name dissipation, irreversibility, viscosity
- The jump does not ask what made it — both name dissipation, irreversibility, viscosity
- The limit that is not the value — both name dissipation, turbulence, viscosity
- The sound that only leaves — both name boundary condition, dissipation, turbulence
- A compression that costs nothing in the end — both name irreversibility, optimisation
Named objects
A dashed tag is an object no other essay names yet.
Boundary conditionDissipationIrreversibilityMinimum-dissipationOptimisationPoiseuille flowStokes flowTurbulenceVariational principleViscosity