The Merton Portfolio Problem
If you can rebalance continuously and your appetite for risk does not change with how rich you are, the optimal fraction of your wealth to hold in a risky asset is a single constant, edge divided by risk aversion times variance, no matter your horizon or your bank balance.
Prerequisites: Geometric Brownian Motion, The Kelly Criterion, The Risk-Return Tradeoff
Markowitz answers a one-shot question: given a pot of money and one period ahead, how should you split it? But nobody invests once. You invest, the market moves, your wealth changes, and you get to decide again, tomorrow and every day after that for forty years. The question a real investor faces is not "what is the best split today" but "what is the best rule for the split, at every moment, for the rest of my life?"
That is a much harder question, and the surprising thing about Robert Merton's 1969 answer is how simple it turns out to be.
An analogy before any symbols
Think of a thermostat. It does not care what the weather did last week, how long until spring, or how big the house is. It has one setting, and it does whatever is needed at each moment to hold that setting. When the room drifts warm it cools; when it drifts cold it heats.
Merton's result is that, under a specific set of assumptions, the best investment policy is a thermostat. You pick one number, the fraction of your wealth that sits in the risky asset, and you hold it there forever. Not a fraction that ramps down as you age, not one that depends on how well you have done so far. One setting, held constant, with continual small adjustments to stay on it.
That is genuinely counterintuitive. Most people's instinct is that a young investor should take more risk, or that after a windfall you should get more aggressive. Merton's model says no, and understanding why it says no is more useful than the formula itself.
The setup
There is one risky asset whose price follows a Geometric Brownian Motion:
In words: over any short instant, the asset drifts up at rate and jitters randomly with size , where is the random shove and is the tiny slice of time. There is also cash paying a certain rate .
You choose , the fraction of your wealth in the risky asset at time . Not a dollar amount, a fraction. If you are 60% invested and 40% in cash; if you have borrowed 50% of your wealth to invest more. Your wealth then evolves as
In words: you always earn the cash rate, plus your share of the extra return the risky asset offers over cash, and you carry your share of its jitter.
Finally you need a way to say what you want. Merton uses constant relative risk aversion (CRRA) utility,
which is a mathematical way of writing "more money is better, but each extra dollar is worth less than the last, and how much less is governed by one number ." Small means you barely flinch at risk; large means you hate it. The word "relative" is the crucial part: your feelings depend on percentage changes in wealth, not dollar changes. Losing half is equally painful whether you started with ten thousand or ten million.
The answer
Maximising expected utility of terminal wealth gives the Merton fraction:
In words: your edge over cash, divided by how much you hate risk times how violently the asset moves. Three readings worth holding onto:
- Double the edge, double the position.
- Double the volatility and you quarter the position, because is squared. Risk punishes you far harder than reward rewards you.
- Nothing on the right-hand side is or . No time, no wealth. That is the constant-thermostat result.
Why does horizon drop out? Because under this model, returns in disjoint time intervals are independent and identically distributed, and CRRA preferences care only about percentage outcomes. A ten-year investor is just a one-year investor repeated ten times, facing the same problem each time, so they make the same choice each time. Why does wealth drop out? Because doubling your wealth doubles every possible outcome proportionally, and "relative" risk aversion is indifferent to that scaling.
With (log utility) the formula becomes , which is exactly the continuous-time The Kelly Criterion. Kelly is Merton with one specific risk appetite baked in.
is a fraction of wealth held constant through time, not a dollar amount set once. Holding it constant is an active policy: it forces you to sell into rallies and buy into declines, automatically.
Growth versus the size of the bet
Plug back into the wealth equation and the long-run growth rate of your money is
In words: you earn cash, plus your share of the edge, minus a variance drag that grows with the square of your bet. That subtraction is why leverage eventually turns on you: the reward term is linear in , the penalty is quadratic, so past some point the penalty wins.
To see the randomness this policy is steering through, run a few sampled paths. Set the drift and volatility, then generate again a few times, notice how wide the spread of outcomes is even when the average drift is comfortably positive:
Worked example: sizing a real position
An equity index with expected return , cash at , volatility . So the edge is 6% and .
Moderately risk-averse investor, :
So 61.7% in the index, 38.3% in cash. On a $200,000 portfolio that is $123,400 in equities and $76,600 in cash. Its growth rate is .
Log-utility (Kelly) investor, : , meaning 185% invested, borrowing 85% of wealth. Growth rate , the highest achievable. But the portfolio's volatility is a year, which very few people can hold through a bad decade.
Cautious investor, : , so 37% invested. Growth . Less growth, far smoother ride.
And the cliff. At the growth rate is , exactly the cash rate, with 67% annual volatility. You have taken on ruinous risk to earn what a savings account pays.
Worked example: what "constant" actually costs you
Start with $100,000 and a target of : $60,000 in the index, $40,000 in cash.
Day 1, the index rises 10%. Your equity sleeve becomes $66,000, cash stays $40,000, total $106,000. Your actual fraction is now , above target. Sixty percent of $106,000 is $63,600, so you sell $2,400 of equities.
Day 2, the index falls 10%. Your $63,600 becomes $57,240, cash is $42,400, total $99,640. Your fraction is , below target. Sixty percent of $99,640 is $59,784, so you buy $2,544 of equities.
Two things to notice. First, the index went up 10% then down 10% and ended below where it started (), and so did you, at $99,640. Second, the policy sold high and bought low without being told to. A constant fraction is mechanically contrarian. That is a virtue in choppy markets and a cost in trending ones, which is exactly the trade-off buy-and-hold makes in reverse.
Also notice the friction: two trades in two days. Merton's model assumes trading is free and continuous. It is neither, which is why the practical version uses tolerance bands, see Turnover and Rebalancing and Transaction-Cost-Aware Portfolio Optimization.
What this means in practice
- Volatility targeting is Merton in disguise. If and are treated as roughly fixed, then , so when volatility doubles you cut exposure to a quarter. That is precisely what a Vol Targeting overlay does. See also Position Sizing.
- Multi-asset version. With many risky assets, : the same shape, with the covariance matrix doing the work. The direction of that vector is the The Tangency Portfolio and the Capital Market Line; only sets its length. Merton and Markowitz agree on the recipe and differ only on the dose.
- Adding consumption. Merton's 1971 extension lets you also spend. The answer keeps its shape: invest a constant fraction of wealth, and spend a constant fraction of wealth per year. This is where the "4% rule" of retirement planning comes from.
- When the constant breaks. If or move predictably over time, the constant-fraction result fails and an extra hedging demand appears, an adjustment for the fact that today's portfolio also insures you against tomorrow's worse investment opportunities. That is the interesting part of modern dynamic asset allocation.
The formula is only as good as , and is the one thing you cannot measure. The position size is directly proportional to the estimated edge, and estimating an equity risk premium to within a percentage point takes many decades of data. Volatility, by contrast, can be estimated well from a few months. So the numerator is a guess and the denominator is nearly a fact, yet the guess drives the answer.
The practical consequence: quants routinely run at a fraction of the Merton or Kelly number, often a half or a quarter. Given the growth curve above, that costs surprisingly little. Half-Kelly gives you about three-quarters of the maximum growth at half the volatility. Overshooting, on the other hand, is punished quadratically. The curve is nearly flat on the left of its peak and falls off a cliff on the right, so when in doubt, be small.
The other frequent misreading is "horizon does not matter, so stocks are no safer over the long run." The model says the optimal fraction does not depend on horizon. It does not say a long horizon reduces risk, that would require returns to mean-revert, which this model explicitly assumes they do not.
Practice
- With , , , what makes exactly 100%? What does that tell you about a fully-invested, unlevered equity investor's implied risk aversion?
- Volatility spikes from 18% to 30% with no change in expected return. By what factor should a Merton investor cut exposure? (Answer: about a third of the original position.)
- Compute for using the numbers above and confirm the peak sits at 1.85.
- Two assets with the same Sharpe ratio but different volatilities. Show that the dollar risk Merton allocates to each is the same, even though the weights differ.
Related concepts
Practice in interviews
Further reading
- Merton (1969), Lifetime Portfolio Selection under Uncertainty
- Merton (1971), Optimum Consumption and Portfolio Rules in a Continuous-Time Model
- Björk, Arbitrage Theory in Continuous Time (Ch. 20)