A Dow Theory rabbit hole · Sep 2026

Does Dow Theory actually work?

I’ve heard about Dow Theory for years — the Industrials and the Transports have to confirm each other before a trend is real — and never once seen anyone simply backtest it. So I turned the rules into code and ran them across 126 years of Dow data. Does confirmation add anything? Does it beat luck? And does it still work today?

Credit where it’s due: this is the grandparent of every trend-following system, and the core idea is genuinely clever. Charles Dow, and later William Hamilton and Robert Rhea, argued that the economy has two halves that have to agree. Factories make things; railroads move them. If the Industrials rally to a new high but the Transports don’t follow, the goods aren’t actually shipping — and the rally is suspect. It’s a real economic intuition, a century before anyone said “cross-asset confirmation.”

The reason nobody backtests it is that Dow Theory is written as judgment, not rules: “secondary reactions,” “lines,” “confirmation.” Any test has to choose a mechanical reading. So I did two things to keep myself honest: I picked a plain, textbook reading before looking at results, and then I ran ten variations of it and report all of them, not the winner.

Return vs buy & hold

−1.0 pt

Average across 10 variants, 1900–2026: 8.8% vs 9.8% a year. Only 3 of the 10 beat buy & hold. The base rule ties it.

Worst drawdown

−54%

Base rule, vs −87% for buy & hold. Most of that is 1929–32. The 10 variants average −48%.

Beat buy & hold since 1953

0 of 10

On return. Drawdown is still ~13 points shallower (−39% vs −52%), so it buys safety, not extra return.

01The rules, as code

This is the most common modern reading of Hamilton’s version. For each index, I track the swing highs and lows using closing prices only. A secondary reaction is a pullback (or bounce) of at least 5% from the running extreme. Then:

Nothing in the rule can see the future: each day only uses swings that were already confirmed. Here is what it did to the Dow over 126 years. Shaded bands are the stretches when it said “cash.”

75 “sell” signals in 126 years

Dow Jones Industrial Average, log scale · shaded = Dow Theory in cash

DJI (price) Dow Theory says cash
The bands cluster around the big ones — 1930–32, 1937, 1973–74, 2008, 2020 — but there are a lot of thin ones too, each a short sell that turned into a re-buy at a higher price. That mix of a few great calls and many false alarms is the whole story of what follows.

02The headline: safer, not richer

Over the full sample the base rule earns 9.8% a year, the same as buy and hold — but with a worst drawdown of −54% instead of −87%, and only 72% of the time in the market. That’s a genuinely better ride. Before getting excited, though, that base rule sits at the high end of my ten variants; the honest number is the average (8.8%), a point below buy and hold.

Dow Theory vs alternatives · 1900–2026 · total return
StrategyCAGRMax DDMARSharpe*In mktTrades
Buy & hold the Dow9.8%−87.0%0.110.61100%—
Dow Theory · base rule (5%, both indexes)9.8%−54.0%0.180.8072%76
Dow Theory · average of all 10 variants8.8%−48.2%0.190.72——
Industrials only, no confirmation (avg of 10)8.3%−61.8%0.140.68——
Transports only, no confirmation (avg of 10)9.0%−44.9%0.210.76——
200-day moving average (Dow)9.1%−54.1%0.170.8365%424
Same 72% exposure, random timing (blend)8.5%−75.1%0.110.6972%—

*Raw return divided by volatility, no cash rate subtracted. Measured against cash, buy & hold is 0.41, Dow Theory 0.51 and the 200-day average 0.50. MAR = CAGR divided by worst drawdown.

Two things stand out. First, the 200-day moving average does about the same job — slightly lower return, the same drawdown, about the same Sharpe — without needing a second index. Second, the “blend” row is a useful reality check: simply holding 72% stocks and 28% cash, with no timing at all, earns 8.5% but still takes a −75% drawdown. Being in cash at the right times is what the Dow rule adds.

Growth of $1 since 1900

Total return, log scale · next-close fills, 0.10% per switch

Dow Theory (base) 200-day average Buy & hold
The Dow line builds a lead of almost 3× over buy and hold by the mid-1930s, holds roughly 2× through 1989, then gives essentially all of it back over the last 25 years. It ends about even, because the return is nearly identical.

The real benefit is the crash

Drawdown from prior peak · 1900–2026

Dow Theory (base) Buy & hold
1929–32 is the whole difference. Buy and hold lost 83% from the September 1929 peak; the rule lost 44% — it flashed its first sell on October 23, 1929, with the Dow at 306, about 20% below the peak. That is late, but far better than holding. After the 1930s the two drawdown curves look much more alike.

03Where it worked, and where it didn’t

Split the 126 years into eras and the picture changes. The rule was excellent in the early decades — +3 points a year over buy and hold from 1900 to 1929 — and roughly break-even after 1950 on return, with shallower drawdowns until 2000. Since 2000, it has lagged buy and hold by two points a year, and its drawdown isn’t any better.

By era · base rule vs buy & hold
PeriodB&H CAGRDow CAGRB&H max DDDow max DDIn mkt
1900-192910.2%13.1%−47.5%−31.3%66%
1930-19494.6%3.9%−83.5%−48.5%58%
1950-19799.2%9.3%−41.1%−35.3%76%
1980-199917.9%17.2%−35.9%−21.4%83%
2000-20268.2%6.2%−51.8%−50.0%77%

The edge faded as the market modernized

Annualized total return by era

Buy & hold Dow Theory
The rule wins clearly in 1900–29 and loses in 2000–26. Rolling 20-year windows tell the same story: the rule was ahead of buy and hold in every window starting before 1930, and behind by roughly 2.6 to 3 points a year in the windows ending 2009 and 2019.

The episode table shows why. In the classic crashes it did what the theory promises. In the last two big drawdowns that were choppy rather than one-way, it did the opposite.

What the rule did in famous drawdowns · total return over the window
EpisodeBuy & holdDow Theory
1929–32 crash (Sep ’29–Jul ’32)−82.8%−43.8%
1937–38 recession−33.0%−6.2%
1987 crash (Aug–Dec)−23.6%−11.3%
2000–02 dot-com bear−23.4%−47.9%
2007–09 financial crisis−42.8%−29.1%
2020 COVID crash (Feb–Mar)−22.0%−4.1%
2022 rate shock (Jan–Oct)−8.4%−19.0%

The 2000–02 bear is the instructive one. The Dow fell nearly 40% peak to trough — but in a series of violent rallies. The rule sold in February 2000, October 2000, March 2001 and September 2001, and each time it re-bought after a bounce, 4% to 16% above where it had sold. Four sells, four higher re-buys. 2022 rhymed: it sold in September near the low and re-bought 15% higher in February. In the 1987 crash, by contrast, it flashed a sell on October 15 and was in cash for Black Monday.

04Does the “confirmation” part matter?

This is the question the whole theory rests on. If you need both indexes to agree, you should do better than watching either one alone. So I ran the same ten parameter variants three ways: Industrials only, Transports only, and both required to confirm.

Confirmation beats the Industrials alone, but not the Transports

Average annualized return across 10 parameter variants · 1900–2026

Requiring confirmation helps relative to the Industrials by themselves: +0.5 points a year and a much shallower drawdown (−48% vs −62% on average). But the Transports on their own do a little better still (9.0% and −45%). Only 2 of the 10 Transports-only variants, and 3 of the 10 confirmed variants, beat buy and hold.

So the “two indexes must agree” rule adds something over the Industrials, but it isn’t clear it adds anything over just watching the Transports. That’s a plausible story — the Transports are the more cyclical, more volatile index and tend to roll over first — but I’d treat it as a hypothesis, not a finding. It could equally be that a faster, noisier index just makes a better trend filter.

The parameter grid shows how much the result depends on the choices. Every variant is below, with the base rule highlighted:

All ten variants · both indexes confirm · 1900–2026
Reaction sizeMin durationCAGRMax DDMARBuys
3%none9.2%−43.4%0.21152
3%10 days10.0%−45.2%0.2291
5%none9.8%−54.0%0.1876
5%10 days10.5%−43.5%0.2455
7.5%none8.5%−59.2%0.1443
7.5%10 days8.9%−39.3%0.2336
10%none7.2%−49.5%0.1530
10%10 days7.8%−47.5%0.1626
15%none7.8%−53.0%0.1512
15%10 days8.4%−47.5%0.1811

Nothing here is a cliff-edge, but the trend is clear: CAGR is about 9–10.5% for 3–5% reactions and drops to 7–8.5% for 10–15%, because bigger thresholds react later. Adding the 10-day minimum duration helps in every row, by 0.4 to 0.8 points. The base rule is not the best cell (that is 5% with the duration filter, 10.5%) but is third best of ten.

05Is it just luck?

A fair worry: with only 76 trades, could any in-and-out pattern with the same amount of cash time look this good? I tested it directly. I took the rule’s exact in/out pattern — same number of trades, same lengths of time in cash — and slid it to a random position along the 126 years, 3,000 times. If the timing were meaningless, the real result would land in the middle.

The real timing beats 99% of random re-timings

Return per unit of worst drawdown (MAR), 3,000 random shifts of the same trade pattern

The rule’s actual MAR (0.18) sits at the 99th percentile. On CAGR it’s the 98.5th percentile, and on max drawdown the 97th. The typical random shift earns 8.0% with a −79% drawdown. So it isn’t luck: the signals really do land near turning points more often than chance.

Two caveats keep that from being a victory lap. First, the random shifts still carry the 1930s inside them, so part of what “real” means here is that the rule reliably caught the big early crashes. Second, individual sell signals are right less often than a coin flip: of 75, only 31 (41%) were followed by lower prices, and 44 by higher. The value comes from size, not frequency. The six best calls account for about half of all the decline the rule avoided.

06What about the published records?

The most-cited modern mechanical version is Jack Schannep’s. His published record claims 13.7% a year for 1953–2025 against 10.9% for buy and hold. Its rules differ from mine: it adds the S&P 500 as a third index, sets a 3% minimum reaction with duration tests, and adds exceptions such as a shorter duration after a capitulation. I can’t audit that record. What I can say is that the plain two-index version doesn’t get there. Its closest cousin in my test (3% reaction, 8-day minimum) earned 8.4% a year since 1953 against 10.7% for buy and hold, with a max drawdown of −29% (vs −52%). Better risk, lower return.

I also tried shorting the sell signals instead of moving to cash. It’s a disaster: 6.9% a year with a −76% drawdown. The bear phases are too often followed by rallies.

07The honest bottom line

Dow Theory is real, but it’s a smaller thing than its reputation. It isn’t a return booster: across ten variants it trails buy and hold by about a point a year over 126 years, and by nearly two since 1953. What it is is a slow, crash-avoiding trend filter — it beat random timing at the 99th percentile, kept you out of most of 1929–32, 1937, 1987 and 2020, and cut the worst drawdown by a third. It works best when markets fall in one long slide, and worst in choppy bears, where it sells low and re-buys high (2000–02, 2022). The 200-day average does nearly the same job with one index instead of two.

For what it’s worth, the rule is currently long: its last signal was a buy on May 2, 2025, with the Dow near 41,300, and no bear confirmation has followed. That’s a description of the model’s state, not a forecast.

Receipts

None of this makes the old idea wrong to respect. Two indexes that have to agree is a good way to make a trend prove itself, and the crash protection in 1930 and 1987 was real. Chasing it one rabbit hole deeper just changes what to expect from it: not a way to beat the market, but a way to sit out some of its worst stretches — at the price of some whipsaws and a good deal of patience.