Analytics / Macro

US Recession Risk

A regime-aware, 12-month monitoring signal combining the yield curve with macro-finance conditions.
Model vintage
2026-06-01
Dashboard generated
2026-08-01 11:27 UTC
Next scheduled release
2026-09-01 10:00 UTC

Current signal

Latest estimate and historical context
12-month recession probability 5.9% Elevated risk Model vintage 2026-06-01
Yield curve 9.9%
Macro-finance 5.1%
Reliability weight 0.83
Weight sensitivity 6.4% → 5.4% Leaning toward macro-finance signals

Read the level. Higher values indicate conditions that historically preceded recessions.

Read the direction. A rising line means risk is building; a falling line means it is easing.

Keep the uncertainty. This is a monitoring signal, not a guaranteed outcome or calendar forecast.

Current signal

How the signal works

Two models and a reliability gate

What it measures

A monitoring signal that combines a classic yield-curve model with broader measures of financial and labor-market stress.

Why regime-aware

The curve can warn for a long time without broad stress. A learned gate shifts weight toward the model that has been more reliable out-of-sample.

What drives it

  • Baseline: yield-curve probit.
  • Augmented: macro-finance elastic-net.
  • Gate: learned mixture weight w.

Indicator Dashboard

What inputs look like right now

What you see: the four key inputs standardized as z-scores for visual comparison.

IndicatorWhat it capturesStress direction
Yield curve spread10Y-3M Treasury rate differenceLower (inverted) = more stress
Credit stress (EBP)Excess bond premiumHigher = more stress
Financial conditions (NFCI)Chicago Fed indexHigher = tighter
Labor market (claims)Initial claims signalHigher = deteriorating
Note: The augmented model combines the term spread with these macro-finance signals in a single elastic-net. See the Model Design panel for how it relates to the baseline.
Indicator Dashboard

Reliability Gate (Mixture Weight)

Which model the system trusts right now

What this is: The reliability gate outputs w, the weight placed on the macro-finance model when forming the hybrid probability.

How to read: When w is near 0, the system is mostly trusting the yield curve. When w is near 1, it is mostly trusting credit and labor-market stress signals.

Why it helps: The two models can genuinely disagree — the curve can invert while broader stress stays low, or stress can build while the curve still looks benign. The gate learns, out-of-sample, which read to weight more heavily.

0 = Yield curve 1 = Macro-finance
Current w: 0.83 • Leaning toward macro-finance signals
Reliability Gate (Mixture Weight)

Model Design

What goes into each model — and why it stays simple

Two models: The baseline uses the yield-curve spread as its only input — the classic Estrella-style recession probit. The augmented model is an additive macro-finance specification: the same spread plus five broader stress measures.

Augmented features: SPREAD, CREDIT_Z, CREDIT_CHG3M, NFCI, NFCI_CHG3M, CLAIMS_SIG

Why additive — and why not interaction terms: An inverted curve plausibly means different things depending on context, so we tested an explicit interaction version — letting the spread's effect bend with credit stress, financial conditions, and labor markets (SPREAD × CREDIT_Z, and so on). In a controlled out-of-sample validation those interaction terms added no distinguishable value: the elastic-net shrank them toward zero and kept the plain spread as the dominant signal, and the variant that dropped the standalone spread was actually less well calibrated. The deployed model therefore stays additive.

What this buys: The additive model is simpler, better calibrated out-of-sample, and matches the specification in the accompanying paper. The reliability gate still has a meaningful job: the baseline reads the curve alone, the augmented model reads the curve in the company of broader stress, and the gate learns which read has been more trustworthy over time.

Regularization: The elastic-net's L1 penalty shrinks unhelpful coefficients toward zero, so the macro features have to earn their place. All inputs are constructed to be leakage-free as of each scoring date.

Known Limitations

What this model cannot do

Why this section exists: Transparency about model limitations is as important as the headline probability.

1. Small sample size

The out-of-sample evaluation covers only 3 recession episodes (roughly 28 recession months out of 356 total). Any statistical analysis of threshold selection, lead times, or calibration is inherently limited by this sample. Results should be treated as suggestive, not definitive.

2. Additive, linear specification

The augmented model combines its features additively on the log-odds scale. It does not capture conditional effects — for example, an inversion meaning something different when credit stress is high versus low. We tested explicit interaction terms (SPREAD × CREDIT_Z, and others) and they added no distinguishable out-of-sample value, so they were left out for parsimony. A tree-based or neural approach could capture richer non-linearities, but at the cost of interpretability and with higher overfitting risk on 3 episodes.

3. Alert threshold interpretation

With only 3 episodes and a low base rate, no single threshold achieves clean separation between true and false alarms. The probability level and direction matter more than whether it crosses a specific line.

4. What this system is not

This is a monitoring tool based on historical statistical relationships. It does not model causal mechanisms, it cannot anticipate unprecedented shocks (e.g., a pandemic), and it should never be used as the sole basis for financial decisions. It is one input among many.