{"site":{"name":"Koji","description":"AI-native customer research platform that helps teams conduct, analyze, and synthesize customer interviews at scale.","url":"https://www.koji.so","contentTypes":["blog","documentation"],"lastUpdated":"2026-08-15T20:17:41.787Z"},"content":[{"type":"documentation","id":"8381e82b-ceec-4086-aece-7a467084f607","slug":"churn-hazard-curve-tenure-analysis","title":"The Churn Hazard Curve: Why One Churn Rate Hides Three Different Problems (2026)","url":"https://www.koji.so/docs/churn-hazard-curve-tenure-analysis","summary":"Explains the difference between a retention (survival) curve and a hazard curve, shows how to compute conditional churn by tenure, imports the Nowlan and Heap 1978 six age-reliability patterns to argue most accounts do not wear out, identifies three churn regimes with distinct research questions, and warns that a falling aggregate hazard may be a heterogeneity sorting effect rather than growing loyalty.","content":"The conditional probability that an account churns in its ninth month is a different number, with a different cause and a different fix, than the conditional probability that it churns in its first. Your monthly churn rate averages them into one figure, and every intervention you fund off that figure is aimed at an average that describes no account in your book. The fix is to plot the hazard - churn as a share of accounts that survived to reach that tenure - instead of the retention curve, and then to run a separate study for each regime the hazard reveals.\n\nAviation learned this in the 1960s. When United Airlines plotted conditional failure probability against operating age for its components, the familiar wear-out story turned out to describe 6 percent of the fleet. This guide ports the analysis, the arithmetic, and the trap that comes with it.\n\n## Retention curve and hazard curve are not the same chart\n\nMost teams have a retention curve. Very few have a hazard curve, and the two answer different questions.\n\nA retention curve is a survival function: of the accounts that started, what share is still here at month t? It only ever goes down. A hazard curve is conditional: of the accounts that reached month t, what share left during month t? It can go down, stay flat, or go up.\n\nReliability engineering defines the second one precisely. It is \"the probability that an item entering a given age interval will fail during that interval\" - a measure also known as \"the hazard rate or the local failure rate\" ([Nowlan and Heap, *Reliability-Centered Maintenance*, 1978](https://www.omdec.com/wikifiles/nowlanHeap.pdf), p. 25).\n\nThe arithmetic is a division you are probably not doing. In the United Airlines analysis of the Pratt and Whitney JT8D-7 engine, an engine had a probability of .692 of reaching 1,000 hours. The probability density of failure in the following 200-hour interval was .053. So the conditional probability of failure for an engine that actually reached 1,000 hours was .053 / .692 = .077 - noticeably higher than the unconditional .053.\n\nTranslate that to accounts:\n\n| Month | Accounts entering the month | Churned during the month | Unconditional (share of original cohort) | Hazard (share of survivors) |\n|---|---|---|---|---|\n| 1 | 1,000 | 120 | 12.0% | 12.0% |\n| 2 | 880 | 70 | 7.0% | 8.0% |\n| 6 | 640 | 19 | 1.9% | 3.0% |\n| 12 | 545 | 16 | 1.6% | 2.9% |\n| 24 | 470 | 24 | 2.4% | 5.1% |\n\nRead the fourth column and the story is \"churn keeps falling, our curve is flattening, we have product-market fit.\" Read the fifth and there is a spike at month 1, a long quiet stretch, and something turning back up at month 24 that the flattening curve cannot show you, because a survival curve that declines slowly looks reassuring whether the underlying risk is falling or climbing.\n\nOur [cohort analysis guide](/docs/cohort-analysis-guide) covers the survival side properly - the three curve shapes and what a plateau implies about your market. This article is the derivative view, and it is the one that tells you what to fund.\n\n## What happened when an airline plotted the hazard instead\n\nNowlan and Heap's study is the most-cited empirical result in maintenance engineering, and the finding is not what anyone expected. Every item United analyzed fell into one of six age-reliability patterns:\n\n| Pattern | Shape | Share of items |\n|---|---|---|\n| A | Bathtub: infant mortality, then flat, then pronounced wear-out | 4% |\n| B | Flat or slowly rising, then pronounced wear-out | 2% |\n| C | Gradually increasing, no identifiable wear-out age | 5% |\n| D | Low when new, quick rise to a constant level | 7% |\n| E | Constant at all ages (exponential) | 14% |\n| F | Infant mortality, then constant or very slowly increasing | 68% |\n\nThe conclusion, verbatim: \"Some 89 percent of the items analyzed had no wearout zone; therefore their performance could not be improved by the imposition of an age limit.\" Only 11 percent (patterns A, B and C) might benefit from a limit on operating age at all, and \"Only 6% of the items studied showed pronounced wearout characteristics.\"\n\nThe most quoted line is about the shape everyone assumes is universal: \"Although it is often assumed that the bathtub curve is representative of most items, note that just 4% of the items fell into this pattern.\"\n\nPattern F - 68 percent, the single largest group - is infant mortality followed by a flat tail, and Nowlan and Heap note it is \"particularly applicable to electronic equipment.\" Complex items with many interacting parts do not wear out. They either fail early because something about the installation was wrong, or they run indefinitely. A B2B software account is a complex item.\n\n## The three regimes, translated\n\n| Regime | What the hazard does | What is actually happening | The research question |\n|---|---|---|---|\n| Infant mortality (months 0-3) | High, falling fast | Wrong-fit accounts, failed implementation, the buyer's problem was misdiagnosed at sale | Why did this account never reach the value it was sold? |\n| Useful life (months 4-18) | Low, roughly flat | Exogenous shocks: budget cuts, the champion leaves, reorgs, acquisition | What happened outside the product? |\n| Wear-out (month 18+) | Rising again | Accumulated workarounds, outgrown the data model, a competitor closed the gap | What did we stop being able to do for you? |\n\nThese are three different studies with three different recruiting screens. A single \"why did you churn\" survey blended across all tenures produces an average of three unrelated causal stories, which is why the answers so often read as bland.\n\nThe regimes also rank your options. Infant-mortality churn is largely a sales-qualification and onboarding problem, and it is the cheapest to fix because the accounts are numerous and the cause is usually recent and recallable. Useful-life churn is mostly not addressable by the product at all - it is the constant-hazard background rate, and spending against it has poor returns. Wear-out churn is the expensive kind: the accounts are large, tenured, and the cause accumulated over a year before anyone noticed.\n\n## The trap: a falling hazard may be sorting, not loyalty\n\nHere is the part that will change how you read your own chart, and it is the reason a falling hazard is weaker evidence than it looks.\n\nSuppose every single customer has a fixed, unchanging monthly churn probability, but customers differ from each other - some are 1 percent per month, some are 20 percent. Nobody becomes more loyal over time. What does the aggregate hazard curve do?\n\nIt falls. Steeply.\n\nThe high-risk customers leave first, so the surviving population is progressively enriched with low-risk customers. The aggregate rate drops even though no individual's rate moved. Fader and Hardie make exactly this point about subscription businesses: increasing cohort-level retention \"is purely due to cross-sectional heterogeneity, with individual customers having a constant propensity to churn. Cohort-level retention rates increase because those customers with high churn propensities drop out early on, leaving an ever-increasing proportion of customers who have low propensities to churn\" ([Fader, Hardie, Liu, Davin and Steenburgh, *Journal of Interactive Marketing*, 2018](http://brucehardie.com/papers/037/BdW_JIM_2018-01-10_rev.pdf)).\n\nThey note this \"flies in the face of conventional wisdom, which assumes that a customer's propensity to churn decreases the longer their tenure with the firm.\" It is a known statistical result that \"unobserved heterogeneity induces spurious negative duration dependence\" - a mixed population of constant-risk individuals imitates a population that is getting safer.\n\nTheir headline finding is blunter still: \"even when aggregate retention rates are monotonically increasing, the individual-level churn probabilities are unlikely to be declining over time.\"\n\nThe practical consequence: **a falling hazard is not evidence that your product is getting stickier.** It is equally consistent with a product that never gets stickier and a customer base that is sorting itself. Distinguishing the two requires segmenting by something you knew at signup - plan, segment, acquisition channel, ICP fit score - and checking whether the hazard still falls within each segment. If it flattens once you segment, you were watching sorting.\n\nThis also means the tempting inference from a flat retention plateau - that the plateau height equals your real market - is one reading among several. Sorting produces the same picture.\n\n## What this is not\n\nThree neighbouring problems get confused with this one, and the distinctions matter because the fixes differ:\n\n- **[Immortal time bias](/docs/immortal-time-bias-retention-analysis)** is a time-alignment error: adopters are credited with time during which they could not have churned. That is a bug in how you assign person-time. The hazard curve is about the shape of risk across tenure once alignment is correct.\n- **[The healthy adherer effect](/docs/healthy-adherer-effect-product-research)** is a selection error: the accounts that complete onboarding differ from those that do not, before the onboarding does anything. That is about who is in the treated group.\n- **[Survivorship bias](/docs/survivorship-bias-customer-research)** is about who you never reached at all.\n\nA hazard curve can be perfectly computed and still be misread, which is what the sorting trap above describes. All four can be present in the same chart.\n\n## Which method belongs in which regime\n\nBecause the causes differ, the evidence has to be collected differently.\n\nInfant-mortality churn is recent, so recall is good and the sample is large - this is where a structured exit study works. Useful-life churn is exogenous, so the useful question is not about your product; it is about what changed at the account, and much of it will be uncontrollable. Wear-out churn is the hardest, because the cause accumulated slowly and the customer often cannot name it. Nobody remembers the month the workarounds became intolerable. That regime needs longitudinal contact with still-active tenured accounts, not exit interviews - by the time they cancel, the reconstruction is post-hoc.\n\nThe awkward truth is that this is three recruiting screens, three discussion guides, and three analyses. Most teams run one, because three was never affordable.\n\n## How Koji changes the economics\n\nRunning a separate study per tenure regime is a scheduling and cost problem before it is a methods problem. Koji removes the constraint that makes teams collapse three studies into one:\n\n- **Segment-matched studies in parallel.** Define three cohorts by tenure band and field all three at once with AI-moderated interviews. The marginal cost of the third study is close to the cost of the first, so the per-regime design stops competing with the deadline.\n- **Identical probing across bands.** A human moderator running month-1 and month-24 interviews will probe the interesting one harder, and the resulting difference is inseparable from the method - the mechanism described in our guide to [surveillance bias](/docs/surveillance-bias-detection-research). An AI moderator applies the same brief and the same follow-up logic to every band, which is what makes cross-regime comparison legitimate.\n- **A detection-independent baseline.** Koji supports six [structured question types](/docs/structured-questions-guide) - open_ended, scale, single_choice, multiple_choice, ranking, and yes_no. The five closed types mean the same thing regardless of how long the conversation ran, so they give you a fixed yardstick across tenure bands. If open-ended themes diverge sharply between month 1 and month 24 but the scale and ranking data does not, you are looking at a probing artifact rather than a real regime difference.\n- **Continuous fielding for the wear-out regime.** Wear-out is the regime that defeats exit research. A standing quarterly study of tenured active accounts catches the accumulation while the customer can still describe it.\n\nTeams adopting AI-assisted research consistently report time-to-insight measured in days rather than the six-to-eight weeks a three-arm qualitative study traditionally consumes - which is the difference between segmenting your hazard curve and rounding it to one number.\n\n## A working checklist\n\n1. Plot the hazard, not just retention: churned in month t divided by accounts that reached month t.\n2. Plot it on a log scale if your month-1 spike compresses everything else flat.\n3. Mark the three regimes and check whether the tail turns back up. A rising tail is the expensive finding.\n4. Re-plot within segments you knew at signup. If the decline disappears, you were watching sorting, not loyalty.\n5. Do not average a churn rate across regimes for any decision that funds work.\n6. Assign one study per regime, with the recruiting screen written from the hazard chart.\n7. For the wear-out regime, interview active tenured accounts, not churned ones.\n8. Re-run the whole analysis quarterly - the regime boundaries move when onboarding or pricing changes.\n\n## Frequently asked questions\n\n### What is the difference between a retention curve and a hazard curve?\n\nA retention curve is a survival function: the share of an original cohort still active at each point in time. It can only decrease. A hazard curve is conditional: among accounts that survived to month t, the share that left during month t. It can rise or fall. Two products with identical retention curves can have completely different hazard shapes, and the hazard is what tells you which intervention will pay.\n\n### How do I calculate a churn hazard rate?\n\nFor each tenure month, divide the number of accounts that churned during that month by the number of accounts that entered that month alive. Do not divide by the original cohort size - that gives the unconditional rate, which understates late-life risk because the denominator includes accounts that already left. The reliability-engineering worked example is .053 / .692 = .077.\n\n### Does a flattening retention curve prove product-market fit?\n\nNot on its own. A flattening curve is exactly what you would see if every customer had a fixed churn probability and the high-risk ones simply left first. Fader and Hardie show that increasing cohort-level retention can be entirely a sorting effect from cross-sectional heterogeneity, with no individual becoming more loyal. Segment by a signup-time attribute and see whether the flattening survives.\n\n### How many tenure buckets should I use?\n\nEnough that each bucket has a stable denominator - as a rule of thumb, at least a few hundred accounts entering the bucket before the rate stops being noise. Monthly buckets for the first quarter, where the action is, then quarterly buckets afterwards is a reasonable default for most B2B books.\n\n### Why does my churn hazard go up again for long-tenured accounts?\n\nThat is the wear-out regime, and it is the pattern worth investigating first because those accounts are usually your largest. Common causes are accumulated workarounds that finally exceed tolerance, outgrowing the data model or permission structure, and a competitor closing a gap that mattered at renewal. It is rarely a single event, which is why churned-customer interviews reconstruct it badly.\n\n### Can I use this for user-level retention, not just accounts?\n\nYes. The arithmetic is identical for users, seats, or any unit that can leave. The regimes tend to compress - user-level infant mortality often plays out over days rather than months - but the three-regime structure and the sorting trap both apply unchanged.\n\n## Related Resources\n\n- [Average Customer Lifetime Is Not 1/Churn](/docs/customer-lifetime-mtbf-ltv-formula) - why the summary statistic hides this shape\n- [Preventive vs Condition-Based Check-Ins](/docs/condition-based-vs-scheduled-customer-research) - what to do once you know the shape\n- [Cohort Analysis: How to Read Retention and Find the Why](/docs/cohort-analysis-guide) - the survival-curve side of the same data\n- [Immortal Time Bias in Retention Analysis](/docs/immortal-time-bias-retention-analysis) - the time-alignment error that fakes a feature effect\n- [The Healthy Adherer Effect](/docs/healthy-adherer-effect-product-research) - why onboarding completers always look better\n- [Structured Questions Guide](/docs/structured-questions-guide) - the six question types and when to use each\n- [Customer Retention Research](/docs/customer-retention-research) - the full retention research program\n- [Case-Control Research for Churn and Lost Deals](/docs/case-control-research-churn-lost-deals) - designing the comparison properly\n","category":"Research Methods","lastModified":"2026-08-15T03:26:02.548835+00:00","metaTitle":"Churn Hazard Curve: Read Churn Risk by Tenure (2026 Guide)","metaDescription":"A single churn rate hides three different problems. How to plot the churn hazard by tenure, read infant mortality, useful life and wear-out regimes, and avoid the heterogeneity sorting trap.","keywords":["churn hazard curve","churn rate by tenure","hazard rate customer churn","retention curve vs hazard","churn analysis by cohort","bathtub curve customers","conditional probability of churn"],"aiSummary":"Explains the difference between a retention (survival) curve and a hazard curve, shows how to compute conditional churn by tenure, imports the Nowlan and Heap 1978 six age-reliability patterns to argue most accounts do not wear out, identifies three churn regimes with distinct research questions, and warns that a falling aggregate hazard may be a heterogeneity sorting effect rather than growing loyalty.","aiPrerequisites":["Basic familiarity with cohort retention tables","Access to account-level churn dates"],"aiLearningOutcomes":["Compute a churn hazard rate correctly by conditioning on survivors","Distinguish a survival curve from a hazard curve","Identify infant mortality, useful life and wear-out regimes in your own data","Detect when a falling hazard is a sorting artifact","Assign the right research method to each regime"],"aiDifficulty":"intermediate","aiEstimatedTime":"12 min"}],"pagination":{"total":1,"returned":1,"offset":0}}