{"site":{"name":"Koji","description":"AI-native customer research platform that helps teams conduct, analyze, and synthesize customer interviews at scale.","url":"https://www.koji.so","contentTypes":["blog","documentation"],"lastUpdated":"2026-09-25T16:08:59.429Z"},"content":[{"type":"documentation","id":"7cedcb0d-54c5-49d0-ab1f-6cd04f2e7e34","slug":"taphonomy-customer-feedback-preservation-bias","title":"The Taphonomy of Customer Feedback: Which Complaints Survive to Reach You (2026)","url":"https://www.koji.so/docs/taphonomy-customer-feedback-preservation-bias","summary":"Taphonomy, the paleontological study of how remains are preserved or destroyed, gives product teams a sharper tool than survivorship bias: the filter that removes customer evidence is class-biased and therefore predictable. Hard parts (tickets, scores, dropdown reasons) preserve readily; soft parts (workarounds, silent tradeoffs, hesitation, the job behind a request) have no container and are destroyed by default. Losses split into a pre-capture phase and a post-capture phase. The correction is to study the filter and to build a conservation deposit, a capture method that preserves soft tissue, which is what an AI-moderated conversation does and a form cannot.","content":"**Most customer feedback is destroyed before it ever reaches you, and the destruction is not random: it removes specific, predictable classes of evidence.** Paleontologists have a name for the study of that destruction. It is called taphonomy, and it is the most useful idea product teams have never borrowed. Once you know which kinds of evidence your collection channels systematically destroy, you can stop treating your feedback pile as a sample of reality and start collecting the missing classes on purpose.\n\nThe practical payoff: your support tickets, NPS comments, and sales-call notes are not a small, faithful sample of what customers experience. They are a biased cast of the hardest, most durable parts of it. The soft parts, which are usually the majority, leave no trace at all unless you use a capture method designed to preserve them.\n\n## What taphonomy is, and why it is not just another bias\n\nThe term was introduced to paleontology in 1940 by the Soviet scientist Ivan Efremov, and it names the study of how organisms decay and become fossilized or preserved in the paleontological record. Taphonomic phenomena are conventionally grouped into two phases: biostratinomy, the events that occur between the death of the organism and its burial, and diagenesis, the events that occur after burial.\n\nThat two-phase split is the part worth stealing. It separates losses that happen *before* the evidence enters your record from losses that happen *after* it is already in there. Most teams only ever audit the second phase.\n\nThe governing principle is broader than any single mechanism: in taphonomy, any factor that affects the likelihood that an organism is preserved as a fossil is a potential source of bias. Note the structure of that claim. It is not that the record is small. It is that the *filter has a shape*, and the shape is knowable.\n\nThis is what separates taphonomic thinking from the more familiar framing in [survivorship bias in customer research](/docs/survivorship-bias-customer-research). Survivorship bias says: you are only hearing from the customers who stayed. True, and important. Taphonomy says something stronger and more actionable: the destruction process selectively removes *particular classes* of evidence, so you can predict exactly which insight types are absent and go collect those specific types directly. Survivorship bias tells you the sample is wrong. Taphonomy tells you which shelf to go and look on.\n\nIt is also distinct from the problem described in [why your repository cannot show you what it lost](/docs/undocumented-deletion-research-repository). That is a diagenetic problem: material that made it into the record and then degraded or vanished without a log entry. Taphonomy asks the prior question, the biostratinomic one: what never made it into the record at all, and why not?\n\n## Hard parts and soft parts\n\nThe central taphonomic fact is that preservation is class-biased. Organisms that contain hard parts have a far greater chance of being represented in the fossil record than organisms consisting of soft tissue only.\n\nThe magnitude of that bias is not a rounding error. Soft-bodied organisms may form 30% to 100% of the biota, but most fossil assemblages preserve none of this unseen diversity. The missing class is routinely the majority.\n\nThe sharpest single measurement comes from the Burgess Shale in British Columbia, a 508-million-year-old deposit famous for the exceptional preservation of the soft parts of its fossils. Only about 14% of genera in the Burgess Shale had possessed biomineralized tissues in life. Read that the other way round: in an ordinary deposit, roughly 86% of that fauna would simply not exist as far as the record is concerned. Not underrepresented. Absent.\n\nYour feedback record has hard parts and soft parts too, and the ratio is not kinder.\n\n| Hard parts (preserve readily) | Soft parts (rarely preserved) |\n| --- | --- |\n| A bug report with reproduction steps | The workaround the user invented and now considers normal |\n| A cancellation reason picked from a dropdown | The three months of quiet erosion that preceded it |\n| A feature request, stated as a request | The underlying job the request was a guess at |\n| An NPS score with a one-line comment | The tradeoff the customer weighed and resolved silently |\n| A complaint escalated by a paying account | The identical problem experienced by someone who never escalates |\n| A quote a salesperson wrote down | The hesitation, the pause, the qualifier before it |\n\nEvery row on the left has a container built for it: a ticket field, a form, a CRM note. Every row on the right has none. The asymmetry is structural, not cultural, and no amount of encouraging customers to \"share more feedback\" changes which container exists.\n\n## The first phase: what dies before burial\n\nBiostratinomy, for feedback, is everything that happens between a customer having an experience and that experience entering a system you can query.\n\nThe attrition at this stage is severe and it has been measured. The 2006 Retail Customer Dissatisfaction Study, run by the Jay H. Baker Retailing Initiative at Wharton with the Verde Group, found that only 6% of shoppers who experienced a problem contacted the company about it, while 31% went on to tell friends, family, or colleagues what happened. The experience was not forgotten. It was preserved, at volume, in a deposit you do not own.\n\nNothing about that 6% is a random draw from the 100%. Contacting a company selects hard on effort tolerance, on how articulate the customer is, on whether they believe complaining works, and on whether the problem was legible enough to describe in a ticket form. Those are precisely the traits that correlate with the *kind* of problem being reported. Sharp, nameable, single-cause failures survive. Diffuse, cumulative, hard-to-articulate friction does not.\n\n## The second phase: what degrades after burial\n\nDiagenesis is the set of changes after the evidence is already in your record, and it is where paraphrase does its damage.\n\nA customer says something specific and slightly awkward. A CSM writes a cleaner version of it in the CRM. A summary of that note reaches a planning document. A slide quotes the summary. At each step the utterance gets more quotable and less informative, and every step is invisible in the final artifact, which looks like a clean fact. This is why [a relayed customer claim is not customer evidence](/docs/hearsay-secondhand-customer-evidence), and why the verbatim matters more than the tidy version of it.\n\n## How paleontologists correct for it, and how you can\n\nPaleontology does not solve this problem by wishing the record were better. It solves it two ways, and both translate.\n\nThe first is to study the filter itself. If you know the mechanism that destroys soft tissue, you know what is missing, and you can reason about the gap instead of ignoring it. Practically: write down which of your insight classes has no container, and stop treating its absence from dashboards as evidence of absence in the world.\n\nThe second is to find, or create, a conservation deposit. These are the deposits known for the exceptional preservation of fossilized organisms or traces, and their value is precise: they preserve lightly sclerotized and soft-bodied organisms or traces of organisms that are not otherwise preserved in the usual shelly and bony fossil record, and so they offer more complete records of ancient biodiversity and behavior. The Burgess Shale is one. The Chengjiang deposits in Yunnan, at 518 million years, are another.\n\nYou cannot wait 500 million years for favorable conditions. You can, however, build the conditions. A conservation deposit for customer evidence is any capture method that records the soft parts: the hesitation, the workaround, the abandoned sentence, the tradeoff reasoned aloud. A conversation does that. A form does not. This is the entire reason a moderated interview outperforms a survey on the same question, and it is the reason an unmoderated survey at ten times the sample size does not close the gap: it is a concentration deposit, dense in hard parts, empty of everything else.\n\n## How Koji handles this\n\nKoji is built to be the conservation deposit in your stack rather than another form that collects hard parts.\n\n- **AI-moderated conversations, not forms.** Koji's AI interviewer asks a real follow-up when an answer is thin, vague, or interesting, which is the only mechanism that gets the unprompted reasoning on the record at all. A static survey cannot probe; it captures the answer the respondent was willing to type first.\n- **Voice and text in the same study.** Voice interviews preserve hesitation, self-correction, and the qualifier before the claim. Koji runs voice and text side by side, so you are not trading modality for reach.\n- **Verbatim transcripts kept under the analysis.** Every theme in a Koji report stays linked to the transcript segment it came from, so the diagenetic chain from utterance to slide is inspectable instead of assumed.\n- **In-vivo theme tags.** Koji's analysis can tag a theme in the participant's own framing rather than forcing it into a house label, which is how a soft, awkwardly-phrased signal survives coding instead of being rounded to the nearest known category.\n- **Structured questions where hard parts are genuinely what you want.** Koji supports six question types: open_ended, scale, single_choice, multiple_choice, ranking, and yes_no. The discipline is to use scale and choice questions for the durable, countable things and open_ended with AI follow-up for the soft tissue, in one session, rather than running two disconnected studies. See [structured questions in AI interviews](/docs/structured-questions-guide).\n- **A quality gate on the deposit.** Koji scores each conversation from 1 to 5 on relevance, depth, and coverage, and only conversations scoring 3 or above consume a credit. A thin interview is a poorly preserved specimen, and you are not billed for it.\n\nThe practical workflow: run the countable questions and the conversational probe in the same Koji study, then treat the resulting transcripts as your primary deposit and the ticket queue as what it actually is, a bone bed of hard parts, useful for counting and useless for anatomy.\n\n## Frequently asked questions\n\n### Is taphonomy just a fancy word for survivorship bias?\n\nNo, and the difference is operational. Survivorship bias says your sample over-represents survivors. Taphonomy says the filter destroys identifiable *classes* of evidence in a predictable way, which means you can name the missing class and go collect it deliberately. Survivorship bias diagnoses; taphonomy prescribes. The two also sit at different stages, and it is worth reading [survivorship bias in customer research](/docs/survivorship-bias-customer-research) alongside this.\n\n### What counts as a soft part in customer feedback?\n\nAnything with no container built for it: an invented workaround the user no longer notices, a tradeoff resolved silently, the job a feature request was a guess at, hesitation before a confident claim, and the reasoning that connects a complaint to its cause. If your systems have no field for it, it is soft tissue, and it is being destroyed by default.\n\n### Does collecting more feedback fix the bias?\n\nNo. More volume through the same channel deepens a concentration deposit: you get more hard parts, more precisely counted, with the same classes missing. Fixing a class-biased filter requires a different capture method, not a larger sample through the existing one. This is why running a bigger survey rarely resolves a question a handful of conversations would answer.\n\n### How do I know which classes are missing if they leave no trace?\n\nStudy the filter rather than the record. List your collection channels, and for each one write down what it structurally cannot capture: a dropdown cannot capture a reason not on the list, a ticket form cannot capture a problem the user does not consider reportable, an NPS comment box cannot capture a tradeoff. The union of those blind spots is your missing-class list, and it is stable enough to plan against.\n\n### Where does the 6% figure come from?\n\nFrom the 2006 Retail Customer Dissatisfaction Study by the Jay H. Baker Retailing Initiative at Wharton and the Verde Group, which found that only 6% of shoppers who had a problem contacted the company, while 31% told friends, family, or colleagues. It is a retail study, so treat the exact number as illustrative rather than as your own conversion rate, and treat the direction as robust.\n\n### How does Koji change the preservation ratio in practice?\n\nBy replacing the form with a conversation at the point of capture. Koji's AI interviewer probes thin answers, runs in voice or text, keeps the verbatim transcript under every theme, and can tag themes in the participant's own words. That combination preserves the classes of evidence that a survey destroys, which is the only intervention that changes the ratio rather than the volume.\n\n## Related Resources\n\n- [Survivorship Bias in Customer Research](/docs/survivorship-bias-customer-research) - the sampling half of the problem, and the right companion to this piece\n- [Structured Questions in AI Interviews](/docs/structured-questions-guide) - when to use each of the six question types, and when to let the AI probe instead\n- [Nobody Logged the Deletion](/docs/undocumented-deletion-research-repository) - the diagenetic failure mode, after evidence is already in your record\n- [Why the Loudest Complaint Hides the Real One](/docs/dominant-complaint-masking-interviews) - what happens when a preserved signal drowns out a fainter one\n- [Hearsay in Product Research](/docs/hearsay-secondhand-customer-evidence) - how paraphrase degrades a quote on its way to a slide\n- [In-Product Research Recruiting](/docs/recruiting-from-your-product) - reaching the customers who would never open a ticket\n","category":"Research Methods","lastModified":"2026-09-25T03:39:27.958283+00:00","metaTitle":"Taphonomy of Customer Feedback: Which Complaints Survive | Koji","metaDescription":"Your feedback channels destroy predictable classes of evidence. How to name what is missing and capture the soft parts a survey cannot preserve.","keywords":["customer feedback bias","feedback preservation bias","why customers do not complain","taphonomy","survivorship bias customer research","voice of customer gaps","unvoiced feedback"],"aiSummary":"Taphonomy, the paleontological study of how remains are preserved or destroyed, gives product teams a sharper tool than survivorship bias: the filter that removes customer evidence is class-biased and therefore predictable. Hard parts (tickets, scores, dropdown reasons) preserve readily; soft parts (workarounds, silent tradeoffs, hesitation, the job behind a request) have no container and are destroyed by default. Losses split into a pre-capture phase and a post-capture phase. The correction is to study the filter and to build a conservation deposit, a capture method that preserves soft tissue, which is what an AI-moderated conversation does and a form cannot.","aiPrerequisites":[],"aiLearningOutcomes":[],"aiDifficulty":"intermediate","aiEstimatedTime":"9 min"}],"pagination":{"total":1,"returned":1,"offset":0}}