{"site":{"name":"Koji","description":"AI-native customer research platform that helps teams conduct, analyze, and synthesize customer interviews at scale.","url":"https://www.koji.so","contentTypes":["blog","documentation"],"lastUpdated":"2026-08-18T09:11:51.644Z"},"content":[{"type":"blog","id":"08484a94-8b5d-4cce-bae5-7715ba7fec68","slug":"ai-brand-misrepresentation-research-2026","title":"When AI Describes Your Product Wrong (2026): Why More Mentions Can Make Your Positioning Worse","url":"https://www.koji.so/blog/ai-brand-misrepresentation-research-2026","summary":"An AI answer is synthesised prose assembled from third-party sources, not a retrieved page, so brand visibility and description accuracy are independent variables. The EBU and BBC found 45% of AI responses contained a significant issue and 31% had sourcing problems, including misattributing claims to the organisations covered. Increasing mentions without controlling the description simply distributes a wrong description more widely, and the fix is an unprompted-description study with real buyers.","content":"## The short answer\n\nAn AI answer is not a retrieved page. It is a paraphrase assembled from many sources, most of which you did not write. That means visibility and accuracy are two independent variables, and almost every team is optimising only the first. If the model has settled on a wrong description of your product, then every additional mention distributes that wrong description more widely. For a brand being described incorrectly, more AI visibility is a larger liability, not a smaller one.\n\nThis is the second of three articles on AI answers and brand measurement. The first covered [how to measure AI search visibility honestly](/blog/ai-search-visibility-measurement-2026). The third covers [why AI answers carry no timestamp](/blog/ai-answers-outdated-product-information-2026).\n\n## What an AI answer actually is\n\nThe mental model most marketers carry over from search is wrong in one specific way, and the error is consequential.\n\nA search result is a pointer. Google decides which of your pages to show; the words the user reads are your words. If the description is wrong, you can edit the page.\n\nAn AI answer is a synthesis. The model produces new prose about you, drawing on training weights, retrieved documents, and whatever third-party pages the retrieval layer happened to surface. Nobody wrote the sentence the buyer reads. It did not exist before the question was asked, and it will be phrased differently for the next person.\n\nThis has an uncomfortable implication that follows directly and that most AI-visibility strategy ignores: **your description in AI answers is a lagging index of what other people have written about you, not of what your product does.** Your own site is one input among many, and it is the input the system is least likely to treat as a neutral account, because it is obviously promotional. Review sites, forum threads, comparison roundups, news coverage and competitors' comparison pages all get a vote, and in aggregate they outvote you.\n\n## The evidence on how often the description is wrong\n\nThe best public evidence is not from a marketing vendor. In October 2025 the European Broadcasting Union and the BBC published *News Integrity in AI Assistants*, a study involving 22 public service media organisations across 18 countries and 14 languages, in which professional journalists evaluated AI responses against 3,113 questions.\n\nThe headline findings:\n\n- 45% of all AI responses contained at least one significant issue.\n- Counting lower-severity problems as well, 81% of responses had an issue of some form.\n- Sourcing was the single biggest cause of significant issues, at 31% of all responses. That category covers \"information in the response not supported by the cited source, providing no sources at all, or making incorrect or unverifiable sourcing claims.\"\n- Accuracy problems affected 20% of responses, and the assistants were closely grouped, all between 18% and 22%.\n- Sourcing performance varied enormously by assistant. Gemini had significant sourcing issues in 72% of responses, three times ChatGPT's 24%, with Perplexity and Copilot both at 15%.\n\nOne line in the EBU report ports to brands so exactly it is worth quoting in full: \"Of particular concern for publishers are sourcing errors that misrepresent them, such as when a response misattributes an incorrect claim to them.\"\n\nSubstitute \"brands\" for \"publishers\" and that is the risk in a sentence. The failure mode is not only that the model says something false in general. It is that the model attaches a claim to *you* that you never made and cannot find the origin of.\n\nForrester's *The State Of Business Buying, 2026* corroborates this from the buyer's side, and it is notable that an analyst firm rather than a critic is saying it. AI search tools, Forrester writes, \"also known as answer engines, offer speed and efficiency, but they often deliver incomplete or unreliable information, creating mistrust. Buyers compensate for this by seeking validation from trusted sources, emphasizing the value of human contact in the buying process.\"\n\nTo be clear about what this evidence is and is not: the EBU study evaluated news questions, not product questions. Nobody has published an equivalent audit of brand descriptions at that rigour. The mechanism, though, is not news-specific. Synthesis from mixed-quality third-party sources with imperfect attribution is how these systems answer every question, including \"is Acme any good for a team of 50?\"\n\n## Why volume and accuracy move independently\n\nThis is the crux, and it is where most AI-visibility programmes have a logical hole.\n\n| What you optimise | What it changes | What it does not change |\n| --- | --- | --- |\n| More third-party mentions and citations | How often you appear in answers | Which description the model has settled on |\n| Better structured data on your site | How parseable your own pages are | How much weight your pages get against third-party accounts |\n| More comparison-page placements | Which consideration set you appear in | Whether the framing in those pages is yours or a competitor's |\n| Higher review volume | Perceived credibility signals | Whether reviews are paraphrased into the claim you want |\n\nNothing in the left-hand column controls the description. The two variables are genuinely orthogonal, which is why a brand can run a successful AEO programme, watch its mention rate climb every month, and end the year with worse-qualified pipeline than it started with. The mentions went up. The sentence the buyer read stayed wrong, and more buyers read it.\n\nThis is not the same question as whether third-party reviews are honest. That question, including what the FTC's Consumer Review Rule now makes illegal, is covered in our analysis of [what software review sites can and cannot tell you](/blog/software-review-sites-b2b-buyers-2026). The problem here survives perfect honesty upstream: entirely accurate reviews, written in good faith, get compressed into a one-paragraph description that no reviewer wrote and nobody checked.\n\n## The four ways a description goes wrong\n\n| Failure mode | What it looks like | What it costs |\n| --- | --- | --- |\n| Category misassignment | You sell a research platform; the model files you under \"survey tools\" | You are compared on the wrong axis, against the wrong rivals, and lose on criteria that do not apply |\n| Attribute error | Wrong pricing, a free tier that no longer exists, \"enterprise only\" when you self-serve | Buyers disqualify you before contact, and you never learn it happened |\n| Misattributed claim | A limitation from a competitor's comparison page, or a criticism from one review, stated as fact about you | You spend the first call defending something you never said |\n| Framing inheritance | The model adopts a reviewer's or competitor's framing as neutral description | Your positioning is written by the party with the most published content, not the best product |\n\nThe first and fourth are the expensive ones, and they are the ones that never show up in a dashboard. A tracker records that you were mentioned. It scores that as a win. It does not notice that you were mentioned as a cheaper alternative to a category you deliberately exited two years ago.\n\nCategory misassignment is particularly costly because it is self-reinforcing. Once the model files you alongside a set of competitors, buyers ask follow-up questions in that frame, comparison content gets written in that frame, and the corpus that trains the next answer contains more of that frame.\n\n## Route around it: you cannot edit the model, but you can find the belief\n\nThere is no correction desk. No major assistant offers brands a mechanism to file a factual dispute about how they are described and have it fixed. Publishing a correction on your own site enters your correction into the same contested corpus as everything else, competing with third-party accounts, on a timeline nobody controls.\n\nSo the tractable question is not \"how do I fix the model?\" It is \"what do buyers currently believe, how wrong is it, and which wrong belief is costing the most?\" That question is fully answerable, and the answer tells you where to aim your corpus effort.\n\nThe instrument is an unprompted-description study, and its design has one rule that decides whether the data is worth anything: **ask what they believe before you tell them anything.** Any positioning statement you show first contaminates every answer after it.\n\n- Open with `open_ended`: \"In your own words, what does this product do, and who is it for?\" Nothing before it. No logo, no tagline, no category label.\n- Follow with `single_choice` on category placement, offering your intended category alongside the three you keep getting confused with. This quantifies misassignment.\n- Use `multiple_choice` for specific attribute beliefs: pricing model, deployment, company size served, whether there is a free tier. This catches attribute errors cheaply.\n- Use `ranking` to order the vendors they consider comparable. This surfaces the consideration set the machine built, which is often not the one your competitive deck assumes.\n- Use a `scale` on confidence in their own description, which separates a firmly held wrong belief from an idle guess. These need different responses.\n- Use `yes_no` on whether they checked any claim against your own site. A high \"no\" rate means the third-party description is the only one operating.\n\nRun the same instrument against buyers who chose you and buyers who did not. The gap between the two unprompted descriptions is your positioning problem, stated in customers' words and quantified.\n\nOur [brand perception survey guide](/docs/brand-perception-survey-guide) covers the unprompted-description technique in depth, and [positioning research](/docs/positioning-research-customer-validation) covers validating category framing with real buyers. For the consideration-set half, [competitive intelligence interviews](/docs/competitive-intelligence-interviews) covers what customers know about your rivals, and the [competitive research guide](/docs/competitive-research-guide) covers the wider market picture.\n\nOne warning specific to 2026. It is tempting to study this with synthetic respondents, since the subject is AI. Do not. When the thing you are measuring is a distortion introduced by a language model, asking a language model to simulate the buyer measures the distortion against itself. Our guide to [synthetic users](/docs/synthetic-users-research-methodology) sets out where AI personas are and are not defensible, and this is squarely on the \"not\" side. The same caution applies to respondent quality generally, covered in [survey fraud and respondent quality](/docs/survey-fraud-respondent-quality).\n\n## What to do in the next quarter\n\n1. Capture the actual description. Run your top 20 buyer questions across four assistants and save the prose, not just whether you were mentioned. This takes an afternoon.\n2. Score each answer on the four failure modes above. Most teams find at least one systematic category or attribute error immediately.\n3. Field the unprompted-description study with 20 to 30 buyers, split between won and lost.\n4. Compare the two. Where the buyers' wrong belief matches the model's wrong description, you have a causal story worth acting on and a specific claim to go correct in the corpus.\n5. Add description accuracy to your reporting next to mention rate. Reporting volume alone is what created this blind spot.\n\n## Where Koji fits\n\nThe study above is easy to design and historically painful to run, because unprompted description requires a moderator who asks the question cold and does not react to the answer. Human moderators leak. They nod at the right answer, follow up harder on the wrong one, and by respondent twelve they are subtly coaching.\n\nKoji's AI interviewer does not. It asks all 30 buyers the same opening question in the same neutral way, probes for specifics without signalling which answer is wanted, and never reveals your positioning because it was not asked to. That is the difference between measuring what buyers believe and measuring what buyers will agree to.\n\nBecause Koji supports all six structured question types inside a conversational interview, the unprompted `open_ended` description, the `single_choice` category placement, the `ranking` of competitors and the `scale` on confidence come from the same respondent in one sitting, which is what lets you say \"the buyers who put us in the wrong category also rank us against the wrong three vendors.\" Thematic analysis runs automatically across every transcript, so the recurring misconceptions surface as themes with quotes attached rather than as a pile of recordings. See the [structured questions guide](/docs/structured-questions-guide) for how the types combine and aggregate.\n\nYou cannot make the model say the right thing. You can find out exactly what it is saying wrong, to whom, and what that costs, and then go change the sources it reads.\n\n[Run an unprompted-description study with Koji](https://www.koji.so) and see your positioning in your buyers' words.\n\n## Related reading\n\n- [AI Search Visibility (2026)](/blog/ai-search-visibility-measurement-2026)\n- [AI Answers Have No Clock (2026)](/blog/ai-answers-outdated-product-information-2026)\n- [G2 vs Capterra vs TrustRadius (2026)](/blog/software-review-sites-b2b-buyers-2026)\n- [Best Brand Tracking Software 2026](/blog/best-brand-tracking-software-2026)\n- [AI Personas vs Real Customer Interviews](/blog/ai-personas-vs-real-interviews-2026)\n\n## Frequently Asked Questions\n\n### Why does AI get my brand wrong?\n\nBecause an AI answer is synthesised, not retrieved. The model composes new prose about you from training data, retrieved pages and third-party accounts, weighting sources it treats as independent above your own marketing site. If the third-party corpus is thin, contradictory or out of date, the answer can be fluent, confident and wrong at the same time. Nobody wrote the sentence the buyer reads, so there is no page to go and edit.\n\n### Can I ask OpenAI or Google to correct information about my company?\n\nThere is no established correction channel for brands at any major assistant comparable to a press correction or a search removal request. The practical levers are indirect: improve the accuracy and clarity of your own pages, and work on the third-party sources the models draw from. Both operate on a timeline you do not control, which is why measuring what buyers currently believe matters more than waiting for a fix.\n\n### How often do AI assistants make significant errors?\n\nIn the EBU and BBC study *News Integrity in AI Assistants*, published in October 2025 and covering 3,113 questions across 22 public service media organisations, 45% of AI responses contained at least one significant issue and 81% had an issue of some form. Sourcing was the largest single category at 31%, and accuracy problems affected 20%. The study covered news questions rather than product questions, so treat it as evidence of mechanism rather than a brand-specific error rate.\n\n### Does more AI visibility always help my brand?\n\nNo. Visibility and description accuracy are independent variables. If the model has settled on a wrong category or a wrong attribute for your product, increasing the number of answers you appear in increases the number of buyers who receive the wrong description. Track what the answers say about you, not only how often they mention you.\n\n### What is category misassignment and why does it matter?\n\nCategory misassignment is when an AI answer files your product under the wrong category, so buyers compare you against the wrong competitors on criteria that do not apply to what you actually sell. It is expensive because it is self-reinforcing: buyers ask follow-up questions in the wrong frame, more content gets written in that frame, and future answers inherit it. It rarely appears in visibility dashboards, because a mention in the wrong category still counts as a mention.\n\n### Should I use synthetic respondents to study how AI describes my brand?\n\nNo. When the phenomenon you are measuring is a distortion introduced by a language model, using a language model to simulate your buyers measures that distortion against itself and will tend to confirm it. This is one of the clearest cases for real respondents. The point of the study is to find the gap between what the machine says and what humans actually took away, and only humans can supply the second half.\n","category":"Research","lastModified":"2026-08-16T03:23:49.944814+00:00","metaTitle":"When AI Describes Your Product Wrong (2026): The Positioning Risk","metaDescription":"AI answers are synthesised from third-party sources, so visibility and accuracy move independently. The EBU/BBC data on AI error rates, the four ways a brand description goes wrong, and the study that finds yours.","keywords":["ai brand misrepresentation","ai gets my brand wrong","llm brand mentions","brand positioning ai","answer engine optimization risk","ai brand perception"],"aiSummary":"An AI answer is synthesised prose assembled from third-party sources, not a retrieved page, so brand visibility and description accuracy are independent variables. The EBU and BBC found 45% of AI responses contained a significant issue and 31% had sourcing problems, including misattributing claims to the organisations covered. Increasing mentions without controlling the description simply distributes a wrong description more widely, and the fix is an unprompted-description study with real buyers.","aiKeywords":["ai brand misrepresentation","category misassignment","llm brand description","unprompted description study","positioning research","ebu bbc ai study"],"aiContentType":"guide","faqItems":[{"answer":"Because an AI answer is synthesised, not retrieved. The model composes new prose about you from training data, retrieved pages and third-party accounts, weighting sources it treats as independent above your own marketing site. If the third-party corpus is thin, contradictory or out of date, the answer can be fluent, confident and wrong at the same time. Nobody wrote the sentence the buyer reads, so there is no page to go and edit.","question":"Why does AI get my brand wrong?"},{"answer":"There is no established correction channel for brands at any major assistant comparable to a press correction or a search removal request. The practical levers are indirect: improve the accuracy and clarity of your own pages, and work on the third-party sources the models draw from. Both operate on a timeline you do not control, which is why measuring what buyers currently believe matters more than waiting for a fix.","question":"Can I ask OpenAI or Google to correct information about my company?"},{"answer":"In the EBU and BBC study News Integrity in AI Assistants, published in October 2025 and covering 3,113 questions across 22 public service media organisations, 45% of AI responses contained at least one significant issue and 81% had an issue of some form. Sourcing was the largest single category at 31%, and accuracy problems affected 20%. The study covered news questions rather than product questions, so treat it as evidence of mechanism rather than a brand-specific error rate.","question":"How often do AI assistants make significant errors?"},{"answer":"No. Visibility and description accuracy are independent variables. If the model has settled on a wrong category or a wrong attribute for your product, increasing the number of answers you appear in increases the number of buyers who receive the wrong description. Track what the answers say about you, not only how often they mention you.","question":"Does more AI visibility always help my brand?"},{"answer":"Category misassignment is when an AI answer files your product under the wrong category, so buyers compare you against the wrong competitors on criteria that do not apply to what you actually sell. It is expensive because it is self-reinforcing: buyers ask follow-up questions in the wrong frame, more content gets written in that frame, and future answers inherit it. It rarely appears in visibility dashboards, because a mention in the wrong category still counts as a mention.","question":"What is category misassignment and why does it matter?"},{"answer":"No. When the phenomenon you are measuring is a distortion introduced by a language model, using a language model to simulate your buyers measures that distortion against itself and will tend to confirm it. This is one of the clearest cases for real respondents. The point of the study is to find the gap between what the machine says and what humans actually took away, and only humans can supply the second half.","question":"Should I use synthetic respondents to study how AI describes my brand?"}],"relatedTopics":["AI search","brand positioning","competitive intelligence","market research"]}],"pagination":{"total":1,"returned":1,"offset":0}}