{"site":{"name":"Koji","description":"AI-native customer research platform that helps teams conduct, analyze, and synthesize customer interviews at scale.","url":"https://www.koji.so","contentTypes":["blog","documentation"],"lastUpdated":"2026-08-19T09:56:42.030Z"},"content":[{"type":"documentation","id":"dc6a29dc-97c1-4497-9067-4bd2c3cbca54","slug":"agreed-upon-procedures-research-findings","title":"Agreed-Upon Procedures: The Research Deliverable That Deliberately Has No Recommendation","url":"https://www.koji.so/docs/agreed-upon-procedures-research-findings","summary":"An agreed-upon procedures study is one where requester and researcher agree the exact procedures in advance, the researcher performs them, and the report states only what was found: no opinion, conclusion or recommendation. ISRS 4400 (Revised), effective for engagements agreed on or after 1 January 2022, defines findings as factual results capable of being objectively verified, meaning different practitioners running the same procedures would reach equivalent results. This mode fits when the decision hinges on commercial context the researcher lacks, when the requester is better informed about the trade-off, when the finding will be contested, or when several teams will use the data differently. Five of Koji's six structured question types produce objectively verifiable results by construction; open_ended does not, which is why it belongs in the interpretive section.","content":"Some research questions should be answered with facts and no recommendation. When the decision depends on commercial context the researcher does not hold, the most useful and most defensible deliverable is a list of pre-agreed procedures, the factual findings each one produced, and nothing else. The accounting profession formalised this decades ago and calls it an agreed-upon procedures engagement. Product research has no equivalent, which is why researchers keep getting drawn into recommendations they are not positioned to make, and then blamed for the outcome.\n\n## The short version\n\nAn agreed-upon procedures study is one where the requester and the researcher agree in advance on exactly what will be done, the researcher does it, and the report states what was found. No conclusion. No recommendation. No \"therefore we should\".\n\nThis runs directly against the standard advice, including most of the advice on this site, which is that a study that does not change a decision was a waste of money. That advice is right most of the time and wrong in a specific, common case: when the researcher would have to supply business judgement they do not have in order to reach a conclusion. In that case a recommendation is not added value, it is added noise wearing a lab coat.\n\n## What the standard actually says\n\nISRS 4400 (Revised), the International Standard on Related Services covering agreed-upon procedures engagements, has been in effect for engagements agreed on or after 1 January 2022. Two of its provisions are worth reading closely, because they are more precise than anything product research has written on the subject.\n\nFirst, on what the engagement is not. The standard states that an agreed-upon procedures engagement is not an audit, review or other assurance engagement, and does not involve obtaining evidence for the purpose of expressing an opinion or an assurance conclusion in any form.\n\nSecond, on what a finding is. The standard defines findings as the factual results of the procedures performed, states that findings are capable of being objectively verified, and says explicitly that references to findings exclude opinions or conclusions in any form as well as any recommendations the practitioner may make.\n\nThe operational test is in the application material: findings are capable of being objectively verified, which means that different practitioners performing the same procedures are expected to arrive at equivalent results.\n\nThat sentence is the whole discipline. If two competent researchers running your stated procedures on your stated sample would produce different write-ups, you have not produced findings. You have produced interpretation, and interpretation needs a different kind of report, a named author, and a stated basis.\n\n## Why product research needs this mode\n\nFour situations come up constantly where the recommendation-shaped report actively damages the decision.\n\n**The decision hinges on numbers you do not have.** \"Should we build the enterprise SSO tier?\" depends on pipeline, contract values, the cost of the build and the opportunity cost of the alternative. A researcher can establish how many interviewed buyers named SSO as a blocker, in what words, at what stage. Turning that into \"we should build it\" requires the commercial half of the equation. Report the half you own.\n\n**The requester is more senior and better informed about the trade-off.** An executive asking a specific factual question usually already has the surrounding context. Supplying a recommendation on top invites them to argue with the recommendation and ignore the facts, which is the worst of both outcomes.\n\n**The finding will be contested.** When research feeds a negotiation, a pricing change, a regulatory filing or a partner discussion, a factual-findings report survives cross-examination and an interpretive one does not. This is the same instinct behind [survey evidence that holds up in court](/docs/survey-evidence-court-daubert): the more the output must withstand an adversary, the more it should confine itself to what was done and what was found.\n\n**Multiple teams will use the same data differently.** One study, four consumers, four legitimate conclusions. Publishing one team's conclusion as \"the finding\" quietly disenfranchises the other three.\n\n## How to run one\n\n**Agree the procedures in writing, before fieldwork.** This is the defining step and the one people skip. The requester must acknowledge that the procedures are appropriate for their purpose; ISRS 4400 makes that acknowledgement a condition of accepting the engagement, and it protects both sides. A procedure is a sentence with no adjectives in it: \"Ask each participant to rank the five proposed capabilities in order of importance to their renewal decision, using a ranking question, and report the mean position of each.\"\n\n**Define the population and how it was drawn.** \"Twenty current customers on the Growth plan, with more than ninety days tenure, invited by email in the order returned by the export, until twenty completed.\"\n\n**Perform exactly those procedures.** Deviations are permitted, but they get reported as deviations. Quietly adding a question because it seemed interesting breaks the objective-verifiability property that gives the report its authority.\n\n**Report findings and exceptions, including the uncomfortable ones.** ISRS 4400 requires the report to include the findings from each procedure performed, including details of exceptions found. An exception is any case where the procedure could not be performed as agreed: four participants declined to rank, two sessions ended early, one respondent misread the scale. Exceptions belong in the report at the same prominence as the results.\n\n**Stop.** No summary paragraph beginning \"This suggests\". No slide titled \"So what\". If the requester wants your view, they can ask for it in the meeting, on the record, as your view.\n\n## The report template\n\n```\nAgreed-Upon Procedures Report: Renewal Blockers, Growth Plan\n\nEngaging party:   VP Customer Success\nDate agreed:      3 March 2026\nFieldwork:        10-12 March 2026\n\nProcedures agreed\nP1  Interview 20 Growth-plan customers with >90 days tenure...\nP2  Ask each to rank five capabilities by importance to renewal...\nP3  Ask each a yes_no question on whether they have evaluated an alternative...\n\nFindings\nP1  22 interviews invited, 20 completed, 2 abandoned mid-session.\nP2  Mean rank position: Reporting 1.8; SSO 2.4; API 3.1; ...\nP3  9 of 20 answered yes. Of those 9, 6 named the same competitor...\n\nExceptions\nP2  3 participants ranked only their top three and declined the rest.\n    Their partial rankings are excluded from the means above.\n\nThis report contains no opinion, conclusion or recommendation.\n```\n\nThat last line is not decoration. It is the sentence that stops a reader from importing a conclusion you did not make, and it is the sentence that protects you when the decision goes badly.\n\n## Structured questions make this mode practical\n\nThe reason agreed-upon procedures have been impractical in product research is that qualitative research does not naturally produce objectively verifiable results. Two analysts coding the same twenty transcripts will not produce equivalent write-ups, which fails the standard's core test.\n\nKoji's structured questions solve this directly, and it is one of the clearest cases where the platform's design maps onto a governance need. Five of the six question types produce results that are objectively verifiable by construction: `scale` yields a distribution, `single_choice` and `multiple_choice` yield frequencies, `ranking` yields mean positions, and `yes_no` yields a count. Anyone re-running the same procedure on the same responses gets the same numbers. The sixth type, `open_ended`, is the one that does not, which is exactly why it belongs in the interpretive part of your work rather than in the findings section of an agreed-upon procedures report.\n\nThis is also where AI-moderated interviews earn their place over a conventional survey tool. A survey platform can give you a clean ranking, but it cannot ask the participant why they put reporting first, and it cannot tell whether the person was engaged or clicking through. Koji runs the structured question and then probes conversationally on the answer, so the same session produces both the objectively verifiable finding and the verbatim material you will need later when someone asks what the number means. Every session carries a quality score from 1 to 5, which is itself an objectively verifiable input to whether a response should be included. See the [structured questions guide](/docs/structured-questions-guide) for how each type is presented and analysed.\n\nPractically: put the agreed procedures on the structured questions, and let the conversational layer collect the colour separately. Report the first. Offer the second on request.\n\n## When this mode is the wrong choice\n\nIt is the wrong choice more often than it is the right one, and using it as a hiding place is a real risk.\n\n- **When the requester wants your judgement and is entitled to it.** If you have run forty studies in this domain and they have run none, withholding a view is not neutrality, it is abdication.\n- **When the question is genuinely exploratory.** You cannot pre-agree procedures for \"why are trials not converting\" because you do not yet know what to ask. Discovery work needs the opposite posture; see [continuous discovery](/docs/continuous-discovery-user-research).\n- **When it is being used to avoid delivering bad news.** A factual-findings report is not a way to avoid saying the thing. If the finding is that the feature nobody wants is the one already being built, the finding says so, plainly, in the findings section.\n\nThe honest test: are you declining to conclude because the conclusion requires context you lack, or because concluding would be uncomfortable? Only the first is a legitimate use.\n\n## The failure mode: the smuggled conclusion\n\nThe most common way this goes wrong is that a conclusion is smuggled in through presentation. Ordering the findings so the intended one comes first. Choosing a chart that makes one bar look decisive. Bolding one number. Titling a section \"The reporting problem\". Each of these is an opinion expressed without accountability, and it is worse than an openly stated recommendation because nobody can argue with it.\n\nThe remedy is a review pass by someone who was not in the study, looking only for implied conclusions. This slots naturally into an existing [peer review gate](/docs/research-peer-review-qa-gate). Give the reviewer one instruction: mark every place where the report tells the reader what to think.\n\n## Frequently asked questions\n\n### Is this just a fancy way of saying \"here is the data\"?\n\nNo. \"Here is the data\" is a dump with no defined procedures, no stated population, no exceptions and no prior agreement about what would be done. The agreed-upon procedures form is a commitment made before fieldwork about exactly what will be performed, which is what makes the findings verifiable and what stops the scope shifting once early results appear.\n\n### Does this let researchers off the hook?\n\nThe opposite. A recommendation can be defended with judgement and seniority. A findings-only report is checkable line by line, so any sloppiness in sampling, question wording or execution is exposed. Teams generally find this mode harder to produce, not easier.\n\n### How do I stop stakeholders from just asking for the recommendation anyway?\n\nGive it to them verbally, in the meeting, clearly labelled as your view rather than as a finding. The distinction that matters is between what the study established and what you personally think, and keeping those in different containers is the entire point. Many researchers find the conversation goes better once the two are separated, because the facts stop being negotiable.\n\n### Can a study be part agreed-upon procedures and part interpretive?\n\nYes, and this is the usual arrangement in practice. Keep them in separate, clearly titled sections, in that order, and never let interpretation appear inside the findings section. Readers will honour the boundary if you do.\n\n### What about the participants' quotes? Are those findings?\n\nA verbatim quote with its source and date is a factual result, so it can sit in the findings. A selection of quotes chosen to illustrate a theme is interpretation, because a different researcher would choose differently. If you include quotes in the findings section, state the selection rule you applied, such as every mention of pricing in order of occurrence.\n\n### Does an AI-moderated study make findings more or less verifiable?\n\nMore, on the mechanical dimension, and it is the main practical argument for running this mode on a platform like Koji. The questions asked are identical across participants unless the AI probes, the full transcript is retained, the probes themselves are on the record, and the structured question results are computed the same way every time. A human moderator introduces unrecorded variation in wording and emphasis that makes exact re-performance impossible.\n\n## Related Resources\n\n- [Structured Questions in AI Interviews](/docs/structured-questions-guide) - the six question types, and which ones produce verifiable findings\n- [Research Peer Review](/docs/research-peer-review-qa-gate) - the QA gate that catches smuggled conclusions\n- [Presenting Research Findings to Stakeholders](/docs/presenting-research-findings) - when interpretation is the deliverable\n- [Survey Evidence in Court](/docs/survey-evidence-court-daubert) - research built to survive cross-examination\n- [How to Write a Research Brief](/docs/how-to-write-research-brief) - where the agreed procedures get recorded\n- [Research Debrief](/docs/research-debrief-guide) - separating what was found from what to do about it\n","category":"Research Operations","lastModified":"2026-08-16T03:22:03.635045+00:00","metaTitle":"Research Findings Without Recommendations: Agreed-Upon Procedures (2026)","metaDescription":"When a decision depends on business context the researcher does not hold, a findings-only report beats a recommendation. ISRS 4400 applied to product research: agreeing procedures, reporting exceptions, and refusing to conclude.","keywords":["research findings vs recommendations","agreed-upon procedures","should researchers make recommendations","factual findings report","objective research reporting","research deliverable template"],"aiSummary":"An agreed-upon procedures study is one where requester and researcher agree the exact procedures in advance, the researcher performs them, and the report states only what was found: no opinion, conclusion or recommendation. ISRS 4400 (Revised), effective for engagements agreed on or after 1 January 2022, defines findings as factual results capable of being objectively verified, meaning different practitioners running the same procedures would reach equivalent results. This mode fits when the decision hinges on commercial context the researcher lacks, when the requester is better informed about the trade-off, when the finding will be contested, or when several teams will use the data differently. Five of Koji's six structured question types produce objectively verifiable results by construction; open_ended does not, which is why it belongs in the interpretive section.","aiPrerequisites":["Experience writing research reports for stakeholders","A research brief or intake process already in place"],"aiLearningOutcomes":["Decide when a findings-only report beats a recommendation","Write procedures that two researchers would perform identically","Report exceptions at the same prominence as results","Use structured question types to produce objectively verifiable findings","Spot a conclusion smuggled in through ordering, charting or section titles"],"aiDifficulty":"intermediate","aiEstimatedTime":"9 min"}],"pagination":{"total":1,"returned":1,"offset":0}}