A pilot should reveal how people interpret an item and whether the form displays it to the right respondents.
Cognitive pilots and survey branch-logic audits
Test interpretation before rates
Recruit a small set of people resembling the intended respondents, including recent handoff cases and people without a handoff. Ask them to complete the form and explain what “next step” meant in their own case. If one respondent thinks it refers to the agent’s next step and another to the customer’s, the item is not yet stable. A pilot of nine people can reveal such problems, but it is not a precision estimate for the full population.
Log the question path
The item “How clear was the next step after the handoff?” should display only when the respondent confirms that a handoff occurred. The form must preserve the branch variable and a not-displayed state. An API payload with handoff=false and a substantive clarity answer is internally inconsistent; reject or quarantine it rather than quietly including it. Response states make this validation explicit.
Record observed problems
Capture issue codes such as ambiguous reference, missing response option, faulty branch, excessive recall demand and inaccessible wording. Keep the respondent’s explanation separately from the researcher’s interpretation. If five of nine pilot respondents misread one item, that is a signal to revise it; it is not an estimate that 55.6% of future respondents will misread it. Qualitative pilots find mechanisms, not population rates.
Retest after revision
Changing the wording can repair one ambiguity and introduce another. Re-run the affected branch and translation with new participants or a second pilot wave. Test a keyboard-only path, mobile layout, the “no handoff” branch and incomplete submissions. Save the old and new item versions. The measurement contract should document why a change was made.
Set a release gate
Before fielding, require no unresolved critical branch defect, a readable codebook, a declared eligible population and a planned response-rate report. Check that the survey platform records form version and assignment. This gate does not certify absence of response bias; nonresponse analysis starts once invitations are sent.
Implementation
def validate_handoff_response(response):
handoff = response["handoff_occurred"]
clarity = response["clarity_answer"]
if handoff is False and clarity != "not_displayed":
return "invalid_branch"
if handoff is True and clarity == "not_displayed":
return "missing_display"
if handoff not in (True, False):
return "unknown_handoff_status"
return "valid"
assert validate_handoff_response({"handoff_occurred": False,
"clarity_answer": "not_displayed"}) == "valid"
assert validate_handoff_response({"handoff_occurred": False,
"clarity_answer": "mostly_clear"}) == "invalid_branch"Performance and operating cost
Validating N responses costs O(N) time and O(1) working space for a streaming checker. Qualitative interview review is a separate human task whose cost depends on participant diversity, form branches and the number of revisions.
Common Mistakes
- Do not use a tiny cognitive pilot as a population prevalence estimate.
- Do not repair an invalid branch by silently recoding the substantive answer.
- Do not skip retesting after a wording or logic change.
