Spoken Instruction Accuracy Assessment for Phone, Radio and Shop-Floor WorkYou heard all of it. The question is how much of it reached your hands.
Sixteen instructions spoken to you, twice at most and never printed, with your confidence scored against whether you were actually right.
Almost nothing at work arrives in writing first
Somebody says it while you are holding something else, and what you do next is whatever survived the gap between hearing and acting. Every version of this on the market is either a free reaction-time toy or an employer-bought typing test priced at twenty to forty dollars a seat, where the candidate never sees a result. This one is yours, it costs less than a coffee, and it tells you which kind of detail you lose.
Sixteen instructions are SPOKEN, twice at most, and the written choices differ only in which detail reached you: a condition, a figure or the order of two steps. Half the conditions deliberately do not fire, so treating every unless as a stop scores no better than ignoring them. That balance is the whole design.
On eight of the sixteen you are also asked how sure you are, and four of those are worded as doubt so that agreeing with everything cannot look like calibration. Your confidence is then scored against whether you were right, as a skill score against somebody equally sure about everything. Being sure and being right come apart, and which one you are is a separate finding with a separate fix.
Every spoken exercise carries its wording in writing, one tap away, and automatically when the device cannot speak. An exercise taken in writing is counted as reading rather than listening and is left OUT of the listening figure, with the report saying how many took that route. A test that scored you on an instruction nobody managed to speak to you would be reporting our own machinery as your mistake.
The headline carries the band that belongs to a sixteen-exercise form, which is wide, and says so. There is no percentile anywhere, because there is no norm group yet and inventing one would be the easiest lie on the page.
What you walk away with
Accuracy on the spoken instructions, corrected against the answers people typically give rather than against a coin.
A skill score on your own confidence, with the calibration error and the discrimination reported apart.
Conditions, figures and order scored separately, because losing an exception and losing a sequence are different problems.
What you said against what you got right, with the line of perfect agreement drawn through it.
Drawn from the shape of the two figures rather than the level of either, and guarded by what the report refuses to claim.
Aimed at whichever kind of detail your own sitting lost most often.
Inside your report
Illustrative sample — your report is generated from your own responses.
Three kinds of detail, scored apart. Losing the exception and losing the order are different problems with different fixes, and one accuracy percentage hides both.
Plotted against the line of perfect agreement, with a table of the same three numbers beside it. Nearly no report in this category prints whether the reader's confidence was worth anything.
Built for
- Contact centre, dispatch, radio and shop-floor workers who take instructions by voice
- Clinical, care and laboratory staff who work from verbal orders and handovers
- Anybody who has ever done the right thing to the wrong quantity
- Team leads deciding where a read-back step would pay for itself
Find out which part of it you lose
35 exercises across five formats · about 22 minutes · the calibration diagram and the error bands printed.
Take free · full report ₹249 (incl. GST)
Frequently asked questions
No. It is not a hearing test and says nothing about hearing acuity. It measures how much of an instruction reaches your action, which is a different thing, and somebody with perfect hearing can lose the exception at the end of a sentence just as easily.
Every spoken exercise carries its wording in writing, shown automatically when the device has no speech synthesiser and available on request at any time. Exercises taken that way are recorded as reading rather than listening and are left out of the listening figure, so nothing about taking that route counts against you.
Because confidence and accuracy come apart, and which way they come apart changes what you should do. Somebody who loses detail and knows it can put a read-back step in; somebody who loses detail and feels certain cannot, because nothing tells them to.
No. The instructions are in plain, short sentences and the difficulty is in the detail rather than the vocabulary. It does not measure accent, fluency or vocabulary, and it is not a language proficiency test of any kind.
About twenty-two minutes for thirty-five exercises. ₹249 in India, inclusive of GST, or US$2.99 elsewhere, one time, for the sitting and the full report.
Each one takes a single capability, puts you inside the situations where it is actually tested, and scores your choices against published evidence — with a report designed for that capability alone, not a template. They span hiring, compliance, education, operations and personal skill.
Browse the catalogue →Methodology: Thirty-five original exercises across five formats: sixteen spoken instructions with written response sets, eight confidence probes on a six-point scale, four select-every-that-applies exercises, four keyed claims and three ordering exercises. The declared response instruction is knowledge and applied accuracy throughout - what the instruction you heard actually requires - and one instruction covers the whole form. Construct statement: it measures how much of a SPOKEN instruction reaches the action, which class of detail is lost when one is lost - a condition, a figure or the order of two steps - and whether the respondent knows which of their own answers are right. It does not measure hearing acuity and is not a hearing test, it does not measure English proficiency or accent, it is not a memory span or a clinical attention measure, and it says nothing about performance in any named job. Scoring is calibration scoring. The accuracy strand is chance-corrected against the authored answer priors carried on every option, and the confidence strand is scored as a Brier skill score against a reference respondent who is equally sure about everything, decomposed into calibration error and discrimination and reported as two separate figures because a person can be well calibrated and undiscriminating, or discriminating and over-confident. Half the spoken conditions deliberately do not fire, so applying every exception and ignoring every exception both land at chance as arithmetic rather than as a threshold, and half the confidence probes are worded as doubt and reverse-keyed so that agreeing with everything cannot look like calibration. Every spoken exercise carries a written equivalent; an exercise delivered in writing is recorded as such and left out of the listening figure rather than scored as listening. Constructs and sources: the phonological loop and the limits of holding spoken material (Baddeley and Hitch's working-memory model; Baddeley's later revisions); the serial position curve and the loss of late-arriving material under a competing task; closed-loop communication and read-back or hear-back as an error control in aviation and clinical handover (the read-back requirement in radiotelephony practice, and the SBAR and check-back literature in health care); the finding that confirmation of understanding is weaker than repetition of wording; gist versus verbatim memory for instructions (fuzzy-trace accounts of gist extraction); the calibration literature on confidence and accuracy coming apart, and the Brier score with Murphy's decomposition into reliability, resolution and uncertainty; acquiescence and balanced keying (Baumgartner and Steenkamp); and item-writing guidance from Haladyna, Downing and Rodriguez. All items are original works written for this instrument. No test, vendor, framework or standard is reproduced, named or implied, and no endorsement or affiliation exists or is suggested.