Standing Instructions and Follow-Through Assessment for Detail-Critical WorkEverybody remembers the rule. The question is whether you remember it at the moment it applies.
Three standing instructions, issued once and never repeated, then sixteen jobs about something else — eight with a real trigger and eight with a near miss.
The hardest part of a rule is not learning it
Almost every avoidable error in detail-critical work has the same shape. Somebody knew the rule perfectly, and the moment it applied arrived inside a job that was about something else, with nothing on the screen to mention it. Ask people about their own forgetting and they will tell you about lists and intentions; that is what the questionnaires in this area actually measure. This one measures what happens.
The first exercise issues three plain instructions. Twenty exercises later, a supplier moves its delivery window for six weeks, a stock figure has been crossed out and rewritten in pen, a colleague is on leave and a client wants their file. Each of those is a genuine trigger. Interleaved with them are eight near misses: a change that is permanent rather than temporary, a figure changed by an overnight system update with no person behind it, a colleague who is present but in interviews all afternoon.
That balance is the whole design. Applying the rule everywhere and applying it nowhere both land at chance, as arithmetic rather than as a threshold, so a habit cannot pay. What is left after any tendency to act, or not to act, has been removed is the thing the report calls separation.
Beside it, and reported separately, is the direction of your error. Missing triggers and firing on look-alikes are different problems: the first is expensive because nothing downstream reports it, the second because every unnecessary confirmation spends somebody else's time. A single accuracy percentage destroys that distinction, and almost every instrument in this category prints one.
Four short memory spans run as a control strand and are deliberately kept out of the headline. Holding a sequence for four seconds and noticing that a rule applies are different abilities, and telling somebody their noticing is poor because their span is short would be a mechanical artefact reported as a personal failing.
What you walk away with
How far apart a real trigger and a look-alike are for you, with the sixty-eight and ninety-five per cent spans drawn and written out.
Missed triggers on one side, wrongly fired look-alikes on the other, both counted in words. Different problems, different fixes.
Separation on one axis, whether you build a check that does not need you on the other, and a specific next step in each corner.
Reported on its own, so a low result is not read as a memory problem when it is a noticing problem.
Catching, holding fire, holding the rule and building the check — each as what it gets you and what it costs when it runs unchecked.
Drawn from your widest gap, naming a specific moment and a specific behaviour rather than a resolution.
Inside your report
Illustrative sample — your report is generated from your own responses.
Every corner is labelled with an action rather than an adjective, and the two cut lines are printed because they are decisions rather than discoveries.
Missing a trigger and firing on a look-alike are different problems with different fixes. One accuracy percentage would have hidden both.
Built for
- Anybody whose work carries standing instructions: care, clinical, laboratory, aviation, rail, finance operations
- Quality, compliance and audit teams who want a measure rather than a briefing
- Team leads deciding where a checklist would pay and where it would be ignored
- Individuals who keep meaning to do the thing later and want to know why it does not happen
Find out whether the rule reaches the moment
40 exercises across six formats · about 30 minutes · a separation score with its band, and the direction of your error named.
₹499 (incl. GST) · assessment and full report, nothing further to pay
Frequently asked questions
No. Four short memory spans are included as a control strand and are deliberately excluded from the headline, precisely so that a low result on the main finding is not read as a memory problem. The main finding is about whether a rule you were given reaches the moment it applies.
No, and that is the point of the design. Half the trials contain a near miss that resembles a trigger and fails one stated condition. Acting on the rule there counts against you exactly as missing a real trigger does, so applying it everywhere and applying it nowhere both land at chance.
It is where you set your threshold for acting. Two people with the same separation, one of whom over-fires and one of whom under-fires, are different people with different fixes. Reporting only accuracy would hide that, which is what almost everything in this category does.
No. It is not a clinical instrument, it says nothing about any condition, and a low result is a finding about a form on one afternoon rather than about a person. It carries no pass mark and should not be used on its own to select anybody.
About thirty minutes for forty exercises. ₹499 in India, inclusive of GST, or US$4.99 elsewhere, one time, for the sitting and the full report.
Each one takes a single capability, puts you inside the situations where it is actually tested, and scores your choices against published evidence — with a report designed for that capability alone, not a template. They span hiring, compliance, education, operations and personal skill.
Browse the catalogue →Methodology: Forty original exercises across six formats: sixteen delayed-cue exercises split evenly between genuine triggers and near misses, two late recall checks, five select-every-that-applies exercises, six keyed claims, four matching exercises, three ordering exercises and four grid memory spans. Construct statement: it measures whether a standing instruction is applied at the moment it applies, when nothing on the screen mentions it, and whether the respondent can tell a genuine trigger from something that resembles one. It does not measure intelligence, does not measure clinical memory, is not a diagnosis of anything, and says nothing about any condition. Declared response instruction: knowledge and rule application, one instruction for the whole instrument. Scoring is signal detection (C6). Acting on the rule on a trigger exercise is a hit; acting on it on a near-miss exercise is a false alarm; d-prime is the discrimination and the criterion is where the respondent sets their threshold, both computed with the log-linear correction of half a trial per cell so an extreme rate cannot make either infinite. Reporting the criterion is not decoration: two people with the same discrimination, one of whom over-fires and one of whom under-fires, are different people to manage and a single accuracy percentage hides it. The two strands that are not signal detection - the memory spans and the exercises on building external checks - are chance-corrected against each item's own authored answer priors and reported apart from the headline, never averaged into it. Every competency carries a three-way placement rather than a number, because none of the four reaches the eight exercises a number would need. McDonald's omega is reported rather than alpha, estimated before live data from the exercise count and a stated assumed inter-item correlation. The trial set is balanced eight and eight by construction and the balance is audited empirically rather than asserted. Constructs and sources: prospective memory and its event-based and time-based forms (Einstein and McDaniel, multiprocess framework; Dismukes 2010 on prospective memory in laboratory, workplace and everyday settings; Loft, Dismukes and Grundgeiger on safety-critical work); implementation intentions and if-then planning (Gollwitzer; Gollwitzer and Sheeran meta-analysis); goal activation and place-keeping after interruption (Altmann and Trafton, memory for goals); signal detection theory and the separation of sensitivity from criterion (Green and Swets; Macmillan and Creelman; Hautus on the log-linear correction); checklist and external-memory design in aviation and surgery (Degani and Wiener; Gawande); the psychology of everyday slips and lapses (Reason, human error taxonomy; Norman on action slips); and item-writing guidance from Haladyna, Downing and Rodriguez. All items are original works written for this instrument. No commercial instrument is reproduced, no trademarked scale name is used, and no affiliation with any test publisher is implied or exists. The self-report questionnaires that exist in this area measure what a person believes about their own forgetting; this measures what they do, which is a different question with a different answer.