Applied Judgment Assessment
Self-insight profile · anybody who argues with anybody · browse the full catalogue

Perspective Accuracy Assessment for Everyday DisagreementsCould the other side sign what you just said about them?

Sixteen real disagreements — a rota, a landlord, a rewrite, a subscription — and four ways of putting the other person's view. One of them is what they actually said.

25 minutes39 scored exercisesEvidence-keyed scoringGlobal · INR & USD

The tests in this category will not tell you their price. The training will not tell you your score.

The two best-known critical-thinking instruments are quote-only: a person cannot buy one, and neither publishes a reliability figure or a norm group on the page a buyer reads. At the other end, accredited mediation training runs to several thousand pounds a delegate and produces no transferable measure at all. Between the two there is nothing.

None of them measures this construct anyway. Reading a supplied passage and answering questions about it is not the same as reconstructing somebody's position so well that they would sign it, and that second thing is what actually decides whether a disagreement can move.

It is measurable, and it has been measured: a peer-reviewed study had six hundred people with strong opinions write the reasons an opponent might give, and twelve hundred independent judges rated those arguments blind. Pass rates ran from 54 to 71 per cent depending on the topic. The criterion is acceptance by somebody who holds the view, which is the only criterion that is not circular.

So the wrong answers here are the three real failure modes rather than inventions: the weakened version that drops the strongest reason, the version with a motive attached that nobody gave, and the version that is accurate right up until it hands you the conclusion. That last one is the hardest, because it is usually the best argument on the page.

On eight exercises you also say how sure you are, and the report scores that separately. Being wrong about somebody is ordinary. Being certain while wrong is the state in which nobody asks, and it is the one worth knowing about.

Nothing here is political. Every disagreement is an ordinary one about a rota, a repair, a price or a rule, because a politics item would measure which side you are on rather than how well you read.

Three parts, one of which is about you rather than about the argument:
Restating a view its holder would acceptNaming what a bad restatement did wrongKnowing when you have got it wrong

What you walk away with

Sixteen restatements

Four ways of putting somebody's view; one is what they said and three are the named ways people get it wrong.

Your own certainty, scored

Eight paired confidence ratings, scored as a Brier skill score against somebody equally sure about everything.

The over-confidence gap

Average confidence minus how you actually did, with the sign named in both directions.

Exercise by exercise

A strip with one cell per exercise, so the texture is visible instead of hidden in an average.

The four failure modes

Named and matched, so a vague sense that a summary was unfair becomes a specific thing to fix.

One if-then change

Chosen from which way your certainty ran, not from a template.

Inside your report

Illustrative sample — your report is generated from your own responses.

Exercise by exercise, not a single average

The texture is the finding. Even means one thing about how you read; spiky means another, and no summary statistic can show either.

How sure you were, against how you did
9261889544739055tall and outlined is the expensive combination

Being wrong is ordinary. Being certain while wrong is the state in which nobody asks, and it is the one this strip is drawn to make visible.

Built for

  • Anybody who has been told they are arguing with something the other person did not say
  • Managers who have to carry two positions into the same room
  • Teachers, moderators and facilitators of anything contested
  • People who want the honest version of a skill usually sold as a course

Find out whether you can state the other side

39 exercises across five formats · about 25 minutes · the strips, the confidence score and one change. Free to start.

Take free · full report ₹249 (incl. GST)

Secure checkout · INR & USDFull report immediately after submission

Frequently asked questions

Is this political?

No. Every one of the sixteen disagreements is an ordinary one — a cleaning rota, a landlord, a subscription, a car repair, a supplier's price. A political item would measure which side you are on, which is a different question and a worse one.

How can a restatement have a right answer?

The criterion is whether the person who holds the view would accept it as a fair statement of what they said. That is the criterion used in the published, judge-rated version of this task, and it is the only one that is not circular. Each exercise's rationale names which failure the three wrong options commit.

Why does my confidence get scored?

Because being wrong is ordinary and being certain while wrong is expensive: it is the state in which nobody asks. Eight exercises carry a paired confidence rating, which is the minimum a calibration figure can honestly be computed from, and below that the report refuses to print one.

Does this mean I should agree with people more?

No, and the report says so. Stating a position accurately and being persuaded by it are different things, and this instrument only has a view about the first.

How long is it and what does it cost?

About twenty-five minutes for thirty-nine exercises. ₹249 in India, inclusive of GST, or US$2.99 elsewhere, one time. At that price the sitting starts without payment.

One of the AssessAll applied-judgment assessments

Each one takes a single capability, puts you inside the situations where it is actually tested, and scores your choices against published evidence — with a report designed for that capability alone, not a template. They span hiring, compliance, education, operations and personal skill.

Browse the catalogue

Methodology: Thirty-nine original exercises across five formats: sixteen single-choice restatements, eight confidence ratings on a six-point scale with half of them worded the other way round, six select-every-that-applies exercises, five binary claims and four matching exercises. One response instruction is declared for the whole instrument and it is a KNOWLEDGE instruction - which restatement the person who holds the view would accept - never what the respondent would say. The confidence ratings carry their own instruction and are scored as a probability, never as an answer. Construct statement: it measures whether somebody can state a position they do not hold in a form its holder would accept, and whether they can tell when they have failed to. It does not measure who is right, political or social opinions, intelligence, verbal ability, empathy, agreeableness, or willingness to change one's mind. No item uses a partisan political subject; every disagreement is an ordinary one about a rota, a repair, a price, a booking or a rule. Scoring is calibration. Accuracy on the sixteen restatements is chance-corrected against an authored answer prior per option rather than against a uniform draw. The eight paired confidence ratings are scored with the Brier score, decomposed into calibration error and discrimination, and reported as a skill score against a respondent who is equally sure about everything - a reference that has perfect consistency and no information in it. A calibration claim from fewer than eight items is noise, which is why there are eight. Reliability is McDonald's omega estimated from item count and a stated assumed inter-item correlation, printed with what it permits; where it will not carry a number, the refusal is printed. Every reported figure carries its standard error band, drawn on the chart rather than described in a footnote. Constructs and sources: the ideological Turing test as a scored, judge-rated behavioural measure of perspective taking, with published pass rates of 54, 64 and 71 per cent across three topics and 1,200 blind judges (Brand, Brady and Stafford 2025); naive realism and why acceptance by the holder is the right criterion (Ross and Ward 1996); the finding across twenty-five experiments that instructing people to imagine another's perspective made them LESS accurate while asking the person directly made them substantially more accurate (Eyal, Steffel and Epley 2018); the illusion of explanatory depth (Rozenblit and Keil 2002), cited for the measurement claim only, because the intervention claim built on it has failed preregistered replication and this report says so on the page; speakers' overestimation of their own clarity (Keysar and Henly 2002); the ordering rule for engaging a position - restate, agree, learn, then object - which is Rapoport's and is normative rather than empirical (Rapoport 1960; Dennett 2013); the Brier score and its decomposition into reliability, resolution and uncertainty (Brier 1950; Murphy 1973); overconfidence and the calibration of subjective probability (Lichtenstein, Fischhoff and Phillips 1982); acquiescence and balanced keying (Baumgartner and Steenkamp 2001); and the Barnum effect (Forer 1949). All items are original works. No commercial critical-thinking instrument, mediation framework, argument-mapping product or report layout is reproduced or implied, and no affiliation with any assessment publisher exists or is claimed.