§ The guide
What a cognitive bias test actually measures.
A cognitive bias test measures how far your judgments drift from a normative standard under specific, well-studied conditions. This one does it with behavioral tasks rather than self-report alone: you make estimates, judge probabilities, evaluate arguments, and rate decisions, and the scoring compares what you actually did against what the research says an unbiased responder would do. It covers seven biases — anchoring, hindsight, overconfidence, belief bias, base-rate neglect, sunk cost, and outcome bias — plus your Bias Blind Spot: the gap between how biased you are and how biased you think you are. It takes about 15 minutes, runs entirely in your browser, and stores nothing.
Why self-report alone does not work
Asking "how susceptible are you to anchoring?" measures your theory of yourself, not your behavior. The methodological move that makes bias measurement possible is to put you in the situation and observe the drift. Anchoring is measured by giving one group a reference number and another none, then comparing estimates. Hindsight bias is measured by asking for a prediction, revealing the outcome, and then asking what you originally predicted. Both designs are built into this assessment, which is why some questions come back a second time.
The seven biases this test measures
1. Anchoring
An arbitrary number pulls your estimate toward it, even when you know it is arbitrary. How it is measured here: you give six numerical estimates with no reference point (Phase A). Later, the same six questions return with a reference number attached (Phase B). Your anchoring score is the movement between them. Research anchor: Kahneman and Tversky's heuristics-and-biases program.
2. Hindsight bias
Once you know how something turned out, your memory of what you expected shifts toward the outcome. How it is measured here: you rate how likely four historical outcomes were (Phase 1). Much later, after the actual outcomes are revealed, you are asked to recall your own earlier ratings (Phase 2). The distance between your real rating and your recalled rating is the bias.
3. Overconfidence
Confidence outruns accuracy. How it is measured here: eight general-knowledge questions, each followed by a confidence rating. Calibration is the relationship between the two — a well-calibrated person is right about 70% of the time when they say 70%.
4. Belief bias
You judge an argument as valid because you agree with its conclusion, not because the logic holds. How it is measured here: eight short logic problems in which validity and believability are deliberately crossed, so agreement and correctness come apart.
5. Base-rate neglect
Vivid, specific information swamps the underlying prior probability. How it is measured here: four probability problems where the base rate is stated plainly and the individuating detail pulls the other way.
6. Sunk cost
Money, time, or effort already spent — and unrecoverable — keeps you committed to a failing course. How it is measured here: four decision scenarios in which the rational choice is to stop, and the prior investment varies.
7. Outcome bias
You rate a decision as good because it worked out, ignoring what was knowable at the time. How it is measured here: eight short decision cases, each rated for how well-reasoned it was, with outcomes varied independently of reasoning quality.
Plus: the Bias Blind Spot
Twelve self-report statements ask how susceptible you believe you are. Your Bias Blind Spot is the discrepancy between that self-assessment and your measured task performance. Scopelliti and colleagues (2015) demonstrated that most people underestimate their own susceptibility while readily spotting bias in others — and that this blind spot is itself measurable and varies between people.
How the assessment works
Twelve steps, seven tasks, roughly 15 minutes. The order is not arbitrary. Two of the tasks are split in half and separated by other work, because anchoring and hindsight can only be measured across a gap: you must answer once without the manipulation, do something else, and answer again with it. If you skip ahead or look back, you break the measurement — which is why the test does not let you.
Some questions will feel repetitive. They are not. The second pass is the experiment.
How your score is calculated
Each of the seven biases is scored on its own task, then expressed as a susceptibility percentile relative to the scoring model. Those scores are summarized in a two-factor interpretation rather than a single number, and your responses also place you in one of six archetypes — The Anchored, The Contrarian, The Escalator, The Certain, The Calibrated, or The Variable — describing the shape of your profile, not its quality.
There is deliberately no single "rationality score." The research is genuinely divided on whether one exists: Stanovich (2016) argues for a rationality factor with caveats; Teovanović and colleagues (2015) find the biases largely independent. Publishing one number would mean picking a side of an open scientific question and hiding that from you.
How to read your results
Three things worth holding onto.
Susceptibility is not intelligence. Twenty-plus years of work by Stanovich, West, and Toplak shows bias susceptibility is largely independent of IQ. Highly intelligent people can be highly susceptible; the two abilities dissociate.
Everyone has all seven. These are features of ordinary human cognitive architecture, not defects. A high score is an invitation to try the interventions listed for that bias, not a diagnosis.
Your score is a snapshot. Susceptibility shifts with practice, mood, time of day, and item set — research suggests movement on the order of 10–15 percentile points. Retest in a month and expect a different number. That is measurement working, not failing.
What you can actually do about it
The literature on debiasing is more modest than the popular literature suggests, but a few strategies have support. Consider the opposite: before committing to a judgment, explicitly generate reasons it could be wrong — one of the better-replicated interventions, effective against overconfidence and hindsight. Take the outside view: ask what usually happens to projects like this one, rather than reasoning from the specifics of yours; this directly counters base-rate neglect. Pre-commit to a stopping rule before you invest, which is the only reliable defense against sunk cost, since the bias operates after the money is spent. Separate decision quality from outcome quality in reviews: ask what was knowable at the time, not what happened.
Limitations
This is an educational instrument, not a clinical or diagnostic one. It uses a fixed item set, so repeated testing with the same items measures partly your memory of them. Scores are percentile positions within a scoring model, not clinical norms. The Bias Blind Spot is a discrepancy measure and inherits the noise of both components. Bias tasks in general show modest test-retest reliability — a well-documented property of the field, not a flaw in this implementation, and a reason to treat any single score as provisional.
The evidence base
The instrument draws on the heuristics-and-biases and rational-thinking literatures: Kahneman and Tversky on heuristics and biases; Frederick on cognitive reflection; Bruine de Bruin and colleagues on adult decision-making competence; Stanovich, West, and Toplak on the independence of rationality and intelligence; Scopelliti and colleagues on the bias blind spot; and Teovanović and colleagues on the structure — or absence of structure — among individual bias measures. Full references and scoring detail are in the methodology.