How the thinking test measures — and where the limits are
What the test measures
44 puzzles in five kinds: pattern matrices, number series, odd-one-out tasks, a memory matrix (visual working memory) and the quick look (counting a briefly shown field). The focus is on what research calls “fluid intelligence” — spotting rules in new material; memory and attention are smaller building blocks. The test is deliberately language-free and identical in every language.
It does NOT measure: knowledge, vocabulary, creativity, practical wisdom, social intelligence — thinking is much broader than any online test. Memory and attention are only touched on (6 puzzles each) and do NOT yield a robust score on their own.
Why a range instead of a number
Every test score carries measurement error: daily form, tiredness, practice, lucky guesses. Serious assessment therefore works with confidence intervals. An online test selling you “IQ 128” as an exact number hides this uncertainty — we show it.
Calibration phase — what that means
This test is new and being calibrated. For now your placement rests on a provisional, uncalibrated model: it shows your hit rate corrected for lucky guessing with a statistical uncertainty range (Wilson interval), not a comparison against a sample — we say so right on the result. Every anonymous run grows the data base; once it carries, we switch to empirical percentile ranks and re-anchor all results for free. Online samples remain self-selected — we will state that too.
Four rounds — and why you can only stop between them
The test runs in four rounds of 11 puzzles. Each round contains all five puzzle kinds and the same difficulty mix, starting easier and ending harder. That is why you can stop at the end of any round: what you worked through is a fair miniature of the whole test. Mid-round stopping is deliberately blocked — you would skip the hard puzzles at the round's end and your result would come out too good.
The “meaningfulness” bar shows how sharp your result already is: the error of a proportion falls with the square root of the number of puzzles — one round gives roughly half the sharpness of all four.
Fairness & limits
- No speed score: the generous per-puzzle limits only discourage outside help — speed never enters your result. This keeps the test fair on phone and computer. In the quick-look task a brief noise pattern follows the image: it ends the after-image so that perception is measured, not lingering review.
- Unsupervised means estimate: nobody checks who takes the test and how. For robust scores there is professional, supervised assessment.
- Not a clinical instrument, no diagnosis, not for hiring or admissions — and no judgement about a person.
- Puzzles are colour-vision-friendly (shape and pattern carry the information, never colour).
The business model, stated openly
Free result, detailed breakdown for a one-time 4.99 € — no subscription, no ads, no data sharing. Purchased breakdowns update for free with every norm improvement.