How psychological measurement actually works
The concepts behind every test in the catalog — reliability, validity, norms, measurement error — explained without the jargon barrier, plus straight answers to the framework questions people actually ask.
How measurement works
What makes a psychological score trustworthy, how much of it is error, and which questions separate a validated instrument from a questionnaire someone wrote in an afternoon.
What is psychometrics?
The discipline that works out whether a psychological measurement is any good — and the reason two tests asking similar questions can differ enormously in what their results are worth.
What is reliability in psychological testing?
Reliability is consistency of measurement, not correctness. A test can be highly reliable and still measure the wrong thing — which is why it is the first question asked and never the last.
What is validity in psychological testing?
Validity is not a property a test has. It is an argument about whether a particular interpretation of a particular score, for a particular purpose, is defensible — and it has to be rebuilt for every new use.
What is construct validity?
The hardest question in measurement: whether the thing you are measuring exists as you have defined it, and whether your instrument reaches it rather than something adjacent.
Convergent and discriminant validity
Two mirror-image requirements: a measure must agree with what it should agree with, and must stay distinct from what it claims not to be. Most weak instruments fail the second one.
What is Cronbach's alpha — and what it is not
The most reported statistic in psychological testing, and the most misread. Alpha does not tell you a scale is unidimensional, and a high value is often a symptom rather than a virtue.
Standard error of measurement: how much your score could be off
Every score carries a margin of error, and for most psychological tests it is wider than people assume. This is the number that tells you when a difference is real and when it is noise.
Norms and percentiles: what your score is compared against
A raw score is uninterpretable on its own. It becomes meaningful only against a reference group — and which group was used is one of the most consequential and least advertised facts about any test.
How a psychometric test is actually built
From construct definition to published norms, the sequence takes years and discards most of what goes into it. Knowing the steps makes it obvious which ones a quick online quiz skipped.
What does it mean when a test is called validated?
The word is unregulated and used freely by products that have done none of the work. Here is what it means when it means something, and the four questions that separate the two cases.
Why most internet personality tests are not reliable
Not because they are dishonest, but because the steps that make a measurement trustworthy are invisible, expensive and easy to skip — and skipping them changes nothing about how the result looks.
What self-report can and cannot measure
Questionnaires ask you to be the observer of yourself, which works better for some things than others. Knowing where the method is strong and where it fails changes how you read any result.
What it means when a test predicts something
A correlation of 0.30 is a strong finding in personality research and a weak basis for a decision about one person. Both statements are true, and the gap between them causes most misuse of test results.
Screening is not diagnosis
A screening questionnaire is built to catch as many possible cases as it can, accepting a high false-positive rate as the price. Reading a screening score as a diagnosis inverts what it was designed to do.
Frameworks compared
MBTI, DISC, the Big Five and the questions that get asked about them. NOESIS publishes none of the first two, which is exactly why these answers can be straight.
Is the MBTI scientific?
The honest answer is more specific than yes or no: the questionnaire is reasonably well constructed, and the type theory it reports results in is not supported. Those two facts are usually collapsed into one argument.
Big Five vs MBTI: what actually differs
They partly measure the same things. The difference is not which traits they cover but what they do with the numbers — and one of those choices throws away information the other keeps.
Is DISC a personality test?
DISC describes behavioural style at work rather than personality structure, which its better publishers say plainly. The trouble starts when a workplace style profile is used as if it measured the person.
Big Five vs DISC: what each one is for
They are not competing measures of the same thing. One is a structural model of personality built from data; the other is a workplace style framework built for facilitation. Comparing them on accuracy misses the point of both.
Personality and temperament: what is the difference?
Temperament is the biologically grounded raw material visible in infancy; personality is what develops on top of it through experience. The distinction is real, and the boundary is fuzzier than textbooks suggest.
Does personality change over your lifetime?
Yes, in a direction most people find reassuring, and more slowly than self-help promises. The evidence for both the change and its pace is unusually good.
Is introversion the same as shyness?
No, and the confusion causes real harm. One is a preference about how you spend energy; the other is anxiety about being judged. They are separate traits that happen to produce similar-looking behaviour.
Can emotional intelligence be measured?
Partly, and the answer depends on which of two very different things you mean by the term. One is measurable as an ability; the other is measurable as a self-perception. They are usually sold as the same thing.
See what a documented test looks like
Every test in the catalog names the published instrument it implements, its authors and year, how it is scored, and what its results do not establish. That last part is the comparison worth making.