- Age and item sets differ.
- On speeded error-prone tasks, rushing can raise attempts and lower quality.
- Only as “I liked this toy.” Units differ.
- If everyone scores 16/16, the raw score stopped informing.
- No.
- Visual counting lite and category fluency lite both emit raw counts.
Why convert at all?
Age and item sets differ. A raw 10 can be typical at one age and low at another on a real battery. Übung, kein klinischer IQ.
Is a higher raw always better?
On speeded error-prone tasks, rushing can raise attempts and lower quality. Read the toy’s rule. Übung, kein klinischer IQ.
Can I compare raw scores across games?
Only as “I liked this toy.” Units differ. No joint metric. Übung, kein klinischer IQ.
Ceiling and floor?
If everyone scores 16/16, the raw score stopped informing. Our lites are short; ceilings happen. Übung, kein klinischer IQ.
Does computer scoring make it scaled?
No. Automatic counting is still raw unless a norm engine is documented. Übung, kein klinischer IQ.
What to try?
Visual counting lite and category fluency lite both emit raw counts. Enjoy the count. Do not file it as IQ. Übung, kein klinischer IQ.
❓ Häufig gestellte Fragen
What is construct validity?
Construct validity is the argument that a score reflects a named idea — working memory, not “being clever in general.” It is built from theory, patterns of correlations, and failed rival explanations. A 20-trial number-Stroop lite can feel like control. Feeling is not a validity coefficient. This site never reports a licensed IQ. Übung, kein klinischer IQ.
What is construct validity? →What is criterion validity?
Criterion validity asks whether a score lines up with something you care about that is not the test itself — a later grade, a supervisor rating, or a Wechsler index. Concurrent means measured at the same time; predictive means later. A visual 2-back lite has no published criterion table. It is not an employment screen and not an IQ. Übung, kein klinischer IQ.
What is criterion validity? →What is content validity?
Content validity is whether the item set covers the skill or knowledge you claimed. A spelling test with only the letter Q is a coverage failure. Shape-span lite samples a tiny visuospatial sequence. It does not cover “all spatial intelligence.” Experts, blueprints and reviews build content arguments. A weekend hackathon does not. Übung, kein klinischer IQ.
What is content validity? →What is a confidence interval?
A 95% confidence interval is a procedure: if you repeated the study under the same model, 95% of such intervals would cover the parameter. It is not the probability that this one interval magically contains the truth after you saw it. Licensed IQ manuals turn a standard error into a range around a scaled score. Digit-span lite does not. Treat any integer from this site as a toy count. Übung, kein klinischer IQ.
What is a confidence interval? →What is item response theory?
Item response theory (IRT) models each item: harder items need more of the trait for a 50% chance of success; discriminating items separate nearby people. Computerized adaptive tests pick the next item from that model. Pattern-matrices lite is a fixed original set. It does not estimate a theta with published item parameters. It is not an IQ. Übung, kein klinischer IQ.
What is item response theory? →