Absolute mastery vs. relative ranking
Criterion-referenced tests measure a student's performance against a fixed standard of competence, regardless of how others perform. A student either demonstrates the ability to multiply fractions or does not, based on a pre-defined mastery threshold. If all students master the standard, all receive high scores. If none do, all receive low scores. The interpretation is always absolute: the score tells you what the student can do, not how they compare to peers.
Norm-referenced tests rank students relative to each other, with the average performance determining the middle of the score distribution. A student's score depends on how many other students perform worse. If everyone improves dramatically, scores stay roughly the same because what matters is relative position. A norm-referenced test inherently produces a spread of scores, even if all students learn the material well.
Uses and implications for interpretation
Criterion-referenced tests align with standards-based education and mastery learning approaches. They answer the question: Does this student meet grade-level standards? Norm-referenced tests answer: How does this student rank compared to peers? Schools increasingly favor criterion-referenced tests for formative feedback and classroom decisions because the feedback is more actionable: teachers know exactly which standards students have not met.
However, norm-referenced tests remain common for college admissions and gifted identification because those decisions require differentiation among high-performing students. A criterion-referenced test might report that 90% of test-takers achieve proficiency, providing no guidance for selective purposes. The choice between frameworks depends on the purpose of assessment: improving instruction argues for criterion-referenced tools, while competitive selection argues for norm-referenced tools. Many high-stakes tests use hybrid approaches that provide both types of information.