Comprehensive Definition
A criterion-referenced assessment measures a student's performance against a fixed standard, criterion, or set of learning objectives, rather than comparing them to other students. The question it answers is: "What can this student do?" or "Has this student mastered this specific skill or content?"
The student's score tells you whether they have met a defined benchmark, not how they rank against peers. Results are typically reported as mastered or not yet mastered, a percentage correct, a stage or level reached, or a list of specific skills achieved.
How It Differs From Norm-Referenced Assessment
This contrast is usually the clearest way to understand criterion-referenced assessment.
Norm-referenced assessments compare a student to a representative sample of peers (the "norm group"), usually of the same age or year level. Results are reported as percentiles, standard scores, or stanines. The question is: "How does this student compare to others?" These assessments are often used to help identify learning disorders or significant deviation from age-level expectations. Examples include the DYMOND, PAT, and most standardised diagnostic tools.
Criterion-referenced assessments compare a student to a pre-set standard. The question is: "Has this student learned what we taught?" or "Where is this student in the scope and sequence?" These assessments are used to guide instructional decisions and monitor progress.
A useful way to think about it: norm-referenced tests tell you where a student sits in a crowd; criterion-referenced tests tell you where a student sits in the curriculum.
When Criterion-Referenced Assessments Are Useful
They are particularly well-suited to:
Placement within a structured literacy program (identifying which stage to start teaching)
Monitoring progress through a defined scope and sequence
Confirming mastery before moving to the next skill
Screening for specific skills (e.g., letter-sound knowledge, phonics elements, phonemic awareness)
Informing day-to-day instructional decisions and targeted teaching groups
They are less useful for identifying learning disorders or establishing how a student compares to age-level expectations, which is where norm-referenced tools are needed.
FAQs
Q: Can the same assessment be both criterion-referenced and norm-referenced?
A: Some tools have features of both, or can be interpreted either way. What matters is the reference point used to make meaning of the score. If the score is compared to a standard or benchmark, it is being used as a criterion-referenced measure. If it is compared to a peer sample, it is being used as a norm-referenced measure.
Q: Are criterion-referenced assessments standardised?
A: Not usually in the same way as norm-referenced assessments. Criterion-referenced tools are standardised in their administration and criteria (so all students are measured against the same standard), but they do not typically have statistical norms derived from a representative sample population.
Q: Can I use criterion-referenced assessments to identify a learning disorder?
A: Generally, no. Identifying a learning disorder usually requires norm-referenced data that shows how significantly a student's performance deviates from age-level expectations. Criterion-referenced tools can flag that a student is well behind expected curriculum progress, which may prompt further investigation, but formal diagnosis requires norm-referenced assessment.
Q: How often should criterion-referenced assessments be administered?
A: This depends on the purpose. Placement tests within a structured literacy program are typically administered termly to monitor movement through the scope and sequence. Mastery checks for individual skills may be administered more frequently (e.g., weekly or at the end of a teaching cycle) as part of formative assessment. Norm-Referenced Test (NRT) Placement Test Screening Assessment Diagnostic Assessment Benchmark Progress Monitoring Scope and Sequence