Education & Measurement
A single number on a report card summarizes performance, a teacher’s expectations, school rules, and a student’s development. It is therefore useful for ongoing communication but risky as the sole basis for comparing children from different classes and schools.
1. What Can Be Substantiated
An IDEA study showed that schools differ in the strictness of their grading even after test results are taken into account. For equally capable students, part of the difference in grades may therefore arise from a school’s rules and context.[1]
TIMSS 2023 provides a standardized view of mathematics and science knowledge. Czech eighth graders scored 518 points in mathematics, but this indicator is neither a grade for any particular child nor a measure of all the goals of schooling.[2]
The Czech Ministry of Education prepared guidance for the transition to narrative assessment in the first two grades. The change does not by itself guarantee the quality of feedback; that depends on the criteria, clarity, and the teacher’s work.[3]
2. How to Read the Claims in Context
A grade is an ordinal category, not a thermometer with the same nationwide scale. A top grade at a strict school and a top grade at a lenient school may reflect different levels of performance. This does not mean that the teacher is acting improperly; it means that a comparison without context has limited explanatory value.
A standardized test addresses part of the problem by giving everyone the same tasks and rules. At the same time, it measures only a defined subset of skills on a particular day. A sound decision therefore combines multiple sources of information: long-term work, a test, a narrative description of strengths and weaknesses, and the conditions under which the child worked.
The most precise question is not whether grades or tests are better. It is what decision the information is needed for and what bias each instrument introduces into it.
- The strictness of grading varies among schools.
- A standardized test and a school grade measure differently defined constructs.
- Narrative assessment will be introduced gradually in the early grades.
- A single universal weighting that would convert a grade and a test result into “true ability.”
- The impact of the change in assessment without subsequent evaluation in schools.
- How a particular child would perform in a different environment.
3. Five Questions to Ask
- What criteria were used to assign the grade?
- Is the comparison within one class, or among children from different schools?
- What subset of skills does the test measure?
- Does the feedback describe the next step?
- What uncertainty remains before a decision about the next educational path?
4. Conclusion
A grade can be a useful message, but it is not a universal unit of ability. Assessment becomes fair only when we understand the number in context and supplement it with further evidence of learning.
— Jiný Kontext
