What Tests Can and Cannot Measure
Tests are among the most familiar features of school life, and few students pass through education without taking many of them. They are often criticized, yet they are rarely abandoned. Understanding both their strengths and their limits is therefore an important task for anyone who cares about education. Otherwise, debates about testing tend to become a contest between those who praise it without reservation and those who condemn it without thought.
The strengths are real. A well-designed test applies the same questions under the same conditions to everyone, which makes comparison possible and reduces the influence of personal favoritism. It gives teachers information about what has been learned, and it gives learners a clear target and feedback. In systems where opportunities are limited, an impartial measure can help protect fairness. Many students from families with few connections have found that a common examination gave them a chance that personal recommendation would not.
The limits, however, are equally real. A test can measure only what can be examined within a short time and a fixed format. Qualities such as curiosity, persistence, creativity, and the ability to cooperate are difficult to capture in this way. When a score becomes the main goal, moreover, students and schools may begin to concentrate on what appears on the test and neglect everything else. A measure that was meant to reflect learning may then start to distort it. Teachers, too, feel this pressure and may teach only what is likely to be tested.
Some people conclude that tests should be reduced or even removed. This response seems too hasty. Without a common standard, decisions about students might rely on impressions and personal connections, which could be more unfair than any examination. The difficulty lies not in measurement as such but in giving one measurement too much power. A ruler is a fine instrument for length, but no one would use it to measure kindness.
A more balanced approach would combine several kinds of evidence: written tests, projects, presentations, and the observations of teachers over time. No single item would be decisive, and each would balance the weaknesses of the others. Learners, in turn, could be encouraged to see a test as a tool that reveals their present position rather than as a verdict on their worth. Such an attitude can reduce anxiety and keep attention on improvement.
Ultimately, the question is not whether to measure, but how to interpret what we measure. A score is a useful piece of information, not a complete portrait of a person, and education is at its best when it remembers the difference.
※この英文は原田英語のオリジナルです。