Standardized testing does more harm than good
What's this about?
People disagree about whether common tests do more harm than good. The answer may depend on how schools use them.
What supporters say
- Test gains may miss lost time, stress, and less learning in art, health, or civic life.
- Test scores show both learning and a child’s family wealth, help, and school tools.
- Big test stakes can make schools teach test tricks instead of deep and wide learning.
- Test stress can hurt focus, joy in learning, and some students’ test scores.
What critics say
- Test rules can push schools to work harder, which may raise student scores.
- The same test can show unfair gaps between groups that schools might hide.
- Test scores can give useful clues about how students may do in later school work.
How to read this
The number of points on each side does not show who is right; check the strength of each point’s proof.
The bottom line
Common tests can help schools spot gaps and track learning. But high-stakes use can narrow lessons, add stress, and hide wider harms.
Both sides have points, but weak proof supports each claim, so the answer is still unclear.
The claim that standardized testing does more harm than good depends heavily on how tests are used. The evidence points to real benefits from common measures, but also serious costs when scores carry high stakes for students, teachers, or schools.
The case for
The strongest argument is that high-stakes testing can reshape education around the test rather than around broad learning. Reviews of accountability systems have found incentives to focus on tested subjects and question formats, reduce time spent on untested areas, and sometimes exclude or manage low-performing students. Gains are often larger in tested mathematics and tested grades than in other subjects or wider educational outcomes, suggesting that schools may be improving their scores more than education as a whole. 1 (see Figure 3)
Standardized scores can also turn unequal opportunity into what appears to be a neutral measure of individual merit. Research finds strong links between test performance and family income, preparation, and access to educational resources. Studies of test-optional college admissions suggest that requiring tests can discourage some qualified students from applying when they lack access to preparation or confidence about their scores. 2
High-pressure testing may create additional problems. Reviews and psychological research link high stakes with anxiety, stereotype threat, weaker intrinsic motivation, and a greater focus on performing rather than learning. The effects differ among students and settings, but experimental and observational evidence indicates that pressure can reduce some students’ demonstrated performance. 3
There is also a broader concern about what test results leave out. A score may show improvement in a narrow area while missing lost teaching time, reduced attention to the arts or civic learning, stress, and other costs. For that reason, measured gains do not necessarily mean that students’ overall education has improved. 4
The case against
Standardized tests provide information that local grades and teacher judgments may not offer in the same way. Common results can reveal differences among schools, regions, and student groups, identify underperforming schools, direct support, and track changes over time. They can make achievement gaps visible rather than allowing them to remain hidden. 5
Accountability can also produce improvements, not merely measure them. Quasi-experimental research from New York associated graduation examinations with more course-taking and higher graduation outcomes. Reviews and meta-analyses find positive average effects from accountability, especially in mathematics, along with some gains for disadvantaged students. These improvements are generally modest and are not consistent across every group or outcome. 6 (see Figure 4)
Test scores can further provide useful comparative and predictive information in college selection. Admissions research finds that scores contain information about academic preparation and first-year college performance, particularly when grading standards differ between schools. But their added value beyond high-school grades is limited and depends on the context. 7
The bottom line
The evidence is balanced rather than strong enough to show that standardized testing as a whole causes more harm than good. Confidence in that overall assessment is high, but the answer changes substantially according to the test’s design and consequences.
The evidence for harm is repeatedly documented when tests are tied to sanctions, promotion decisions, or school ratings based on narrow measures. The benefits of common measurement and accountability are also credible, but they are usually modest and uneven. The best-supported conclusion is therefore conditional opposition to high-stakes testing, not a categorical rejection of standardized assessment.
Low-stakes monitoring and accountability systems using several measures may preserve comparability while reducing curriculum distortion and pressure. Still, researchers cannot yet combine learning gains, stress, lost instruction, unequal opportunity, and predictive value into one reliable estimate of overall harm or benefit. The balance ultimately depends on how decision-makers value narrow achievement gains against broader learning and distributional costs.
Figures & data
All contributions are reviewed for clarity, balance, and evidence. The strongest insights are elevated into the argument graph — with credit to you.
Help improve this analysis →

