Call it reductive, but the fundamental to understanding educational assessment might just lie with a simple question.
It has plagued philosophers from Plato and Spinoza to Russell. It spurred psychologists, such as Alfred Binet and Theodore Simon, to develop new instruments. It has even haunted Ms. Whitney Houston herself.
That question: How do I know?
However, if this question is applied to contemporary, large-scale summative assessments, the question blossoms to: How does the American education system know that students are learning and based on these results, what decisions can be made to hold schools accountable and allocate resources for students?
That simple question got really big, really fast. The history of standardized summative testing, like DC CAPE Next, its application, and contemporary considerations about its future are equally large and complicated. So, pick up those pencils or log in to those secure computer applications! We’re about to break down the history and application of standardized summative assessments into four blog posts. First, we’ll lay out some key terms and definitions. Then, we’ll look at the history of educational testing, followed by its contemporary application. Finally, we’ll consider what the future of assessment might look like, especially considering advances like Artificial Intelligence (AI).
What is Educational Assessment?
According to the Institute of Education Sciences, “educational assessment is a systematic process of documenting and using evidence to improve educational programs and student learning” (Regional Educational Laboratory Northeast & Islands, 2025). Broadly, there are two categories of educational assessment: Assessment for Learning and Assessment of Learning.
Assessment for Learning
Educators assess student learning frequently to identify students’ strengths, diagnose gaps in learning and make more informed instructional decisions. These assessments are generally diagnostic and formative assessments used to help educators and students make in-time decisions to facilitate learning. They help a student establish goals and may help a teacher revise lesson plans to provide more targeted instruction for their students. The feedback from these assessments is frequent and ongoing to help make day-to-day decisions.
Assessment of Learning
Alternately, there are assessments of student learning. These are interim or summative assessments that are administered infrequently to track learning achievement against specific targets, such as learning standards. For the purposes of this blog series, we will focus summative assessments or those that evaluate long-term student learning and achievement.
These are assessments that are administered at the end of a specific period, such as the end of a school year, to summarize students’ growth and achievement. While a summative assessment may be administered by a teacher or department to evaluate students’ achievement at the end of a unit or semester, these are not necessarily standardized summative assessments, like annual statewide summative assessments or college entrance exams.
All summative assessments evaluate student learning and achievement at the end of a specific period of time, but not all are standardized. The word standardized is critical here. It refers to statistical measures in the design and interpretation of an assessment to ensure its fairness and better identify opportunity gaps to create a more equitable learning experience for all students. The key distinction is that statewide, standardized summative assessments are developed to ensure reliable and valid results across groups of students, providing educators, schools, districts and states the ability to make comparisons across student groups to analyze progress and inform policy decisions (Gregory, 2026).
A Simple Premise, A Complicated History
The premise of standardized summative assessments is deceptively simple: If every student answers the same questions in the same conditions, then results will be more efficiently generated, more objective, and provide a clearer understanding of how students are learning (Gregory, 2026). Standardized summative testing appears to be incredibly egalitarian and democratic, even. However, the advancement of standardized testing that began in the late 19th century is predicated on complex population change, a bourgeoning and relatively undefined academic field, and a limited set of methodology which led misapplication of assessments to confirm biases and promote inequity.
The next installment of this series will trace the origins of standardized summative testing to establish how we arrived at the perennial assessment question—How do I know?—in the first place.
References
Gregory, M. (2026, April 1). Student testing history: How standardized testing shaped K–12 education in the U.S. [Blog post]. Education Advanced.
National Education Association. (2020, June 25). History of standardized testing in the United States [Blog post]. NEA.
Regional Educational Laboratory Northeast & Islands. (2025, January). Making sense of educational assessment [Fact sheet]. U.S. Department of Education, Institute of Education Sciences. [ies.ed.gov]
United States Congress, Office of Technology Assessment. (1992). Lessons from the past: A history of educational testing in the United States (Chapter 4 in Testing in American schools: Asking the right questions, OTA-SET-519). U.S. Government Printing Office. [princeton.edu]

