The Literacy Racket: Twenty-Five Years of Treating Children Like Data and Teachers Like Suspects
An opinion essay
There is a peculiar kind of stupidity that only the well-funded can afford, and American education reform has been marinating in it for a quarter century. Ask yourself a simple question, one that a ten-year-old could pose and that no consultant has yet been paid to answer honestly: if the last twenty-five years of federal mandates, billionaire philanthropy, standardized testing regimes, and educational technology were actually working, why do two entire generations now read worse, on average, than the generation that came before No Child Left Behind was ever signed into law?
This is not a rhetorical flourish. It is the plain arithmetic sitting inside the 2024 Nation's Report Card. Twelfth-grade reading scores are the lowest ever recorded since the assessment began in 1992 — lower than 2019, lower than 1992 itself, with the bottom of the distribution in a free-fall that predates the pandemic by more than a decade. Only thirty-five percent of graduating seniors read at a level associated with readiness for entry-level college coursework. A record share of fourth graders now score below even the "Basic" threshold — the floor beneath which a child cannot reliably infer why a character in a story did what she did. Reform, on its own terms, has failed. Not "underperformed." Not "produced mixed results requiring further study." Failed — while consuming, along the way, staggering sums of public and philanthropic money that could have gone almost anywhere else and probably done less damage.
I want to walk through exactly how this happened, because the how matters more than the outrage, and because the people responsible have spent twenty-five years making sure the how stays comfortably vague. Vagueness is the fog in which bad reform survives long enough to be replaced by the next bad reform.
No Child Left Behind: measuring the patient to death
George W. Bush's signature education law promised that standardized testing and sanctions for underperforming schools would close achievement gaps and lift the floor. It is worth being precise about what the actual research — not the marketing around the bill, the research — eventually found. The most rigorous national evaluation, by Thomas Dee and Brian Jacob, did find some gains in elementary math. It found none in reading, at either the fourth- or eighth-grade level. Zero. A separate, earlier analysis by Jaekyung Lee at Harvard, using two decades of NAEP data as a baseline, concluded that NCLB did not improve achievement overall and did not narrow the gaps it was built to close. Researchers studying individual states found that schools facing sanctions responded rationally to the incentives they were given: they triaged. They devoted instructional time to "bubble kids" hovering near a proficiency cutoff, and evidence from at least one state showed outcomes for the most disadvantaged students actually deteriorating under the pressure. This is what you get when you attach the survival of a school to a single number: you do not get better teaching, you get better test management. The law's genius, if you can call it that, was in making an entire profession answerable for outcomes while stripping away almost all of its control over the conditions that produce those outcomes — home literacy environment, poverty, class size, curriculum mandates handed down from above. NCLB weighed the hog obsessively and never once fed it.
Race to the Top and the Common Core: standardizing our way to stagnation
The Obama-era answer to NCLB's failure was, naturally, more of the same mechanism with a fresh coat of paint: competitive grants, new tests, and a single set of standards imposed with the enthusiastic backing of the Gates Foundation, which spent hundreds of millions of dollars helping write and disseminate the Common Core State Standards. Even sympathetic researchers examining Common Core implementation found that whatever instructional strategies schools adopted to raise math scores under the standards showed little corresponding benefit for reading and English language arts — the two subjects were not, it turned out, interchangeable, and no one had bothered to check before rolling the standards out nationally. Later analysis using NAEP data and teacher surveys found that Common Core's narrow tested subjects crowded out instructional time, materials, and the quality of teacher-student interaction in the subjects the standards did not directly target. The reform, in other words, cannibalized the very breadth of instruction — the reading aloud, the discussion, the untested subjects that build the background knowledge on which comprehension depends — that a literate citizenry actually requires.
Bill Gates: the billionaire who ran the largest uncontrolled experiment in American schooling and admitted it failed
I want to dwell on this one, because it deserves dwelling on. No single private citizen has shaped American classroom policy over the last twenty-five years more than Bill Gates, and to his considerable credit, he has been more honest about the results than almost anyone else in this essay. The foundation's first major bet, roughly $650 million to break up large high schools into smaller ones, was abandoned in 2009 after the Gateses concluded in their own annual letter that most of the small schools they funded produced no meaningful gains in student achievement. The foundation pivoted to teacher evaluation, funding a seven-site, $575 million "Intensive Partnerships for Effective Teaching" initiative between 2009 and 2016. The RAND Corporation's evaluation — 587 pages, plus nearly 200 pages of appendices, because failure this expensive requires a great deal of paperwork to properly document — found the initiative did not achieve its goals for students, particularly the low-income and minority students it was explicitly designed to help. Outcomes were, in the report's own bloodless language, "null to negative" across a range of measures. It found no evidence the reform even succeeded at its narrower goal of getting schools to hire more effective teachers. And in the middle of all this the foundation was simultaneously bankrolling Common Core.
Add it up and you get more than a billion dollars spent by a single foundation on interventions that, by the foundation's own admission and by independent research, did not improve outcomes for the students they targeted. Gates deserves some credit for saying so publicly — most reformers when they fail simply rebrand and move on to a new grant cycle — but the candor does not undo the two decades of instructional time, teacher energy, and public trust those experiments consumed while classrooms were told to fall in line with the science of the moment. A man who has never taught a classroom of twenty-eight seven-year-olds through a Tuesday afternoon got to run, essentially, a national pilot program on other people's children, at public-school scale, for a decade, and when it did not work, the bill was paid by the students who had already lost the most ground and would go on to lose more.
The ed-tech grift: selling the cure that the data does not support
Now the ed-tech companies, screaming — Sean's word, and it's the right one — that they have the fix. Screens as pedagogy, adaptive software as personalized learning, an app for every deficiency a testing regime has just spent two decades manufacturing. Here the evidence is not even ambiguous in the way education research usually is. A 2009 U.S. Department of Education study of reading and math software products used with genuine randomized assignment found the overall effect on student achievement was, in plain terms, zero. A comprehensive 2017 review of the ed-tech research literature, published by the National Bureau of Economic Research, reached similarly deflating conclusions about most classroom technology interventions on their own. More recent studies tracking screen time and reading achievement in elementary-age children have found negative associations between heavy non-academic screen use and reading scores, and researchers looking at the timing of the NAEP's most recent declines have noted, cautiously, that American students have never had more digital access to text and have simultaneously never scored lower. Nobody serious claims this proves causation on its own — the honest researchers are careful to say declines have many contributing causes, the pandemic among them — but that only makes it stranger that an entire industry spent the same twenty-five years insisting the correlation should run the other way, that more devices in more hands would be the intervention that finally worked. It is difficult to overstate how convenient it is, for a company selling licenses by the seat, that the recommended treatment for declining literacy is always more of the product being sold.
The publishers: selling the same horse in different tack
Meanwhile the curriculum publishers — the Pearsons and their descendants — have spent this same period cycling through "the next great program" with the reliability of a metronome, each new basal series arriving with a glossy binder and a promise, most of them never independently validated by anyone without a financial stake in the sale, and most of them abandoning any serious commitment to systematic phonics and oral language development in favor of whatever balanced-literacy or three-cueing fashion was ascendant that decade — approaches the cognitive science of reading has been quietly and thoroughly discrediting since long before "the science of reading" became a marketing phrase itself, which tells you something about the industry's capacity to absorb even its own critique and resell it.
The scapegoat: what we did to the people actually in the room
And through all twenty-five years of this — the testing, the standards churn, the philanthropic pilot programs, the software licenses, the curriculum adoptions — there has been one constant, reliable object of blame whenever the numbers came in bad: the teacher. Not the architecture. The person standing at the front of the room, holding a salary that RAND's most recent survey puts roughly thirty thousand dollars below what a comparably educated adult earns elsewhere, working measurably more hours per week than that comparable adult, and reporting burnout, depression symptoms, and inability to cope with job stress at multiples of the general working population. Researchers at Brown University, examining fifty years of data, describe teaching as being in its worst structural condition in half a century. This is the workforce that twenty-five years of reform decided was the variable to fix through sanctions, scripted curricula stripped of professional judgment, and evaluation systems tied to test scores whose own architects, per RAND's own evaluation of the Gates initiative, could not make work. You do not professionalize a workforce by treating it as the point of failure in a system it did not design and does not control. You do not recruit the next generation of talented, literate adults into a profession that pays them less and blames them more with each passing reform cycle. It should not require a Ph.D. in labor economics to notice that the years of declining literacy are also the years of declining teacher autonomy, and to at least entertain the possibility that these are not unrelated facts occurring in parallel by coincidence.
What actually got starved
Here is the part that should make you angriest, if you are not already there. While all of this was happening — the testing, the sanctions, the standards, the software, the scapegoating — the two things a literate, discerning citizenry actually requires were left to wither by neglect. The first is oracy: the rich, extended, dialogic talk between adults and children that builds the vocabulary and background knowledge on which all later reading comprehension depends, and that gets systematically squeezed out of a school day organized around test preparation. The second is the capacity to evaluate what you read once you can, technically, decode it. Stanford's History Education Group tested nearly eight thousand students, from middle school through college, on their ability to assess the credibility of information online, and its own researchers described the results in one word: bleak. Eighty percent of students could not distinguish sponsored content from a news article. A follow-up study found more than half of high schoolers considered a grainy, unverified video "strong evidence" of a claimed act of voter fraud. Ninety-six percent gave no consideration to whether a source's funding might bias its content. This is not a coincidence sitting next to the literacy collapse. It is the same collapse, viewed from a different angle — the predictable result of twenty-five years spent teaching children to bubble in an answer sheet rather than interrogate a claim, to comply with a pacing guide rather than reason through a text, to fear a test rather than love a book.
What would actually have to change
None of this is a mystery, and none of it required twenty-five years and tens of billions of dollars to discover. Give teachers the professional autonomy, the smaller caseloads, and the compensation that treats them as the load-bearing wall of the entire enterprise rather than as replaceable line workers to be evaluated by a number they do not control. Fund the boring, unglamorous, well-replicated science of how children actually learn to read — systematic phonics instruction paired with vast amounts of vocabulary-building talk and text, not whichever cueing fashion a publisher's marketing department has rebranded this fiscal year. Stop importing corporate logic — key performance indicators, market disruption, scalable solutions — into an enterprise that is fundamentally relational, slow, and human, and that has never once been fixed by a man who made his fortune in an entirely different industry deciding he knows better than the people who have spent their careers in the room. And admit, finally, out loud, in public, that twenty-five years of top-down reform imposed on the people actually doing the work has produced exactly what you would expect from any system designed by people insulated from its consequences: worse outcomes, a demoralized workforce, and a generation increasingly unable to read a claim critically enough to know when it is being lied to.
That last part should worry you most of all. A citizenry that cannot discern propaganda is not merely an educational statistic. It is a vulnerability, and it was manufactured, one testing mandate and one grant cycle at a time, by people who were very well paid to get it wrong.
Billionaire-funded reforms, particularly those spearheaded by the Gates Foundation, have failed to improve reading scores and have coincided with a period of significant decline in student literacy since 1992,,.
The impact of these reforms on reading scores can be broken down into the following key findings from the sources:
Historically Low Performance
The 2024 Nation's Report Card reveals that twelfth-grade reading scores are the lowest ever recorded since assessments began in 1992. This decline is not merely a post-pandemic phenomenon; the bottom of the score distribution has been in "free-fall" for more than a decade. Currently, only 35% of graduating seniors read at a level ready for college coursework, and a record number of fourth graders score below the "Basic" threshold.
Specific Billionaire-Funded Initiatives
The Bill & Melinda Gates Foundation has spent over one billion dollars on education interventions that, by their own admission and independent research, did not improve student outcomes,.
- Small High Schools: Approximately $650 million was spent to break large high schools into smaller ones, a project abandoned in 2009 after it produced no meaningful achievement gains.
- Teacher Evaluation: A $575 million initiative known as "Intensive Partnerships for Effective Teaching" (2009–2016) resulted in outcomes described as "null to negative". A RAND Corporation evaluation found it failed to achieve goals for the low-income and minority students it was designed to help.
- Common Core: The Gates Foundation spent hundreds of millions supporting the Common Core State Standards. Research found these standards provided little benefit for reading and English language arts and actually "crowded out" essential instructional time for reading aloud and discussion.
Structural Failures of Reform Logic
The sources argue that the billionaire-led "reform" movement shifted the focus of education in ways that actively harmed literacy:
- Instructional Cannibalization: The focus on narrow, tested subjects under the Common Core led to the neglect of untested subjects and oral language development (oracy), which are critical for building the background knowledge necessary for comprehension,.
- Test Management over Teaching: Federal mandates like No Child Left Behind (NCLB), which shared the reform movement's emphasis on standardized testing, found zero gains in reading at the fourth- or eighth-grade levels. Instead of better teaching, these policies encouraged "test management" and the triaging of students.
- The Ed-Tech "Grift": Despite massive investment in classroom technology and software, a 2009 U.S. Department of Education study found the effect on student achievement was zero. Furthermore, heavy non-academic screen use has been negatively associated with reading scores.
Ultimately, the sources suggest that twenty-five years of top-down, corporate-style reform has replaced human-centric, relational learning with a data-driven regime that has left a generation increasingly unable to discern propaganda or interrogate a claim,,.
The Common Core "cannibalized" reading instruction by narrowing the focus of the classroom to only those subjects and skills that were directly tested, which effectively crowded out the foundational elements required for true literacy.
According to the sources, this process of instructional cannibalization occurred in several specific ways:
- Crowding Out Non-Tested Content: The focus on narrow, tested subjects led schools to neglect untested subjects like social studies, science, and the arts. These subjects are critical because they build the background knowledge that students need to comprehend complex texts.
- Starving "Oracy" and Discussion: Essential instructional time previously used for reading aloud and rich, extended discussion—known as "oracy"—was systematically squeezed out in favor of test preparation. This dialogic talk between adults and children is what builds the vocabulary and knowledge base that comprehension depends on.
- Prioritizing Compliance Over Reasoning: The pressure to adhere to standardized tests and "pacing guides" encouraged students to "bubble in an answer sheet" rather than engage deeply with a text. This shift focused on technical decoding and test management rather than teaching children how to reason through a text or interrogate a claim.
- Erosion of Critical Thinking: By cannibalizing the breadth of instruction, the reforms left students less capable of evaluating the credibility of information. One study noted that 80% of students could not distinguish sponsored content from news, a result the sources link to 25 years of teaching children to fear a test rather than critically engage with information.
Ultimately, the sources argue that by standardizing instruction to improve test scores, the Common Core eliminated the human-centric, relational learning—such as reading aloud and debating ideas—that actually produces a literate and discerning citizenry.
Testing mandates, specifically those under No Child Left Behind (NCLB), led schools to adopt a "triage" approach that focused heavily on students near the proficiency cutoff, often referred to as "bubble kids",.
The sources describe several consequences of this focus:
- Instructional Triaging: Schools facing sanctions for underperformance responded rationally to the incentives provided by the law by prioritizing students who were just below or at the edge of passing standardized tests.
- Neglect of Other Students: Because schools focused their limited instructional time and resources on these "bubble kids," evidence from at least one state showed that outcomes for the most disadvantaged students—those furthest from the cutoff—actually deteriorated under the pressure.
- "Test Management" vs. Teaching: This focus shifted the goal of education from quality instruction to "test management". Instead of improving overall teaching practices, schools concentrated on narrow strategies designed to nudge students over a specific numerical threshold to ensure the school's survival.
- Instructional Cannibalization: This triage occurred within a broader context where "untested" subjects and critical oral language development were squeezed out of the school day to make room for intensive test preparation,.
Ultimately, the sources argue that attaching a school's survival to a single proficiency number did not "lift the floor" as promised, but instead encouraged a data-driven regime that managed scores rather than educating children,.
Testing mandates over the last twenty-five years have fundamentally altered the teaching profession by transforming teachers from relational educators into "test managers" operating within a data-driven regime.
The impact on the profession can be categorized into several key areas:
Loss of Professional Autonomy and Control
Federal mandates like No Child Left Behind (NCLB) made the entire teaching profession answerable for student outcomes while simultaneously stripping away their control over the conditions that produce those outcomes, such as poverty, class size, and home literacy environments. Teachers have been forced to adhere to "scripted curricula" and rigid "pacing guides" that remove their professional judgment in favor of standardized compliance. The sources suggest there is a direct correlation between this decline in teacher autonomy and the overall decline in student literacy.
Increased Scapegoating and Structural Decline
Through decades of reform, the teacher has been the "constant, reliable object of blame" whenever test scores were low. Rather than addressing the architecture of the education system, reforms treated the person at the front of the room as the primary "variable to fix" through sanctions and evaluation systems. Consequently, researchers describe the teaching profession as being in its "worst structural condition in half a century".
Psychological and Economic Strain
The pressure of testing mandates has led to significant well-being and recruitment issues:
- Mental Health: Teachers report symptoms of burnout, depression, and job stress at multiples of the general working population.
- Compensation Gap: Teachers currently earn roughly $30,000 less than comparably educated adults in other fields while working measurably more hours per week.
- Recruitment: The sources argue that the cycle of lower pay and increased blame makes it increasingly difficult to recruit the next generation of talented adults into the profession.
Shift to "Instructional Triaging"
Testing mandates forced a rational but harmful shift in how teachers allocate their time. To ensure school survival, teachers were encouraged to "triage" their students, focusing instructional energy on "bubble kids"—those hovering just below proficiency cutoffs—while the most disadvantaged students often saw their outcomes deteriorate under the pressure.
Devaluation of Relational Teaching
The focus on standardized data has replaced the "human-centric, relational learning" that defines the profession. Teachers are now often evaluated as "replaceable line workers" based on numerical targets they do not control, rather than being treated as the "load-bearing wall" of the educational enterprise. This shift has squeezed out the time for rich, extended discussion and reading aloud—elements of "oracy" that teachers previously used to build foundational knowledge.
Selected sources: NAEP/Nation's Report Card, 2024 reading assessments (nationsreportcard.gov, nagb.gov); Dee & Jacob, NBER Working Paper No. 15531; Lee (2006), Harvard; Cato Institute policy analyses of NCLB; RAND Corporation evaluation of the Gates Foundation's Intensive Partnerships for Effective Teaching Initiative; NEPC (National Education Policy Center) reporting on Gates Foundation education spending; U.S. Department of Education/Institute of Education Sciences 2009 study of reading and math software; Escueta, Quan, Nickow & Oreopoulos, NBER Working Paper No. 23744 (2017); RAND State of the American Teacher surveys, 2024–2025; Kraft & Lyon, Brown University working paper on teacher morale; Stanford History Education Group, "Evaluating Information: The Cornerstone of Civic Online Reasoning" (2016) and follow-up studies.


























.png)
