Public School Quality Rankings: A Policy Guide for Public Administrators

An analysis of state education rankings, limitations, and how administrators can drive improvement.

By Max SheltonReviewed by PAP Editoral TeamUpdated July 24, 202623 min read

What you’ll learn in this article…

  • New York tops one 2026 ranking while Massachusetts leads another.
  • Popular metrics like class size poorly predict actual student achievement.
  • Growth models and long-term outcomes outperform raw test scores for policy.

In 2026, two major rankings of state public schools produced starkly different conclusions: World Population Review placed New York at the top and Arizona at the bottom, while WalletHub crowned Massachusetts the best and New Mexico the worst. For public administrators, these rankings are tempting shorthand, but they reward socioeconomic privilege, not school effectiveness. State-level averages obscure dramatic within-state disparities, making it easy to overlook the schools that need the most support. As Stanford professor Sean Reardon notes, such ratings "are relatively arbitrary and reflect the influence of many factors outside of schools' control."1

How State School Rankings Are Built: A 2026 Methodology Deep-Dive

State school rankings have become a fixture of the education policy conversation, yet few policymakers or administrators pause to examine how those rankings are constructed before citing them in legislative debates or strategic plans. Understanding the methodology behind two of the most widely referenced 2026 rankings reveals both their utility and their limitations.

WalletHub's 32-Metric Scoring System

WalletHub's 2026 ranking, which placed Massachusetts at the top followed by Connecticut, New Jersey, New Hampshire, and Wisconsin, evaluates all 50 states and the District of Columbia across 32 weighted metrics. These metrics span categories such as test scores, pupil-teacher ratios, dropout rates, and even the number of school shootings. The breadth of inputs is noteworthy: by combining academic performance indicators with safety and resource measures, WalletHub produces a composite score that captures more than classroom outcomes alone. For administrators seeking to replicate or examine that analysis, WalletHub publishes its full methodology and data sources alongside the ranking, making it a reasonable starting point for state-level benchmarking.1

At the bottom of WalletHub's list sit New Mexico, Alaska, Oklahoma, Oregon, and West Virginia, states that share structural challenges such as lower per-pupil spending, higher poverty rates, and geographic barriers. These patterns underscore that rankings reflect policy choices and demographics, not just school quality.

World Population Review's Four-Category Framework

World Population Review takes a different structural approach. Its 2026 ranking named New York the top state, with Connecticut, Massachusetts, New Jersey, and Illinois completing the top five, organizing data into four categories: K-12 performance, school funding and resources, higher education quality, and school safety.1 By folding higher education into the equation, this ranking captures a broader educational ecosystem, though it also introduces variables that may have little to do with the K-12 systems most state education agencies directly govern. Arizona, Alaska, Nevada, Oklahoma, and New Mexico fell to the bottom. The divergence from WalletHub's top five -- Massachusetts leads one list but sits third on the other, and New York appears only once -- illustrates how weighting, category inclusion, and data selection produce different results.

Where to Cross-Check and Validate

Public administrators should treat any single ranking as a conversation starter. Cross-reference with: - NAEP: standardized K-12 performance data comparable across states. - NCES: school safety, enrollment, and expenditure tables to verify ranking inputs. - State department of education websites: district-level data to expose intra-state disparities. - NEA and AFT analyses: practitioner-oriented perspectives on ranking metrics.

Why Methodology Matters for Policy

For MPA and MPP professionals, the core takeaway is that methodology is not a technical footnote; it is a choice rooted in public policy making. Deciding to weight safety incidents equally with test performance, or to include higher education alongside K-12 metrics, embeds normative assumptions about what 'quality' means. Administrators who rely on rankings without interrogating those assumptions risk designing interventions that optimize for the wrong outcomes. Before using a ranking, ask: what exactly is being measured, and whose definition of success does it serve?

Why Rankings Often Mislead: What the Experts Are Saying

State-level school rankings often treat education quality as a simple, single-number score, but that simplicity masks deep methodological flaws. When a ranking claims Arizona has the “worst” schools while New York has the “best,” it is often measuring wealth and demographics far more than instructional effectiveness.1 For public administrators, these rankings can become dangerous shortcuts, making it vital that rethinking MPP and MPA curricula focuses on analytical rigor.

The Demographic Confound

Stanford University professor Sean Reardon has been blunt about the limits of these systems, calling them “relatively arbitrary and reflect the influence of many factors outside of schools’ control.” His research shows that roughly two-thirds of the variation in test scores across districts can be explained by student background, family income, parental education, and neighborhood resources, while less than one-third is attributable to what happens inside schools.2 When raw test scores or graduation rates dominate a ranking, affluent states inevitably float upward, and high-poverty states sink, regardless of how much growth their students achieve. The ranking ends up telling you more about the ZIP codes of the students than about the quality of the educators working with them, a problem that fixing bias in state K-12 education rankings aims to solve.

Inputs Are Not Outcomes

Compounding the problem, many rankings lean heavily on input metrics that have little consistent relationship with student learning. Paul Bruno, an assistant professor at the University of Illinois Urbana-Champaign, warns that measures such as student-teacher ratios and teacher qualifications are not reliable proxies for school effectiveness. Research bears this out: holding a master’s degree, for example, shows little or no consistent impact on student achievement, yet credential counts frequently appear in ranking formulas. Similarly, per-pupil funding data, while easy to obtain, often reflects regional cost differences rather than strategic investment in instruction. A school can have small classes and highly credentialed staff but still deliver mediocre learning gains if those resources are not deployed effectively. When rankings reward states for these input metrics, they create a false picture of quality and can incentivize administrators to chase credentials or class-size targets rather than focus on deeper instructional improvement in areas like career and technical education policy.

The Equity Cost

Georgetown University research professor Marguerite Roza argues that failing to adjust for demographics like poverty and race fundamentally distorts comparisons. Rankings that ignore those factors “penalize schools serving disadvantaged students,” as ratings from U.S. News and GreatSchools have been shown to do,1 steering families away from exactly the schools that need community support. Over time, these seeming “failing” labels can reinforce housing segregation and drain enrollment from urban and rural districts.2 For public administrators, the message is clear: rewarding states for raw achievement without accounting for the challenges their students face punishes the communities that need the most investment and makes it politically harder to direct resources toward equity.

A state that looks “low-performing” on a simplistic ranking may well be producing larger academic gains for its most vulnerable students than a top-ranked state where privileged students are merely maintaining an advantage they arrived with. Until rankings routinely incorporate growth measures and demographic adjustments, they remain more a reflection of socioeconomic sorting than of educational excellence.

Ranking systems are relatively arbitrary and reflect the influence of many factors outside of schools' control.

What Top-Performing States Get Right, and Why It’s Complicated

If a state ranks among the best for public schools, does that automatically mean all its students are thriving? The reality is far more complicated. Even in states celebrated for their education systems, deep disparities often separate wealthy suburbs from underfunded urban and rural districts. For public administrators, the lesson is not to reject rankings outright but to know how to go beyond the headline numbers and uncover the true state of educational equity.

The Aggregate Trap: How Averages Mask Inequality

State rankings like those from WalletHub or World Population Review depend on aggregated metrics: average test scores, overall graduation rates, pupil-teacher ratios. While these provide a snapshot, they smooth over critical variation. In Massachusetts, for example, the state’s strong performance is largely driven by its affluent suburban districts. Urban centers like Boston or Springfield often post significantly lower achievement levels, yet those gaps are invisible in a single state rank. Public administrators need to ask: which students are being served well, and which are being left behind?

Your Toolkit for Independent Analysis

To move past rankings, start with the National Center for Education Statistics (NCES). Its Common Core of Data tracks enrollment, staffing, and fiscal data by district and school. The NAEP Data Explorer lets you compare state performance by demographic subgroup, revealing proficiency gaps that averages conceal. For even more granular insight, Every Student Succeeds Act (ESSA) required state report cards, accessible on each state's department of education website, detail per-pupil expenditures, chronic absenteeism, and college readiness indicators.

The U.S. Department of Education’s Civil Rights Data Collection is another essential resource, offering data on course access, discipline disparities, and advanced placement participation. These datasets empower administrators to identify not just academic outcomes but also opportunity gaps that drive them.

Contextualizing with Workforce and Funding Data

Effective public administration also requires understanding the inputs. The Bureau of Labor Statistics (BLS) provides occupational data on teacher salaries, employment trends, and projected demand, key inputs for public sector pay transparency and diagnosing staffing inequities across districts. Meanwhile, the Census Bureau’s Annual Survey of School System Finances sheds light on how funding varies within a state. By marrying this financial data with student outcomes, you can pinpoint where investment is insufficient or misaligned.

Engaging Professional Networks and Best Practices

Professional associations like the National Education Association (NEA) or the Association for Supervision and Curriculum Development (ASCD) publish policy analyses, host conferences, and maintain communities of practice. These platforms help administrators learn how peers in similar communities are tackling disparities and function as a vital channel for public administration education.

From Data to Action

Ultimately, the goal is to design policies that lift all students, not just the averages. By triangulating state rankings with local, disaggregated data, public administrators can craft targeted interventions that advance equity and ensure every school truly performs at a high level.

Equity in Education: The Missing Metric

The temptation is to crown states with the highest average test scores as the best school systems. A more rigorous approach asks which states close the widest gaps between historically advantaged and disadvantaged students.

Equity vs. Equality: A Foundational Distinction

In education policy, equality means distributing the same resources to every student. Equity requires directing resources where they are needed most to overcome systemic barriers linked to race, income, language, and disability. The Every Student Succeeds Act (ESSA), a federal-state partnership, requires states to report achievement data for specific subgroups, including racial/ethnic groups, economically disadvantaged students, English learners, and students with disabilities, but many popular state rankings treat those subgroups as an afterthought rather than a central yardstick.5

The Equity Metrics That Go Unnoticed

Standard league tables prioritize average proficiency rates and graduation numbers. Equity measures, by contrast, track gap sizes between the highest- and lowest-performing subgroups, the rate at which gaps close over time, and whether school funding is distributed progressively or regressively. The Education Law Center's School Funding Fairness report, for example, scores states across three dimensions: funding level, funding distribution, and fiscal effort.1 The OECD's Education Equity Dashboard includes 35 indicators spanning learning outcomes, social mobility, and resource allocation.2 UNESCO's Equitable Financing Index uses three measurement dimensions to gauge how well a nation's financing system supports disadvantaged learners. These frameworks rarely appear in simplified rankings, yet they capture the dimensions that public administrators care about most: whether a system serves all children, not just those born into advantage.

When Bottom-Ranked States Outperform on Equity

The disconnect between overall rank and equity performance becomes stark when you look beneath the surface. New Mexico, routinely placed among the bottom five states in overall school quality rankings, shows stronger-than-expected results on certain equity-adjusted indicators, such as narrowing the graduation rate gap for Hispanic and Native American students. Marguerite Roza, a research professor at Georgetown University, emphasizes that student demographics like poverty and race heavily influence outcomes and must be adjusted for in any meaningful comparison. When experts control for socioeconomic status, many lower-ranked states achieve greater relative progress than their absolute numbers suggest. Sean Reardon of Stanford University captures the risk of raw rankings bluntly: they "are relatively arbitrary and reflect the influence of many factors outside of schools' control."

  • Achievement gap size: The distance in test scores or graduation rates between, for instance, White and Black students or low-income and higher-income peers.
  • Gap closure rate: How quickly those differences are shrinking year over year, a truer measure of system improvement.
  • Funding fairness: Whether per-pupil spending reaches high-need districts more generously than wealthy ones, a critical equity lever.
  • Subgroup accountability: ESSA requires states to set improvement targets for each subgroup, but those results rarely make headlines in state rankings.5

Making Equity Primary, Not Supplemental

For public policy professionals and MPA graduates designing accountability systems, the lesson is clear: equity indicators must sit at the core of evaluation frameworks, not at the margin. The Pell Institute's 2026 Equity Indicators report articulates the principle of "differentiated support, not equal resource distribution,"4 meaning that simply equal inputs will perpetuate gaps. A state that raises overall test scores while leaving its most vulnerable students behind has not succeeded by any equity-informed standard. When administrators advocate for dashboards that equally weight gap closure, subgroup growth, and funding fairness alongside traditional proficiency metrics, they shift the conversation from a beauty contest of affluent suburbs to a genuine assessment of whether public education expands opportunity for every child.

Metrics like student-teacher ratios and teacher qualifications appear in nearly every state school ranking, yet research by Paul Bruno of the University of Illinois Urbana-Champaign shows these inputs explain surprisingly little variation in student achievement once socioeconomic factors are accounted for. Rankings built on them may reward wealth, not effective teaching.

Policy Lessons for Public Administrators: Moving Beyond Rankings

State rankings provide a seductive shortcut, but public administrators know that improving school quality demands a far more granular strategy that targets resources where they are needed most. The path forward involves moving from simplistic comparisons to concrete policy actions that rewire funding, strengthen the educator workforce, expand early learning, and redesign accountability around growth and equity. The following strategies draw from real state-level reforms already showing promising results.

Implementing Weighted Student Funding

Progressive funding formulas allocate additional dollars for students facing poverty, English language barriers, or disabilities, rather than treating every pupil as costing the same. Massachusetts' Student Opportunity Act, passed in 2019, committed $1.5 billion in new annual funding by phasing in increases over seven years.1 In the first three years alone, low-income districts saw per-pupil funding rise by roughly 30 percent,2 with supplements ranging from $4,600 to $8,800 for each economically disadvantaged student.3 This shift empowered districts to make targeted investments: Brockton doubled its full-day pre-K classrooms3, New Bedford hired additional counselors and nurses and modernized aging buildings4, and Revere expanded summer learning, credit recovery, and extended learning time2. Each district was required to submit a public, three-year evidence-based plan5, ensuring that new dollars were tied to measurable strategies rather than absorbed into general budgets. The hybrid model preserved local control while aligning spending with state equity goals, a blueprint other states can adapt.

Investing in Teacher Quality

Money alone won't close gaps if classrooms are not staffed with skilled, supported educators. Tennessee's Grow Your Own initiative addresses this by creating pipelines from local communities into teaching careers, pairing candidates with mentor teachers and subsidizing certification costs. Administrators can pair such recruitment efforts with high-quality induction programs and leadership pathways that keep experienced teachers in high-need schools. Massachusetts complemented its funding overhaul with CURATE, an initiative that rates instructional materials for alignment to standards2, reducing the curriculum variability that often undermines even well-funded classrooms. When districts use financial flexibility to hire and retain staff, as Holyoke and Worcester did under the Student Opportunity Act3, the combination of stable teams and coherent materials lifts instruction more reliably than across-the-board class size reductions.

Expanding Early Childhood Education

The long-term payoff of high-quality pre-K is among the strongest findings in education research. States that view early learning as part of the K-12 system, not a separate add-on, see compounding benefits in reading proficiency and grade-level readiness. Under the Massachusetts act, several districts channeled new funds into full-day, aligned pre-K programs2, creating a seamless pipeline into elementary school. Structural commitments matter more than boutique pilots; embedding pre-K in the school funding formula guards against the year-to-year grant cycles that disrupt continuity.

Redesigning Accountability for Growth and Equity

Finally, public administrators must replace accountability frameworks that simply report proficiency rates with metrics that track student growth over time and reveal performance for underserved subgroups. Proficiency-based rankings reward districts serving affluent families and mask progress in high-poverty schools. A growth-oriented system, by contrast, asks whether each student made a year's worth of academic gain, regardless of starting point. Massachusetts' public outcome tracking linked to Student Opportunity Act plans is an early example of tying resource accountability to growth and equity, not just absolute scores.5 When combined with the expert critiques earlier in this article, the message is clear: administrators must advocate for evaluation systems that measure what schools actually control and that spotlight gaps so resources can be redirected to the students who need them most.

Questions to Ask Yourself

Unadjusted comparisons often just measure community wealth rather than school effectiveness, which can misdirect funding, staffing, and intervention decisions toward or away from the wrong districts.

Point-in-time proficiency rewards schools with affluent, well-prepared students, while growth measures reveal which schools are actually adding value, regardless of where students start.

If a different methodology would flip your conclusions about which schools need support or recognition, the ranking itself may be too fragile a basis for policy.

Experts caution that resource metrics correlate weakly with effectiveness, so policies built solely on spending comparisons may miss what truly drives student achievement.

Toward Better Accountability: Value-Added and Growth Models

Status accountability versus growth accountability: the first asks whether students hit a proficiency bar in a given year; the second asks how much a school moved them from where they started. Rankings almost universally rely on the first. Serious accountability systems increasingly rely on the second.

What Value-Added Measurement Actually Does

Value-added measurement (VAM) attempts to isolate a school's or teacher's contribution to student learning by comparing a student's observed growth against the growth of statistically similar peers. Instead of penalizing a school for enrolling students who arrive below grade level, or rewarding one whose students arrive already proficient, VAM asks a narrower question: given where these students started, did they progress more or less than expected?

The appeal for public administrators is straightforward. A school serving high-poverty neighborhoods can post low absolute scores and still be one of the most effective growth engines in the state. A wealthy suburban school can post high proficiency rates while its students actually progress at below-average rates. Rankings miss both cases. Growth models catch them.

Tennessee's TVAAS: The Pioneer

The Tennessee Value-Added Assessment System (TVAAS) is the longest-running example in the country and has shaped federal thinking on growth-based accountability. TVAAS uses a mixed-model (random-effects) multivariate longitudinal analysis to estimate teacher, school, and district effects from student performance on TCAP and end-of-course exams.12 Each student serves as their own control across years, and effectiveness is scored on a 1 to 5 scale against statewide peers with similar prior achievement.3

The results inform multiple decisions. TVAAS is one of three key components in Tennessee's teacher evaluation framework3, it informs school and district accountability classifications1, and it is also used to evaluate the effectiveness of graduates from educator preparation programs4. The policy signal has been unambiguous: Tennessee shifted state accountability away from static proficiency and toward growth, and other states followed.5

California's Multi-Measure Dashboard

California took a different path with its School Dashboard, which pairs academic growth with a broader set of indicators: chronic absenteeism, suspension rates, English learner progress, graduation rates, and college and career readiness. Each indicator is reported by student subgroup, so a district cannot hide disparities behind a strong average.

Why This Matters for Equity

Both approaches share a design principle that administrators should note: they reward improvement anywhere on the distribution, not just movement across a single proficiency cutoff. Under status models, schools have a perverse incentive to focus on "bubble" students near the cutoff while ignoring both the furthest-behind and the already-proficient. Growth and multi-measure systems dismantle that incentive, which is why they align far better with the equity goals most public agencies claim to pursue.

Beyond Test Scores: Long-Term Student Outcomes as the Real Metric

Education accountability is shifting away from snapshot assessments and toward a harder, more consequential question: what happens to students after they leave high school? If the purpose of K-12 schooling is to prepare young people for college completion, career readiness, and civic participation, then test scores at age 17 capture only one early signal in a much longer trajectory. Public administrators and policymakers who rely on conventional state rankings risk optimizing for the wrong endpoint.

Why Postsecondary Outcomes Matter More

A state where students post strong reading and math scores but fail to persist through college or enter the workforce at competitive wages has not succeeded in any meaningful sense. Postsecondary enrollment, completion, and early-career earnings offer a more direct measure of whether K-12 systems are doing their job. The National Student Clearinghouse, which covers roughly 95 percent of postsecondary enrollments nationwide,5 now makes it possible to track students from high school graduation into and through higher education at scale. As of spring 2026, total postsecondary enrollment nationally stood at approximately 18.6 million students, with about 15.5 million at the undergraduate level,4 providing a massive research dataset.

The Complicated Picture Across States

State-level data complicate the tidy narratives that rankings produce. Florida, for example, reported a postsecondary enrollment rate within one year of high school graduation of about 54 percent, paired with a postsecondary attainment rate of roughly 55 percent, according to the Florida College Access Network. North Carolina's postsecondary enrollment rate among 18-to-24-year-olds sits around 40 percent,2 a figure the state's myFutureNC initiative aims to raise. Texas illustrates another wrinkle: among students who started at four-year institutions, the six-year completion rate was about 63 percent, but for those who began at two-year colleges, completion dropped to approximately 17 percent.3 These figures, drawn from an earlier cohort, highlight how different starting points within the same state yield vastly different results.

Some Midwest states that land in the middle of conventional K-12 test-score rankings show relatively strong postsecondary persistence and workforce outcomes, suggesting that school systems in those states may be doing more effective work than their ranking position implies. Conversely, states that score well on standardized assessments sometimes see significant attrition between high school graduation and college completion, a gap that rankings built on test scores alone will never reveal.

Building Systems That Track the Full Arc

California's Cradle-to-Career data system offers one model worth studying. By linking K-12 records to postsecondary enrollment, degree completion, and wage data, the system allows policymakers to evaluate whether a school or district is actually launching students toward economic stability, not just shepherding them through graduation. This kind of longitudinal accountability infrastructure gives administrators the ability to identify which interventions, curricula, and support structures produce durable gains rather than test-score bumps that fade after a few years.

For MPA and MPP professionals working in education policy, the practical takeaway is clear:

  • Advocate for linked data systems that follow students from kindergarten through early career, connecting K-12, postsecondary, and workforce databases.
  • Interpret rankings cautiously when they rely solely on point-in-time test scores or input measures like spending and class size.
  • Prioritize completion and earnings metrics alongside enrollment rates, since access without completion is an incomplete victory.
  • Account for institutional variation within states, because aggregate figures can mask enormous differences between community college and four-year outcomes, as the Texas data illustrates.

Ultimately, a state's education quality should be judged by where its students end up in their twenties and thirties, not solely by how they perform on a standardized exam at seventeen. Rankings that ignore this longer arc may look authoritative, but they measure the wrong finish line.

Recent News

Recent Articles