School Information System

Four AI assistants score the experts behind the NYT’s school-fix rankings

On Sept. 28, 2026, The New York Times asked 37 education researchers to rate 30 ideas for reversing a decade-long slide in reading and math scores. Their verdict: small-group tutoring and the science of reading work best. School accountability, high-quality curriculum and wider access to advanced classes give the most for the money. Private school vouchers and smaller class sizes cost a lot for little academic gain. Above all, the experts agreed that how well an idea is carried out matters more than which idea is chosen, as Mississippi and Louisiana show.

But who are these experts? We asked four AI assistants (Claude, Gemini, Grok and Perplexity) to profile each panelist and score how well their research positions them to advise on raising K–12 reading and math achievement. The table below shows all four scores side by side.

What stands out

  • They agree on the top. All four put Tom Kane, Matthew Kraft, Thomas Dee and Eric Hanushek at or near the highest score.
  • They agree on the bottom, too. Dominique Baker, whose research focuses on colleges rather than K–12 schools, is scored lowest, or tied for lowest, by all four.
  • Gemini and Perplexity grade generously. Gemini gave 18 of 37 experts a perfect 10; Perplexity gave 22 a perfect 5 of 5. Claude (7.4 average) and Grok (7.9) spread their scores more widely, so their rankings separate the experts more.
  • The biggest disagreements are over Prudence Carter (6–10), Sarah Lubienski (6.5–10), Ilana Horn (6.5–10) and Huriya Jabbar (6.5–10). These are marked “Wide” in the table.

All 37 scores, side by side

Sorted by average score. All scores are out of 10. Perplexity scored out of 5, so its scores are doubled here. “Gap” is the difference between the highest and lowest score; “Wide” marks a gap of 3 or more.

RankExpertClaudeGeminiGrokPerplexity (×2)AverageGap
1Tom Kane9.51010109.90.5
2Matthew Kraft91010109.81
3Thomas Dee9109109.51
4Eric Hanushek9109109.51
5Dan Goldhaber8.5109109.41.5
6Douglas Harris8.5109109.41.5
7Heather Hill8.5109109.41.5
8Morgan Polikoff8.5109109.41.5
9Katharine Strunk8.5109109.41.5
10James Kim8109109.22
11Sean Reardon9108109.22
12Phil Capin7.5109109.12.5
13Marguerite Roza8.5108109.12
14Martin West8.599109.11.5
15Cory Koedel8108109.02
16Beth Schueler8.598108.92
17Lori Taylor7108108.83 Wide
18Sarah Lubienski6.5108108.63.5 Wide
19Ilana Horn6.5107108.43.5 Wide
20Sarah Cohodes89888.21
21Sarah Winchell Lenhoff797108.23 Wide
22Huriya Jabbar6.597108.13.5 Wide
23Brendan Bartanen79888.02
24Anna J. Egalite79888.02
25James Soland710788.03 Wide
26Joseph Cimpian6.59887.92.5
27Sarah Novicoff6.59887.92.5
28Prudence Carter696107.84 Wide
29Susan Dynarski78887.81
30Harry Anthony Patrinos78887.81
31Jonathan Schweig69887.83 Wide
32Patrick Wolf79787.82
33Chris Torres6.59787.62.5
34Jack Schneider68687.02
35Jeremy Singer69767.03 Wide
36Robert Maranto5.58686.92.5
37Dominique Baker5.57565.92
Average of all 377.49.37.99.18.5


grok

claude

gemini

perplexity

———-

The State of Reading Around the World.

Recent literacy history, 1998-

How to read these scores

Each score is an AI assistant’s judgment of how directly a person’s research bears on raising K–12 test scores. It is not a measure of the quality of their scholarship; several lower-scored panelists are leading experts on college access, attendance or equity. The assistants drew on different information and sometimes disagree on basic facts, such as a person’s current job, so treat the scores as a starting point and check individual profiles before relying on any one of them.

Source: “Expert” Analysis & “Students simply can’t read or do math as well as they used to”; scores too!

Share

Fast Lane Literacy™ by sedso

i

Explore teaching tips and learn more about the word i.