AI Launches Two-Year School Trial of Khanmigo Tutoring
| Source: HN | Original article
A two-year cluster randomized trial in 18 Tennessee middle schools is testing Khanmigo AI tutoring, providing early large‑scale evidence of its impact.
A two‑year cluster‑randomised trial is now under way to test the impact of Khan Academy’s AI tutor, Khanmigo, on middle‑school mathematics. The experiment involves 18 Tennessee middle schools, where students assigned to the treatment arm use Khanmigo during their daily remedial math sessions. Unlike a simple answer‑generator, the system is configured to “coach” learners, prompting them to work through problems rather than handing over solutions.
The study provides some of the first large‑scale, school‑level evidence on generative‑AI tutoring. Proponents have long argued that AI could deliver personalised instruction at scale, but empirical data from real classrooms have been scarce. By embedding the tutor in existing remedial schedules and randomising at the school level, the trial aims to isolate Khanmigo’s contribution to learning outcomes while controlling for typical classroom variables.
Why this matters is twofold. First, the research follows earlier work that showed AI‑driven tutoring can outperform traditional active‑learning formats in controlled settings – a finding we reported on 1 October. Demonstrating comparable gains in a natural school environment would bolster the case for wider adoption and could shape district‑level budgeting, especially as policymakers grapple with the rapid rollout of AI tools in education. Second, the “coach‑rather‑answer” design addresses growing concerns about over‑reliance on AI for direct answers, a point highlighted in recent coverage of AI risks to children.
Watch for the trial’s first‑year results, which should reveal whether Khanmigo improves math proficiency, engagement or equity metrics. Stakeholders will also be keen on how the findings influence state education agencies, funding decisions and future research on AI‑assisted instruction. If the evidence proves positive, the study could become a benchmark for evaluating other generative‑AI tutoring platforms in schools worldwide.
Sources
Back to AIPULSEN