The tutor just matched the expert at one nine-hundredth of the price. a studentbench study of 2,383…
the tutor just matched the expert at one nine-hundredth of the price. a studentbench study of 2,383 students found ai tutors produced gre gains statistically equal to human tutors, at about 918 times lower cost per percentage point, per oct 2 results.
the premium tutor is now a pricing problem.
Context
A Sep 23 preprint, StudentBench, reports AI tutoring produced GRE learning gains statistically equivalent to expert human tutoring. It covered 2,383 adult participants and 2,469 usable sessions, was funded by Handshake AI Research, and used a one-hour session between a pre-test and a matched post-test.
The pooled AI-minus-human adjusted gain was -0.58 percentage points. Ground Truth stresses the result is narrow and does not show chatbots can replace schooling.
The 2,383 participants and the GRE equivalence match the sources. The preprint is dated Sep 23, not Oct 2. The 918 times figure appears as 'up to 918 times lower cost'; the exact per-point basis was not seen, so it is unsupported here, not refuted.
'The premium tutor is now a pricing problem' is the author's opinion.
Related work
- Wins Solutions: AI tutoring matches human tutoring gains ↗Study summary.
- Ground Truth: StudentBench finds equivalent immediate gains ↗Limits of the result.
- Reuters, Oct 5, 2026 ↗Other regulation news.
Watch next
- Read the preprint's cost method.
Sources
- Wins Solutions, Sep 25, 2026winssolutions.org
- Ground Truth, Sep 25, 2026groundtruth.day
- Alibaba Cloud Community, Sep 25, 2026alibabacloud.com
Provenance
The note above is reproduced unedited from the original post, first published on Threads on 10 October 2026 at 14:17 IST. Sources are the papers and datasets the note draws on.
View the original post ↗Embed this note
More notes
The air is now being asked to keep its own ledger
the air is now being asked to keep its own ledger: ecmwf’s aifs compo becomes the first ai model to forecast atmospheric composition globally every three hours, cleanair simulates 365 days of pm2.5 over china in ten seconds, and a unified framework maps six pollutants at one kilometer across the whole country. the air now files its own composition report.
read the note →The current is now being asked to draw its own map
the current is now being asked to draw its own map: china’s langya 2.0 predicts six ocean phenomena including internal waves and mesoscale eddies, a deep net called wenhai resolves eddies globally with air sea flux formulas built in, and scripps infers surface currents from the way temperature patterns deform in satellite images. the ocean now files its own circulation report.
read the note →The soil is now being asked to report its own carbon
the soil is now being asked to report its own carbon: a nix color sensor paired with generative data augmentation predicts soil organic carbon without a lab, random forest drives 74 percent of soil health mapping studies, and sentinel 2 tracks five year carbon change across france and italy from 922 samples. the dirt now files its own carbon account.
read the note →