AI Learning Evidence

Not guidance. A record of published research on AI-assisted learning, and our assessments of it. The limits

Secondary students with low prior knowledge gain less, or are harmed more, by AI assistance during practice than high-prior-knowledge peers.

Current status
Preliminary signalEvidence unclear · what the nine statuses mean
Why
Two independent weak analyses point the claim's way (retention gains confined to high-prior learners; a positive treatment-by-baseline interaction), while the strongest trial's preregistered heterogeneity analysis — gated for stance because its assessment conditions are not reported — found no detectable ability moderation. An early, internally inconsistent signal.
What would change it
Prespecified prior-knowledge moderation analyses on unassisted outcomes in the larger randomized studies.
Linked evidence
3 links · 2 supports · 1 consistent, but doesn't test the claim
Last updated
Sep 18, 2026

The status is the collection's own judgment on its nine-label scale — not a certainty rating, not a GRADE level, and not advice. Every status change is dated, reasoned, and kept below under History.

01The evidence

The evidence, sorted by what it shows

Supports (2)

Consistent, but doesn't test the claim (1)

02Certainty

Certainty by outcome

A claim can be broken into separate bodies of evidence — one per outcome, split by whether AI was available at assessment and when the outcome was measured. Each body carries four separate judgments: how confident we are (certainty), what the evidence points to (the conclusion), how directly it speaks to this claim (applicability), and who has stood behind the judgment. Confidence and conclusion are never merged into one word.

UA-MOD — Prior-knowledge moderation, unassisted outcomes · AI at assessment: no · immediate post

Field Assembly certainty: Weak (not a GRADE rating — what our scale means) · Conclusion: Inconsistent · Applicability: Partial · AI: two passes agreed

Rated against: differential favoring high-prior learners

Two weak signals that gains concentrate among higher-prior learners; the strongest trial detected no ability moderation

DomainJudgment and reasoning
Risk of biasserious · unprespecified figure-derived subgroup split; working paper with unsettled attrition
Inconsistencyserious · two weak supporting analyses against a preregistered no-detectable-heterogeneity result (gated for stance, informative here)
Indirectnessserious · different subjects, tools, and moderator definitions
Imprecisionvery serious · subgroup analyses of small samples
Reporting and publication biasnot serious · not assessable

Population: secondary students (programming camps; Nigerian secondary schools)
Comparator: same practice/schooling without AI
Assessed by software, two independent passes · search: seed/corpus.yaml (verified inventory, proposal section 7) + data/searches/ · method: EVIDENCE-MODEL.md v2 + fa-certainty-scale v2

2 studies in this body. 1 further record was considered and left out:

03History

How this assessment has changed

  • Preliminary signalSep 18, 2026 · Evidence unclear

    initial assessed status (Q-009, owner-accepted IN-018)

  • Not yet assessedSep 18, 2026 · Evidence unclear

    initial curated status