← What we believe Evidence library

Home · Sources · AI-011

The Relative Effectiveness of Human Tutoring, Intelligent Tutoring Systems, and Other Tutoring Systems

Added to the Kunskapsbank on 2026-09-07 but not on the old published site.

Authors: VanLehn, K. · Year: 2011 · Tier: meta · DOI/URL: Educational Psychologist, 46(4), 197–221. https://doi.org/10.1080/00461520.2011.611369

Across ages AI and learning Foundations of optimal learning

Links to claims

Supports (1)

ClaimBasis
K28 (candidate)Note carries the bank kill sentence behind K28 (group K28); no change in review

Review state: first-pass draft, pending human review.

Source note (Swedish, verbatim from the Kunskapsbank)

Syntes: VanLehn (2011) — human vs ITS vs CAI (AI-011)

Källa: VanLehn, K. (2011). The relative effectiveness of human tutoring, intelligent tutoring systems, and other tutoring systems. Educational Psychologist, 46(4), 197–221. https://doi.org/10.1080/00461520.2011.611369
Original: 01-originaldokument/vanlehn-2011-relative-effectiveness-tutoring/ (ABSTRACT; KÖP fulltext)
ID: AI-011
Tier: meta
Closed-book follow-up: delvis (post-test i underliggande studier; pre-GenAI)
HE vs K–12: övergripande
Evidensstyrka: hög som klassisk ITS-review; begränsad av studieurval före GenAI och STEM-tyngd.

Metadata

Finding (+ comparator)

Vs no tutoring (samma innehåll utan tutor): adult human tutoring d≈0,79; step-based ITS d≈0,76; answer-based CAI lägre (~0,3-klassen i äldre tro). Human ≈ ITS — inte Blooms mytiska 2σ. Comparator: no tutoring / answer-based / step-based / substep / human.

Mechanism

Finare interaktionsgranularitet (stegvis diagnostik + just-in-time scaffolding) ökar lärande upp till step-based; därefter interaction plateau — substep/human ger inte tillförlitligt större vinst. Pedagogik i mjukvaran (modell av elevfel) ≠ skärmen i sig.

Limits

Pre-2011; mest STEM; många små experiment; GenAI-chatbots ingår inte. Effektstorlekar aggregerar olika outcome-mått.

This claim dies if …

Storskaliga K–12-metas visar att designad step-based ITS systematiskt underpresterar human tutoring med >0,3 d och att answer-giving GenAI matchar ITS på closed-book utan guardrails.

Swedish implication

Stop likställa “AI/skärm” med ITS-evidens; start kravspec: stegvis tutor som håller inne facit (jfr Bastani Tutor, AI-003); measure closed-book efter digital övning — inte bara practice-score.

Links

Metadata as recorded

ID
AI-011
Tier
meta
Closed-book follow-up
delvis (post-test utan tutor i underliggande studier; pre-GenAI ITS)
HE vs K–12
övergripande (K–12 + HE; STEM-tungt)
Titel
The Relative Effectiveness of Human Tutoring, Intelligent Tutoring Systems, and Other Tutoring Systems
Författare
VanLehn, K.
År
2011
Källa/DOI
Educational Psychologist, 46(4), 197–221. https://doi.org/10.1080/00461520.2011.611369
URL
https://doi.org/10.1080/00461520.2011.611369
Typ
meta-analytisk review (experimentella jämförelser human/ITS/CAI/no tutoring)
Åldersgrupp
övergripande
Teman
ai-och-larande; optimalt-larande-grunder; uppmarksamhet-och-minne
Sparade fil
ABSTRACT.md (KÖP fulltext Taylor & Francis)
Hämtad
2026-09-07
Paywall
ja (T&F)
Obs
Steelman-ankare för designad digital tutor — step-based ITS ≈ human tutoring (d≈0,76 vs 0,79 vs no tutoring).

Original files held in the project (not published): ABSTRACT.md