Explore the evidence
32 of 88 decisions · clear everything
Early gains fade by default — halving every 12–18 months, faster when bigger. What persists is trajectory (placement, graduation), not ability.
strong supportconf: highgc: lowFadeout — why early gains disappear, and what actually persists · early-childhood · ages 4–18 · structure
Timed retrieval practice builds arithmetic automaticity — beating identical untimed tutoring head-to-head — and the anti-timed-test harm claims have no causal evidence.
strong supportconf: highgc: lowMath-fact fluency and timed practice (and the anti-timed-test claims) · math · ages 5–12
Testing yourself beats rereading — robust in real classrooms; honest durable size ~0.1–0.3 SD, biggest after delay, thinnest on transfer.
strong supportconf: highgc: lowRetrieval practice (the testing effect) · practice · ages 8–18 · method
Spacing beats massing at equal total time — the most robust finding in learning science. Space repetitions at ~10–20% of how long you need to remember.
strong supportconf: highgc: lowSpaced (distributed) practice · practice · ages 6–18 · method
Systematic phonics beats whole language for word reading — largest in K-1 and for at-risk readers; near-null for comprehension and past grade 3.
strong supportconf: highgc: lowSystematic/explicit phonics vs whole language and balanced literacy · reading · ages 4–9
Physical capacity is about as heritable as cognitive ability (~60%, no shared environment). Whether trainability is a stable trait remains unproven — the skeptics are winning.
strong supportconf: highgc: lowTalent and trainability — what is heritable, and what that does not license · talent · ages 4–18 · input
Teacher quality is the largest within-school lever — 1 SD of teacher ≈ 0.10–0.15 SD/yr, worth more than ten fewer students — and credentials predict none of it. Select; don't workshop.
strong supportconf: highgc: lowTeacher quality — selection over credentials and workshops · teachers · ages 4–18 · structure
High-dosage tutoring is education's most reliable lever: ~0.29 SD in trials, ~0.2 well-scaled — not Bloom's 2σ. Groups of 3–4 work; 1:1 is unnecessary.
strong supportconf: highgc: lowTutoring — the honest effect, the Bloom 2-sigma myth, and what survives scale · tutoring · ages 5–18 · method
Group within classes or across grades by subject: modest, nearly free wins. Whole-school streaming does nothing, and early between-school tracking harms the bottom.
moderate supportconf: highgc: lowAbility grouping and tracking — four practices, four verdicts · grouping · ages 5–18 · structure
Accelerate ready kids: they keep pace with older classmates, bank a year, and show no social-emotional harm at 50. The gifted label itself does nothing; the content does.
moderate supportconf: highgc: mediumAcceleration and gifted programs — the label vs the content · grouping · ages 5–18 · structure
Minimal-guidance math trailed every rival in the one multi-curriculum RCT; explicit instruction for strugglers is math's most replicated result.
moderate supportconf: highgc: lowExplicit vs reform/constructivist math instruction · math · ages 5–18
Word-problem solving is its own skill: computation fluency doesn't produce it. Teaching problem schemas explicitly does (~0.25–0.45 SD on trained content).
moderate supportconf: highgc: lowWord problems, schema instruction, and the conceptual-vs-procedural question · math · ages 6–14
Later bells buy adolescents 40+ measured minutes of sleep — the best-identified positive effect in the archive; the achievement payoff is real but far smaller.
strong supportconf: mediumgc: lowDo later school start times increase adolescent sleep, and does that raise achievement? · health · ages 11–18 · input
Yes, on their own terms: lottery-assigned museum and theatre trips move blind-rated analysis of art, plot knowledge and tolerance by 0.08-0.18 SD weeks later, and a film of the same play moves nothing.
moderate supportconf: mediumgc: lowAre museum and theatre trips worth the day out of school? · arts · ages 8–18
Comprehension is knowledge — but vocabulary teaching moves standardized comprehension only d≈0.10. The big content-knowledge bet (Core Knowledge lottery, 0.24) is real and unreplicated.
moderate supportconf: mediumgc: lowBackground knowledge and vocabulary as drivers of reading comprehension · reading · ages 5–14
Selective CTE high schools raise male graduation 8-10pp and early-career earnings 17-35% on lottery and cutoff designs; test scores, degrees, and every outcome for girls are flat.
moderate supportconf: mediumgc: lowDo career and technical education tracks help students — and which students? · practical · ages 14–18
Immersion costs nothing in English and buys a little — a lottery puts English reading 0.13-0.22 SD ahead by grades 5 and 8 — but no lottery has ever measured how much of the second language students actually learn.
moderate supportconf: mediumgc: lowDoes dual-language immersion work, and is the gain in the second language or in English? · foreign-language · ages 4–18
Teaching handwriting works and transfers: freeing the hand frees composition, with gains still present at six months — even taught in groups of three.
moderate supportconf: mediumgc: mediumDoes handwriting and transcription instruction matter, and should young children type instead? · writing · ages 5–12
Sentence combining improves writing where grammar teaching fails — the same meta-analyses score them +0.50 and −0.32.
moderate supportconf: mediumgc: lowDoes sentence combining improve writing? · writing · ages 5–18
Strategy instruction is writing's best-supported method — at roughly a fifth of its advertised size once measures are independent (d≈0.8 → ~0.15).
moderate supportconf: mediumgc: lowDoes strategy instruction (SRSD, the writing process) improve writing? · writing · ages 5–18
Yes — sight-reading, performance and aural skill move about half a standard deviation under instruction — but the evidence is quasi-experimental and thinner than the transfer literature built on top of it.
moderate supportconf: mediumgc: mediumDoes teaching music produce musical skill? · music · ages 4–18
Teaching coding teaches coding: the one clean school RCT gives g=0.47-0.68. Which approach you pick barely matters, and almost every effect size rests on an instrument the developers built.
moderate supportconf: mediumgc: lowDoes teaching programming actually teach programming — and does the method matter? · computer-science · ages 5–18
Formal spelling instruction works (ES 0.54) and more formal beats less formal — explicit wins again, and it transfers beyond spelling itself.
moderate supportconf: mediumgc: lowDoes teaching spelling work, and what does it transfer to? · writing · ages 5–18
Unassisted discovery loses to explicit teaching; well-scaffolded guided discovery beats both. The operative variable is guidance, not ideology.
moderate supportconf: mediumgc: lowExplicit instruction vs discovery/inquiry — how much guidance? · instruction-style · ages 4–18 · method
Mixing problem types helps where confusion is the enemy — discriminating similar categories, mixed math practice — and is useless or worse for facts and prose.
moderate supportconf: mediumgc: lowInterleaving (mixing problem types vs blocking) · practice · ages 10–18 · method
Manipulatives help when bland, guided, and aged ~7–11 — a guided-representation effect, not 'hands-on learning.' Rich, toy-like objects hurt transfer.
moderate supportconf: mediumgc: lowManipulatives and concrete-representational-abstract (CRA) sequences · math · ages 5–14
Phonemic awareness transfers to reading only with letters attached: speech-only training peaks near 10 hours and moves reading d=0.19; with print, 0.66.
moderate supportconf: mediumgc: lowPhonemic awareness training (and whether it needs letters) · reading · ages 4–7
Diagnose what a child already knows: teachers cut 40–50% of curriculum for high-ability children and achievement ROSE. The payoff is skipping, not monitoring.
moderate supportconf: mediumgc: mediumPlacement and mastery diagnosis — deciding what to teach next from evidence of current skill · assessment · ages 5–18 · structure
Guided oral reading works, but the ingredient is volume, not repetition: at equal exposure, re-reading has no edge over wide reading.
moderate supportconf: mediumgc: lowReading fluency — guided repeated oral reading, and whether repetition matters · reading · ages 6–12
Content-rich history teaching raises history knowledge (g 0.19-0.46) but not standardized reading in under three years; generic 'historical thinking' has no adequately controlled positive result.
moderate supportconf: mediumgc: lowShould history be built on content knowledge or on generic historical-thinking skills? · history-civics · ages 5–18
Outdoor time prevents myopia from starting (not progressing); screening plus free glasses raises test scores in children who need them. Two claims, both real.
moderate supportconf: mediumgc: lowVision — outdoor time against myopia, and screening and correction for achievement · health · ages 4–18 · input
Novices learn faster studying solutions than solving problems — then the effect reverses with expertise. Use worked examples early; fade them.
moderate supportconf: mediumgc: lowWorked examples (studying solutions vs solving problems) · instruction-style · ages 10–18 · method
The URL is the state — any view you build here is a permalink you can hand to someone mid-argument.