Level 1 Criteria
-
C Class-level RCT
- Randomisation occurred at the school level (193 schools), which is a stronger unit than class-level and therefore automatically satisfies the class-level RCT requirement.
- "Following recruitment, schools were randomly allocated, within geographical area, to either a 20-week oral language intervention group or a business-as-usual control group." (p. 1426)
- Relevant Quotes: 1) "A cluster randomized controlled trial (RCT) was conducted in 193 state primary schools (containing 238 Reception classrooms) from 13 geographical areas in the UK..." (p. 1426) 2) "Following recruitment, schools were randomly allocated, within geographical area, to either a 20-week oral language intervention group or a business-as-usual control group." (p. 1426) 3) "After completion of t0 and t1 testing, schools were randomized to intervention or control group by an independent evaluator. Randomization was stratified by geographical area and the number of classes participating in each school (dichotomized: 1, or more than 1, class)." (p. 1427) Detailed Analysis: The paper explicitly and repeatedly states that the unit of randomisation was the school, not individual classes or students. Entire schools (193 of them) were allocated to intervention or control, with stratification by geographical area and number of participating classes. Under the ERCT Standard, a school-level RCT is a stronger design than a class-level RCT and, where present, automatically satisfies the weaker class-level criterion, since randomising whole schools eliminates any risk of within-class or within-school contamination between treatment and control pupils. Criterion C is met because randomisation was conducted at the school level, which exceeds the class-level requirement.
-
E Exam-based Assessment
- The study used well-established standardized language assessments (CELF Preschool, Renfrew APT), not custom-built instruments.
- "As outlined in the trial preregistration, the primary outcome measures were the four standardized tests of language ability administered at t1 and t2 to children identified as eligible for the NELI programme." (p. 1427)
- Relevant Quotes: 1) "As outlined in the trial preregistration, the primary outcome measures were the four standardized tests of language ability administered at t1 and t2 to children identified as eligible for the NELI programme." (p. 1427) 2) "Language skills were assessed with the Expressive Vocabulary subtest from the Child Evaluation of Language Fundamentals (CELF) Preschool IIUK (Semel, Wiig, & Secord, 2006), The Renfrew Action Picture Test (APT; Renfrew, 2003; information and grammar scores) and the Recalling sentences subtest from the Child Evaluation of Language Fundamentals (CELF) Preschool IIUK (Semel et al., 2006)." (p. 1427) Detailed Analysis: The primary outcome measures are drawn from the CELF Preschool (Semel, Wiig, & Secord, 2006) and the Renfrew Action Picture Test (Renfrew, 2003), both long-established, commercially published, psychometrically validated standardized language assessments used widely in research and clinical practice, not instruments created for this study. The authors themselves label these "standardized tests of language ability." Criterion E is met because the primary outcome measures are recognised, published standardized assessments rather than custom-built tests.
-
T Term Duration
- Outcomes were measured roughly 7-8 months after the intervention began, far exceeding the one-term minimum.
- "Assessments took place before the start of the intervention at screening (t0)...and at pretest (t1)...and immediately following the intervention (post-test, t2)." (p. 1426)
- Relevant Quotes: 1) Figure 1 timeline: "Screening (t0, Sept 2018)", "In-depth pretest (t1, Oct 2018)", "2-day training (Nov 2018)" marking the start of intervention delivery, through to "Concurrent screening & in-depth posttest (t2, June - July 2019)." (p. 1427) 2) "Assessments took place before the start of the intervention at screening (t0) for all children in participating classrooms and at pretest (t1) for children selected via screening...and immediately following the intervention (post-test, t2). The timeline is presented in Figure 1." (p. 1426) Detailed Analysis: Intervention delivery (TA training and NELI sessions) began around November 2018, and the final outcome measurement (t2) took place in June-July 2019, an interval of roughly 7-8 months. This comfortably exceeds the minimum one academic term (approximately 3-4 months) required by this criterion. Criterion T is met because the interval from intervention start to outcome measurement substantially exceeds one academic term.
-
D Documented Control Group
- The control group's size, demographics, baseline scores and the provision it received are documented in detail in the text, CONSORT diagram and Table 1.
- "Schools in the control group delivered their usual school provision and received payment to purchase the programme at the end of the trial if they wished." (p. 1426)
- Relevant Quotes: 1) "Schools in the control group delivered their usual school provision and received payment to purchase the programme at the end of the trial if they wished." (p. 1426) 2) "Allocated to Control group: School n = 96; mean cluster size = 7.20; cluster variance = 9.07; children n = 592" (Figure 2, CONSORT diagram, p. 1428) 3) Table 1 reports, for the Control Group (n = 592), baseline and post-test means (SD) for age, LanguageScreen subtests, CELF, APT and YARC word reading measures alongside the intervention group. (p. 1429-1430) 4) "Critically, there were no significant differences at pretest in gender...age...or language factor scores derived from standardized language tests...between children who completed the study and those who dropped out at post-test." (p. 1429) Detailed Analysis: The control condition (business-as-usual teaching, with deferred access to the programme) is clearly described, and the control group's size, cluster structure, demographic composition and detailed baseline test scores are reported in Table 1 and the CONSORT diagram, alongside confirmation that completers and dropouts did not differ significantly at baseline. Criterion D is met because the control group is thoroughly documented with size, baseline characteristics and the treatment it received.