Level 1 Criteria
-
C Class-level RCT
- Randomisation was conducted at the individual student level, not at the class or school level, and the intervention is not a one-to-one tutoring exception.
- "To investigate the causal impact of corequisite remediation on student outcomes, we conducted a student-level RCT." (p. 88)
- Relevant Quotes: 1) "To investigate the causal impact of corequisite remediation on student outcomes, we conducted a student-level RCT. All first-time-in-college students scoring within a predetermined range on the state's college readiness exam at each participating institution were recruited to participate in the study during orientation or initial advising sessions." (p. 88) 2) "Students who opted to participate were randomized to corequisite remediation or to the highest level standalone IRW DE course." (p. 88) 3) "The study focused on first-time-in-college students... Within this common eligibility range, students were recruited and randomized into either corequisite remediation or traditional DE." (p. 84) Detailed Analysis: The paper explicitly and repeatedly states that the unit of randomisation was the individual student ("we conducted a student-level RCT"), not the class or school. Students within the same eligibility pool at each college were individually assigned to either the corequisite (treatment) or standalone IRW (control) pathway. The ERCT standard requires randomisation at the class level or stronger (school level) to avoid contamination, unless the intervention is a one-to-one tutoring/personal-teaching intervention, in which case student-level randomisation is acceptable. This intervention is a course-placement reform (corequisite English course plus a developmental-education support component), not a personal one-to-one tutoring intervention, so the tutoring exception does not apply, even though two of the five colleges implement the DE support as individual tutoring rather than group instruction. The core treatment is which course sequence/section a student is placed into, and contamination risk from mixing treatment/control students in overlapping classes is a real concern the paper does not address by randomising at the class or school level. Criterion C is not met because randomisation occurred at the individual student level without a valid personal-tutoring exception.
-
E Exam-based Assessment
- Outcomes are based on ordinary course grades (A-F) assigned by instructors in the courses being studied, not a standardised, widely recognised exam.
- "To assess course passing, grades were recorded on a standard A-F basis in THECB data, and we defined passing as receiving a grade of C or better." (p. 90)
- Relevant Quotes: 1) "To assess course passing, grades were recorded on a standard A-F basis in THECB data, and we defined passing as receiving a grade of C or better. We focus on the completion of English Composition I because it is a key gateway course required for all academic degree programs..." (p. 90) 2) "Given that some stakeholders had expressed concerns that instructors may 'water down' content or artificially inflate grades in English Composition I under a corequisite model in which many or all students had lower test scores, we wanted to assess whether corequisites improved pass rates in courses that build upon the content of English Composition I." (pp. 90-91) 3) "Our key outcomes of interest for this paper included passing English Composition I by the end of the first and second academic year and one- and two-year persistence." (p. 90) Detailed Analysis: The primary outcomes in this study are instructor-assigned course grades (pass/fail at grade C or better) in English Composition I and II and a college reading course, drawn from administrative transcript records, plus credit accumulation and persistence. These are not scores from a standardised, widely recognised exam-based assessment (e.g., a state or national standardised test); they are ordinary course grades assigned by the instructors teaching the very courses being evaluated. The paper itself flags the concern that instructors could inflate or "water down" grades, which is precisely the bias problem the ERCT "E" criterion is designed to guard against by requiring a standardised exam external to the course. No standardised exam (comparable to EGRA/EGMA, a state assessment, or similar) is used as the outcome measure anywhere in the paper; only the TSIA placement exam is standardised, and that is used solely for eligibility/placement, not as an outcome measure. Criterion E is not met because the outcome measures are ordinary course grades, not a standardised exam-based assessment.
-
T Term Duration
- The intervention runs a full semester and outcomes are tracked for one and two years after it begins, far exceeding one academic term.
- "Our key outcomes of interest for this paper included passing English Composition I by the end of the first and second academic year and one- and two-year persistence." (p. 90)
- Relevant Quotes: 1) "We recruited and consented 1,276 newly enrolling students over three semesters (fall 2016, spring 2017, and fall 2017) from five Texas community colleges." (p. 79) 2) "The course ran for 16 weeks for all colleges except College E." (p. 85) 3) "Our key outcomes of interest for this paper included passing English Composition I by the end of the first and second academic year and one- and two-year persistence." (p. 90) 4) "To assess one- and two-year persistence, we examined the effect of corequisite remediation on enrollment one and three semesters after the semester in which the randomization and initial course enrollment took place." (p. 90) Detailed Analysis: The intervention itself (the corequisite course plus DE support) begins at the start of a semester and runs the length of that semester (a full academic term, 16 weeks for most colleges). Primary outcomes (passing English Composition I, credit accumulation, persistence) are measured at one year and two years after the semester in which the intervention began, which is far longer than the minimum one-term interval required by criterion T. Criterion T is met because the interval from intervention start to outcome measurement (one to two years) greatly exceeds one academic term.
-
D Documented Control Group
- The paper documents the control group's composition, baseline covariates, and the specific standalone DE course it received in detail, including a formal covariate-balance table.
- "Table 5 presents findings on the balance of predetermined covariates across treatment and control students... The means are quite similar for treatment and control students, consistent with successful randomization." (p. 92)
- Relevant Quotes: 1) "We refer to the set of students randomized to corequisite remediation as the 'treatment group' and those randomized to the stand-alone IRW course as the 'control group.'" (p. 88) 2) "Prior to the introduction of the intervention, students in our eligible score ranges would have been placed into the highest level of stand-alone DE courses, which consisted of an Integrated Reading and Writing (IRW) course that ranged from three to five credit hours depending on the college... the traditional DE course in which control students enrolled was the same across the participating colleges." (p. 84) 3) "Table 5. Covariate means by assignment status with mean equality test... Age, Black, Economic disadvantage, Female, Hispanic, Limited English, White, Part-time, Bachelor's intent, Associate's intent, Certificate intent, High school diploma, GED, First language English, First generation." (p. 92) 4) "Consent rates varied from 67% to 91% across colleges (Table 3)." (p. 89) Detailed Analysis: The paper provides extensive documentation of the control condition: it clearly describes what the control group received (the standalone Integrated Reading and Writing DE course, a mandated state-wide course of three to five credit hours), and it reports a detailed covariate-balance table (Table 5) comparing treatment and control students on 15 baseline demographic and academic characteristics, both for the full randomized sample and for enrolled students, along with sample sizes by college (Table 3). This level of detail allows readers to confirm the comparability of the control group and understand exactly what "business as usual" consisted of. Criterion D is met because the control group's characteristics, size, and the treatment they received are thoroughly documented.