This methods article describes an 8-week wait-list controlled classroom protocol for implementing, monitoring, and evaluating self-regulated English learning among university students.
Method Article
This methods article describes an 8-week wait-list controlled classroom protocol for implementing, monitoring, and evaluating self-regulated English learning among university students.
Self-regulated learning is an important component of successful language learning, but classroom-based interventions are often difficult to reproduce when instructional activities, student outputs, control conditions, fidelity procedures, and assessment rules are insufficiently described. This methods article presents an 8-week wait-list controlled classroom protocol for implementing and evaluating self-regulated English learning among university students. Students are assigned to either a structured self-regulated English learning protocol or ordinary English instruction with delayed access to the intervention materials. Although each week has a specific instructional focus, the protocol repeatedly embeds the cyclical self-regulated learning sequence of goal review, strategy use, monitoring, reflection, and adjustment. Weekly activities include self-diagnosis, measurable goal setting, strategic planning, vocabulary and reading strategy practice, writing revision, learning diaries, peer feedback, and final transfer planning. Each session generates a predefined student output, allowing fidelity, attendance, worksheet completion, adherence, protocol deviations, and contamination-control procedures to be documented. The protocol combines questionnaire-based outcomes with performance-based measures, including English achievement and writing performance, and specifies baseline assessment, stratified block randomization, blinded writing scoring, parallel testing procedures, missing-data rules, and adjusted statistical analysis. Representative results demonstrate how the protocol can be implemented, monitored, and analyzed in a classroom setting. These results are presented as illustrative implementation and outcome data, not as generalizable evidence of intervention efficacy. This protocol is suitable for researchers and instructors who require a practical, ethically acceptable, and reproducible procedure for studying self-regulated English learning in higher education.
Self-regulated learning has become an important framework for understanding why students exposed to similar teaching conditions often show different learning trajectories. In this framework, learners are not treated as passive recipients of instruction, but as participants who set goals, select strategies, monitor their progress, and revise their behavior when learning conditions change. Zimmerman described self-regulated learners as students who actively use forethought, performance control, and self-reflection to guide learning behavior rather than depending only on external instruction1. Pintrich further emphasized that self-regulation involves the monitoring and control of cognition, motivation, behavior, and context, which makes the construct particularly relevant to classroom-based language learning where students must coordinate effort, time, strategy use, and feedback2.
The relevance of self-regulation is especially clear in university English learning. Many students know that they need to improve vocabulary, reading, writing, or communication, but their learning actions remain fragmented. A student may intend to “study harder,” yet this intention often fails to become a measurable goal, a scheduled task, a monitored learning behavior, or a revised strategy after feedback. Language-learning strategy research has long shown that learners differ not only in how much they study, but also in how they choose, combine, and evaluate strategies across tasks3. However, language classes do not always provide a reproducible structure through which such strategy use can be taught, observed, documented, and compared with ordinary classroom instruction. This creates a methodological problem: researchers may claim to study self-regulated learning, but the instructional procedure may remain too general for another classroom to reproduce.
Existing studies have clarified the measurement and instructional importance of self-regulated language learning. Validated work on foreign-language self-regulated learning strategy measurement has shown that the construct can be organized into identifiable strategy dimensions rather than treated as a single vague learning habit4. Research on second-language writing has also shown that self-regulated strategy instruction can support writing development by making planning, drafting, revising, and reflection more explicit in classroom practice5. Higher-education intervention research further suggests that the effect of self-regulated learning support depends on how the intervention is implemented, not only on whether the term “self-regulated learning” appears in the course design6. A classroom protocol should therefore specify what students do, when they do it, what they produce, how the control condition differs, and how researchers verify that the procedure was delivered as intended.
Several limitations remain in classroom-based self-regulated learning research. Some studies rely heavily on post-intervention questionnaires, which can show perceived changes but provide limited information about the weekly learning behaviors that produced those changes. Others describe strategy instruction but provide insufficient detail about lesson timing, worksheet structure, control-group activities, fidelity monitoring, adherence thresholds, or contamination prevention. In language-learning settings, this problem is particularly important because strategy use, motivation, anxiety, writing performance, and achievement may change at different rates. Reviews of self-regulated learning in online and higher-education environments have highlighted the need for stronger alignment between intervention design, learning-process evidence, and outcome measurement7. Without such alignment, it is difficult to determine whether a weak or inconsistent result reflects an ineffective theoretical approach, poor implementation, inadequate measurement, or contamination between study groups.
The present methods article addresses this gap by translating self-regulated English learning into an 8-week wait-list controlled classroom protocol with predefined weekly outputs. The protocol is not organized as a single linear sequence in which students complete one self-regulatory phase and then move permanently to the next. Instead, each week has a dominant instructional focus while preserving a recurring self-regulated learning cycle: students review or refine a goal, apply a strategy to an English-learning task, monitor their performance, reflect on the outcome, and make an adjustment for the next learning activity. This structure follows the practical implication of self-regulated strategy instruction: learners need repeated opportunities to connect a goal with a strategy, an action, feedback, and revision8. The weekly focus gradually shifts from self-diagnosis and goal setting to planning, strategy use, writing revision, monitoring, feedback use, and independent transfer, but the underlying cycle is repeated throughout the 8-week period.
The protocol includes a wait-list control condition to support ethical and methodologically interpretable classroom implementation. Students in the control condition continue ordinary English instruction during the 8-week period and receive access to the self-regulated learning worksheets after posttest assessment. This design allows the structured protocol to be compared with ordinary instruction while reducing the ethical concern of withholding potentially useful learning materials. Because participants may be randomized within the same classroom, the protocol includes contamination-control procedures, including separate activity periods or workspaces, restricted worksheet distribution, weekly control-class content logs, and delayed release of intervention materials. These procedures do not eliminate all risk of informal communication between students, but they make contamination visible and documentable when interpreting the representative results.
The protocol uses both self-report and performance-based outcomes. Self-regulated learning, self-efficacy, motivation, and anxiety are measured through structured questionnaires, while English achievement and writing performance provide task-based evidence of learning change. This combination is deliberate. Self-efficacy is closely linked to students’ willingness to initiate and sustain learning behavior, but confidence alone is not the same as improved performance9. Writing performance is also useful because it requires students to plan, organize ideas, select language, revise errors, and respond to task demands. Studies of writing strategy instruction in college settings have shown that self-regulation can be embedded in planning and revising processes rather than treated as a separate study-skills lesson10. The current protocol therefore connects self-regulated learning activities with observable English-learning tasks.
The method strengthens reproducibility through procedural controls. Baseline assessment is completed before randomization. Allocation is stratified by class and baseline proficiency. The intervention group receives a standardized weekly worksheet package, whereas the wait-list control group continues ordinary instruction without access to the structured self-regulated learning materials before posttest assessment. Writing samples are blinded before scoring. Fidelity, attendance, worksheet completion, diary completion, protocol deviations, and contamination-control indicators are recorded throughout the 8-week period. These elements are consistent with broader reporting expectations for randomized trials, where participant flow, allocation, follow-up, and analytic inclusion must be transparent11. In classroom research, these controls are not merely administrative. They help distinguish the implementation of a structured learning protocol from ordinary teacher feedback, student motivation, or informal sharing of materials across groups.
The overall goal of this article is to present a reproducible classroom protocol for implementing, monitoring, and evaluating self-regulated English learning among university students. The protocol is designed for researchers and instructors who need a practical method that can be ethically implemented in real classes while still producing interpretable implementation and outcome data. It is most appropriate for university English courses in which students can complete weekly learning worksheets, baseline and posttest questionnaires, English achievement tests, and writing tasks. The protocol can also be adapted for blended or online delivery if worksheet access, timing, fidelity monitoring, contamination control, and data security are preserved. Current evidence on technology-supported self-regulated language learning suggests growing interest in such adaptable procedures, but also shows the need for clearer intervention structures and more consistent outcome reporting12.
The classroom implementation described in this article was reviewed and approved by the Foreign Languages Ethics Committee of Zhengzhou Shengda University (Approval No. 20260606) on June 3, 2026. Written informed consent was obtained from all participants before data collection. Participation was voluntary, and students could withdraw at any time without affecting their course grades. The wait-list control group received access to the self-regulated learning materials after the posttest assessment. The course-grading instructor did not access individual questionnaire responses, diary entries, consent decisions, or participant-level research data before finalizing course grades.
1. Prepare the classroom trial
2. Recruit participants and obtain consent
3. Conduct baseline assessment
4. Randomize participants and conceal allocation
5. Deliver the self-regulated English learning protocol
6. Deliver the wait-list control condition
7. Conduct posttest assessment
8. Score outcomes and process data
9. Monitor fidelity, adherence, and protocol deviations
10. Construct the analytic dataset
11. Analyze the data
12. Manage withdrawal, confidentiality, and delayed access
Participant flow, baseline characteristics, and posttest completion
The representative dataset included 180 students assessed for eligibility. Twenty students were excluded before randomization: 8 did not meet the inclusion criteria, 9 declined participation, and 3 were unavailable for the full 8-week follow-up period. The final randomized sample included 160 students, with 80 assigned to the self-regulated learning protocol group and 80 assigned to the wait-list control group. Posttest assessment was completed by 76 students in the protocol group and 75 students in the wait-list control group, yielding an available-case analytic sample of 151 students, or 94.4% of the randomized sample. Posttest missingness was 5.0% in the protocol group and 6.3% in the wait-list control group. In the protocol group, 72 students attended at least 6 of 8 sessions and were included in the per-protocol sensitivity sample (Figure 1).
Baseline characteristics were similar between groups. The mean age was 20.1 ± 1.2 years in the protocol group and 20.0 ± 1.1 years in the wait-list control group. Female students accounted for 58.8% and 57.5% of the two groups, respectively. Baseline English proficiency was also similarly distributed: 22.5% of students in the protocol group were classified as A2, 53.8% as B1, and 23.8% as B2; the corresponding proportions in the wait-list control group were 23.8%, 52.5%, and 23.8%. Baseline self-regulated English learning scores were 3.12 ± 0.44 and 3.10 ± 0.45, baseline English achievement scores were 68.4 ± 8.7 and 68.7 ± 8.5, and baseline writing performance scores were 12.6 ± 2.1 and 12.7 ± 2.0 in the protocol and wait-list control groups, respectively (Table 4).
Protocol fidelity, adherence, and control-condition monitoring
All 8 planned weekly sessions were delivered in the protocol group. The mean fidelity rating was 4.5 ± 0.3 out of 5, and all sessions received a fidelity rating of 4 or higher. Mean attendance was 7.1 ± 1.0 sessions out of 8, and 72 students, representing 90.0% of the protocol group, met the predefined exposure threshold. The mean worksheet completion count was 6.8 ± 1.2 out of 8. Completion rates were 97.5% for the Week 1 learning profile, 95.0% for the Week 2 goal card, 92.5% for the Week 3 study-planning worksheet, 91.3% for the Week 4 reading and vocabulary strategy sheet, 90.0% for the Week 5 writing cycle worksheet, 87.5% for the Week 6 learning diary, 88.8% for the Week 7 peer-feedback sheet, and 87.5% for the Week 8 final reflection sheet.
The wait-list control group received ordinary English instruction, ordinary homework, and ordinary teacher feedback during the 8-week period. The control-class content log confirmed that structured self-regulated learning worksheets, learning diaries, goal cards, strategy checklists, peer-feedback templates, and reflection sheets were not released to the control group before posttest assessment. Two informal cross-group discussions about study habits were recorded in the contamination-control log, but no early worksheet release, template sharing, or control-group access to intervention materials was confirmed. Delayed access to the full self-regulated learning package was provided after posttest data collection (Table 5).
Questionnaire-based outcomes
In this representative dataset, the self-regulated English learning score increased from 3.12 ± 0.44 to 3.68 ± 0.42 in the protocol group and from 3.10 ± 0.45 to 3.16 ± 0.44 in the wait-list control group. After adjustment for baseline score and class, the between-group difference in posttest self-regulated English learning score was 0.49, with a 95% confidence interval of 0.36 to 0.62 and a p value < 0.001 (Table 6). English self-efficacy increased from 3.18 ± 0.48 to 3.54 ± 0.46 in the protocol group and from 3.20 ± 0.47 to 3.31 ± 0.45 in the wait-list control group. The adjusted between-group difference was 0.24, with a 95% confidence interval of 0.10 to 0.38 and a p value of 0.001. English learning motivation increased from 3.42 ± 0.46 to 3.62 ± 0.43 in the protocol group and from 3.40 ± 0.45 to 3.45 ± 0.44 in the wait-list control group. The adjusted between-group difference was 0.14, with a 95% confidence interval of 0.02 to 0.26 and a p value of 0.023. English learning anxiety decreased from 3.01 ± 0.52 to 2.72 ± 0.49 in the protocol group and from 3.00 ± 0.50 to 2.99 ± 0.51 in the wait-list control group. The adjusted between-group difference was −0.26, with a 95% confidence interval of −0.41 to −0.11 and a p value of 0.001 (Table 6).
Performance-based outcomes
English achievement scores increased from 68.4 ± 8.7 to 77.3 ± 8.1 in the protocol group and from 68.7 ± 8.5 to 71.9 ± 8.3 in the wait-list control group. The adjusted between-group difference at posttest was 5.3 points, with a 95% confidence interval of 3.1 to 7.5 and a p value < 0.001. Writing performance scores increased from 12.6 ± 2.1 to 14.8 ± 2.0 in the protocol group and from 12.7 ± 2.0 to 13.2 ± 2.1 in the wait-list control group. The adjusted between-group difference at posttest was 1.5 points, with a 95% confidence interval of 0.8 to 2.2 and a p value < 0.001 (Table 6). Pretest-to-posttest changes in questionnaire-based and performance-based outcomes are summarized in Figure 3A–F.
Data quality, scoring reliability, and sensitivity analyses
Writing samples were scored by two independent raters blinded to group allocation and assessment time point. The intraclass correlation coefficient was 0.86 for the total writing score, 0.82 for content relevance, 0.80 for organization, 0.78 for vocabulary use, 0.84 for grammatical accuracy, and 0.81 for task completion. Eleven total-score discrepancies greater than 3 points on the 0-20 rubric were identified, including 6 in the protocol group and 5 in the wait-list control group. All discrepancies were resolved before final score calculation. Model checks did not indicate major violations of the ANCOVA assumptions. The relationship between baseline and posttest scores was approximately linear, residual plots did not show strong heteroscedasticity, and no influential observation changed the direction of the adjusted group estimate. The group-by-baseline score interaction for the primary model was not statistically significant (p = 0.42).
The primary available case analysis included 151 students with valid baseline and posttest data. Because overall posttest missingness exceeded 5%, a multiple-imputation sensitivity analysis was conducted under a missing-at-random assumption. The imputed estimate for the adjusted between-group difference in self-regulated English learning was 0.47, with a 95% confidence interval of 0.34 to 0.60. The per-protocol sensitivity analysis showed an adjusted between-group difference of 0.52, with a 95% confidence interval of 0.38 to 0.66. Five range-check errors were identified before the dataset lock, including 3 in the protocol group and 2 in the wait-list control group. All were corrected against the original records. No protocol deviation affecting outcome assessment was recorded, and no participant withdrawal due to discomfort was recorded (Table 7).

Figure 1: Participant flow through the randomized classroom protocol. The flow diagram shows eligibility screening, consent, baseline assessment, randomization, allocation to the self-regulated learning protocol group or wait-list control group, posttest completion, and analytic inclusion. The wait-list control group continued ordinary English instruction during the 8-week period and received delayed access to the learning materials after posttest assessment. Please click here to view a larger version of this figure.

Figure 2: Eight-week self-regulated English learning classroom workflow. The workflow summarizes baseline assessment, randomization, weekly protocol delivery, posttest assessment, and delayed access for the wait-list control group. The 8-week sequence includes orientation and self-diagnosis, goal setting, strategic planning, vocabulary and reading strategy practice, writing strategy practice, monitoring and self-recording, peer-feedback calibration, and final reflection. Please click here to view a larger version of this figure.

Figure 3: Pretest-to-posttest changes in representative questionnaire-based and performance-based outcomes. (A) Self-regulated English learning. (B) English self-efficacy. (C) English learning motivation. (D) English learning anxiety. (E) English achievement. (F) Writing performance. Please click here to view a larger version of this figure.
Table 1: Enrollment, intervention, control-condition, and assessment schedule. This table summarizes the timing, duration, main procedure, group exposure, key record, and responsible person or team for each stage of the randomized classroom protocol. Please click here to download this Table.
Table 2: Outcome definitions, scoring rules, and dataset variables. This table defines the primary outcome, secondary outcomes, writing rubric domains, process indicators, data-completion variables, dataset variable names, scoring rules, missing-data rules, and data sources. Please click here to download this Table.
Table 3: Fidelity, adherence, contamination-control, data-quality, and analysis-readiness indicators. This table defines the indicators used to monitor protocol delivery, student exposure, worksheet completion, control-condition integrity, contamination risk, writing-score reliability, data cleaning, ANCOVA assumption checking, protocol deviations, ethics monitoring, and reproducibility. Please click here to download this Table.
Table 4: Baseline characteristics of randomized participants. This table presents demographic characteristics, baseline English proficiency categories, and baseline outcome scores for participants assigned to the self-regulated learning protocol group and the wait-list control group. Values are presented as mean ± SD or n (%). Self-regulated English learning, self-efficacy, motivation, and anxiety were scored from 1 to 5. English achievement was scored on a scale of 0 to 100. Writing performance was scored from 0 to 20. Please click here to download this Table.
Table 5: Protocol fidelity, adherence, control-condition integrity, and safety indicators. This table summarizes session delivery, fidelity ratings, attendance, worksheet completion, diary completion, ordinary instruction in the wait-list control group, contamination-control status, delayed access, protocol deviations, and withdrawal due to discomfort. Fidelity was rated on a 1–5 scale, with higher scores indicating stronger protocol delivery fidelity. Attendance in the wait-list control group refers to ordinary English class attendance. Structured self-regulated learning worksheets, learning diaries, goal cards, strategy checklists, peer-feedback templates, and reflection sheets were not provided to the wait-list control group before posttest assessment. Please click here to download this Table.
Table 6: Representative pretest-posttest outcomes and adjusted between-group differences. This table presents pretest scores, posttest scores, mean changes, adjusted between-group differences, 95% confidence intervals, standardized effect sizes, and p-values for the primary and secondary outcomes. Adjusted between-group differences were estimated using analysis of covariance with the corresponding baseline score and class as covariates. Positive values favor the protocol group for self-regulated English learning, self-efficacy, motivation, English achievement, and writing performance. Negative values for anxiety indicate lower posttest anxiety in the protocol group. Please click here to download this Table.
Table 7: Data quality, posttest completion, and sensitivity-analysis indicators. This table summarizes posttest completion, available-case analytic inclusion, posttest missingness, per-protocol sensitivity sample inclusion, writing-score rater agreement, discrepancy resolution, range-check correction, ANCOVA assumption checking, sensitivity-analysis estimates, protocol deviations affecting outcome assessment, and withdrawal due to discomfort. Completed posttest components included the posttest questionnaire battery, parallel English achievement test, and posttest writing task. The available case analysis included participants with valid baseline and posttest data for the primary outcome. The per-protocol sensitivity sample included protocol-group participants who completed posttest assessment and attended at least 6 of 8 sessions. ICC indicates the intraclass correlation coefficient. Dashes indicate indicators summarized at the overall dataset or model level rather than separately by group. Please click here to download this Table.
Supplementary Table 1: Weekly self-regulated learning worksheet package. This supplementary table specifies the worksheet code, week, protocol component, student output, required fields, completion rule, and study record for each week of the 8-week self-regulated learning protocol. The worksheet package was used only in the self-regulated learning protocol group prior to the posttest. The wait-list control group received the full worksheet package after posttest data collection. Completion records were used for adherence monitoring and per-protocol sensitivity classification.Please click here to download this file.
This methods article presents a reproducible classroom protocol for implementing self-regulated English learning among university students under randomized, wait-list controlled conditions. The protocol translates self-regulated learning from a broad theoretical construct into observable classroom actions by converting goal setting, strategic planning, strategy use, monitoring, feedback uptake, and reflection into weekly student outputs. This structure is important because classroom-based intervention research requires visible evidence of what students actually did during the learning process, rather than relying only on students’ post-intervention self-reports. Recent higher-education evidence suggests that self-regulated learning interventions are more interpretable when the intervention content, implementation process, and measurement strategy are explicitly aligned19,20,21.
The 8-week sequence follows a repeated self-regulated learning cycle rather than a one-time study-skills lesson. The early sessions establish learning awareness and measurable goal setting, the middle sessions translate goals into scheduled English-learning tasks, strategy practice, writing revision, and monitoring, and the final sessions focus on feedback use, reflection, and transfer planning. This order allows reflection to be based on actual learning evidence and allows strategy use to remain connected to a goal, an action, and an adjustment. The weekly worksheets, therefore, function as both instructional supports and process records. This design is consistent with foreign-language self-regulated learning research emphasizing repeated strategy practice, teacher support, and opportunities for learners to monitor their own progress22.
The representative results show that the protocol can be implemented and documented in a university English classroom. All planned weekly sessions were delivered, fidelity ratings were high, attendance remained stable, and worksheet completion was maintained across the intervention period. The wait-list control condition was also documented through ordinary instruction logs, worksheet distribution records, delayed access records, and contamination-control monitoring. The outcome pattern should be interpreted as representative protocol output rather than definitive evidence of generalizable intervention efficacy. In this dataset, the protocol group showed larger adjusted posttest gains in self-regulated English learning, with similar directional patterns for self-efficacy, motivation, English achievement, and writing performance, while English learning anxiety decreased. These results demonstrate how implementation, scoring, missingness, adjusted analysis, and sensitivity checks can be reported in a randomized classroom protocol. The use of baseline assessment before randomization, stratified allocation, blinded writing scoring, protected wait-list materials, and transparent analytic inclusion is consistent with broader randomized-trial reporting principles23.
Several procedural elements are central to replication. Baseline assessment should be completed before randomization, and the randomization list should be generated only after baseline data are cleaned and locked. Writing samples should be blinded before scoring because writing assessment is more vulnerable to rater expectation than fixed-response testing. The wait-list control condition also requires active protection: goal cards, learning diaries, strategy checklists, peer-feedback templates, and reflection sheets should not be released to control-group students before posttest assessment. The protocol can be adapted for shorter, writing-focused, blended, or online courses, but the core cycle of goal setting, strategy use, monitoring, feedback, reflection, and adjustment should be preserved. The protocol should not be reduced to a single goal-setting activity, because repeated prompts, monitoring opportunities, and structured feedback are central to self-regulated learning support24.
This protocol has limitations. Individual-level randomization within the same classroom may increase the risk of contamination, particularly when students communicate through informal course groups. Controlled worksheet distribution, separate activity arrangements where feasible, and contamination-control logs reduce this risk, but they cannot eliminate peer-to-peer communication. In settings where separation is not feasible, class-level randomization may be more appropriate, although this design requires different reporting and analytic procedures because students are nested within teaching groups25. Measurement is another limitation. Questionnaire scores capture students’ reported self-regulated learning, self-efficacy, motivation, and anxiety, but these constructs are not fully observable through self-report alone. For this reason, the protocol combines questionnaires with English achievement scores, writing scores, attendance, worksheet completion, diary completion, fidelity ratings, and data-quality indicators. Adaptations should preserve the core self-regulated learning cycle while adjusting materials to learner proficiency, course length, and instructional context26.
The main contribution of the protocol is its reproducible classroom structure. It provides a practical way to connect self-regulated learning theory, weekly English-learning activities, process monitoring, control-condition documentation, and outcome assessment. The wait-list design also supports classroom feasibility by allowing the control group to receive ordinary instruction during the study period and delayed access to the structured materials after posttest assessment. Future applications may link the weekly worksheets to learning analytics, teacher feedback records, writing portfolios, or learning-management-system data. The adherence indicators may also be used to examine whether completion of monitoring and reflection activities is associated with stronger learning changes, although such analyses should remain exploratory because adherence is not randomized. Recent higher-education and technology-enhanced language-learning research continues to emphasize the need for clearer procedures and more consistent measurement across contexts, and this protocol provides one structured workflow for that purpose27,28.
The authors declare that they have no competing financial or non-financial interests related to this work.
The authors thank the university students who participated in this study for their time and consistent engagement throughout the 8-week classroom protocol. Gratitude is also extended to the course instructors and research assistants for their assistance with protocol delivery, data collection, and result verification. This research was supported by the 2025 Henan Provincial Philosophy and Social Sciences Project: "Research on the Reconstruction of Generative AI-Driven Foreign Language Education Paradigm Based on Intersubjectivity Theory" (Project No.: 2025BYY004).
| Name | Company | Catalog Number | Comments |
|---|---|---|---|
| Item name | Source / responsible unit | Specification | Purpose in the protocol |
| Analytic writing rubric | Research team | 0–20 total score; five 0–4 domains | Used to score content relevance, organization, vocabulary use, grammatical accuracy, and task completion. |
| Baseline and posttest English achievement tests | Research team or course teaching unit | Two parallel 40–50 min tests; scored from 0 to 100 | Used to assess course-aligned English achievement. The posttest form is matched to the baseline form by item type, length, topic coverage, and expected difficulty. |
| Baseline and posttest writing prompts | Research team | Two parallel 20 min writing tasks; expected length 120–180 words | Used to assess writing performance. The posttest prompt does not repeat the baseline prompt but uses the same genre, response time, and expected difficulty. |
| Blinded writing-sample folder | Research team | Writing files labeled by participant ID and time point only | Used to store baseline and posttest writing samples without group labels before independent scoring. |
| Control-condition and contamination-control logs | Research team | Control-class content log, worksheet distribution log, and contamination-control log | Used to document ordinary instruction in the wait-list control group, delayed access, restricted worksheet distribution, and possible cross-group exposure. |
| Data-entry and analysis dataset template | Research team / data manager | Spreadsheet with fixed variable names, valid ranges, coding rules, and missing-value rules | Used for data entry, range checking, scale scoring, missing-data review, and construction of the analytic dataset. |
| Demographic questionnaire | Research team | Participant-code-based form | Used to collect age, gender, major, year of study, class ID, prior English-learning experience, and baseline English proficiency category. |
| Eight-week self-regulated learning worksheet package | Research team | W1 learning profile, W2 goal card, W3 study plan, W4 reading/vocabulary strategy sheet, W5 writing cycle worksheet, W6 learning diary, W7 peer-feedback sheet, and W8 final reflection sheet | Used to deliver the structured self-regulated learning protocol and generate weekly student outputs for adherence monitoring. |
| Eligibility and enrollment form | Research team | Screening and enrollment record | Used to confirm eligibility, consent status, class, demographic characteristics, baseline proficiency category, and participant identification code. |
| English learning anxiety scale | Research team; adapted from language-learning anxiety measures | 8 items; 1–5 Likert scale | Used to assess English learning anxiety. Reverse-coded items are defined before data collection, and higher final scores indicate higher anxiety. |
| English learning motivation scale | Research team; adapted from learning motivation measures | 8 items; 1–5 Likert scale | Used to assess effort, persistence, perceived value, and willingness to continue English learning. |
| English self-efficacy scale | Research team; adapted from English-learning self-efficacy measures | 8 items; 1–5 Likert scale | Used to assess confidence in vocabulary learning, reading comprehension, writing, task completion, and independent English study. |
| Fidelity and adherence monitoring forms | Research team | Fidelity checklist, attendance record, worksheet completion record, and diary completion record | Used to record session fidelity, attendance, worksheet completion, diary completion, exposure level, and per-protocol eligibility. |
| Participant information sheet and written consent form | Research team | Ethics-approved study information and consent documents | Used before data collection to explain study purpose, voluntary participation, confidentiality, withdrawal rights, delayed access, and the statement that participation does not affect course grades. |
| Protocol deviation and posttest completion logs | Research team | Deviation record and posttest completion record | Used to document deviations, corrective actions, posttest questionnaire completion, posttest achievement-test completion, and posttest writing-task completion. |
| Randomization list | Statistician / data manager | Stratified block randomization by class and baseline proficiency category | Used to assign participants 1:1 to the self-regulated learning protocol group or wait-list control group after baseline data cleaning. |
| Rater training packet | Research team | Rubric, scoring guide, anchor samples, example scores, and discrepancy rules | Used to train independent writing raters before formal scoring and to support consistency in writing assessment. |
| Secure electronic storage | Institution | Encrypted institutional project folder | Used to store consent forms, restricted participant code files, raw data, anonymized data, writing samples, scoring files, logs, statistical scripts, and archived reproducibility materials. |
| Self-regulated English learning questionnaire | Research team; adapted from published self-regulated learning instruments | 24 items; 1–5 Likert scale | Used to assess goal setting, strategic planning, strategy use, monitoring, and reflection. Scale scores are calculated by averaging valid item responses. |
| Statistical packages | R package set | tidyverse, emmeans, effectsize, mice, lme4, psych, irr, gtsummary or flextable | Used for data management, adjusted group means, effect sizes, multiple imputation, mixed-effects sensitivity analysis, ICC calculation, and manuscript table generation. |
| Statistical software | R Foundation for Statistical Computing | R version 4.4.2 or later | Used for descriptive statistics, reliability analysis, ANCOVA, sensitivity analysis, multiple imputation, and table export. |
| Weekly lesson script package | Research team | Eight weekly lesson scripts | Used to standardize session objective, timing, instructor explanation, student task, required output, and fidelity monitoring across Weeks 1–8. |
Request permission to reuse the text or figures of this JoVE article
Request Permission