From the K–12 edition of September 7, 2026
Classroom trial finds limited benefits from specialized self-regulation prompts
Tim Fütterer and colleagues · K–12 secondary education · Baden-Württemberg, Germany
- Publisher
- Enhancing School Students’ Self-Regulated Learning through Generative AI Support: A Randomized Controlled Trial
- Original publication
- April 14, 2026; classroom study April–May 2025
- Source retrieved
- 2026-09-08
What happened
Specialized prompts produced no significant advantage over the AI control in knowledge, effort or strategy use. Utility-value trajectories favored motivational over strategy prompting, not over control.
Why it matters
Newly archived counterevidence for fall tutoring decisions; German secondary-school findings require local curriculum and learner validation.
Evidence and measured results
371 Grade 7–9 students; individual randomization; six 45-minute sessions including pre/posttests and four learning sessions. Mixed-effects analyses compared GPT-4o conditions. Posttest attrition was 35%.
Limitations and uncertainty
Convenience sample, immediate outcomes, attrition and measurement limitations. Control also used AI; no AI-versus-no-AI estimate. Supplementary data were not independently reanalyzed.
Put this evidence to work
Lighthouse Advisory interpretation, grounded in this source. Enriched 2026-09-08; this does not change the original publication date. Labels below come from the analysis itself.
Sales
Role takeaway
A curriculum director considering a study-support assistant needs evidence that students learn, beyond finding the interaction appealing. Ask which independent assessment will be used, whether the comparison includes ordinary instruction, and who supports disengaged learners. Offer a bounded evaluation design for one subject and grade. The value hypothesis is a clearer continuation decision. The trial's limited contrasts cannot support a general claim that AI improves self-regulation or that it never works. Treat the German setting as a reason to validate locally, and confirm purchasing needs directly rather than inferring an opportunity.
Pre-sales engineering
Role takeaway
For a classroom proof of value, isolate task materials, prompts and learner identity from unrelated student records. Establish a controlled release process and a teacher fallback before adding a chatbot to the learning platform. Prerequisites include approved curriculum, testable access rules and permission to evaluate the configured model. Use synthetic cases to test answer leakage and cross-student disclosure. Compare independent pre/post and delayed assessments with an ordinary-instruction baseline, while recording missing responses and review effort. These proposed tests address uncertainty; they do not reproduce the reported study or establish suitability for autonomous instruction.
Delivery
Role takeaway
Assign the curriculum evaluation lead as owner, with classroom teachers, IT, accessibility and privacy staff responsible for defined checks. Reserve lesson time, prepare accessible alternatives and train teachers to identify frustration without taking over assessment responses. Approve the analysis and stopping rules before collecting data.
- Proposed acceptance
- every participant has a permitted route through the activity, missing assessments are reported, and delayed performance and staff effort are reviewed before expansion. These are proposed criteria, not observed results. Risks include selective reporting, attrition and deploying a changed model under the original evidence.
Implementation considerations
Lighthouse Advisory interpretation across the operating dimensions a public-sector buyer must settle before this evidence becomes a design. Each note answers the question under its heading for this specific source.
Architecture and integration
What must connect, and where does the AI sit in the workflow?
Version prompts, model and task content together; evaluate the complete learner workflow. Autonomous agents, local hosting and hybrid infrastructure were not compared.
Governance
Who approves, reviews and stays accountable for outcomes?
Define knowledge and motivation outcomes separately before the pilot; avoid selecting exploratory subgroups after results.
Security and privacy
What data, permissions and controls need testing?
Minimize learner context, restrict transcript access, document retention and prohibit unrelated reuse; local privacy review remains necessary.
Accessibility and workforce
Who is affected, and what skills or accommodations follow?
Test reading demands, accommodations and teacher intervention time; text interaction may impose uneven burdens.
Procurement
What should contracts, pricing and exit terms secure?
Require evidence for the actual configured service and exportable evaluations; no savings or retention guarantee is supported.
Operating model
Which teams own the service once it runs?
Curriculum leads evaluation, teachers handle learner support, and IT controls versions and access.
What changed
Not an archive repeat: absent from all 20 K12 records and the full-library candidate search. Newly added context for fall 2026 decisions; not represented as new September 7 publication.
Publication history
- 2026-09-07K–12 · Issue 023 resources
Stable resource ID: german-secondary-srl-genai-rct-2026