Conditions of AI Engagement in Learning: An Online Experiment

Last registered on August 10, 2026

Pre-Trial

Trial Information

General Information

Title
Conditions of AI Engagement in Learning: An Online Experiment
RCT ID
AEARCTR-0019322
Initial registration date
August 07, 2026

Initial registration date is when the trial was registered.

It corresponds to when the registration was submitted to the Registry to be reviewed for publication.

First published
August 10, 2026, 4:48 PM EDT

First published corresponds to when the trial was first made public on the Registry after being reviewed.

Locations

There is information in this trial unavailable to the public. Use the button below to request access.

Request Information

Primary Investigator

Affiliation
Bowdoin College

Other Primary Investigator(s)

PI Affiliation
Bowdoin College
PI Affiliation
Bowdoin College
PI Affiliation
Bowdoin College

Additional Trial Information

Status
In development
Start date
2026-08-06
End date
2027-08-06
Secondary IDs
Prior work
This trial does not extend or rely on any prior RCTs.
Abstract
This study examines how the incentives participants anticipate, and the time available to them, shape the way they engage with AI assistance while studying, and how those engagement patterns relate to later unaided performance. Adults recruited from Prolific complete a two-phase task. In the first phase, participants answer multiple-choice questions drawn from several academic subjects areas, with optional assistance available on request for each question. In the second phase, they answer related questions with no assistance available and with accuracy incentivized by a bonus payment. Participants are randomly assigned with equal probability to one of five arms, which differ the task conditions surrounding the first phase. Outcomes include the depth and patterns of assistance use in the first phase, time spent per item, and accuracy in the unassisted phase. The design is intended to distinguish assistance use that supports independent performance from assistance use that substitutes for it.
External Link(s)

Registration Citation

Citation
Abel, Martin et al. 2026. "Conditions of AI Engagement in Learning: An Online Experiment." AEA RCT Registry. August 10. https://doi.org/10.1257/rct.19322-1.0
Experimental Details

Interventions

Intervention(s)
Participants complete a computer-administered two-phase task. In the study phase, they answer four multiple-choice questions, one from each of four academic domains, presented one per page in individually randomized order. In four of the five arms, an optional on-demand menu of AI-generated hints is available beside each question; participants choose freely whether and how much to consult it, and no correctness feedback is given. In the fifth arm no hints are available. In the test phase, all participants answer a further set of questions on related material with no assistance available.

Participants are randomly assigned with equal probability to one of five arms. The arms differ in the conditions under which the study phase is completed. Those conditions are specified in the non-public intervention field and in the pre-analysis plan, and will be released when the trial is complete.

All hint content is pre-generated by the research team using an AI language model, reviewed for accuracy, and embedded as fixed text in the survey instrument; no participant data are transmitted to any external AI system at any point. Correct answers earn a per-item bonus in both phases, in every arm.
Intervention Start Date
2026-08-06
Intervention End Date
2027-08-06

Primary Outcomes

Primary Outcomes (end points)
Learning test score: total correct answers on the eight-item no-AI knowledge test (0-8; two questions per domain: one recall and one near-transfer). Main learning outcome for all primary confirmatory tests.

Mode 4 usage: the share of the four main-task questions on which the participants used Mode 4 at least once. Defined only for AI-available arms.

Mode 3 or 4 as first hint use: the share of the four main-task questions on which the participants used Mode 3 or 4 first. Defined only for AI-available arms and estimated at the question level.

Time duration: total seconds from question display to answer submission.
Primary Outcomes (explanation)

Secondary Outcomes

Secondary Outcomes (end points)
Immediate task accuracy: number of correct answers on the four main-task questions (0-4). Captures in-task performance, which may diverge from learning if AI use substitutes for rather than supports understanding.

Time measures: total seconds from question display to answer submission; time till first click; seconds between a hint click and the participant's next action (another click or answer submission). Time measures winsorized at the 99th percentile.

Perceived learning / pre-test confidence: participant's domain-by-domain prediction of whether they will answer each knowledge test question correctly, collected after the main task but before the test outcome is known. Serves as a proxy for perceived learning.

Hint-mode sequence: ordered list of mode buttons clicked within a question; analyzed for escalation patterns and repeat clicks.

Recall sub-score: correct answers on the four recall/recognition items referencing content from the original task question (0-4).

Near-transfer sub-score: correct answers on the four novel-scenario items applying the same underlying concept to a new context (0-4).

Mode 4 exclusive use: binary indicator per question (= 1 if Mode 4 was the only hint used on that question), averaged across questions. Captures the purest form of offloading -- answer elimination without any other hint engagement.

Number of distinct hint modes per question: count of unique mode buttons clicked, averaged across questions (0-4).
Secondary Outcomes (explanation)

Experimental Design

Experimental Design
Between-subjects randomized experiment conducted online with approximately 2,000 participants recruited through Prolific, targeting approximately 400 per arm. Eligibility is restricted to United States residents aged 18–25 with at least some college education, prescreened by Prolific and re-verified within the survey instrument.

Participants first complete a baseline block comprising an instructed-response attention check and self-rated confidence in each of the four task domains. They are then randomly assigned with equal probability to one of five arms using the survey platform's randomizer. Randomization is at the individual level; assignment is stored as an embedded field and governs all subsequent survey branching.

One arm serves as a pure control in which no assistance is available. The remaining four arms, in all of which assistance is available, form a 2×2 between-subjects factorial over two manipulated task conditions. The identity of those two conditions is specified in the non-public design field and in the pre-analysis plan, and will be made public when the trial is complete. The control arm identifies the overall effect of assistance availability relative to working unaided; the factorial identifies the main effect of each manipulation within assistance-available arms, and their interaction.

Following the study phase, participants complete brief post-task ratings and report a domain-by-domain expectation of their own performance on the forthcoming test, then complete the unassisted test phase, followed by a validity check and a demographic block, both administered after all outcomes are measured.

Analysis is intent-to-treat. Person-level outcomes are estimated by OLS on arm indicators with the control condition as the reference category, using heteroskedasticity-robust standard errors, and reported both with and without a vector of baseline covariates. Question-level process outcomes are estimated with question fixed effects and standard errors clustered by participant. The factorial contrasts are estimated on assistance-available arms only, in both a saturated specification including the interaction and a restricted specification omitting it, always reported together. Pre-specified exclusions cover the attention check, the eligibility criteria, incomplete responses, duplicate submissions from a single platform identifier, and a post-outcome validity check; results including participants failing the validity check are reported as a robustness check.
Experimental Design Details
Not available
Randomization Method
After the baseline survey and task introduction, participants are randomly assigned with equal probability to one of five experimental arms using Qualtrics embedded data. The assignment is stored as an embedded field and used to condition all subsequent survey branching.
Randomization Unit
Individual participant
Was the treatment clustered?
No

Experiment Characteristics

Sample size: planned number of clusters
2000 individuals
Sample size: planned number of observations
2000 individuals
Sample size (or number of clusters) by treatment arms
400 participants in each of five arms: 400 in the no-assistance control and 400 in each of the four assistance-available arms.
Minimum detectable effect size for main outcomes (accounting for sample design and clustering)
Supporting Documents and Materials

There is information in this trial unavailable to the public. Use the button below to request access.

Request Information
IRB

Institutional Review Boards (IRBs)

IRB Name
Bowdoin College Institutional Review Board
IRB Approval Date
2026-07-27
IRB Approval Number
IRB-2026-54
Analysis Plan

There is information in this trial unavailable to the public. Use the button below to request access.

Request Information