Experimental Design
Between-subjects randomized experiment conducted online with approximately 2,000 participants recruited through Prolific, targeting approximately 400 per arm. Eligibility is restricted to United States residents aged 18–25 with at least some college education, prescreened by Prolific and re-verified within the survey instrument.
Participants first complete a baseline block comprising an instructed-response attention check and self-rated confidence in each of the four task domains. They are then randomly assigned with equal probability to one of five arms using the survey platform's randomizer. Randomization is at the individual level; assignment is stored as an embedded field and governs all subsequent survey branching.
One arm serves as a pure control in which no assistance is available. The remaining four arms, in all of which assistance is available, form a 2×2 between-subjects factorial over two manipulated task conditions. The identity of those two conditions is specified in the non-public design field and in the pre-analysis plan, and will be made public when the trial is complete. The control arm identifies the overall effect of assistance availability relative to working unaided; the factorial identifies the main effect of each manipulation within assistance-available arms, and their interaction.
Following the study phase, participants complete brief post-task ratings and report a domain-by-domain expectation of their own performance on the forthcoming test, then complete the unassisted test phase, followed by a validity check and a demographic block, both administered after all outcomes are measured.
Analysis is intent-to-treat. Person-level outcomes are estimated by OLS on arm indicators with the control condition as the reference category, using heteroskedasticity-robust standard errors, and reported both with and without a vector of baseline covariates. Question-level process outcomes are estimated with question fixed effects and standard errors clustered by participant. The factorial contrasts are estimated on assistance-available arms only, in both a saturated specification including the interaction and a restricted specification omitting it, always reported together. Pre-specified exclusions cover the attention check, the eligibility criteria, incomplete responses, duplicate submissions from a single platform identifier, and a post-outcome validity check; results including participants failing the validity check are reported as a robustness check.