Secondary Outcomes (explanation)
Before the estimation tasks, participants state the smallest number, the largest number, and their best guess for how many correct answers the system would give in 100 comparable tasks. The distance between the smallest and largest figures measures how much ambiguity the participant perceives about the source. This elicitation is incentivized: a participant is paid if the system's verified accuracy falls inside the interval they state, minus an amount proportional to how wide that interval is. The penalty for width means participants gain nothing by stating a very wide interval simply to be safe.
These beliefs allow us to estimate two quantities that describe how disclosed information is absorbed. The first is the weight participants place on the midpoint of the statement they read. The second is the weight they place on its width. The control arm identifies what participants believe when they are told
nothing, and each disclosure arm then reveals how far beliefs move toward what was stated. Because the point arm and the interval arm each provide an estimate of the same width parameter, the two estimates can be compared, and disagreement between them would indicate that the framework does not describe belief formation well. We regard this as a strength of the design, since it allows the framework to be rejected rather than only confirmed.
Three further measures serve as checks on the intervention rather than as outcomes of interest in themselves. We ask how clear the accuracy statement was, how certain the participant feels about the system's accuracy, and how much they trust the system. The design requires that the interval statement lowers certainty without lowering clarity. If both fall together, the range was confusing rather than ambiguous, and the results would need to be interpreted differently. We report these measures before the main results so that readers can judge whether the intervention worked as intended.
After the estimation tasks we also ask how accurate participants believe such systems to be. Because this question comes after treatment, it is affected by the treatment and is used only descriptively. Its purpose is to confirm that participants regarded the system as capable of error. If participants believe the system is rarely wrong, a statement about its accuracy carries little information, and the intervention has nothing to work with.