Citation timeline · 1999–2025 · 19 studies · 8 recurring authors

Does explaining the wrong answers actually help?

Twenty-six years of argument about multiple-choice testing, feedback, and what learners retain. The premise under examination: MC may be the best format for knowledge transfer, but only if wrong answers are explained, and that authoring cost is why it's rare. Three separable claims. The evidence treats them very differently.

Claim A · supported, with a boundary

MC may be best for transfer. True for related, untested material. On directly tested material, short answer with feedback still wins.

Claim B · needs revision

Only if wrong answers are explained. The one direct test found no effect of feedback type. What works is making the learner reason, not handing them prose.

Claim C · untested

Rare because authoring is expensive. No study has examined this. The likelier cause is structural: Anki has no MC note type at all.

Verdict on each claim

Where the timeline lands, claim by claim.

Supported

Claim A: MC may be the best format for knowledge transfer

Supported, with a precise boundary. Little et al. (2012) is real and has held up, but the advantage is specific to related, untested information. On directly tested material, short answer with feedback wins: Kang et al. (2007), and the synthesis in Üner, Tekin & Roediger (2021).

For an exam of novel passages testing adjacent knowledge, related-untested is arguably the right target. So the claim survives for the MCAT specifically, not in general.

Revise

Claim B: only if wrong answers are explained

This is the weak link. Three problems:

  1. Butler, Karpicke & Roediger (2007) tested feedback type on MC directly and found no difference. Timing mattered; type did not.
  2. Finn & Metcalfe (2010) found the winning elaboration was learner-generated: scaffolded hints. Provided elaboration tied with plain correct-answer feedback.
  3. Alamri & Higham (2022) found corrective feedback actively impairs related-item performance, the exact outcome the Little benefit is measured on.

Revised: MC works for transfer when the learner is made to reason about the distractors, not when the explanations are handed to them.

Untested

Claim C: rare because it's more effort to author

No research supports this, because nobody has studied it. What exists is adjacent: LLM distractor generators still produce invalid distractors needing expert review, and automated explanatory-feedback pipelines report meaningful error rates.

The likelier explanation is structural. Anki ships no multiple-choice note type, no distractor field, and no way to record confidence. The format is absent from the tooling, not merely expensive, and a bad distractor is worse than none, since the entire effect is conditional on competitiveness.

What the evidence actually specifies

Not "MC with explanations." Five properties, each traceable to a result.

  1. Competitive distractors

    Required, not optional: the effect vanishes with implausible alternatives. Little, Bjork, Bjork & Angello (2012).

  2. An elimination prompt before reveal

    Ask the learner to rule out what they can, and why. This was the Experiment 3 manipulation in Alamri & Higham (2022), the one that strengthened the beneficial controlled pathway.

  3. Confidence captured with the response

    Improves learning on its own (Sparck, Bjork & Bjork, 2016) and fixes the 25% guess floor that otherwise corrupts scheduler grades.

  4. Scaffolded rather than expository feedback

    Hints first, answer last. The only elaboration format that beat plain correct-answer feedback at a delay: Finn & Metcalfe (2010).

  5. Delayed feedback where the interface allows

    The one variable that did move the needle in the direct test: Butler, Karpicke & Roediger (2007).