Citation timeline · 1999–2025 · 19 studies · 8 recurring authors

Does explaining the wrong answers actually help?

Twenty-six years of argument about multiple-choice testing, feedback, and what learners retain. The premise under examination: MC may be the best format for knowledge transfer, but only if wrong answers are explained — and that authoring cost is why it's rare. Three separable claims. The evidence treats them very differently.

Claim A · supported, with a boundary

MC may be best for transfer. True for related, untested material. On directly tested material, short answer with feedback still wins.

Claim B · needs revision

Only if wrong answers are explained. The one direct test found no effect of feedback type. What works is making the learner reason, not handing them prose.

Claim C · untested

Rare because authoring is expensive. No study has examined this. The likelier cause is structural: Anki has no MC note type at all.

Verdict on each claim

Where the timeline lands, claim by claim.

Supported

Claim A — MC may be the best format for knowledge transfer

Supported, with a precise boundary. Little et al. (2012) is real and has held up, but the advantage is specific to related, untested information. On directly tested material, short answer with feedback wins — Kang et al. (2007), and the synthesis in Üner, Tekin & Roediger (2021).

For an exam of novel passages testing adjacent knowledge, related-untested is arguably the right target. So the claim survives for the MCAT specifically, not in general.

Revise

Claim B — only if wrong answers are explained

This is the weak link. Three problems:

  1. Butler, Karpicke & Roediger (2007) tested feedback type on MC directly and found no difference. Timing mattered; type did not.
  2. Finn & Metcalfe (2010) found the winning elaboration was learner-generated — scaffolded hints. Provided elaboration tied with plain correct-answer feedback.
  3. Alamri & Higham (2022) found corrective feedback actively impairs related-item performance — the exact outcome the Little benefit is measured on.

Revised: MC works for transfer when the learner is made to reason about the distractors — not when the explanations are handed to them.

Untested

Claim C — rare because it's more effort to author

No research supports this, because nobody has studied it. What exists is adjacent: LLM distractor generators still produce invalid distractors needing expert review, and automated explanatory-feedback pipelines report meaningful error rates.

The likelier explanation is structural. Anki ships no multiple-choice note type, no distractor field, and no way to record confidence. The format is absent from the tooling, not merely expensive — and a bad distractor is worse than none, since the entire effect is conditional on competitiveness.

What the evidence actually specifies

Not "MC with explanations." Five properties, each traceable to a result.

  1. Competitive distractors

    Required, not optional — the effect vanishes with implausible alternatives. Little, Bjork, Bjork & Angello (2012).

  2. An elimination prompt before reveal

    Ask the learner to rule out what they can, and why. This was the Experiment 3 manipulation in Alamri & Higham (2022) — the one that strengthened the beneficial controlled pathway.

  3. Confidence captured with the response

    Improves learning on its own (Sparck, Bjork & Bjork, 2016) and fixes the 25% guess floor that otherwise corrupts scheduler grades.

  4. Scaffolded rather than expository feedback

    Hints first, answer last. The only elaboration format that beat plain correct-answer feedback at a delay: Finn & Metcalfe (2010).

  5. Delayed feedback where the interface allows

    The one variable that did move the needle in the direct test — Butler, Karpicke & Roediger (2007).