AI · Flashcard

Asked whether it is sure, Claude changes its answer. What has that change actually demonstrated?

  • AThat the model reacted to pressure in the prompt, which is no evidence that either answer is right
  • BThat the first answer was wrong, since a correct answer would have survived the challenge intact
  • CThat the second answer is verified, since re-checking triggers a stricter internal review pass
  • DThat confidence was low the first time, which the model reveals by revising when it is questioned

Why this is the answer

A challenge in the prompt is itself an input, and the model can revise a correct answer or defend a wrong one — which is why grounding the answer in a source, rather than re-interrogating the model, is the recommended way to reduce fabrication. Revision does not identify which answer was wrong, no stricter review pass is triggered by the question, and there is no confidence level being disclosed.

Official docs
Study in Gnoseed →