Probing finds a direction in activations separating correct from incorrect answers; it can be used for detection or steering.
Models often 'know' internally that they are wrong, yet this information is not directly available in the output.