Pretraining does not monotonically improve distinguishing truth from plausibility: performance on surprising truths and common misconceptions mode-hops in both OLMo3 and Apertus. · CiteArk
Pretraining does not monotonically improve distinguishing truth from plausibility: performance on surprising truths and common misconceptions mode-hops in both OLMo3 and Apertus.