Loading page…
After general post-training, the 4.5T OLMo3-32B checkpoint is more robust to prefilling attacks than the 4.9T checkpoint (53% versus 21%). The expanded comparison also favors the selected intermediate checkpoint over the final pretraining and mid-training checkpoints. · CiteArk