Loading page…
C1+C2 allocations were stable under disjoint corpora and fourfold smaller or larger calibration sets, with downstream LongBench scores in a narrow band; a Python-source calibration also changed HumanEval pass@1 by at most 1.6 points. · CiteArk