Loading page…
The evidence is limited to one training corpus, modest token budgets, models no larger than 350M parameters, no throughput measurements, and mostly few seeds; the 350M tier has one seed per arm, and natural incidence of the projection gate’s adversarial direction blind spot was not measured. · CiteArk