Loading page…
On ImageNet, the 34-layer plain network had higher training error throughout training and higher validation error than the 18-layer plain network, demonstrating optimization degradation with increased depth rather than overfitting. · CiteArk