Are Straight-Through gradients and Soft-Thresholding all you need for Sparse Training?

Benchmark Model Rank Results
network-pruning-on-imagenet-resnet-50-90-sparsityST-3#3Top-1 Accuracy: 76.03