ByNobleID
    A proof of convergence for the gradient descent optimization method with\n random initializations in the training of neural networks with ReLU\n activation for piecewise linear target functions | NobleID