1. Bishop, C. M. (1995). Training with Noise is Equivalent to Tikhonov Regularization. Neural Computation.

  2. Srivastava, N., Hinton, G., Krizhevsky, A., et al. (2014). Dropout: A Simple Way to Prevent Neural Networks from Overfitting. Journal of Machine Learning Research.

  3. Jacot, A., Gabriel, F., & Hongler, C. (2018). Neural Tangent Kernel: Convergence and Generalization in Neural Networks. Advances in Neural Information Processing Systems.

  4. Dwork, C., Feldman, V., Hardt, M., et al. (2015). Preserving Statistical Validity in Adaptive Data Analysis. Proceedings of the forty-seventh annual ACM symposium on Theory of Computing.