Deep Learning: Why does increase batch_size cause overfitting and how does one reduce it?

Open in new window