What is the vanishing gradient problem in deep neural networks?
-
A
Gradients become too large and cause unstable training
-
B
Gradients shrink to near zero as they propagate back through layers, halting learning
-
C
The model forgets earlier training data
-
D
Weights converge too quickly to a local minimum