Which of the following best describes the purpose of a calibration curve in model evaluation?
-
A
To measure the trade-off between precision and recall at various thresholds
-
B
To compare predicted probabilities against actual outcome frequencies
-
C
To visualize the decision boundary of a classifier
-
D
To plot training loss against validation loss across epochs