Question 1
Choose
, since it gives the lowest training loss and allows the model to fit the data best.Choose
, as it provides a reasonable compromise between model complexity and generalization.Choose
, since it yields the lowest validation loss and achieves the best bias–variance trade-off.Choose
, as it produces a highly regularized model that minimizes the risk of overfitting.
